rtrvr.ai
Browser ExtensionStart in Chrome, on the page you're on.CloudA thousand browsers, on your schedule.RoverThe AI customer engineer for your product.Data & EvalsExpert trajectories for AI labs.
ACCESSAPI + MCPCLI & SDKTemplatesIntegrationsWhatsApp
Use cases
Vibe ScrapingLead EnrichmentWeb MonitoringForm FillingJob ApplicationsSocial MediaAI Web ContextAgentic CheckoutAll use cases
Pricing
BlogLaunches, benchmarks, deep divesDocsExtension, Cloud, API, MCP, CLIModelsWhich model runs your taskCase StudiesReal teams, real runsVideosNew runs every weekChangelogWhat just shippedNewslettersProduct releases and real runs
Docs
Log inBook DemoAdd to Chrome
Log in
Menu
Add to ChromeBook a demo
ProductsBrowser ExtensionCloudRoverData & EvalsExploreUse casesPricingBlogDocs

MODELS

Pick the right model for the job.

Every model on rtrvr.ai, compared on the dimensions that matter for a browser agent: how capable it is on complex sites, how fast it streams, what a task actually costs, and how much page it can hold. Numbers come from recent production traffic.

Most capable

GPT-5.6 Terra

The strongest pick for complex sites and extensive tool calling. Premium rates — bring it in when the task earns it.

Everyday default

DeepSeek V4 Flash

The platform default and its most production-proven model: fast, cheap, and reliable across multi-step runs.

Biggest saver

GLM 5.3 Flash

The lowest cost per task, with native vision. It reasons lightly — great for straightforward tasks, not the hardest pages.

Head to head

Capability is our editorial ranking from quality batteries and external benchmarks; speed, affordability, and context come from measured throughput and current credit rates.

Compare up to 5 models

DeepSeek
ZAI GLM
OpenAI
Gemini
Thinking Machines
Xiaomi MiMo

Every model, side by side

Per-task costs price the same three canonical workloads under each model’s current rates, so columns compare apples to apples.

ModelBest for$ / 1M tokensin · cached · outQuickcredits/taskEverydaycredits/taskLong runcredits/taskSpeedCache hitContextVision
DeepSeek
DeepSeek FlashFree Mode
Everyday default — fast, cheap, reliable0.112 · 0.0224 · 0.2240.020.371.4117 tok/s72%1M—
DeepSeek Pro
Deep multi-step reasoning · ~10× Flash cost1.32 · 0.044 · 3.960.324.013116 tok/s60%128K—
ZAI GLM
GLM 5.3 FlashFree Mode
Lowest cost · reads screenshots (vision)0.075 · 0.015 · 0.250.020.260.99~110 tok/s—1M✓
GLM 5.3
Flagship open-weights · strong on hard pages1.4 · 0.26 · 4.40.344.818~65 tok/s—1M—
OpenAI
GPT-5.6 LunaFree Mode
Frontier quality · quick tool calls · vision0.22 · 0.022 · 1.320.070.802.971 tok/s27%400K✓
GPT-5.6 Terra
Most capable — complex sites, heavy tool use2.2 · 0.22 · 13.20.738.029~55 tok/s—400K✓
Gemini
Gemini Flash LiteFree Mode
Snappiest replies · quick lookups & summaries0.3 · 0.03 · 2.50.121.24.4169 tok/s42%1M✓
Gemini Flash
Multimodal all-rounder · media-heavy pages0.75 · 0.075 · 3.750.222.69.4102 tok/s34%1M✓
Thinking Machines
Inkling SmallFree Mode
Compact open-weights reasoner · low cost0.3 · 0.06 · 1.20.081.14.1~90 tok/s—256K—
Inkling
Frontier open-weights reasoning · premium1 · 0.17 · 4.050.273.513~60 tok/s—256K—
Xiaomi MiMo
MiMo V2.5Free Mode
Budget pick for huge pages · slower output0.119 · 0.00238 · 0.2380.020.341.134 tok/s70%1M—
MiMo V2.5 Pro
Deeper reasoning on huge pages0.3045 · 0.00252 · 0.6090.060.862.7~30 tok/s—1M—

1 credit = $0.01. Per-task credits price the same canonical token bundle under every model’s rates — see the task shapes below. Speed and cache-hit are medians from recent production traffic; ~values are estimates for newer tiers.

What a “task” means here

We aggregated every LLM round of every production task into per-task token bundles. These three shapes are that fleet’s 25th percentile, median, and 75th percentile.

Quick lookup

One round — read a page, answer a question

Rounds
1
Input
1.5K tokens
Cached
0%
Output
300 tokens

Everyday task

A short plan plus an extraction or a few actions

Rounds
1–2
Input
50K tokens
Cached
50%
Output
1.5K tokens

Long agent run

Multi-step browsing with tool calls across pages

Rounds
5+
Input
250K tokens
Cached
70%
Output
6.5K tokens

Try them on your own tasks

Switch models any time from the composer in the Chrome extension or on rtrvr.ai/cloud — value tiers run free in Free Mode.

rtrvr.ai

Make every site
work for you.

Launches first, roadmap early, and the occasional trick we only share by email.

Products

Browser ExtensionCloudRoverData & Evals

Use cases

Vibe ScrapingLead EnrichmentForm FillingWeb MonitoringSocial MediaJob ApplicationsData MigrationAI Web ContextAgentic Checkout

Resources

DocsBlogModelsData for AI LabsCase StudiesVideosNewslettersChangelogPricingAppSumoDemoAffiliate

Company

TeamContactGCP PartnerWhat We BelieveSecurityPrivacyTerms

Developers

APIMCPCLI & SDKTemplatesIntegrationsWhatsApp

Compare

ApifyBardeenBrowserbaseBrowser UseClayClaudeCometFirecrawl
Products
Browser ExtensionCloudRoverData & Evals
Use cases
Vibe ScrapingLead EnrichmentForm FillingWeb MonitoringSocial MediaJob ApplicationsData MigrationAI Web ContextAgentic Checkout
Resources
DocsBlogModelsData for AI LabsCase StudiesVideosNewslettersChangelogPricingAppSumoDemoAffiliate
Company
TeamContactGCP PartnerWhat We BelieveSecurityPrivacyTerms
Developers
APIMCPCLI & SDKTemplatesIntegrationsWhatsApp
Compare
ApifyBardeenBrowserbaseBrowser UseClayClaudeCometFirecrawl
BACKED BYNVIDIA InceptionGoogle Cloud for StartupsBright DataNEC XSalesforce LaunchpadElevenLabs GrantsGMI CloudComposioSmallest.ai Grants
DISCOVERYllms.txtllms-full.txtagents.mdDocumentation indexSitemapOpenAPIAI Catalog
© 2026 Retriever AI · rtrvr.ai
DiscordYouTubeInstagramTikTokLinkedInXGitHub
support@rtrvr.ai