PerchAI
How it works

Models

The models available on each Perch plan, and how Roost chooses between them.

Every frontier model, zero switching — Perch routes your work to the right one, automatically.

Perch runs on a range of models rather than a single one. Roost picks between them for you on every task, and you can pin a specific model by hand: Starter inside the included pool, Pro across the full registry. This page lists what each plan includes. The Pro plan guide explains the separate Roost and manual-model usage limits.

This list evolves

The model lineup changes as we add and retire models, so treat the names below as the current registry rather than a fixed contract. What stays constant is the structure: Starter runs the included pool, Pro adds the premium models on top, and any plan can run your own.

The rates below are the published base price for each model, per million tokens: input, output, and cached input (prompt-cache hits). Cached input applies only to tokens the provider reuses from cache; everything else uses the input rate. OpenAI charges more for GPT-6 and GPT-6.1 prompts above 272K input tokens. Rates change from time to time; we reflect changes here. Rates shown are current as of September 29, 2026. Your account's Usage view is the source of truth for the PT recorded by each turn.

Starter

Starter has a rotating list of frontier models, subject to availability. Roost picks from the current pool for you on every task, and you can also hand-pick any current Starter model by name on Perch Desktop and the CLI. Starter runs the Roost Standard and Standard Max tiers and includes up to $20 of hosted usage each month. GPT-6.1 Sol, GPT-6 Sol and Luna, GPT-5.6 Sol and Terra, the rest of the premium-only registry, and the Pro and Pro Max tiers stay on Pro.

The Starter pool and its published rates:

The Starter lineup rotates

The Starter pool changes as our provider deals change: we add models when we can keep them free and retire them when a promo ends. What stays constant is the structure: Starter runs a rotating pool of capable hosted models, Pro adds the premium models on top, and any plan can run your own.
ModelInput / 1MOutput / 1MCache / 1M
DeepSeek V4.1 Flash (1M context)$0.30$1.20$0.03
GPT-5.6 Luna$0.20$1.20$0.02
Qwen 3.6$0.25$1.25$0.05
Kimi K2.5$0.60$3.00$0.10
GLM 5$1.00$3.20$0.20
Qwen3 Coder$0.45$1.80$0.09
Nemotron Super$0.15$0.65$0.03
Gemma 4 E2B$0.04$0.08$0.04
Gemma 4 31B$0.14$0.40$0.14

Pro

Every model in the registry, left to Roost or pinned by hand, across all four Roost tiers (Standard, Standard Max, Pro, Pro Max). Pro includes the full Starter pool plus the premium models below. Roost has no fixed monthly cap and uses rolling fair-use windows of 5,000 PT per five hours and 20,000 PT per seven days. Manual selection has its own 15,000 PT five-hour limit, 37,500 PT seven-day limit, and 75,000 PT billing-period allowance. See the Pro plan guide for how the two lanes work.

The premium-only models Pro adds on top of the Starter pool, with their published rates. GPT-6.1 Sol and GPT-6 Sol and Luna are on Pro. Pro also includes every model in the current Starter pool, including DeepSeek V4.1 Flash:

ModelInput / 1MOutput / 1MCache / 1M
GPT-6.1 Sol$2.00$10.00$0.10
GPT-6 Sol$2.00$10.00$0.20
GPT-6 Luna$0.10$0.50$0.01
GPT-5.6 Sol$4.00$20.00$0.40
GPT-5.6 Terra$2.00$12.00$0.20
GLM 5.3$1.75$5.50$0.325
Kimi K3$3.30$16.50$0.33
Grok 4.6$2.20$6.60$0.55
Qwen 3.8 Flash$0.16$0.47$0.016
GLM 5.3 Flash$0.15$0.50$0.05
GLM 5.2$1.40$4.40$0.26
DeepSeek V4 Flash 0731 (262K context)$0.14$0.28$0.007
Kimi K2.6$0.95$4.00$0.16
Kimi K2.7 Code$0.94$4.00$0.19
MiniMax M3$0.29$1.20$0.06
Nemotron Ultra$0.75$2.75$0.15
Nemotron 3.5 Lightning$0.10$0.25$0.05
Qwen 3.7 Plus$0.40$1.60$0.08
Qwen 3.8 27B$0.40$3.00$0.15
DeepSeek V4 Pro$1.74$3.48$0.14
Inkling$1.00$4.05$0.17

Bring your own key

Available on every plan, including the free one. Add a key from any OpenAI-compatible provider, or point Perch at your own endpoint and name any model you want. Perch usage limits do not apply to inference you pay for yourself, so your included allowance stays untouched while you use it. A cloud key works on Web, Desktop, and CLI; a model running on your own machine is CLI only. See Bring your own key.

At a glance

StarterPro
Model selectionIncluded pool, automatic or pinnedEvery model, automatic or pinned
GPT-6.1 accessNoSol
GPT-6 accessNoSol and Luna
GPT-5.6 accessLunaSol, Terra, and Luna
Premium-only registryNoYes
Manual model pinIncluded pool onlyAny model
Roost tiersStandard, Standard MaxAll four (adds Pro, Pro Max)
Included usageUp to 20,000 PT / billing periodSeparate Roost fair-use windows + 75,000 PT manual / billing period
Bring your own keyYesYes
Images and scans on any modelYesYes

Model choice is not only for the model you talk to. On Perch Desktop you can also set which model the helpers Perch spawns run on, which is how you have one model check another's work. See Sub-agents.

For prices and plans, see Pricing and the Pro plan guide. For how the routing works, see Perch Roost.