PerchAI
How it works

Models

The models available on each Perch plan, and how Roost chooses between them.

Every frontier model, zero switching — Perch routes your work to the right one, automatically.

Perch runs on a range of models rather than a single one. Roost picks between them for you on every task, and you can pin a specific model by hand: Starter inside the included pool, Pro across the full registry. This page lists what each plan includes.

This list evolves

The model lineup changes as we add and retire models, so treat the names below as the current registry rather than a fixed contract. What stays constant is the structure: Starter runs the included pool, Pro adds the premium models on top, and any plan can run your own.

The rates below are the published list price for each model, per million tokens: input, output, and cached input (prompt-cache hits). Perch meters your usage against these rates and does not mark them up. Cached input is billed only for tokens the provider reuses from cache; everything else bills at the input rate. Published rates change from time to time; we reflect changes here. Rates shown are current as of August 2026.

Promo rates through September 1, 2026

GLM 5.2 and Kimi K2.6 / Kimi K2.7 Code show prior list struck through with the promo price beside it. Metering uses the promo price through September 1, 2026, then returns to prior list. MiniMax and everything else are unchanged.

Starter

The included pool of capable hosted models. Roost picks from it for you on every task, and you can also hand-pick any model in the pool by name on Perch Desktop and the CLI. Starter runs the Roost Standard and Standard Max tiers, and includes up to $20 of hosted usage each month. The premium models and the Pro and Pro Max tiers stay on Pro.

The Starter pool and its published rates:

ModelInput / 1MOutput / 1MCache / 1M
Qwen 3.6$0.25$1.25$0.05
DeepSeek V4 Flash$0.14$0.28$0.0028
Kimi K2.5$0.60$3.00$0.10
GLM 5$1.00$3.20$0.20
Qwen3 Coder$0.45$1.80$0.09
Nemotron Super$0.15$0.65$0.03
MiniMax M2$0.30$1.20$0.03
Gemma 4 E2B$0.04$0.08$0.04
Gemma 4 31B$0.14$0.40$0.14

Pro

Every model in the registry, left to Roost or pinned by hand, across all four Roost tiers (Standard, Standard Max, Pro, Pro Max). Pro includes the full Starter pool plus the premium models below, and up to $150/mo of total usage: unlimited Roost fair use plus up to $75 of manual model selection.

The premium models Pro adds on top of the Starter pool, with their published rates:

ModelInput / 1MOutput / 1MCache / 1M
GLM 5.2$1.40 $1.00$4.40 $3.50$0.26 $0.19
DeepSeek V4 Pro$1.74$3.48$0.14
Kimi K2.6$0.95 $0.75$4.00 $3.75$0.16 $0.13
Kimi K2.7 Code$0.94 $0.75$4.00 $3.75$0.19 $0.15
MiniMax M3$0.29$1.20$0.06
Nemotron Ultra$0.75$2.75$0.15
Nemotron 3.5 Lightning$0.10$0.25$0.05
Grok 4.3$1.25$2.50$0.20
Qwen 3.7 Plus$0.40$1.60$0.08
Inkling$1.00$4.05$0.17

Bring your own key

Available on every plan, including the free one. Add a key from any OpenAI-compatible provider, or point Perch at your own endpoint and name any model you want. Perch usage limits do not apply to inference you pay for yourself, so your included allowance stays untouched while you use it. A cloud key works on Web, Desktop, and CLI; a model running on your own machine is CLI only. See Bring your own key.

At a glance

StarterPro
Model selectionIncluded pool, automatic or pinnedEvery model, automatic or pinned
Premium modelsNoYes
Manual model pinIncluded pool onlyAny model
Roost tiersStandard, Standard MaxAll four (adds Pro, Pro Max)
Included usageUp to $20 / monthUp to $150 / month (unlimited Roost + $75 manual)
Bring your own keyYesYes
Images and scans on any modelYesYes

Model choice is not only for the model you talk to. On Perch Desktop you can also set which model the helpers Perch spawns run on, which is how you have one model check another's work. See Sub-agents.

For prices and plans, see Pricing. For how the routing works, see Perch Roost.