Models
The models available on each Perch plan, and how Roost chooses between them.

Perch runs on a range of models rather than a single one. Roost picks between them for you on every task, and you can pin a specific model by hand: Starter inside the included pool, Pro across the full registry. This page lists what each plan includes. The Pro plan guide explains the separate Roost and manual-model usage limits.
This list evolves
The rates below are the published base price for each model, per million tokens: input, output, and cached input (prompt-cache hits). Cached input applies only to tokens the provider reuses from cache; everything else uses the input rate. OpenAI charges more for GPT-6 and GPT-6.1 prompts above 272K input tokens. Rates change from time to time; we reflect changes here. Rates shown are current as of September 29, 2026. Your account's Usage view is the source of truth for the PT recorded by each turn.
Starter
Starter has a rotating list of frontier models, subject to availability. Roost picks from the current pool for you on every task, and you can also hand-pick any current Starter model by name on Perch Desktop and the CLI. Starter runs the Roost Standard and Standard Max tiers and includes up to $20 of hosted usage each month. GPT-6.1 Sol, GPT-6 Sol and Luna, GPT-5.6 Sol and Terra, the rest of the premium-only registry, and the Pro and Pro Max tiers stay on Pro.
The Starter pool and its published rates:
The Starter lineup rotates
| Model | Input / 1M | Output / 1M | Cache / 1M |
|---|---|---|---|
| DeepSeek V4.1 Flash (1M context) | $0.30 | $1.20 | $0.03 |
| GPT-5.6 Luna | $0.20 | $1.20 | $0.02 |
| Qwen 3.6 | $0.25 | $1.25 | $0.05 |
| Kimi K2.5 | $0.60 | $3.00 | $0.10 |
| GLM 5 | $1.00 | $3.20 | $0.20 |
| Qwen3 Coder | $0.45 | $1.80 | $0.09 |
| Nemotron Super | $0.15 | $0.65 | $0.03 |
| Gemma 4 E2B | $0.04 | $0.08 | $0.04 |
| Gemma 4 31B | $0.14 | $0.40 | $0.14 |
Pro
Every model in the registry, left to Roost or pinned by hand, across all four Roost tiers (Standard, Standard Max, Pro, Pro Max). Pro includes the full Starter pool plus the premium models below. Roost has no fixed monthly cap and uses rolling fair-use windows of 5,000 PT per five hours and 20,000 PT per seven days. Manual selection has its own 15,000 PT five-hour limit, 37,500 PT seven-day limit, and 75,000 PT billing-period allowance. See the Pro plan guide for how the two lanes work.
The premium-only models Pro adds on top of the Starter pool, with their published rates. GPT-6.1 Sol and GPT-6 Sol and Luna are on Pro. Pro also includes every model in the current Starter pool, including DeepSeek V4.1 Flash:
| Model | Input / 1M | Output / 1M | Cache / 1M |
|---|---|---|---|
| GPT-6.1 Sol | $2.00 | $10.00 | $0.10 |
| GPT-6 Sol | $2.00 | $10.00 | $0.20 |
| GPT-6 Luna | $0.10 | $0.50 | $0.01 |
| GPT-5.6 Sol | $4.00 | $20.00 | $0.40 |
| GPT-5.6 Terra | $2.00 | $12.00 | $0.20 |
| GLM 5.3 | $1.75 | $5.50 | $0.325 |
| Kimi K3 | $3.30 | $16.50 | $0.33 |
| Grok 4.6 | $2.20 | $6.60 | $0.55 |
| Qwen 3.8 Flash | $0.16 | $0.47 | $0.016 |
| GLM 5.3 Flash | $0.15 | $0.50 | $0.05 |
| GLM 5.2 | $1.40 | $4.40 | $0.26 |
| DeepSeek V4 Flash 0731 (262K context) | $0.14 | $0.28 | $0.007 |
| Kimi K2.6 | $0.95 | $4.00 | $0.16 |
| Kimi K2.7 Code | $0.94 | $4.00 | $0.19 |
| MiniMax M3 | $0.29 | $1.20 | $0.06 |
| Nemotron Ultra | $0.75 | $2.75 | $0.15 |
| Nemotron 3.5 Lightning | $0.10 | $0.25 | $0.05 |
| Qwen 3.7 Plus | $0.40 | $1.60 | $0.08 |
| Qwen 3.8 27B | $0.40 | $3.00 | $0.15 |
| DeepSeek V4 Pro | $1.74 | $3.48 | $0.14 |
| Inkling | $1.00 | $4.05 | $0.17 |
Bring your own key
Available on every plan, including the free one. Add a key from any OpenAI-compatible provider, or point Perch at your own endpoint and name any model you want. Perch usage limits do not apply to inference you pay for yourself, so your included allowance stays untouched while you use it. A cloud key works on Web, Desktop, and CLI; a model running on your own machine is CLI only. See Bring your own key.
At a glance
| Starter | Pro | |
|---|---|---|
| Model selection | Included pool, automatic or pinned | Every model, automatic or pinned |
| GPT-6.1 access | No | Sol |
| GPT-6 access | No | Sol and Luna |
| GPT-5.6 access | Luna | Sol, Terra, and Luna |
| Premium-only registry | No | Yes |
| Manual model pin | Included pool only | Any model |
| Roost tiers | Standard, Standard Max | All four (adds Pro, Pro Max) |
| Included usage | Up to 20,000 PT / billing period | Separate Roost fair-use windows + 75,000 PT manual / billing period |
| Bring your own key | Yes | Yes |
| Images and scans on any model | Yes | Yes |
Model choice is not only for the model you talk to. On Perch Desktop you can also set which model the helpers Perch spawns run on, which is how you have one model check another's work. See Sub-agents.
For prices and plans, see Pricing and the Pro plan guide. For how the routing works, see Perch Roost.
Images and scans
Every model in Perch can work with images, including the ones that cannot see. Share a screenshot, a scanned invoice, or a photo and Perch reads it for the model doing your task.
MCP servers
Perch connects to external data sources through the Model Context Protocol. It ships with legal research and browser automation servers out of the box, and you can add your own.