Announcements

What’s new at Perch

Product updates, launches, and changes — newest first.

October 6, 2026

Cloud Agents now live in Perch AI

Watch the launch film · 20 seconds · Sound available

Hand Perch the job. Pick it up from anywhere. Cloud Agents run in a private cloud workspace and keep working after you close the laptop. Start a task from your browser or phone, follow its progress, and review the result when you return.

Need the files and tools on your own computer? Turn on Remote Control in Perch Desktop and send tasks from anywhere. They run on your computer while it stays awake with Perch open.

Cloud Agents and Remote Control are in beta for Pro, with availability enabled account by account.

Free Launch Celebration Promo

Celebrate the launch with GPT-6.1 Sol, Qwen 3.8 Max 0902, and Qwen 3.8 Flash. The offer runs through October 13, 2026, at 2:30 p.m. Eastern.

  • Pro only

    GPT-6.1 Sol

    /model gpt-6.1-sol
  • Starter + Pro

    Qwen 3.8 Max 0902

    /model qwen-3.8-max
  • Starter + Pro

    Qwen 3.8 Flash

    /model qwen-3.8-flash

Starter is free, with usage counting toward its included allowance. On Pro, these promotional selections do not use your manual-model allowance; they count toward Roost’s fair-use limits at half price. The usual five-hour and weekly limits apply.

September 29, 2026

GPT-6.1 Sol is already in Perch

GPT-6.1 Sol joins Perch Pro on day one. The featured lineup runs through Qwen 3.8 Flash; all models share a $15 rolling five-hour limit and $75 billing-period manual-model allowance.

OpenAI's newest Sol model is now available on Perch Pro, from day one. It brings stronger coding, complex document understanding, and multi-step agent work to your existing workspace.

In OpenAI's evaluations, GPT-6.1 Sol approaches Astra's performance across several demanding tasks while keeping GPT-6 Sol's base input and output prices. Cached input is now half the price.

Select it in Perch Web, Desktop, or CLI with /model gpt-6.1-sol. It joins the full Pro lineup and uses your existing shared model-selection allowance.

New model. Same Perch. Ready when you are.

September 26, 2026

Kimi K3 is free on Starter for a limited time

Kimi K3 is now available on Perch's free Starter plan for a limited time. Pin it in Perch Desktop or the CLI with /model kimi-k3-starter.

Space Bunny Alpha and DeepSeek V4.1 Flash are also available on the free Starter plan. Pin them with /model space-bunny-alpha and /model deepseek-flash-starter.

No paid plan or credit card is required.

September 22, 2026

A copper sunrise and ivory crescent moon above a dark Perch-inspired landscape.
PerchAI

OPENAI GPT-6 · DAY ONE IN PERCH

GPT-6 Sol and Luna are live in Perch on day one

Go deep with Sol. Move fast with Luna. Both are ready in Perch today.

Explore the models

GPT-6 Sol

Pro

For complex coding and agentic workflows.

GPT-6 Luna

Pro

For focused, high-volume work.

GPT-6 Sol and GPT-6 Luna are available in Perch on launch day, both on Pro. Sol is built for complex coding and agentic workflows; Luna is OpenAI's most efficient GPT-6 model for focused work you need to do at scale. Choose either in Perch AI Web, Desktop, or the CLI with /model gpt-6-sol or /model gpt-6-luna.

Both models have a 1.05M-token context window, image input, streaming, tool calls, structured output, and adjustable reasoning. Published base rates are $2 input / $10 output per million tokens for Sol, and $0.10 input / $0.50 output for Luna; cached input is $0.20 and $0.01 respectively. OpenAI applies higher rates to prompts above 272K input tokens. Full lineup and rates: https://perchai.app/docs/concepts/models.

September 19, 2026

DeepSeek V4.1 Flash is live in Perch

DeepSeek V4.1 Flash is now available in Perch as a Pro model. Pin it in Perch Desktop, the web app, or the CLI with /model deepseek-v4.1-flash. DeepSeek V4 Flash 0731 (Starter) is also available free to Starter users; pin it with /model deepseek-v4-flash-0731-starter.

Update, September 22: DeepSeek V4.1 Flash is now also available on Starter. The Starter pool rotates, subject to availability.

V4.1 Flash has a 1M-token context window, native multimodal input, tool use, and graded reasoning you can set to Low, High, or Max. It is built for long-context research, coding, and multi-step agent work.

Published rates are $0.30 / $1.20 per million tokens (input / output), with cached input at $0.03. Full model list and rates: https://perchai.app/docs/concepts/models.

September 16, 2026

Three powerful models are now free on Starter

GLM 5.3 Flash, GPT-5.6 Luna, and the brand-new stealth model Union Alpha are now available on Perch's free Starter plan. No paid plan or credit card is required. Let Roost choose for you, or pin one in Perch Desktop, the web app, or the CLI with /model glm-5.3-flash-starter, /model gpt-5.6-luna, or /model union-alpha.

GLM 5.3 Flash brings a 1M-token context window, native image input, tool use, and graded reasoning for fast coding and long-context work. GPT-5.6 Luna pairs a 1.05M-token context window with native vision, tool use, and adjustable reasoning. Union Alpha is a brand-new stealth preview with a 262K-token context window, multimodal input, and tool use.

All three are included in Starter's 20,000 PT monthly hosted-usage allowance. Usage counts against that allowance, and the rotating Starter lineup may change as provider arrangements change.

September 14, 2026

The GPT-5.6 family has landed in Perch

The GPT-5.6 family in Perch: Sol and Terra on Pro, and Luna on Starter and Pro.

GPT-5.6 Sol, Terra, and Luna are now available in Perch. Sol and Terra are available on Pro, while Luna is available on every plan, including Starter. Pin them in Perch Desktop, the web app, or the CLI with /model gpt-5.6-sol, /model gpt-5.6-terra, or /model gpt-5.6-luna.

Choose Sol for the most demanding professional work, Terra for a balance of intelligence and cost, or Luna for fast, efficient, high-volume tasks. All three have a 1.05M-token context window, native image input, streaming, tool calls, structured output, and adjustable reasoning.

Published rates are $4.00 / $20.00 per million tokens (input / output) for Sol, $2.00 / $12.00 for Terra, and $0.20 / $1.20 for Luna. Cached input is $0.40, $0.20, and $0.02 respectively. Full model list and rates: https://perchai.app/docs/concepts/models.

September 10, 2026

GLM 5.3 and Kimi K3 are live in Perch

GLM 5.3 and Kimi K3 are now available in Perch as Pro models, joining the premium lineup at the top of the list. Both have a 1M-token context window, streaming, tool calls, and graded reasoning you can set to Low, High, or Max. Pin either by hand in Perch Desktop, the web app, or the CLI with /model glm-5.3 or /model kimi-3.

GLM 5.3 is Zhipu's newest GLM, built for strong reasoning and agentic coding across long context. Kimi K3 is Moonshot's newest and largest Kimi, with native image input, and it sits at the top of the lineup for the hardest multi-step work.

Published rates are $1.75 / $5.50 per million tokens (input / output) for GLM 5.3, with cached input at $0.325, and $3.30 / $16.50 for Kimi K3, with cached input at $0.33. Full model list and rates: https://perchai.app/docs/concepts/models.

September 5, 2026

GLM 5.3 Flash is live in Perch

GLM 5.3 Flash is now available in Perch as a Pro model, joining the premium lineup alongside GLM 5.2. It has a 1M-token context window, native image input, streaming, tool calls, and graded reasoning you can set to Low, High, or Max. It is built for fast, low-cost work: quick coding turns, long-context reads, and high-volume agent runs. Pin it by hand in Perch Desktop, the web app, or the CLI with /model glm-5.3-flash.

Published rates are $0.15 / $0.50 per million tokens (input / output), with cached input at $0.05, among the cheapest in the premium lineup. Full model list and rates: https://perchai.app/docs/concepts/models.

August 29, 2026

Grok 4.6 is live in Perch

Grok 4.6 is now available in Perch as a Pro model, replacing Grok 4.3 in the premium lineup. It brings a 500K-token context window, native image input, streaming, structured output, and adjustable reasoning that can be switched off or set from Low through XHigh. Pin it by hand in Perch Desktop, the web app, or the CLI with /model grok-4.6.

Published rates are $2.20 / $6.60 per million tokens (input / output), with cached input at $0.55. Full model list and rates: https://perchai.app/docs/concepts/models.

August 27, 2026

Qwen 3.8 Flash is live in Perch

Qwen 3.8 Flash is now available in Perch — the newest Qwen Flash model, with a 1M-token context window. It is a Pro model built for fast, low-cost work: quick coding turns, long-context reads, and high-volume agent runs where you want speed without spending much. Pin it by hand in Perch Desktop, the web app, or the CLI with /model qwen-3.8-flash.

Published rates are $0.16 / $0.47 per million tokens (input / output), with cached input at $0.016 — among the cheapest in the lineup. Full model list and rates: https://perchai.app/docs/concepts/models.

August 24, 2026

MiniMax M3 is free in Perch

MiniMax M3 is a 1M-context model, and for a limited time it is free to run on every plan, Starter included. Pinning it does not draw down your monthly usage, so you can lean on it for long-context and coding work and keep your allowance for everything else. Normal fair use still applies.

Pin it in Perch Desktop, the web app, or the CLI with /model minimax-m3-free. With a 1M-token context window, it is a strong fit for large codebases, long document review, and multi-step agent runs. Want maximum reliability? Pro users can switch to the standard MiniMax M3 at any time with /model minimax-m3.

This is a limited-time offer. When it ends, the free MiniMax M3 leaves the free lineup; the standard MiniMax M3 stays available on Pro. Full model list and rates: https://perchai.app/docs/concepts/models.

August 22, 2026

Ox Alpha is free in Perch

Ox Alpha is a new 1M-context model, and for a limited time it is free to run on every plan, Starter included. Pinning it does not draw down your monthly usage, so you can lean on it for long-context and coding work and keep your allowance for everything else. Normal fair use still applies.

Pin it in Perch Desktop, the web app, or the CLI with /model ox-alpha. With a 1M-token context window and multimodal input, it is a strong fit for large codebases, long document review, and multi-step agent runs.

This is a limited-time offer. When it ends, Ox Alpha leaves the free lineup. Full model list and rates: https://perchai.app/docs/concepts/models.

August 12, 2026

Kimi K2.7 Code is free this week

Through Wednesday, August 19, pinning Kimi K2.7 Code is free on Pro. Runs on it do not draw down your monthly usage, so you can make it your default for the week and keep your allowance for everything else. Normal fair use still applies.

Pin it in Perch Desktop or the CLI with /model kimi-2.7. Kimi K2.7 Code is a premium Pro model with a 256K context window, native vision, and strong agentic coding, which makes it a good fit for multi-file code work, long document review, and agent runs that go many steps.

After August 19, Kimi K2.7 Code goes back to metering at $0.75 input and $3.75 output per million tokens, with cached input at $0.15. Those promo rates hold through September 1. Full lineup and rates: https://perchai.app/docs/concepts/models.

July 30, 2026

GLM 5.2 is free this week

Through Wednesday, August 5, pinning GLM 5.2 is free on Pro. Runs on it do not draw down your monthly usage, so you can make it your default for the week and keep your allowance for everything else. Normal fair use still applies.

Pin it in Perch Desktop or the CLI with /model glm-5.2. GLM 5.2 is a premium Pro model with a 200K context window and strong tool use, which makes it a good fit for long document review, multi-file code work, and agent runs that go many steps.

After August 5, GLM 5.2 goes back to metering at $1.00 input and $3.50 output per million tokens, with cached input at $0.19. Those promo rates hold through September 1. Full lineup and rates: https://perchai.app/docs/concepts/models.

July 26, 2026

Gemini 3.6 Flash is live in Perch

Gemini 3.6 Flash is now available in Perch — Google's newest Flash model, day one. Let Roost route to it automatically on Pro, or pin it by hand in Perch Desktop and the CLI with /model gemini-3.6-flash.

Published rates are $1.50 / $7.50 per million tokens (input / output). Gemini 3.5 Flash stays in the lineup. Full model list: https://perchai.app/docs/concepts/models.

July 3, 2026

Perch is live on Product Hunt today

We launched Perch on Product Hunt today. If Perch has been useful to you, come say hello. We're reading and replying to every comment.

Take a look and tell us what you think: https://www.producthunt.com/products/perch-ai. Thanks for being here early.

June 30, 2026

Gemini 3.5 Flash is live, plus usage tracking fixes

Gemini 3.5 Flash is now available in Perch. It's a fast Google model that's well suited to quick tasks and high-volume work. Let Roost route to it automatically, or hand-pick it in Perch Desktop and the CLI.

We also fixed usage reporting. The /usage command now shows your correct plan and Perch Token balance in the terminal, and Pro accounts see their full allowance. If anything still looks off, reach out at support@perchai.app.

June 28, 2026

Perch moved to faster, more reliable infrastructure

We migrated Perch to new infrastructure built for our streaming, agentic workloads.

The result: faster responses, steadier long-running sessions, and headroom to grow. Thanks for being here early.