VibePager · Models

The models are your machine’s, not ours

Anthropic, OpenAI, Google DeepMind and xAI on one side; DeepSeek, Moonshot AI, Z.ai, Alibaba and MiniMax on the other. VibePager hosts none of them — it reads the model list out of the agent running on your own computer, which is the only way a list stays true in a field that turns over every few weeks.

Updated September 2026

Why the list is not in the app

Any model list written into a phone app is wrong within a month. Between them, the open-weight labs shipped several generations in 2026 alone, and the closed labs are no slower. A hardcoded list fails in the worst way: it offers you a model your machine cannot actually run, and the conversation fails somewhere you cannot see.

So the picker asks the agent. Whatever OpenCode, Kilo Code, Kimi, Gemini CLI, Codex or Claude Code offers on your machine, with your keys, is what shows up — with search, because some of those lists are long.

The families, and what they are called right now

Read the version column carefully: it is filled in differently for the two halves of the table, on purpose. The open-weight labs publish versioned checkpoints you can name, pin and download, so naming them is useful. The closed labs rename and re-tier constantly — while this page was being written, three reference sites disagreed about which OpenAI and Google flagship was current that week. So the closed families get tiers, not numbers.

FamilyWho builds itWeightsNamed versions (Sept 2026)Usually reached through
DeepSeekDeepSeek AI — Hangzhou, ChinaOpenDeepSeek V4, in Pro and Flash variantsTheir API, a gateway, or self-hosting
KimiMoonshot AI — Beijing, ChinaOpenKimi K3, and K2.5 before itKimi CLI, self-hosting, or a gateway
GLMZ.ai (Zhipu AI) — Beijing, ChinaOpenGLM-5.3 and GLM-5.2The Z.ai GLM coding plan, a gateway, or self-hosting
QwenAlibaba Cloud — Hangzhou, ChinaOpenQwen 3.8 Max and Qwen 3.6Qwen Code, a gateway, or self-hosting
MiniMaxMiniMax — Shanghai, ChinaOpenMiniMax M3Their API or a gateway
ClaudeAnthropic — San Francisco, USAClosedOpus / Sonnet / Haiku tiersClaude Code, or an Anthropic key in another agent
GPTOpenAI — San Francisco, USAClosedflagship / mini tiersCodex CLI, or an OpenAI key
GeminiGoogle DeepMind — Mountain View, USAClosedPro / Flash tiersGemini CLI, or a Google key
GrokxAI — Palo Alto, USAClosedGrok and Grok Code tiersAn xAI key in an agent that accepts one
MistralMistral AI — Paris, FranceOpen and closedCodestral and the Large tiersAn API key in a model-agnostic agent

The labs, briefly

Worth knowing who is who, because the split is not subtle and it decides what you can actually do with a model.

The Chinese labs are the reason this page has a version column at all. Between them, DeepSeek, Moonshot, Z.ai, Alibaba and MiniMax have shipped the bulk of the serious open-weight coding models — at prices that make it reasonable to leave an agent running on a long job, which is exactly the case a phone is for.

Running a specific model through a specific agent

This is the question most people are really asking — not "which model is best" but "can I run this model in that agent". The answer is almost always yes, and it has nothing to do with VibePager: it is the agent’s provider configuration.

Whatever you end up with, the phone shows that list — read from your machine, with search. VibePager never decides which models you may use, and never adds one you cannot reach.

A dated snapshot — and the point of dating it

Everything in the version column above is from September 2026. Six months earlier the same column would have said Kimi K2, GLM-4.6, DeepSeek V3 and Qwen3 Coder — different names, same products, same companies. Nothing about the tools changed; only the numbers did.

That is measurable, not a hunch. Searching for the old names and the new ones side by side in September 2026, the new ones had already overtaken: glm 5.2 free was ahead of glm 4.7 free, and the strongest model queries of all were for versions that did not exist half a year before — minimax m3 price, deepseek v4 pro pricing, glm 5.3 pricing.

The same goes for benchmark rankings and prices. Whichever model tops a coding leaderboard this month will not next month, and this page deliberately quotes no prices: they move faster than it can be republished, and a stale price is worse than no price. Check the provider.

Cheap model, long job, walked away

There is a pattern worth naming, because it is where a phone earns its place. Open-weight models cost little enough that you will happily set one on a long, boring refactor — the kind of job you start and leave. Which means the agent will stop to ask permission at minute nine, while you are somewhere else, and sit there until you come back.

The cheaper the model, the more likely you are not watching. That is the case the notification exists for.

What VibePager does not do

Questions

Which models can I use from the phone?
Whichever ones the agent on your machine offers. If OpenCode on your laptop can reach a model, the phone lists it. Nothing is hardcoded into the app.
Can I run open-weight models like Kimi, GLM, DeepSeek, Qwen or MiniMax?
Yes, through an agent that can reach them — OpenCode and Kilo Code are model-agnostic, and Kimi CLI runs the Kimi family directly. Self-hosted endpoints work the same way: if the agent can reach it, so can the phone.
Do I give VibePager my API keys?
No. Keys live in the agent’s configuration on your own machine, exactly where they are now.
Is there a markup on tokens?
No. VibePager is a flat subscription for the remote control; you pay your model provider directly, as you do today.
Which model is best for coding?
It changes every few weeks, which is the honest answer and also the reason the picker reads from your machine. The practical split most people land on: a cheap open-weight model for long or repetitive work, a premium closed model for the hard parts.

VibePager runs the agents on your own machine and puts the controls on your phone.