<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Models on Vibe Coding</title><link>https://vibecoding.rest/models/</link><description>Recent content in Models on Vibe Coding</description><generator>Hugo</generator><language>en</language><atom:link href="https://vibecoding.rest/models/index.xml" rel="self" type="application/rss+xml"/><item><title>Claude Opus 5</title><link>https://vibecoding.rest/models/claude-opus-5/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>https://vibecoding.rest/models/claude-opus-5/</guid><description>&lt;p&gt;Claude Opus 5 is Anthropic&amp;rsquo;s top-tier model, tuned specifically for agentic coding — multi-file features, large refactors, and end-to-end implementation work rather than single-shot completions. Give it a full task specification up front and it tends to run the whole way to a finished result instead of leaving stubs behind.&lt;/p&gt;</description></item><item><title>Claude Sonnet 5</title><link>https://vibecoding.rest/models/claude-sonnet-5/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>https://vibecoding.rest/models/claude-sonnet-5/</guid><description>&lt;p&gt;Claude Sonnet 5 closes most of the gap to Opus-tier quality on coding and agentic tasks while staying meaningfully cheaper — a solid default for the coding work that makes up most of a normal week, not just the hardest 20%.&lt;/p&gt;</description></item><item><title>GPT-5.6 Sol</title><link>https://vibecoding.rest/models/gpt-5-6-sol/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>https://vibecoding.rest/models/gpt-5-6-sol/</guid><description>&lt;p&gt;GPT-5.6 Sol is the top tier of OpenAI&amp;rsquo;s three-model GPT-5.6 lineup (Luna, Terra, Sol), positioned as the strongest option for coding, enterprise work, and cybersecurity tasks, with a claim on efficiency alongside raw capability.&lt;/p&gt;</description></item><item><title>Gemini 3.1 Pro</title><link>https://vibecoding.rest/models/gemini-3-1-pro/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>https://vibecoding.rest/models/gemini-3-1-pro/</guid><description>&lt;p&gt;Gemini 3.1 Pro is Google&amp;rsquo;s flagship coding and reasoning model — Gemini 3.5 Pro hasn&amp;rsquo;t shipped yet, so 3.1 Pro remains the top tier of the Gemini line for complex, multi-step work.&lt;/p&gt;</description></item><item><title>Grok 4.6</title><link>https://vibecoding.rest/models/grok-4-6/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>https://vibecoding.rest/models/grok-4-6/</guid><description>&lt;p&gt;Grok 4.6 is xAI&amp;rsquo;s frontier model for coding, agentic tasks, and knowledge work, and the current flagship behind Grok Build — xAI&amp;rsquo;s dedicated coding agent and terminal UI.&lt;/p&gt;&#10;&lt;h2 id="why-it-stands-out"&gt;Why it stands out&lt;/h2&gt;&#10;&lt;ul&gt;&#10;&lt;li&gt;&lt;strong&gt;Aggressive tiered pricing&lt;/strong&gt; — a lower rate below 200K prompt tokens and a scaled-up rate above it, so short-context coding sessions stay cheap.&lt;/li&gt;&#10;&lt;li&gt;&lt;strong&gt;Grok Build&lt;/strong&gt;, xAI&amp;rsquo;s open-source coding agent and TUI, gives it a purpose-built harness rather than relying only on third-party tools.&lt;/li&gt;&#10;&lt;li&gt;&lt;strong&gt;500K-token context window&lt;/strong&gt;, comfortably large for most real-world coding sessions.&lt;/li&gt;&#10;&lt;/ul&gt;&#10;&lt;h2 id="good-for"&gt;Good for&lt;/h2&gt;&#10;&lt;p&gt;Fast, iterative agentic coding workflows — especially teams already using Grok Build or wanting a cheaper alternative to the largest frontier models for everyday agent loops.&lt;/p&gt;</description></item><item><title>DeepSeek V4</title><link>https://vibecoding.rest/models/deepseek-v4/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>https://vibecoding.rest/models/deepseek-v4/</guid><description>&lt;p&gt;DeepSeek V4 ships in two MIT-licensed variants — V4-Pro and V4-Flash — with weights published on Hugging Face. It&amp;rsquo;s a genuinely open model: download it, fine-tune it, or ship it in a product, not just call an API.&lt;/p&gt;</description></item><item><title>Qwen3-Coder-Next</title><link>https://vibecoding.rest/models/qwen3-coder-next/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>https://vibecoding.rest/models/qwen3-coder-next/</guid><description>&lt;p&gt;Qwen3-Coder-Next is built for self-hosted coding agents — an 80B-total, 3B-active mixture-of-experts model under an Apache 2.0 license, small enough to run on a single well-equipped workstation instead of a data-center cluster.&lt;/p&gt;</description></item><item><title>Kimi K2.7 Code</title><link>https://vibecoding.rest/models/kimi-k2-7-code/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>https://vibecoding.rest/models/kimi-k2-7-code/</guid><description>&lt;p&gt;Kimi K2.7 Code is Moonshot AI&amp;rsquo;s coding-focused release in the Kimi K2 line, built to carry end-to-end programming tasks across long, multi-turn agent sessions rather than one-shot completions.&lt;/p&gt;&#10;&lt;h2 id="why-it-stands-out"&gt;Why it stands out&lt;/h2&gt;&#10;&lt;ul&gt;&#10;&lt;li&gt;&lt;strong&gt;Always-on thinking mode&lt;/strong&gt; that preserves full reasoning content across multi-turn conversations, instead of resetting context between steps.&lt;/li&gt;&#10;&lt;li&gt;&lt;strong&gt;Native multimodal mixture-of-experts architecture&lt;/strong&gt;, accepting text and image input in the same coding session.&lt;/li&gt;&#10;&lt;li&gt;&lt;strong&gt;Frontier-competitive coding benchmarks&lt;/strong&gt; at a fraction of the per-token cost of comparable Western frontier models.&lt;/li&gt;&#10;&lt;li&gt;&lt;strong&gt;Steep cache discount&lt;/strong&gt; (roughly 80% off cache-miss input pricing), which rewards long, iterative agent loops that keep re-reading the same context.&lt;/li&gt;&#10;&lt;/ul&gt;&#10;&lt;h2 id="good-for"&gt;Good for&lt;/h2&gt;&#10;&lt;p&gt;Teams running long, agentic coding sessions who want frontier-adjacent reasoning quality without frontier-model pricing.&lt;/p&gt;</description></item><item><title>GLM-5.2</title><link>https://vibecoding.rest/models/glm-5-2/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>https://vibecoding.rest/models/glm-5-2/</guid><description>&lt;p&gt;GLM-5.2 is Z.ai&amp;rsquo;s (formerly Zhipu AI) flagship open-weight model — a 753-billion-parameter mixture-of-experts design released under the MIT license, with a 1M-token context window and full downloadable weights.&lt;/p&gt;&#10;&lt;h2 id="why-it-stands-out"&gt;Why it stands out&lt;/h2&gt;&#10;&lt;ul&gt;&#10;&lt;li&gt;&lt;strong&gt;MIT-licensed, 753B-parameter MoE weights&lt;/strong&gt;, fully self-hostable with no usage restrictions.&lt;/li&gt;&#10;&lt;li&gt;&lt;strong&gt;1M-token context window&lt;/strong&gt; with up to 131K tokens of output in a single response.&lt;/li&gt;&#10;&lt;li&gt;&lt;strong&gt;Beats GPT-5.5 on FrontierSWE&lt;/strong&gt; at roughly a sixth of the cost, per Zhipu&amp;rsquo;s own benchmarking.&lt;/li&gt;&#10;&lt;li&gt;&lt;strong&gt;Selectable High/Max reasoning modes&lt;/strong&gt;, letting you trade latency for depth on harder multi-step coding tasks.&lt;/li&gt;&#10;&lt;/ul&gt;&#10;&lt;h2 id="good-for"&gt;Good for&lt;/h2&gt;&#10;&lt;p&gt;Teams that want a genuinely open, frontier-tier coding model to self-host, fine-tune, or run at scale without per-token lock-in.&lt;/p&gt;</description></item><item><title>MiniMax M3</title><link>https://vibecoding.rest/models/minimax-m3/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>https://vibecoding.rest/models/minimax-m3/</guid><description>&lt;p&gt;MiniMax M3 is Shanghai-based MiniMax&amp;rsquo;s frontier open-weight release, combining a 1M-token context window with native multimodal input — text, image, and video — aimed at full-repository code understanding rather than file-by-file work.&lt;/p&gt;</description></item></channel></rss>