<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"><channel><title>Model Intelligence</title><description>Tracking AI model releases, inference engines, benchmarks, and local deployment news.</description><link>https://ai-updates.pages.dev/</link><item><title>RISC-V Inference Lands, AI Middle Class Debate Heats Up</title><link>https://ai-updates.pages.dev/posts/2026-08-12-model-intelligence-2/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-08-12-model-intelligence-2/</guid><description>Key updates include DeepSeek-V4-Pro and FLUX.1-dev trending, llama.cpp optimizations for RISC-V, and significant community discussion on AI&apos;s impact on software engineering.</description><pubDate>Wed, 12 Aug 2026 00:00:00 GMT</pubDate></item><item><title>DeepSeek-V4-Pro Surges, FLUX Takes Over Image Generation</title><link>https://ai-updates.pages.dev/posts/2026-08-12-model-intelligence/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-08-12-model-intelligence/</guid><description>Qwen3.8-2.4T-A95B launches to immediate buzz; llama.cpp hits b10375 with Qwen optimizations; vLLM 0.27 brings Kimi K3 support; Grok 4.6 benchmarks drop.</description><pubDate>Wed, 12 Aug 2026 00:00:00 GMT</pubDate></item><item><title>Qwen 3.8: The 27B That Doesn&apos;t Exist Yet (and What Does)</title><link>https://ai-updates.pages.dev/posts/2026-08-12-qwen-38-27b-distills/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-08-12-qwen-38-27b-distills/</guid><description>Breaking down the Qwen 3.8 landscape — the official 2.4T Max, the missing 27B, and community distillation variants.</description><pubDate>Wed, 12 Aug 2026 00:00:00 GMT</pubDate></item><item><title>Model Intelligence — 2026-06-20</title><link>https://ai-updates.pages.dev/posts/2026-06-20-model-intelligence/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-06-20-model-intelligence/</guid><description>llama.cpp surges with 6 builds today; DeepSeek-R1 climbs 13,403 likes; vLLM 0.23.0 brings major DeepSeek-V4 hardening and Gemma 4 Unified support.</description><pubDate>Sat, 20 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Model Intelligence — 2026-06-19</title><link>https://ai-updates.pages.dev/posts/2026-06-19-model-intelligence/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-06-19-model-intelligence/</guid><description>Massive llama.cpp activity with 23 builds today including Eagle3 spec for Qwen3.6; Noam Shazeer joins OpenAI; DeepSeek-R1 maintains the top spot with 13,400 likes.</description><pubDate>Fri, 19 Jun 2026 00:00:00 GMT</pubDate></item><item><title>AI Benchmark Update — June 19, 2026</title><link>https://ai-updates.pages.dev/posts/2026-06-19/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-06-19/</guid><description>Claude Opus 4.8 leads LLM Stats at 67.9 overall score. GPT-5.5 hits 84.9% on GDPval and 82.7% on Terminal Bench 2.0. Gemini 3 Pro tops Arena with 1501 Elo. Qwen3 Coder Next just landed on June 18. DeepSeek V4-Pro reaches 80.6% on SWE-bench Verified at $3.48/M output tokens. The open-weight gap keeps narrowing.</description><pubDate>Fri, 19 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Model Intelligence — 2026-06-18 (Afternoon Update)</title><link>https://ai-updates.pages.dev/posts/2026-06-18-model-intelligence/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-06-18-model-intelligence/</guid><description>Noam Shazeer joins OpenAI, llama.cpp b9704 lands with router hardening, DeepSeek-V4-Pro surges to 4,952 likes — local AI narrative hardens on HN</description><pubDate>Thu, 18 Jun 2026 00:00:00 GMT</pubDate></item><item><title>AI Benchmark Update — June 18, 2026</title><link>https://ai-updates.pages.dev/posts/2026-06-18/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-06-18/</guid><description>Claude Mythos 5 goes GA, GPT-5.6 looms, Kimi K2.7-Code dominates coding benchmarks, Qwen 3.6 Plus leads open models — the frontier race is tighter than ever.</description><pubDate>Thu, 18 Jun 2026 00:00:00 GMT</pubDate></item><item><title>AI Benchmark Update — June 17, 2026</title><link>https://ai-updates.pages.dev/posts/2026-06-17-benchmark-update/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-06-17-benchmark-update/</guid><description>Claude Fable 5 maintains its #1 position with 95% on SWE-bench Verified and 87% on FrontierMath. GPT-5.1 High leads LiveBench at 72.04. Kimi K2.7 Code — a 1T-parameter MoE with 32B active — scores 71.89 on LiveBench, just 0.15 behind. Qwen 3.6 Plus hits 70.85 on LiveBench. Chatbot Arena Elo shows frontier convergence within 25 points.</description><pubDate>Wed, 17 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Model Intelligence — 2026-06-17 (Updated)</title><link>https://ai-updates.pages.dev/posts/2026-06-17-model-intelligence/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-06-17-model-intelligence/</guid><description>Ollama v0.30.10 stable ships with Cohere2MoE; llama.cpp b9692 adds Metal rope_back + server management API; DeepSeek-V4-Pro surges past 4,926 likes.</description><pubDate>Wed, 17 Jun 2026 00:00:00 GMT</pubDate></item><item><title>AI Benchmark Report — June 17, 2026</title><link>https://ai-updates.pages.dev/posts/2026-06-17/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-06-17/</guid><description>Claude Fable 5 holds the composite crown at 100/100. GPT-5.5 Thinking xHigh leads LiveBench at 81.04. Qwen 3.7 Max ships with 1M context. Gemini 3.2 Flash leaks ahead of Google I/O. Kimi K2.7 Code&apos;s 1T-parameter MoE holds at #2 on LiveBench. The frontier gap has shrunk to 25 Elo points.</description><pubDate>Wed, 17 Jun 2026 00:00:00 GMT</pubDate></item><item><title>AI Benchmark Update — June 16, 2026 (Evening Refresh)</title><link>https://ai-updates.pages.dev/posts/2026-06-16-benchmark-update/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-06-16-benchmark-update/</guid><description>Evening refresh: Claude Fable 5 dominates FrontierMath at 87.8%; GPT-5.1 High leads LiveBench at 72.04; Kimi K2.7 Code surges to 71.89; Qwen 3.6 Plus hits 70.85. Community pushback on Fable 5&apos;s pricing and safety filters; GPT-5.5 still preferred for terminal coding. Arena AI shows Claude Fable 5 at 100/100 on its leaderboard.</description><pubDate>Tue, 16 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Model Intelligence — 2026-06-16</title><link>https://ai-updates.pages.dev/posts/2026-06-16-model-intelligence/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-06-16-model-intelligence/</guid><description>llama.cpp pushes NVFP4 quantization and eagle3 spec decoding; DeepSeek-V4-Pro surges in popularity; Qwen-Robot Suite gains HN traction.</description><pubDate>Tue, 16 Jun 2026 00:00:00 GMT</pubDate></item><item><title>AI Benchmark Report — June 16, 2026</title><link>https://ai-updates.pages.dev/posts/2026-06-16/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-06-16/</guid><description>Claude Fable 5 leads Chatbot Arena at 1510 Elo and dominates SWE-bench Pro at 80.3%; GPT-5.5 leads LiveBench at 80.71 with near-perfect math (96.32%); DeepSeek V4 Pro sets open-weights records at 80.6% SWE-bench Verified; Gemini 3.2 Pro and Llama 4.5 Scout launch with 2M and 10M context respectively. Chinese frontier converges into a four-horse race.</description><pubDate>Tue, 16 Jun 2026 00:00:00 GMT</pubDate></item><item><title>AI Benchmark Report — June 15, 2026</title><link>https://ai-updates.pages.dev/posts/2026-06-15-ai-benchmark-report/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-06-15-ai-benchmark-report/</guid><description>Claude Fable 5 leads Chatbot Arena at 1510 Elo; GPT-5.5 dominates terminal coding; Gemma 4 31B sets open-source records at 85.2% MMLU-Pro; GLM-5.1 hits 1530 Elo on Code Arena; Gemini-3.1-Pro breaks into the top five. LiveBench shows Kimi K2.6 leading at 72.17.</description><pubDate>Mon, 15 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Benchmark Update — June 15, 2026</title><link>https://ai-updates.pages.dev/posts/2026-06-15-benchmark-update/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-06-15-benchmark-update/</guid><description>GPT-5.6 Pro takes Arena Hard #1 at 1465 Elo; Claude Mythos 5 dominates SWE-bench at 95.5%; DeepSeek V4.1 holds the crown for open-weight coding at 93.5% LiveCodeBench. The top eight models cluster within a record-tight ~55 Elo spread.</description><pubDate>Mon, 15 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Model Intelligence — 2026-06-15</title><link>https://ai-updates.pages.dev/posts/2026-06-15-model-intelligence/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-06-15-model-intelligence/</guid><description>vLLM v0.23.0 lands with TRTLLM kernel for DeepSeek-V4; llama.cpp pushes b9660 with chat/toolcall hardening; Ollama v0.30.9-rc1 drops; Kokoro-82M hits 11.7M downloads.</description><pubDate>Mon, 15 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Model Intelligence — 2026-06-14</title><link>https://ai-updates.pages.dev/posts/2026-06-14-model-intelligence/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-06-14-model-intelligence/</guid><description>llama.cpp hits b9637 with Cohere2MoE parser; Rio de Janeiro&apos;s &apos;homegrown&apos; LLM exposed as a merge on HN; DeepSeek-V4-Pro climbs to #14 trending with 3.07M downloads.</description><pubDate>Sun, 14 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Model Intelligence — 2026-06-13</title><link>https://ai-updates.pages.dev/posts/2026-06-13-model-intelligence/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-06-13-model-intelligence/</guid><description>vLLM 0.23.0 lands with DeepSeek-V4 hardening and Model Runner V2; SGLang adds Nemotron 3 Ultra and 7 diffusion models; Ollama improves prompt caching and recurrent model support.</description><pubDate>Sat, 13 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Model Intelligence — 2026-06-12</title><link>https://ai-updates.pages.dev/posts/2026-06-12-model-intelligence/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-06-12-model-intelligence/</guid><description>FLUX.1-dev surges +13 likes in a day, closing in on DeepSeek-R1 — while Anthropic&apos;s Fable apology tops 400 HN points and llama.cpp fires off three builds in a single day.</description><pubDate>Fri, 12 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Model Intelligence — 2026-06-11</title><link>https://ai-updates.pages.dev/posts/2026-06-11-model-intelligence/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-06-11-model-intelligence/</guid><description>Anthropic&apos;s Fable guardrails and data retention policy spark HN backlash — both stories top 380 points — while FLUX.1-dev continues closing in on DeepSeek-R1 at #1.</description><pubDate>Thu, 11 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Model Intelligence — 2026-06-10</title><link>https://ai-updates.pages.dev/posts/2026-06-10-model-intelligence/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-06-10-model-intelligence/</guid><description>FLUX.1-dev is 238 likes from overtaking DeepSeek-R1 at #1, llama.cpp pushes three same-day builds, and Claude Desktop&apos;s runaway VM story tops HN.</description><pubDate>Wed, 10 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Model Intelligence — 2026-06-09</title><link>https://ai-updates.pages.dev/posts/2026-06-09-model-intelligence/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-06-09-model-intelligence/</guid><description>AI model trends, inference engine updates, and research insights for local LLM deployment.</description><pubDate>Tue, 09 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Model Intelligence — 2026-06-06</title><link>https://ai-updates.pages.dev/posts/2026-06-06-model-intelligence/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-06-06-model-intelligence/</guid><description>llama.cpp shipping at 18 builds/day, Qwen3.6 and Gemma-4 families gaining strong traction, FLUX.1-dev approaching #1 on HuggingFace.</description><pubDate>Sat, 06 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Model Intelligence — 2026-06-03</title><link>https://ai-updates.pages.dev/posts/2026-06-03-model-intelligence/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-06-03-model-intelligence/</guid><description>Ollama v0.30.2 patch drops; llama.cpp hits b9488 with 5 more daily builds; Qwen3.6-35B-A3B and Gemma-4-E4B-it gaining strong traction.</description><pubDate>Wed, 03 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Model Intelligence — 2026-06-02</title><link>https://ai-updates.pages.dev/posts/2026-06-02-model-intelligence/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-06-02-model-intelligence/</guid><description>Ollama v0.30.0 stable release with llama.cpp rewrite; llama.cpp pushing 5+ daily builds; Qwen3.6-35B-A3B continues gaining traction.</description><pubDate>Tue, 02 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Model Intelligence — 2026-06-01</title><link>https://ai-updates.pages.dev/posts/2026-06-01-model-intelligence/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-06-01-model-intelligence/</guid><description>llama.cpp continues its extraordinary release pace with 5 builds in 24 hours, Qwen3.6 models grow steadily, and Gemma 4 family gains traction.</description><pubDate>Mon, 01 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Model Intelligence — 2026-05-31</title><link>https://ai-updates.pages.dev/posts/2026-05-31-model-intelligence/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-05-31-model-intelligence/</guid><description>vLLM v0.22.0 ships with major performance gains, llama.cpp drops 3 releases in one day, and Qwen3.6 family continues climbing HuggingFace trending.</description><pubDate>Sun, 31 May 2026 00:00:00 GMT</pubDate></item><item><title>Model Intelligence — 2026-05-28</title><link>https://ai-updates.pages.dev/posts/2026-05-28-model-intelligence/</link><guid isPermaLink="true">https://ai-updates.pages.dev/posts/2026-05-28-model-intelligence/</guid><description>SGLang v0.5.12 adds full DeepSeek V4 support, Ollama v0.30 re-architects around llama.cpp, and vLLM v0.21 deprecates transformers v4. Qwen3.6 and Gemma 4 dominate trending.</description><pubDate>Thu, 28 May 2026 00:00:00 GMT</pubDate></item></channel></rss>