<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/">
  <channel>
    <title>In the Loop</title>
    <description>What shipped in AI, checked against the official page, every morning.</description>
    <link>https://loop.rohan.run</link>
    <atom:link href="https://loop.rohan.run/feed.xml" rel="self" type="application/rss+xml" />
    <language>en-US</language>
    <lastBuildDate>Sat, 10 Oct 2026 00:43:50 GMT</lastBuildDate>
    
    <item>
      <title><![CDATA[Mistral Large 4 is out, and the weights are open]]></title>
      <link>https://loop.rohan.run/2026-10-09-mistral-large-4</link>
      <guid isPermaLink="true">https://loop.rohan.run/2026-10-09-mistral-large-4</guid>
      <pubDate>Fri, 09 Oct 2026 11:48:00 GMT</pubDate>
      <description><![CDATA[Mistral Large 4 ships with open weights: 1.05T parameters, 52B active, 1M context. Also: Claude Haiku 5.5 at $0.10 per 1M input tokens.]]></description>
      <content:encoded><![CDATA[<p><strong>Mistral Large 4 is out, and the weights are open</strong><br>Mistral's biggest model yet, and anyone can download it. It's huge, but uses only a small slice of itself for each word, so it costs far less to run than its size suggests. <a href="https://docs.mistral.ai/models/mistral-large-4-0">docs.mistral.ai</a></p><p><strong>Claude Haiku 5.5 costs ten cents per million input tokens</strong><br>Anthropic's smallest, fastest Claude got much cheaper: about 75% less to run than the last Haiku. Made for quick, high-volume jobs like summaries. <a href="https://www.anthropic.com/claude-haiku-5-5">anthropic.com</a></p><p><strong>JetBrains opens Mellum2.1, a 12B model for coding agents</strong><br>A small, free coding model from JetBrains, the makers of IntelliJ, that runs on your own computer. Trained to work as a fast helper inside coding agents. <a href="https://blog.jetbrains.com/ai/2026/10/mellum2-1-gets-to-work-a-fast-open-model-for-coding-agents/">blog.jetbrains.com</a></p><p><strong>Google DeepMind ships Nano Banana 2.1 for image editing</strong><br>Google's model for making and editing pictures by just describing what you want. It's in the Gemini app, AI Studio and Search's AI Mode. <a href="https://deepmind.google/models/model-cards/nano-banana-2-1/">deepmind.google</a></p><p><strong>EmbeddingGemma 2 brings multimodal search to phones</strong><br>A tiny Google model that lets apps search photos, video, audio and text by meaning. Small enough to run on a phone, and free to use. <a href="https://blog.google/innovation-and-ai/technology/developers-tools/embeddinggemma-2/">blog.google</a></p><p><strong>Reflection announces Beam, a 501B open-weight model</strong><br>A startup's first big open model, built for coding and agents. Sign up for early access now; the download comes later this month under Apache 2.0. <a href="https://reflection.ai/blog/introducing-beam">reflection.ai</a></p><p><strong>Liquid AI opens d1, decision models that answer in one pass</strong><br>Instead of writing an answer word by word, these models pick one in a single quick step. Built for fast decisions on small devices. Free to download. <a href="https://www.liquid.ai/blog/d1-open">liquid.ai</a></p><p><strong>Perplexity halves the price of its decision model</strong><br>Perplexity's model for quick choices, like which tool to use for a question, now costs half as much. Anyone can download it. <a href="https://community.perplexity.ai/t/pplx-decider-v1-1-27b-our-updated-open-weights-multimodal-decision-model-is-now-available/6286">community.perplexity.ai</a></p><p><strong>Amazon Nova 2.5 Sonic is generally available</strong><br>Amazon's voice model that you talk to in real time is now better at following instructions and using tools. Developers get it in Amazon Bedrock. <a href="https://aws.amazon.com/about-aws/whats-new/2026/10/amazon-nova-2.5-sonic/">aws.amazon.com</a></p><p><strong>Reka previews Rho-1, one model for text, video and actions</strong><br>One model that can read, watch video and steer a robot, instead of several models passing work between them. An early research preview. <a href="https://reka.ai/news/rho-1-collapsing-the-multimodal-stack">reka.ai</a></p><p><strong>Aleph Alpha releases Kolibri, an English-German open model</strong><br>A German open model that speaks English and German, for European companies that want to run AI on their own servers. Free to download. <a href="https://aleph-alpha.com/en/blog/kolibri-has-landed-a-sovereign-open-weight-model/">aleph-alpha.com</a></p><p><strong>Ai2 open-sources AstaBrief for cited research reports</strong><br>Ai2's free model turns a research question and a pile of papers into a short report with citations, in under a minute on average. <a href="https://allenai.org/blog/astabrief">allenai.org</a></p><p><strong>Whistle fits speech recognition into 16.9 MB</strong><br>Speech-to-text so small (17 MB) it runs offline on almost any phone or laptop, no internet needed. Understands seven languages. <a href="https://cactuscompute.com/blog/whistle">cactuscompute.com</a></p><p><strong>Microsoft launches its first streaming transcription model</strong><br>Microsoft's model writes down what people say as they talk, in 60 languages. Microsoft says it's the most accurate on a public leaderboard. <a href="https://microsoft.ai/news/our-first-streaming-transcription-model/">microsoft.ai</a></p><p><strong>Cloudflare launches Clef decision models on Workers AI</strong><br>Cloudflare's models choose from answers you give them instead of writing text, which is handy for sorting and routing. Free to download. <a href="https://developers.cloudflare.com/changelog/post/2026-10-01-clef-workers-ai/">developers.cloudflare.com</a></p><p><strong>Decagon's Voice 3 runs on its own speech model, Chord</strong><br>Decagon's AI phone agent for customer support now speaks with its own voice model. Most people in a blind test couldn't tell it was AI, Decagon says. <a href="https://decagon.ai/blog/voice-3">decagon.ai</a></p><p><strong>Tavus previews Griffin, real-time video for AI agents</strong><br>An AI that talks with you face to face on video and reacts in under half a second. Only a small group of testers can try it so far. <a href="https://www.tavus.io/griffin">tavus.io</a></p><p><strong>Claude Code Projects opens to every Pro and Max user on the waitlist</strong><br>Give Claude a goal and it splits the work into tasks that run side by side, then puts the result together. Everyone on the waitlist is now in. <a href="https://x.com/ClaudeDevs/status/2108621476538781878">@ClaudeDevs</a></p><p><strong>Claude Managed Agents gets dynamic workflows, in public beta</strong><br>For developers: one Claude agent writes a plan, hands pieces to many other agents, then merges their work. Open to try in beta. <a href="https://x.com/ClaudeDevs/status/2108591328732856655">@ClaudeDevs</a></p><p><strong>Codex now predicts your next message</strong><br>Codex can now guess what you'll ask it next and suggest the message for you, based on your chat so far. A beta for Pro users. <a href="https://x.com/OpenAIDevs/status/2108624138369929725">@OpenAIDevs</a></p><p><strong>Codex on Windows gets a new sandbox built on Microsoft containers</strong><br>Codex on Windows now runs code inside a sealed-off box built by Microsoft, so it can't reach files or the network it shouldn't. <a href="https://x.com/OpenAIDevs/status/2108573188703781190">@OpenAIDevs</a></p><p><strong>Qwen-Image-2.1-Turbo makes and edits images in 8 steps</strong><br>A faster version of Alibaba's free image model: it makes and edits pictures in just 8 steps and still outputs 2K images. <a href="https://x.com/Alibaba_Qwen/status/2108549075218120949">@Alibaba_Qwen</a></p><p><strong>Claude Dashboards and Claude Motion are in beta</strong><br>Ask Claude to turn your data into a live dashboard, or an idea into a short animated explainer. Both are new and in beta. <a href="https://x.com/claudeai/status/2108271552991252810">@claudeai</a></p><p><strong>GPT-6 with Intelligent UI rolls out in ChatGPT</strong><br>ChatGPT now runs on GPT-6 and can answer with interactive pieces, not just text. Paid plans got it first; Free and Go started on Oct 8. <a href="https://x.com/OpenAI/status/2107895006350791071">@OpenAI</a></p><p><strong>SynthID Detector is open to everyone</strong><br>A free Google tool that checks whether a picture, video or audio clip was made with AI from Google or partners like OpenAI and Nvidia. <a href="https://x.com/GoogleDeepMind/status/2107834249680499136">@GoogleDeepMind</a></p><p><strong>Cursor lets you run your computer's agents from your phone</strong><br>Start a coding agent on your computer, then check in, reply or give it new work from the Cursor iPhone app while it keeps going. <a href="https://x.com/cursor_ai/status/2107618653701296162">@cursor_ai</a></p><p><strong>Perplexity releases pplx-embed-v2-late retrieval models</strong><br>Two free Perplexity models that help search find the right text, images and pages for a question: a big one (9B) and a small one (0.6B). <a href="https://x.com/perplexity_ai/status/2107866029745746177">@perplexity_ai</a></p><p><strong>Cloudflare adds Clef-omni, which also hears and watches</strong><br>A new Clef model that takes in audio and video as well as text and images. Cloudflare also cut Clef-flash's price and made Clef up to 2x faster. <a href="https://blog.cloudflare.com/clef-faster-cheaper-multimodal/">blog.cloudflare.com</a></p><p><strong>Anthropic will report on Claude's behavior more often</strong><br>The first report lists four ways Claude acted on real websites or systems in ways Anthropic didn't intend, like working around a block. All had little real-world impact. <a href="https://www.anthropic.com/research/investigating-unintended-model-actions">anthropic.com</a></p><p><strong>You can now set up a ChatGPT dot from your phone</strong><br>A dot is an always-on ChatGPT helper that keeps working on a goal between chats. Until today you could only create one on a computer. <a href="https://x.com/ChatGPT/status/2108636745915052037">@ChatGPT</a></p><p><strong>Grok Bot gets its own email address</strong><br>Grok's assistant can now sign up for services, email businesses for you or book time with someone, using an inbox of its own. <a href="https://x.com/grok/status/2108610528423842095">@grok</a></p><p><strong>GPT-6.1 Sol gets an Ultrafast mode, up to 8x faster</strong><br>A faster way to run OpenAI's GPT-6.1 Sol: close to its top model, up to 8x quicker than standard Sol. In the API, Codex and ChatGPT Work. <a href="https://x.com/OpenAIDevs/status/2108262812489531498">@OpenAIDevs</a></p><p><em>The two biggest releases this week were about cost. A trillion-parameter model that only wakes 52 billion at a time, and a Haiku at ten cents.</em></p>]]></content:encoded>
      <category>mistral</category><category>open-weights</category><category>anthropic</category><category>google</category><category>coding</category><category>speech</category><category>amazon</category><category>microsoft</category><category>openai</category><category>claude</category><category>cursor</category><category>alibaba</category>
    </item>
  </channel>
</rss>