<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Models on Tokenise</title><link>https://tokenise.rosvetic.com/categories/models/</link><description>Recent content in Models on Tokenise</description><generator>Hugo</generator><language>en-us</language><lastBuildDate>Mon, 05 Oct 2026 23:18:00 +0000</lastBuildDate><atom:link href="https://tokenise.rosvetic.com/categories/models/index.xml" rel="self" type="application/rss+xml"/><item><title>Moving to Sonnet 5.5? Four requests that now return a 400</title><link>https://tokenise.rosvetic.com/posts/sonnet-5-5-migration-breaking-changes/</link><pubDate>Mon, 05 Oct 2026 23:18:00 +0000</pubDate><guid>https://tokenise.rosvetic.com/posts/sonnet-5-5-migration-breaking-changes/</guid><description>&lt;p&gt;Claude Sonnet 5.5 shipped on September 28 at the same price as Sonnet 5: $2 per million input tokens and $10 per million output. Anthropic says it runs 30% or more faster. If you call the API directly, though, changing &lt;code&gt;claude-sonnet-5&lt;/code&gt; to &lt;code&gt;claude-sonnet-5-5&lt;/code&gt; isn&amp;rsquo;t the whole migration. Several requests that worked yesterday now fail.&lt;/p&gt;</description></item><item><title>That 2-point benchmark lead might just be a bigger server</title><link>https://tokenise.rosvetic.com/posts/benchmark-gaps-infrastructure-noise/</link><pubDate>Sat, 03 Oct 2026 12:00:00 +0000</pubDate><guid>https://tokenise.rosvetic.com/posts/benchmark-gaps-infrastructure-noise/</guid><description>&lt;p&gt;Every model launch comes with a table of coding benchmark scores, and the gaps are often a few points. Anthropic&amp;rsquo;s engineering team published a study that should make you read those gaps more carefully. In its &lt;a href="https://www.anthropic.com/engineering/infrastructure-noise" target="_blank" rel="noopener noreferrer"&gt;infrastructure noise analysis&lt;/a&gt;&#10;, it found that how much compute an eval runs on can move agentic coding scores by more than the leaderboard gap between top models.&lt;/p&gt;</description></item><item><title>Claude Opus 5.5 costs 20% less per token and Anthropic says it codes better</title><link>https://tokenise.rosvetic.com/posts/claude-opus-5-5-pricing-and-coding/</link><pubDate>Fri, 02 Oct 2026 12:00:00 +0000</pubDate><guid>https://tokenise.rosvetic.com/posts/claude-opus-5-5-pricing-and-coding/</guid><description>&lt;p&gt;Anthropic released Claude Opus 5.5 on September 22. The model id is &lt;code&gt;claude-opus-5-5&lt;/code&gt;, it&amp;rsquo;s on the Claude Platform, AWS, Google Cloud and Azure, and it&amp;rsquo;s cheaper than the model it replaces. The &lt;a href="https://www.anthropic.com/claude-opus-5-5" target="_blank" rel="noopener noreferrer"&gt;announcement&lt;/a&gt;&#10; says it does most work at the level of Claude Fable 5.1 while costing 40% less to run than Opus 5.&lt;/p&gt;</description></item></channel></rss>