{
  "id": 4985039,
  "title": "Claude Fable 5.1 made me a nice animated pelican",
  "url": "https://urgent.news/2026/09/02/claude-fable-5-1-made-me-a-nice-animated-pelican",
  "topic": "ai",
  "section": "AI",
  "published": "2026-09-02T01:05:52.000Z",
  "source": {
    "name": "Hacker News",
    "slug": "hacker-news",
    "url": "https://simonwillison.net/2026/Sep/1/claude-fable-5-1/"
  },
  "original_language": "en",
  "account": "On September 1st, 2026, a report by Simon Willison detailed the impressive capabilities of Claude Fable 5.1, an Anthropic model. The model demonstrated remarkable performance in coding, knowledge work, and long-running problem-solving tasks, achieving a new standard according to Anthropic's announcement. Among the various benchmarks, the Terminal-Bench-Science 0.1 showed the most impressive improvement, with Fable 5.1 scoring a remarkable 52.6% compared to previous models.\n\nWillison focused on the pelican benchmark, a task used to evaluate how well the models could handle a specific prompt. While the benchmark's connection to overall model performance had become less clear, Willison found value in comparing model performance within the same family and at different reasoning effort levels.\n\nHe tested the model with various reasoning settings, from low to max. In low and medium settings, the model did not seem to execute reasoning at all, as the output token count remained relatively low and the reasoning text was missing. However, at higher reasoning levels, the model produced significantly more detailed reasoning traces, taking longer and costing more.\n\nThe \"xhigh\" setting produced an SVG of a pelican riding a bicycle, complete with intricate details like the bird's long neck, orange beak, and the bicycle's components. The \"max\" setting yielded the best result to date, with a pelican that was not only well-drawn but also included thoughtful additions like a blue hat, a basket with a fish, and a charming helmet.\n\nWillison noted that while the result was impressive, it was not as elegant as Gemini 3.7 Flash's output. However, he maintained that the pelican was still a testament to the model's capabilities. The model's ability to create such detailed and well-designed SVGs, even at higher reasoning levels, was enough to make Willison question if the animated version could be created without incurring additional costs.",
  "summary": null,
  "key_points": [
    "Claude Fable 5.1 scored 52.6% on Terminal-Bench-Science 0.1 benchmark",
    "Pelican benchmark showed model's ability to create detailed SVGs",
    "Fable 5.1 produced animated pelican with bicycle, hat, basket, and helmet"
  ],
  "editors_take": null,
  "illustration": null,
  "coverage": {
    "outlets": 1,
    "also_reported_by": []
  },
  "ai_generated": true,
  "disclaimer": "Summaries, key points and the editor’s take are written by software from other outlets’ reporting and may contain errors — always check the linked original."
}