latentbrief
Back to news
Launch1w ago

Anthropic's Claude Fable 5.1 Breaks New Ground in Scientific Reasoning

Simon Willison1 min brief

In brief

  • Anthropic has unveiled Claude Fable (and Mythos) 5.1, a significant upgrade in artificial intelligence capabilities.
    • This version excels particularly in scientific reasoning, achieving an impressive 52.6% score on the Terminal-Bench-Science 0.1 benchmark-a remarkable jump from its previous score of 24.7%.
    • This improvement positions Claude Fable 5.1 as a standout tool for complex problem-solving and knowledge work.
  • The update introduces five distinct reasoning levels-low, medium, high, xhigh, and max-allowing users to tailor the AI's response depth to their needs.
  • While lower levels like "low" produce outputs without detailed reasoning, higher levels like "xhigh" provide more thorough analysis, enhancing the AI's versatility across various tasks.
  • Looking ahead, Claude Fable 5.1 sets a new standard for AI in scientific research and problem-solving.
  • As Anthropic continues to refine these capabilities, we can expect further advancements that push the boundaries of what AI can achieve in understanding and solving complex real-world challenges.

Terms in this brief

Claude Fable
Claude Fable is a version of Anthropic's AI model that has been upgraded to excel in scientific reasoning. It achieved a high score on the Terminal-Bench-Science benchmark, showing its ability to solve complex problems and perform knowledge work.
Terminal-Bench-Science
A benchmark used to evaluate AI models' performance in scientific reasoning. Claude Fable 5.1 scored significantly higher on this test compared to previous versions, demonstrating improved problem-solving capabilities.

Read full story at Simon Willison

More briefs