Testing Mythos and Fable, Moving Beyond SWE-bench, Nvidia's Open Contender
The Batch ● Covered by 2 sources
Anthropic restricted Claude Fable 5's access to AI researchers and refused certain technical questions, while the U.S. government imposed export controls on the model, prompting independent evaluators to report difficulty assessing its true capabilities due to safety filters routing 8-35% of flagged tasks to weaker models. Claude Fable 5 ranked highest on benchmarks when its fallback mechanisms were included, but dropped significantly in standing when refusals were counted as failures, making true performance impossible to measure independently. These restrictions have accelerated global interest in open-source AI alternatives and raised concerns among developers about the stability of building on proprietary model providers.
Why it matters
The Batch AI News and Insights: Over the last two weeks, both the U.S. Government and Anthropic took significant actions that demonstrated...