Dan Lahav
@dan_lahav
Happy to see @Irregular's benchmark used by @AnthropicAI to test 𝗖𝗹𝗮𝘂𝗱𝗲 𝗠𝘆𝘁𝗵𝗼𝘀 5 𝗮𝗻𝗱 𝗖𝗹𝗮𝘂𝗱𝗲 𝗙𝗮𝗯𝗹𝗲 5!
Offensive capabilities are moving fast, and CyScenarioBench is one of the few benchmarks still keeping pace and providing meaningful signal.
Unlike
Offensive capabilities are moving fast, and CyScenarioBench is one of the few benchmarks still keeping pace and providing meaningful signal.
Unlike
Claude@claudeai · Jun 9Introducing Claude Fable 5: a Mythos-class model that we’ve made safe for general use.
Its capabilities exceed those of any model we’ve ever made generally available.
Its capabilities exceed those of any model we’ve ever made generally available.
0:20 · Edited
1 14