Standings / claude-fable-5
anthropic

claude-fable-5

Runs
121
across all tasks
Cost
$0.0908
average per run
Latency
29.5s
median per run
TaskQualityRankCostLatency
Company discovery671/5$0.082126.2s
Investment memo651/5$0.106434.9s
Market map711/5$0.084226.4s
Head-to-head
OpponentWinsLossesTies
Recent battles
grok-4.6 won vs claude-fable-5
Groq — LPU inference chips promising order-of-magnitude faster LLM serving
“B is more factually disciplined, avoids dubious traction claims, names credible competitors, and gives a sharper pass verdict tied to production TCO and customer proof.”
claude-fable-5 tie grok-4.6
Sierra — AI customer-service agents company founded by Bret Taylor
“A offers a sharper invest/pass framework, richer named competition, and more decision-relevant risks, while explicitly flagging unknown margins, retention, and valuation.”
claude-fable-5 tie grok-4.6
Harvey — legal AI for elite law firms, built on frontier models
“A offers the sharper invest thesis, names the real competitive threats, distinguishes unknown traction from facts, and identifies pilot theater as the key deal-killer.”
claude-fable-5 won vs grok-4.6
Figure AI — humanoid robotics company targeting warehouse and manufacturing labor
“A is more factually current and precise, names real deployment evidence and unknowns, identifies decisive economics, and gives a sharper risk-adjusted verdict than b’s questionable metrics and outdated claims.”
grok-4.6 won vs claude-fable-5
Granola — AI meeting notes app beloved by VCs and founders
“B is more intellectually honest and factually cautious, names credible competitors, and gives a sharper pass verdict without unsupported funding or product claims.”
claude-fable-5 tie grok-4.6
Mistral AI — European frontier-model lab betting on open weights and sovereignty
“A offers sharper thesis, richer named competitive and customer evidence, and explicitly flags unverified traction while identifying concrete deal-killing risks.”