Category / T-CAT / RESEARCH
Research
Research / T-2026-0498
Anthropic raises $65B at $965B valuation, betting compute is the moat
Anthropic's $65B Series H, the largest AI round ever, reveals a strategic bet on compute ownership over model differentiation.
SourceAnthropic raises $65B in Series H funding at $965B post-money valuation
Research / T-2026-0774
The Hy3 Mystery: An Unknown Model Is Dominating OpenRouter
An unknown model called Hy3 is crushing benchmarks on OpenRouter, raising questions about benchmarking culture, API economics, and what 'state of the art' even means anymore.
SourceThe mysterious Hy3 LLM is topping OpenRouter Model Rankings by a large margin
Research / T-2026-3733
Why CORE matters: reasoning improvement without the rollout tax
CORE lets language models improve reasoning with minimal rollouts by generating insights from contrasting success and failure.
SourceCORE: Contrastive Reflection Enables Rapid Improvements in Reasoning
Research / T-2026-6117
Five Frontier LLMs Disagree on 67% of Fact-Check Claims
Five frontier LLMs disagree on 67% of 1,000 real-world fact-check claims, raising questions about their reliability as knowledge tools.
SourceFive frontier LLMs disagree on 67% of 1k real-world fact-check claims
Research / T-2026-0167
Obra's 'Superpowers' Is a Rare Thing: An Agentic Framework That Admits What It Doesn't Know
Obra's open-source agentic skills framework offers a grounded alternative to the industry's sprawling agent benchmarks, focusing on composable, testable capabilities.
Sourceobra/superpowers