Peter H. Diamandis
July 26, 2026
TL;DR
Frontier Labs' K3 architecture uses standard transformer innovations like mixtures of experts and linearized attention to nearly match GPT-4.5, raising questions about why major AI labs spend billions when recognizable transformer designs achieve competitive performance.
“If you can just use a transformer to get this close, what the heck are the American labs spending all of their money on?”
“I derive great comfort in at minimum knowing that the transformer architecture is still alive and kicking.”
1. K3 Architecture Overview
K3 uses standard transformer design with mixtures of experts and linearized attention techniques, not novel architectures.
2. Performance vs. Cost
K3 nearly matches GPT-4.5 max on task cost frontier benchmarks despite using recognizable transformer innovations.
3. Questioning Big AI Spending
The competitive performance of transformer-based designs raises fundamental questions about billion-dollar R&D investments in AI labs.