Search and filter through extracted claims from AI researchers.
Showing 161-180 of 494 claims of type "critique"
Critics often don't understand the difference between an LLM and LLM with tools or a harness
Half of the Astra problems can be solved by Fable, and OpenAI did not have a control group
People went completely nuts over Astra without asking for a control group
The cybersecurity test was run without any meaningful supervision
There was a total failure of alignment training at OpenAI, which is the failure that matters most
LLMs aren't close to doing real discovery, according to yet another paper
Astra being impressive at math alone does not qualify it as AGI
Game-based robot training approaches ignore human factors in the environment
People who enthusiastically promote new AI models without trying them first appear foolish
Current AI systems fail to demonstrate genuine general intelligence
Astra has not yet solved significant open-world problems outside of formal verification
Pipeline data may be stale or degraded.
Last synthesis: 2026-09-20. 8,951 pending.