HypeDelta
DigestTopicsClaimsPredictionsReliabilityResearchers
Admin
DigestTopicsClaimsPredictionsReliabilityResearchers

HypeDelta - AI Research Intelligence

Claims Browser

Search and filter through extracted claims from AI researchers.

Search & Filters
All
agents
benchmarks
general
infrastructure
interpretability
multimodal
other
policy
All
critique
fact
hint
opinion
prediction
7d
14d
30d
90d

Showing 141-160 of 494 claims of type "critique"

general
critique
Bearish
critic

Elon Musk's timeline predictions are notoriously bad and should not be taken as gospel

Gary Marcus
8/8/2026
Confidence: 95%Source
safety
Previous
179
critique
Bearish
critic

Current approaches to aligning LLMs are not working

Gary Marcus
8/8/2026
Confidence: 80%Source
safety
critique
Bearish
critic

Constitutional AI is not working as an alignment approach

Gary Marcus
8/8/2026
Confidence: 80%Source
general
critique
Bearish
independent

Google had all the advantages as the incumbent but was not able to get going in AI

Nathan Lambert
8/8/2026
Confidence: 75%Source
agents
critique
Bearish
academic

Current agent evaluations are limited to narrow, verifiable tasks and cannot assess open-ended AI research capabilities

Narayanan & Kapoor
8/8/2026
Confidence: 85%Source
general
critique
Bearish
critic

There is a discrepancy between Dwarkesh's prediction of Anthropic making over 100B in revenue this year and actual market data

Gary Marcus
8/8/2026
Confidence: 70%Source
general
critique
Bearish
critic

Astra is excellent at some problems but is not AGI or ASI and is vastly oversold

Gary Marcus
8/8/2026
Confidence: 90%Source
general
critique
Bearish
critic

There is a fundamental conflict in how the term AGI is being used: one group claims AGI will be superintelligent enough to kill humanity, while another claims we've already achieved AGI because LLMs beat average people at some tasks

Gary Marcus
8/8/2026
Confidence: 90%Source
reasoning
critique
Bearish
critic

ChatGPT produces impossible chess positions and inconsistent piece sizes, demonstrating reasoning limitations

Gary Marcus
8/8/2026
Confidence: 90%Source
reasoning
critique
Bearish
critic

Commercial LLMs still can't play chess anywhere near as well as serious players, except by calling external tools, three years after initial concerns were raised

Gary Marcus
8/8/2026
Confidence: 80%Source
scaling
critique
Bearish
critic

Scaling AI without theoretical understanding is futile despite massive investment

Gary Marcus
8/8/2026
Confidence: 90%Source
general
critique
Bearish
critic

Current AI software is unreliable despite companies hoping it will improve and seeking IPOs

Gary Marcus
8/8/2026
Confidence: 85%Source
general
critique
Bearish
critic

OpenAI's Astra is not the quantum leap that OpenAI implied it would be

Gary Marcus
8/8/2026
Confidence: 70%Source
general
critique
Bearish
critic

OpenAI Astra may not be the breakthrough that OpenAI wants people to think it is

Gary Marcus
8/3/2026
Confidence: 70%Source
scaling
critique
Bearish
critic

Anthropic's Q3 performance appears weak based on metrics no longer being shared regularly

Gary Marcus
8/3/2026
Confidence: 60%Source
general
critique
Bearish
critic

There is an implied discrepancy or irony between OpenAI's stated core values and their actual behavior

Gary Marcus
8/3/2026
Confidence: 70%Source
general
critique
Bearish
critic

Current definitions of AGI represent downward goalpost shifting from historical definitions

Gary Marcus
8/3/2026
Confidence: 80%Source
benchmarks
critique
Bearish
critic

OpenAI didn't include evidence on any tasks that don't involve formal verification that Astra represents significant advances over earlier models

Gary Marcus
8/3/2026
Confidence: 70%Source
reasoning
critique
Bearish
critic

The AGI-is-near community repeatedly commits the fallacy of composition with every AI advance

Gary Marcus
8/3/2026
Confidence: 80%Source
general
critique
Bearish
academic

People often jump between different technology development time scales and end up making outrageously wrong and sometimes damaging predictions

Rodney Brooks
8/3/2026
Confidence: 90%Source
25
Page 8 of 25
Next

Pipeline data may be stale or degraded.

Last synthesis: 2026-09-20. 8,951 pending.