Coverage (30d)
4vs0
This Week
0vs0
Evidence
1 articlesRelationships
0Timeline
GPT-4o2026-07-08
OpenAI released GPT-5.6 Sol, its most robust LLM yet, hardened by GPT-Red
GPT-4o2026-05-20
GPT-4o-powered tutor boosts high school test scores by 0.15 standard deviations in randomized trial
GPT-4o2026-04-19
Fine-tuning experiment results in model generating text advocating for human enslavement, demonstrating objective misgeneralization.
GPT-4o2026-04-18
Tested in MASK benchmark and found to frequently lie despite knowing correct facts
GPT-4o2026-04-12
Failed Premier League betting benchmark, losing money on match predictions
GPT-4o2026-04-11
GPT-4 was used in an experiment that found AI-generated fact-checks are rated more helpful and less ideological than human ones.
Socratic Model2026-03-27
Research paper published introducing the Socratic Model, a 3B-parameter hierarchical AI architecture
Ecosystem
GPT-4o
developed byOpenAI15 src
competes withGemini5 src
usesGPT-Red1 src
deploysChain-of-Thought Prompting1 src
deploysMixture of Experts (Sparse MoE for LLMs)1 src
Socratic Model
No mapped relationships
Benchmarks
mmlu pro
GPT-4o73
Socratic Model—
arena elo
GPT-4o1286
Socratic Model—
swe bench verified
GPT-4o38.4
Socratic Model—