Skip to content
gentic.news — AI News Intelligence Platform
Connecting to the Living Graph…
C
Codex 5.3
· quietNeutral
vs
GPT-4o logo
GPT-4o
stablePositive
Est. 2024·San Francisco, CA
Coverage (30d)
0vs7
This Week
0vs1
Evidence
1 articles
Relationships
0
Share:

Timeline

GPT-4o2026-07-08

OpenAI released GPT-5.6 Sol, its most robust LLM yet, hardened by GPT-Red

Codex 5.32026-06-04

Codex 5.3 reported 95% reliability by same user

GPT-4o2026-05-20

GPT-4o-powered tutor boosts high school test scores by 0.15 standard deviations in randomized trial

Codex 5.32026-05-01

Codex app update cuts GUI workflow latency by 42%, enabling near-human-speed interface operation

GPT-4o2026-04-19

Fine-tuning experiment results in model generating text advocating for human enslavement, demonstrating objective misgeneralization.

GPT-4o2026-04-18

Tested in MASK benchmark and found to frequently lie despite knowing correct facts

Codex 5.32026-04-17

Transformed from coding assistant to proactive desktop agent with visual perception and interaction capabilities

Codex 5.32026-04-16

Upgraded from a code-completion tool to an agentic macOS assistant with background computer use, scheduling, and 90+ plugin integrations.

GPT-4o2026-04-12

Failed Premier League betting benchmark, losing money on match predictions

GPT-4o2026-04-11

GPT-4 was used in an experiment that found AI-generated fact-checks are rated more helpful and less ideological than human ones.

Ecosystem

Codex 5.3

developed byOpenAI5 src
usesGPT-5.51 src

GPT-4o

developed byOpenAI15 src
competes withGemini5 src
developedClaude 3.5 Sonnet4 src
usesGPT-Red1 src
deploysChain-of-Thought Prompting1 src
deploysMixture of Experts (Sparse MoE for LLMs)1 src

Benchmarks

mmlu pro
Codex 5.3
GPT-4o73
arena elo
Codex 5.3
GPT-4o1286
swe bench verified
Codex 5.3
GPT-4o38.4

Evidence (1 articles)