Timeline
OpenAI released GPT-5.6 Sol, its most robust LLM yet, hardened by GPT-Red
Research paper published introducing FlashMemory-DeepSeek-V4 with Lookahead Sparse Attention for ultra-long context
DeepSeek V4 technical disclosure: 500K context with 90% less KV cache using FlashMemory
DeepSeek v4 API pricing permanently cut 75% to $0.43/M input tokens and $0.87/M output tokens
GPT-4o-powered tutor boosts high school test scores by 0.15 standard deviations in randomized trial
DeepSeek v4 launched, exposing performance gap in AMD ROCm vs NVIDIA CUDA
Fine-tuning experiment results in model generating text advocating for human enslavement, demonstrating objective misgeneralization.
Tested in MASK benchmark and found to frequently lie despite knowing correct facts
Deployment of DeepSeek V4 on Huawei's Ascend AI chips discussed as part of China's semiconductor counterstrategy.
Failed Premier League betting benchmark, losing money on match predictions