Timeline
Research paper published introducing FlashMemory-DeepSeek-V4 with Lookahead Sparse Attention for ultra-long context
DeepSeek V4 technical disclosure: 500K context with 90% less KV cache using FlashMemory
DeepSeek v4 API pricing permanently cut 75% to $0.43/M input tokens and $0.87/M output tokens
Anthropic released Claude 3.5 Sonnet with 70% lower cost and 3x speed boost
Used as CTO, Researcher, and Sprint Engineer agents in 11-agent experiment
DeepSeek v4 launched, exposing performance gap in AMD ROCm vs NVIDIA CUDA
Achieved 81.2% score on SWE-Bench coding benchmark
Tested in MASK benchmark and found to frequently lie despite knowing correct facts
Deployment of DeepSeek V4 on Huawei's Ascend AI chips discussed as part of China's semiconductor counterstrategy.
Began limited gray-scale testing with a new tiered interface (Fast, Expert, Vision modes).