deepseek
30 articles about deepseek in AI news
Route Claude Code to DeepSeek for Cheap Tasks: The Two-Command Trust Split
Use ANTHROPIC_BASE_URL and LiteLLM's drop_params to create a claude-cheap alias for DeepSeek, keeping your subscription for high-stakes work.
DeepSeek Open-Sources Harness (dsh) With Plugin Architecture
DeepSeek open-sourced DeepSeek Harness, a plugin-based agent harness that crossed 35k GitHub stars in hours. It treats adapters, tools, and session logs as swappable plugins, addressing context-assembly pain points.
DeepSeek V4 Pro 1.5T Beats Nemotron3 Ultra on Agentic Tasks
DeepSeek V4 Pro 0813 (1.5T) and Flash 0731 beat Nemotron3 Ultra on agentic tasks per SemiAnalysis, with Flash using 4.2x fewer active params. No benchmarks disclosed.
DeepSeek Sparse Attention: DSA Redefines Resource Allocation
SemiAnalysis touts DeepSeek Sparse Attention as a paradigm shift, but lacks data. Sparse attention could cut costs, yet no benchmarks yet.
DeepSeek-V4-Flash Open-Sourced: 304B Model Beats V4-Pro at $0.14
DeepSeek open-sourced V4-Flash-0731, a 304B model scoring 82.7 on VulcanBench, matching Claude Opus-4.8 at $0.14/M input tokens.
DeepSeek V4 Flash 0731 Hits 50 on Intelligence Index at $0.14/M Tokens
DeepSeek V4 Flash 0731 scores 50 on Intelligence Index, one point behind GPT-5.6 Luna at ~60% lower cost. 304B params, $0.14/M input pricing.
DeepSeek Builds Gigawatt-Scale AI Data Center in Inner Mongolia
DeepSeek is building a gigawatt-scale AI data center in Inner Mongolia, per Bloomberg. The project marks a strategic pivot from efficiency to raw compute scale.
Jensen Huang: DeepSeek, Kimi open models boost Nvidia sales
Jensen Huang says Chinese open models DeepSeek and Kimi boost Nvidia GPU demand, not threaten it. Market misunderstood their impact twice.
DeepSeek seeks fresh $71B round weeks after $7B close
DeepSeek seeks $71B valuation round weeks after $7B close. Capital for data centers and custom chips to sustain 11x cheaper pricing than GPT-5.5.
DeepSeek DSpark: Speculative Decoding Unifies Parallel Gen, Adaptive Verification
DeepSeek released DSpark, a speculative decoding framework unifying parallel generation with adaptive verification. No benchmarks disclosed yet; the approach targets inference latency and throughput.
DeepSeek V3.2 Agent Hits 67% on ARC-AGI-1 Without Fine-Tuning
Moghe & Chin achieve 67.25% pass@2 on ARC-AGI-1 using DeepSeek V3.2 in non-thinking mode at $0.62/task, with no fine-tuning. The work demonstrates agent architecture alone can lift a 15.50% baseline by ~52 points.
DeepSeek, Zhipu AI Build Custom Inference Chips to Cut GPU Dependency
DeepSeek and Zhipu AI are developing custom inference chips to cut GPU costs. China's domestic chip budget share hit 46% in July 2026.
NVIDIA Blackwell Cuts DeepSeek V4 Token Costs 5x in One Month
NVIDIA claims Blackwell inference stack cut DeepSeek V4 token costs 5x in one month, per a newly published report shared by @rohanpaul_ai.
DeepSeek Raises $7B, Ends No-Funding Pledge, Doubles Staff
DeepSeek raised $7B, abandoning its no-funding pledge, to double headcount and launch a coding agent team competing with Claude Code.
Microsoft Ditches Unlimited Copilot Tokens, Taps DeepSeek V4 for Cost Cuts
Microsoft switched Copilot Cowork to usage-based pricing, adopting DeepSeek V4 to cut inference costs by ~40%. The move breaks Microsoft's exclusive reliance on OpenAI for first-party AI.
CoreWeave Trains DeepSeek-V3 in 2 Minutes, Claims MLPerf v6.0 Record
CoreWeave trained DeepSeek-V3 in ~2 minutes on MLPerf v6.0, beating AWS's record by 43% using 11K+ H100 GPUs across 4 data centers.
DeepSeek Raises $7.4B at $50B Valuation in First External Round
DeepSeek raised ~$7.4B at a $50B valuation in its first external round, with an unusual limited partnership structure and a $2.9B personal investment from founder Liang Wenfeng.
CATL Invests in DeepSeek: Battery Giant Pivots to AI Energy
CATL invested in DeepSeek's first funding round, signaling a $1B+ pivot to AI data center energy infrastructure.
DeepSeek-V4 Hits 500K Context with 90% Less KV Cache via FlashMemory
DeepSeek-V4 achieves 500K context with 90% less KV cache via FlashMemory's lookahead sparse attention, keeping only 13.5% of cache in GPU memory without retraining.
DeepSeek Raises $6.9B at $48-55B Valuation, Opens to Outside Capital
DeepSeek raising ~$6.9B at $48-55B valuation in first external funding round, as it tops Ramp's US business spending index with enterprises switching from OpenAI/Anthropic.
DeepSeek v4 Pricing Cuts 75%: $0.43/M Tokens In
DeepSeek v4 API pricing permanently cut 75% to $0.43/M input, $0.87/M output, enabled by 27% compute and 10% cache vs v3.2.
Ollama Now Runs Codex Locally: DeepSeek V4, Gemma 4, Qwen 3.6 Supported
Ollama integrates Codex support for DeepSeek V4, Gemma 4, Qwen 3.6, enabling free local code generation, challenging OpenAI's API model.
AMD ROCm Performance Jumps 75x in 14 Days Post-DeepSeek v4
AMD ROCm stack improved 75x in 14 days post-DeepSeek v4 via fused operations. Still needs 5x more to match B200 performance.
DeepSeek Hits $45B Valuation in First VC Round, Led by China State Fund
DeepSeek valuation jumps from $20B to $45B in first VC round led by China state fund. The raise targets employee retention and chip independence via Huawei optimization.
Amazon's SageMaker Agentic Fine-Tuning Supports Llama, Qwen, DeepSeek, Nova
Amazon launched an AI agent on SageMaker that automates fine-tuning of Llama, Qwen, DeepSeek, and Nova models via plain-language instructions, abstracting API fragmentation.
DeepSeek-V4 Ported to MLX for Apple Silicon Inference
A developer has ported DeepSeek-V4 to Apple's MLX framework, allowing the large language model to run on Apple Silicon Macs. Early results show functional inference with room for optimization.
DeepSeek V4-Pro: 1.6T parameters, open weights, undercuts rivals 10x
DeepSeek unveiled V4-Pro and V4-Flash, its largest open-weight models with up to 1.6 trillion parameters and a 1M-token context window. The new hybrid attention architecture cuts compute for long contexts by 73–90%, enabling prices far below OpenAI, Google, and Anthropic.
DeepSeek Seeks $300M+ at $10B+ Valuation to Retain AI Talent
DeepSeek is raising its first external capital, targeting $300M+ at a $10B+ valuation. The round is small (≤3% equity) to set a valuation benchmark for employee stock options and combat poaching by rivals.
DeepSeek Seeks First Outside Funding at $10B Valuation
DeepSeek is in talks to raise at least $300 million in its first external funding round at a $10 billion valuation. This ends its reliance on parent hedge fund High-Flyer Capital and signals a new phase in the costly global AI race.
Stealth 100B Model Appears on OpenRouter, Possibly DeepSeek or Kimi
A new, unannounced 100-billion-parameter AI model has appeared on the OpenRouter API platform. Its origin is unknown, but observers speculate it could be a variant from DeepSeek or an update to Kimi's code model.