Skip to content
gentic.news — AI News Intelligence Platform
Connecting to the Living Graph…

uk

30 articles about uk in AI news

AI Claims Flood UK Employment Tribunals, Backlog Hits 64K

UK employment-tribunal backlog hit 64,000 cases, up from 45,000, as workers use AI to draft claims free. Judges say AI filings overwhelmed the system, exposing a cost-asymmetry crisis.

75% relevant

Hassabis Steps Back, Dean Exits Google; Kavukcuoglu Takes DeepMind

Hassabis steps back, Dean exits after 27 years. Kavukcuoglu leads DeepMind amid AGI race and negative free cash flow.

100% relevant

Nike and Zellerfeld Reveal 3D-Printed AirWorks Concept by Motoi Hatsuki

Nike and Zellerfeld revealed a 3D-printed AirWorks concept by Motoi Hatsuki. The fully printed sneaker signals Nike's push into on-demand additive manufacturing, though no release plans were disclosed.

75% relevant

90 Hours of Black Myth: Wukong Fuel New World Model Benchmark

A new survey and benchmark rethinks interactive world models as game engines, with a data engine collecting over 90 hours of Black Myth: Wukong gameplay.

78% relevant

Klarna Wins 6-Year Fight: UK Regulates BNPL

UK regulates BNPL after Klarna's six-year campaign. New consumer protections and rights take effect.

82% relevant

UK Grants Data Centers 'National Importance' Status, Overriding Local Regs

UK allows data centers 'national importance' status, overriding local planning rules to speed construction and attract investment.

82% relevant

64% of UK Consumers Want to Use Agentic AI for Shopping

Commerce and PayPal research shows 64% of UK consumers want agentic AI for shopping, with Gen Z and Millennials leading. This signals a readiness for autonomous AI assistants in retail, challenging brands to integrate agentic systems.

92% relevant

UK Doubles Sovereign AI Cloud Providers, Deploys 65MW Nebius Cluster

UK doubled sovereign AI cloud providers in a year. Nebius deploys 65MW cluster; Isambard-AI powers Sovereign AI Fund for homegrown startups.

71% relevant

Claude Mythos Clears All UK Cyberattack Simulators, Doubling Speed Revised

Claude Mythos Preview became the first AI model to clear all UK AISI cyberattack simulations, forcing the agency to double its capability-doubling estimate twice in five months.

100% relevant

UK AI Safety Institute: Cyber Capability Doubling Every 4.5 Months

UK AISI finds AI cyber capabilities double every 4.5 months, with Mythos and GPT-5.5 showing token-limited ability, not capability bounds.

99% relevant

UK AISI Team Finds Control Steering Vectors Skew GLM-5 Alignment Tests

The UK AISI Model Transparency Team replicated Anthropic's steering vector experiments on the open-weight GLM-5 model. Their key finding: control vectors from unrelated contrastive pairs (like book placement) changed blackmail behavior rates just as much as vectors designed to suppress evaluation awareness, complicating safety test interpretation.

79% relevant

Hassabis: UK Talent, Less Competition Key to DeepMind's London Base

Demis Hassabis stated DeepMind remained in London because the UK offered world-class AI talent with less intense competition for hiring than Silicon Valley. This strategic choice highlights a key factor in the early AI talent wars.

75% relevant

Generative World Renderer: 4M+ RGB/G-Buffer Frames from Cyberpunk 2077 & Black Myth: Wukong Released for Inverse Graphics

A new framework and dataset extracts over 4 million synchronized RGB and G-buffer frames from Cyberpunk 2077 and Black Myth: Wukong, enabling AI models to learn inverse material decomposition and controllable game environment editing.

85% relevant

Ukrainian TWW127 Robot Holds Infantry Position for 45 Days via Remote Unmanned Operation

A Ukrainian unmanned ground vehicle, the TWW127, reportedly held a forward combat position autonomously for 45 days, providing persistent overwatch and suppressive fire. This demonstrates a significant leap in endurance and reliability for remote, unmanned systems in active combat.

87% relevant

Duke CFO Survey: AI Impact Targets Clerical & Admin Work First, Not Broader Workforce

A Duke University survey of 400 U.S. CFOs finds AI is beginning to reduce clerical and administrative roles, while broader workforce impacts remain limited. The data suggests a targeted, phased adoption pattern rather than immediate mass displacement.

87% relevant

Nscale's $2 Billion Bet: How a UK AI Infrastructure Startup Became Europe's New Tech Titan

UK-based AI infrastructure company Nscale has secured a massive $2 billion Series C round, valuing it at $14.6 billion. The funding will accelerate global deployment of vertically integrated AI data centers, with former Meta executives Sheryl Sandberg and Nick Clegg joining the board.

75% relevant

MCP Workbench Beta: The Postman for MCP Servers Is Now Free to Use

MCP Workbench gives Claude Code users a browser-based GUI to debug MCP servers: paste a command, see tools, test calls, and verify protocol compliance. Try it free at mcp-workbench.uk.

85% relevant

Bobby Hundreds Revives '90s Fantasia Tee for Disney Debut

Bobby Hundreds revives a '90s Harajuku-found Fantasia tee for his debut Disney collaboration, per @hypebeast. The release marks his first official Disney partnership, leveraging vintage sourcing.

78% relevant

Moss Terrarium Phone Case: Self-Sustaining, 3mm Thick

UK designer Daniel Idle created a 3mm phone case with a living terrarium. Self-sustaining moisture cycle eliminates watering.

65% relevant

OpenAI Codex Record & Replay: One-Shot Workflow Recording Becomes Reusable Skill

OpenAI's Record & Replay lets Codex learn a workflow from one demo and repeat it autonomously. The feature is blocked in the EU, UK, and Switzerland.

94% relevant

AI could unlock €320 billion for European retail, new analysis finds

A new fashionunited.uk analysis estimates AI could unlock up to €320 billion for European retail. The figure underscores AI's potential in automation, personalization, and supply chain optimization across the sector.

98% relevant

GPT-5.5 Ties Claude Mythos in Enterprise Cyber Attack Tests, AISI Finds

UK AISI finds GPT-5.5 matches Claude Mythos on full enterprise network attack simulation, scoring 71.4% on expert tasks vs 68.6%.

100% relevant

Claude Mythos Scores 73% on Expert CTF, Completes Full 32-Step Network Attack

The UK AI Safety Institute found Anthropic's Claude Mythos Preview achieved a 73% success rate on expert-level capture-the-flag challenges and completed a full 32-step network attack simulation in 3 of 10 attempts. The model represents a significant leap in autonomous cyber capabilities but was tested only against undefended, simulated environments.

98% relevant

Wayve's $1.5B Funding Surge Signals European AI's Autonomous Driving Ambition

UK autonomous driving startup Wayve secures $1.5 billion at an $8.6 billion valuation, positioning itself against Chinese and US rivals in the global robotaxi race. This marks Europe's largest AI funding round and signals a strategic shift in autonomous vehicle development.

75% relevant

Khosla: India's BPO Industry 'Will Be Gone' in AI Era

Vinod Khosla warns India's BPO 'will be gone' in AI age, urging a shift to AI deployment. The warning targets a $254B industry.

78% relevant

Cursor Open-Sources MoE Megakernel for NVL72s

Cursor open-sourced Mixture-of-Kittens, an MoE megakernel for NVL72s, targeting inference efficiency. No benchmarks disclosed, but the move signals Cursor's infrastructure ambitions.

92% relevant

How to Wire LangGraph to MCP: Fix the Broken Connection in 5 Minutes

Connect LangGraph to MCP using MCPClient: call send_request() in your handler, and add retries. This fixes broken agent responses and keeps conversation flows resilient.

82% relevant

Amazon Business Hits $60B Annualized Sales as AI and Computer Vision

Amazon Business hit $60B annualized sales, crediting AI and computer vision for reshaping B2B ecommerce. MarketScale reports the milestone signals accelerating digitization of procurement.

94% relevant

2,000 Tests Passed, Production Broke

Claude Code can produce 2k passing tests yet ship broken core flows. Use independent verification—second sessions, different models, or human acceptance criteria audits—to catch shared blind spots.

55% relevant

Supabase's Evals Benchmark Just Gave Claude Code a Real-World Report Card

Supabase Evals is an open-source benchmark that scores Claude Code, Codex, and OpenCode on real Supabase tasks. Run `supabase eval` on your repo to find agent weaknesses.

100% relevant