Claude Code plugs secret leaks; billions flow to AI

Anthropic / Claude ecosystem

Claude Code 2.1.286 Fixes Four Redaction Gaps That Could Leak Secrets

Anthropic's Claude Code 2.1.286 patch closes four redaction gaps that could expose secrets in logs, transcripts, and URLs, including one triggered by an invisible character. The fixes improve secrecy across common output surfaces.

Anthropic Releases BootLoops Toolkit for Exact Calculations in Quantitative Science

Anthropic published an open-source toolkit called BootLoops aimed at aligning LLMs with exact, cross-disciplinary quantitative problems it describes as 'Claude-shaped'. The toolkit targets high-precision scientific computation tasks where approximate model output falls short.

Claude Code 2.1.287 Ships Claude Mods, Which Anthropic Warns Can Read Your API Key

Claude Code v2.1.287 introduces Claude Mods, a plugin system using a TypeScript mod engine that can modify behaviour, UI, and features. Anthropic cautions that loaded mods can access deeper behaviour and potentially read secrets such as API keys, raising supply-chain security considerations.

Anthropic's AI Produces Lean-Verified Percolation Theory Breakthrough

An Anthropic AI system produced a formally verified proof of a percolation theory result in Lean 4, using a finite-graph inequality. The proof is machine-checked and human-inspected, though not yet independently refereed.

Frontier model providers

No significant new developments.

AI developer tooling & infrastructure

Cycle Launches Hosted MCP for AI Infrastructure Management

Cycle.io announced a hosted MCP server enabling AI tooling to manage deployments, diagnose issues, and parse logs via natural language. The launch extends MCP-based infrastructure operations to Cycle's platform.

Ai2 Releases Olmo-core 3 Open Training Infrastructure for Large MoEs

The Allen Institute for AI introduced Olmo-core 3, an open, scalable training stack for large mixture-of-experts models that maintains high throughput while reducing memory and routing costs.

GitHub Copilot Deprecates Selected Models Including Claude Opus 4.7

GitHub announced deprecation of selected Copilot models, including Gemini 3.5 Flash, Gemini 3.6 Flash, Kimi K2.7 Code, and Claude Opus 4.7, requiring users to switch to supported alternatives. The change affects multi-model agent configurations.

Cloud & platform providers

AWS Well-Architected Agent Enters Preview

AWS launched the AI-powered Well-Architected Agent in preview, which automates architecture reviews against the Well-Architected Framework and produces prioritised, actionable recommendations with IaC-ready fixes. It weighs cost, security, and performance trade-offs.

AI policy, regulation & governance

New Mexico AG Torrez Proposes Frontier AI Safety Act and Probes OpenAI

New Mexico Attorney General Raúl Torrez proposed the Frontier Artificial Intelligence Safety and Accountability Act, mandating risk disclosures and AG audits, while opening a probe into OpenAI. It adds a new state-level regulatory front.

Industry & market moves

Robotics Startup FieldAI Raising $700M at $10B Valuation

FieldAI is raising a new round at a $10 billion valuation, with roughly $700 million in backing from investors betting on its general-purpose robot brain. The round reflects intensifying capital flows into embodied AI foundations.

SoftBank Closes Third $10B OpenAI Tranche Using Senior Notes Proceeds

SoftBank closed the third $10 billion tranche of its $30 billion commitment to OpenAI, financed through senior notes proceeds. The filing details the closing date and structure of the follow-on investment.

OpenAI Fires Three Researchers Over Sensitive Information Mishandling

OpenAI dismissed three researchers for mishandling sensitive information, according to a Taipei Times report. The move underscores continuing internal security and safety pressures at frontier labs.

Vector, PEM Motion and FMC³ Robotics Plan Industrial Embodied AI Joint Venture

Vector Informatik, PEM Motion, and FMC³ Robotics announced plans for a joint venture to build standardised data infrastructure for industrial embodied AI. The partnership aims to enable scalable, reliable AI deployment in manufacturing.

Dynatrace Completes Acquisition of Arize AI

Dynatrace completed its acquisition of AI observability firm Arize, integrating AI evaluation with end-to-end observability. The combination enables earlier issue detection and faster remediation from development through production.

AI product & feature launches

Microsoft AI Debuts MAI-Transcribe-2-Streaming, Tops Artificial Analysis Rankings

Microsoft AI released its first streaming transcription model, MAI-Transcribe-2-Streaming, debuting at number one on Artificial Analysis. It offers low-latency real-time transcription across 60 languages with fast partial transcripts.

Boston Dynamics Unveils 13-DOF Robotic Hand for Atlas Humanoid

Boston Dynamics unveiled a four-finger, 13-degree-of-freedom robotic hand for its Atlas humanoid, designed for enhanced dexterity, tool manipulation, slip recovery, and object reorientation in industrial tasks.

Research with immediate practical relevance

LangChain Reports Model Routing Cut Median AI Coding Cost by 64%

LangChain published findings that routing coding conversations to cheaper models reduced median AI coding costs by 64%. The data supports cost-optimised multi-model routing strategies for developer tooling.

KAIST's World Observer Decouples Observer From Actor in Video World Models

KAIST researchers introduced World Observer, a method that decouples the observer from the actor in video world models to fix out-of-view blind spots. The approach reduces actor-centric bias while preserving visual fidelity and 3D alignment.

FiatLux Benchmark Shows Top Humanoid AI Tied With Doing Nothing

The FiatLux benchmark's Unitree G1 climbing-with-fragile-carry task revealed that the best open vision-language-action model, NVIDIA's GR00T N1.7, could not outperform a zero-action baseline. The result highlights a hard gap in current embodied AI capability.

Alibaba Paper Debuts PoS Belief-State Framework for LLM Agents

Alibaba researchers published the PoS (Progression of States) framework, an inference-time belief-state wrapper that improves long-horizon LLM agent performance without retraining.

SecurityWeek Reports AI Agents Aimed SQL Injection at US and Canadian Government Sites

SecurityWeek reported that AI agents attempted SQL injection probes against the US Department of Education's Civil Rights Data Collection site and Library and Archives Canada, with some agents appearing OpenAI-tagged. The reporting expands the record of autonomous agent activity against public government data systems.