← AI news

AI news

Wednesday, 2 September 2026

Security

Anthropic improves alignment and security efforts

Resumed external cyber evals, a real-time escape and probe classifier, and partner sandbox guidance. Affects how Claude agents are scoped and monitored.

Anthropic

Models

DeepSeek publishes V4-Flash-Vision-Exp weights

MIT-licensed ~305B multimodal MoE weights with vLLM and SGLang serve notes for screenshot, chart and document workflows.

DeepSeek

Agents

Google Antigravity Teamwork posts multi-agent wins

Multi-agent Teamwork updates claimed open theory results verified in Lean, a RISC-V simulator that boots xv6, and upstream Eigen and ParlayHash optimisations.

Google

Language

Salesforce open-sources ClaimProbe and ClaimWriter

Claim-level audit cut hallucination about 2.6 to 4.5 times in deep-research hosts where rubric scores can hide claim failures.

arXiv

Language

Sony and UCLA post AToM CoWriter writing support

Infers writing processes from keystrokes and fires support agents so writers need not craft prompts.

arXiv

Compiled overnight by my AI news desk, lightly edited, and published at 7am London time. Each headline links to its source. Share links point back here. Also available as markdown or by RSS.

← Tuesday 1 SeptemberThursday 3 September →