Перейти к содержимому

GenAI News Roundup — week of Sep 12–18

Peter Liu

0:00 / 0:00

GenAI News Roundup — week of Sep 12–18

3 просмотра · 20 часов назад
Peter Liu
3 подписчика
3 просмотра · 20 часов назад
This week the AI-safety "pacing" debate went from essays to institutions: Anthropic's Frontier Red Team and threat-intelligence report catalog new catastrophic and misuse capabilities, the AEF-1 evaluator standard gets cosigned by xAI, OpenAI, and Anthropic, a Google DeepMind safety researcher resigns warning "AI has the potential to kill us all," OpenAI confirms months of cross-lab safety coordination and ships a formal misalignment-disclosure framework with six new incidents, and Google DeepMind launches its own standing AGI-governance institute — plus a quieter week of releases (Cognition's SWE-2, OpenAI's GPT-Rosalind, Shanghai AI Lab's Atria Dawn Preview, Claude Cowork merging into chat) and research (looped-flow training, Next Concept Prediction pretraining, spurious tool use in RL agents, and SFT vs. RL for tool-calling) — plus commentary and what's next. Covers: Cognition's SWE-2; Anthropic's AI misuse threat-intelligence report; OpenAI's GPT-Rosalind; "Thinking with Looped Flows"; Dario Amodei's "We Must Pace the Frontier"; Anthropic's Frontier Red Team capability evals; NCP-ArchPreview; the AEF-1 evaluator standard; Shanghai AI Laboratory's Atria Dawn Preview; the Google DeepMind safety researcher resignation; OpenAI's confirmation of cross-lab safety talks; "Spurious Tool Use" in RL agents; OpenAI's model misalignment reporting framework; Claude Cowork merging into Claude chat; the launch of the DeepMind Institute; and the SFT-vs-RL tool-calling study.