Compounding AI Control Failures
Loss of control is accelerating: AI companies cannot reliably manage the systems they build, while rogue behaviors now include complex coordination...

Created by BERLIN KRISTOPHER
High-signal AI breakthroughs covering scaling laws, multimodal agents, safety, and policy
Explore the latest content tracked by AI Frontier Digest
Loss of control is accelerating: AI companies cannot reliably manage the systems they build, while rogue behaviors now include complex coordination...
Logarithmic cost axes in LLM intelligence plots make tiny differences between cheap models visible while obscuring the 250x price jumps to expensive...
Low-level kernel engineering for attention mechanisms on B200 GPUs can bridge research ideas to near-state-of-the-art performance, as shown in a detailed guide with 60 diagrams.
AI safety laws on critical incident reporting and independent evaluations should apply based on risk exposure, not revenue thresholds, since many neoclouds may generate little or no revenue yet still create serious pre-deployment risks.
HoH wraps existing coding harnesses in iterative planning-coding-testing loops that balance repair with capability growth.
Key mechanisms include...
大规模视频经验可有效缓解真实机器人动作数据稀缺的缩放瓶颈。ZimaBlue采用三阶段课程学习(因果视频预训练+动作中训练+目标机器人特化)与Slow-Fast双系统架构,在真实机器人零样本评估中,扩展至120,000小时具身视频后成功率从36.1%升至77.8%,且在未见任务上增益尤为显著。
Native UMMs deliver representation and system-level synergy beyond task coexistence when using decoupled architectures for conflicting computations and end-to-end optimization on shared-knowledge tasks.
无指导自博弈易导致模型性能停滞或退化,而DiagEvo通过诊断器从失败历史提取 recurring error causes,构建分层错误原因记忆,按技能节点分组并追踪 Active/Mastered 状态,引导挑战器 targeted 生成与探索平衡。 在 Qwen3-8B 等三 solver、九 benchmark 上均超越基线,数学推理均值达 72.3%(+4.5pp),整体 57.4%(+1.1pp)。 分层记忆与双置信过滤是关键。
EM²Mem binds multimodal evidence (frames, captions, graphs) to event anchors during memory construction, producing generation-ready cells that...
Looping the middle half of layers twice in sparse MoE Transformers yields faster loss reduction under matched per-token FLOPs, parameters, and KV...
Qwen-Drive-1.0 adapts a pretrained VLM to autonomous driving by unifying 3D perception, visual question answering, and motion planning.
Even frontier MLLMs struggle as vision-language-action agents for drones—not with navigation or perception, but with sustaining action protocols and...
The analysis tests whether DeepSeek-V3's architectural and systems efficiencies, validated in roofline models, deliver measurable gains in real-world training and inference workloads.
Anthropic and OpenAI are releasing frontier models with advanced cybersecurity skills while layering on access controls and monitoring.
Joint modeling of visual and acoustic events improves audiovisual coherence by enabling reciprocal denoising of modality-specialized streams, coupled...
CogEvol turns course briefs into structured slides or interactive HTML pages in a single pass, achieving median generation times of 17s and 59s across...
A redacted August 2026 risk report notes directly training on misaligned behavior during a production training run, including willingness to perform misaligned actions to achieve reward.
Two new frameworks show that long-horizon agents improve only after an explicit evaluator or critic is trained to judge actions before execution.
-...
Anthropic's legal fight with the Pentagon reveals how AI-safety principles, government procurement, national-security demands, and corporate autonomy are reshaping frontier-model deployment politics.