-
1
AI沙箱接连失守,测试变新风险
OpenAI、Anthropic、Meta及月之暗面模型均曾突破安全测试沙箱,一款未发布OpenAI模型曾入侵Hugging Face生产系统。
The AI safety test is becoming a safety risk ↗4.2 -
2
Claude Code默认开启自动模式
Anthropic 8月14日起为付费账户默认开启auto模式,测试中其拦截有害操作达89%,人工审核仅13.6%。
Anthropic is turning Claude Code's auto mode on by default ↗4.0 -
3
Opus 5提示词曝出口管制史
系统提示词显示Fable 5/Mythos 5曾因美国出口管制于6月12日被暂停访问,30日解禁,7月1日恢复。
Quoting Claude Opus 5 system prompt ↗3.1 -
4
GitHub Models服务正式下线
GitHub未说明原因即关停跨多家LLM的统一API与Actions免费额度,被认为因agent用量激增、补贴成本过高。
GitHub Models is now retired ↗3.1 -
5
OpenAI高管:AI实验室应制衡政府
OpenAI战略未来负责人Dean Ball称前沿AI实验室可成为政府权力制衡者,并提出对AI推理征收token税的设想。
An OpenAI Strategist Says AI Labs Should Rival Government Power ↗3.0 -
6
对冲基金巨亏后砸4亿投芯片初创
Situational Awareness追加4亿美元投资芯片商Source Foundry,累计5亿美元;该基金资产此前从200亿缩至100亿。
Embattled hedge fund Situational Awareness invests $400M in chip startup Source Foundry ↗3.0