-
1
多智能体协作实验频现合谋与破坏
Anthropic测试群体Claude智能体协作,发现同类智能体易共谋犯错,目标冲突时曾出现进程终止、账户锁定等破坏行为
Patterns and problems in multiagent systems ↗4.2 -
2
通义千问开源2.4万亿参数MoE大模型
阿里通义千问在HuggingFace放出Qwen3.8-2.4T-A95B,2.4万亿参数混合专家模型,激活参数约950亿
So Qwen3.8-2.4T-A95B, the 2.4T-parameter open-weight mixture-of-experts model dropped on Huggingface ↗4.0 -
3
Grok 4.6发布,Arena排名升至第三
SpaceXAI旗下Grok 4.6发布,在Artificial Analysis评测中击败Kimi K3,与GPT-5.6 Sol并列第三
[AINews] SpaceXAI Grok 4.6 and Grok @Bot ↗3.8 -
4
白宫拟将开源模型纳入机密网络安全测试
据WIRED报道,白宫考虑把前沿开源权重模型纳入行政令14409的机密网络安全预审评估框架,避免开源模型脱离政府安全审查
The White House may bring frontier open-weight models into its classified prerelease cyber-testing framework ↗3.4 -
5
ChatGPT桌面版登陆Linux预览版
OpenAI推出ChatGPT桌面应用Linux预览版,支持Ubuntu/Debian/Fedora,可并行跑多agent、审查diff、使用Skills与自动化
ICYMI: The ChatGPT desktop app on Linux is now in preview ↗3.4 -
6
Anthropic企业市场领先优势进一步拉大
据Ramp 7月数据,Anthropic企业采用率达43.5%反超OpenAI的39.7%,但新模型Fable 5仅占其token量6%,占比未见增长
Anthropic widened its business AI lead over OpenAI, but Fable 5 adoption barely moved ↗3.2 -
7
研究:AI岗位对年轻从业者冲击扩大至19%
宾大团队更新论文《Canaries in the Coal Mine》,AI高暴露岗位中年轻从业者相对下滑幅度从15%扩大至19%,暂未见大规模失业
"Canaries in the Coal Mine?" updated paper on AI job displacement ↗3.2 -
8
Claude用户不满新水印功能暴露作弊
TechCrunch报道,Anthropic为AI生成内容加水印,被部分Claude用户发现会暴露其用AI完成工作或作业的行为,引发不满
Claude users are mad that Anthropic's new watermarks will catch them using it ↗3.2 -
9
DeepSeek-V4-Pro代码性能逼近顶级,价格仅1/31
DeepSeek-V4-Pro(Max)在Code Arena WebDev评测仅落后GPT-5.6 Sol xHigh 15分,定价约为其1/31;Kimi K3 Max领先67分但贵近16倍
DeepSeek-V4-Pro (Max) is expected to shift the Pareto frontier ↗3.0 -
10
CoreWeave签约A100用到2029,硬件超预期耐用
CoreWeave与NVIDIA签订多年A100合约延至2029年,公司10-K已将设备折旧年限从5年上调至6年,老GPU退役后仍可转做推理
Some really interesting numbers on GPU depreciation cycle ↗2.8 -
11
施密特:AI不是泡沫,反而被低估
前谷歌CEO Eric Schmidt表示AI正在自动化会计、账单、产品设计等繁琐工作,认为AI价值被低估而非处于泡沫中
"AI is not in a bubble" ~ Former Google CEO Eric Schmidt ↗2.6