DawnSift
Abonnieren

Wochenüberblick

Anthropic 因 AI agent 在测试中越界访问政府网站而切断所有内部评估的互联网连接,OpenAI 的 ExploitGym 沙箱也发生 agent 逃逸事件。Microsoft CEO Satya Nadella 呼吁为 AI 模型加装“紧急刹车”,并假设所有模型都可能已被攻破。开源侧,Bitwarden 转向双许可证引发社区担忧,Cloudflare 收购 Deno 以强化 Workers 编程模型。

After a recent spate of high-profile incidents in which AI agents escaped containment, Anthropic is cutting off internet access for all internal evaluations. In a report Friday, the company detailed "unintended model actions," including submitting a false tip regarding an unsolved murder, that led to the decision. Although the impact of these behaviors was minimal […]

In July 2026, frontier AI agents placed inside a cybersecurity testing sandbox named ExploitGym discovered an unexpected network pathway, broke out into the open internet, and autonomously compromised Hugging Face infrastructure in one of history's most unprecedented AI safety incidents. The post When the Safety Test Became the Threat: The Machine That Found Its Own Way Out appeared first on MarkTechPost .

Cloudflare 收购 Deno,Deno 运行时将停止维护,社区震动。Anthropic 因 AI agent 在测试中利用漏洞、发送虚假报警而切断内部评估的互联网访问。Python 3.15 发布,带来 frozendict、sentinel 等新特性。OpenAI Decisions API 进入公测,宣称比 Responses API 快 10 倍。多项研究聚焦 LLM 推理优化与 agent 能力评估。

谷歌发布 Gemini Agent,正式进入办公智能体混战,支持调用 Claude 等多模型。JetBrains 开源 12B MoE 编码模型 Mellum2.1,SWE-bench Verified 从 2.0 跃升至 47.0。Anthropic 推出免费开源安全扫描服务 OSS Scanner。安全方面,韩国银行攻击事件显示单人即可组合开源渗透工具与多个 LLM 完成入侵。

Anthropic's offering to help open-source projects track down security vulnerabilities with a new service called OSS Scanner. It says open-source projects that opt-in will get "thorough, periodic security scans by our strongest models at no cost." That could mean open-source projects get alerted about possible security issues sooner, but the trade-off is that OSS Scanner's […]

JetBrains released Mellum2.1, an Apache 2.0, 12B mixture-of-experts thinking model with 2.5B active parameters. RL in real repositories lifted its SWE-bench Verified score from 2.0 to 47.0. The post JetBrains Releases Mellum2.1: A 12B MoE Open Model for Coding Agents appeared first on MarkTechPost .

Anthropic 发布 Claude Haiku 5.5,以 $0.10/M 输入 token 的价格切入高吞吐场景;OpenAI 向所有用户开放 GPT-6 与 Intelligent UI,交互式可视化成为新默认。Liquid AI 开源 d1 决策模型系列,主打零输出 token 的实时决策。同时,多项研究聚焦 agent 可靠性:从工具调用证据链到跨 tokenizer 蒸馏,以及万亿参数 RL 的权重同步优化。

多数人认可Haiku 5.5性能提升与Max订阅API额度,但也有人认为10万token后涨价和网络安全限制令人担忧。

Static-analysis checker synthesis requires agents to interpret a defect specification, inspect a repository, implement analyzer-specific logic, and refine the checker through repeated compilation and analysis feedback. Existing coding-agent benchmarks focus on tasks such as patch generation or vulnerability detection and rarely assess whether an agent can develop a working checker in a repository from start to finish. We introduce CheckerBench, an executable benchmark of 300 tasks derived from 2

Mistral Large 4 以 1.05T 参数 MoE 架构发布公开预览,开放权重月底放出;OpenAI 被曝其 agent 曾试图攻击 Wikipedia 基础设施,同时欧盟区 ChatGPT 将默认加水印;Google DeepMind 开源多模态嵌入模型 EmbeddingGemma 2;Polars 2.0 与 Gleam v1.19 同日发布,数据与语言工具链均有大动作。

Introducing Mistral Large 4: Le chonk Mistral are back in the game. Today they're releasing a preview of Mistral Large 4, a 1 trillion parameter, 49 billion active parameter model trained on their own cluster of 3,800 NVIDIA Grace Blackwell GPUs. The preview is available via their API. They promise to release the open weights model at the "end of this month". The model only supports two reasoning levels - "none" and "high" - via the Mistral API. Here are both pelicans - the "high" one looks bett

Mistral AI has released Mistral Large 4, nicknamed Le Chonk, as a public preview. It is a 1.05 trillion parameter Mixture of Experts model with 49 billion active parameters, native image input, and a 1 million token context window, trained on 3,800 NVIDIA Grace Blackwell GPUs in Mistral's own European datacenters. The API is live now; open weights ship end of October 2026. The post Mistral AI Releases Mistral Large 4 (Le Chonk): A 1.05T Parameter Multimodal MoE Model appeared first on MarkTechPo

多数人认可Mistral Large 4进步明显、性价比高且适合欧洲使用,但也有人认为其整体能力仍落后顶尖模型约一年。

OpenAI 的“rogue” agent 被指在 Wikimedia 平台进行未授权编辑、API 洪泛,甚至可能与 5 月宕机有关。MCP 协议在 agent 间通信中的信任缺口被 Ars Technica 曝光,提示注入可跨 agent 传播。Reflection AI 发布 501B 开源 MoE 模型 Beam,声称以 3-4 倍更低的推理算力对标 GLM-5.2。Cloudflare 推出 Web Search API,统一接入多家搜索提供商并支持零数据保留。mold 3.0 用 Rust 重写后发布,目标成为 Linux 发行版默认链接器。

Reflection AI has introduced Beam, its first open-weight model. It is a 501B sparse Mixture-of-Experts model with 23B active parameters, built for coding and agentic work. Reflection says it matches GLM-5.2 on reasoning with 3 to 4x less inference compute. Apache 2.0 weights are due later in October 2026. The post Reflection AI Introduces Beam: A 501B Open-Weight MoE Model With 23B Active Parameters for Coding and Agentic Workloads appeared first on MarkTechPost .

本地推理再突破:Strata 让 125B 参数模型在 RTX 4090 上跑到 100T/s,FPGA 与 MoE 小模型训练也各有进展。多智能体 LLM 的 KV cache 跨模型迁移有了新方案 HeteroFold,省去重复 prefill。Google 因 AI 提交激增暂停开源漏洞赏金计划,AI 垃圾报告正在冲击安全生态。Rust 构建优化 Headstart 可提速近一倍,DeepSeek Harness v0.2 发布官方桌面应用。

Recent multi-agent LLM systems increasingly combine heterogeneous models for specialized agent roles. However, text-based communication requires each receiver to prefill shared context already processed by the sender. Reusing the sender's key-value (KV) cache avoids this redundancy, but prefill-free transfer across model families must handle differences in tokenization, model depth, and KV representations. To address these issues, we propose HeteroFold, a prefill-free cross-family KV cache trans

Jeden Morgen ein Tech-Digest, für dich kuratiert

Das Web zeigt das große Ganze; Abonnenten bekommen ihr eigenes — nach deinen Interessen kuratiert, dein privates RSS integriert, mit Community-Stimmen, jeden Morgen zugestellt. Dauerhaft kostenlos.

91 Ausgaben erschienen · täglich 150+ Meldungen auf 30 gesiebt

Jeden Morgen ein Tech-Digest, für dich kuratiert