2026-09-03 · AI 资讯日报
今日新收录 158 条公开资讯,按模型 / 产品 / 行业 / 论文 / 观点 自动归类汇编(非 AI 生成,点击可溯源原文)。
今日精选6 条
- 1.
LLM-based evaluators of natural language generation (NLG) quality are widely deployed as scoring tools and as automated training signals, yet the internal procedure by which they assign a rating remai…
- 2.
Evaluating software engineering agents on realistic benchmarks is costly, since each task may require multi-step code exploration, modification, and test execution. Existing efficient evaluation metho…
- 3.
The repository-level code generation task requires synthesizing code that satisfies task requirements while remaining consistent with the target repository context. Since real-world repositories often…
- 4.
Dynamic agent harnesses let language models change the software that shapes their own execution. This flexibility brings a new reasoning burden: a local plugin change can propagate through dependencie…
- 5.
Natural language is emerging as a primary feedback channel for improving language agents, capable of conveying intent, preferences, and causal structure in forms interpretable by both humans and moder…
- 6.
Real-world robotic assembly at sub-millimeter tolerances demands spatial precision, compliant interaction, and robustness to contact failures. We present Facet-0, a robotic foundation model that predi…
模型发布/更新6 条
- 1.
The Fairwind Program is a limited access program for governments and trusted partners to use our cyber defense tools.
- 2.
ATV Big Air Tour uses ChatGPT Work to speed up marketing, merchandising, and more. It even turned merchandise photos into an inventory website in 15 minutes.
- 3.
Anthropic has released Claude Fable 5.1 and Claude Mythos 5.1, the same model behind two different safeguard layers. Fable 5.1 is generally available on the Claude API, AWS, Google…
- 4.
Google released Gemini 3.8 Flash and Gemini 3.8 Flash Cyber on September 2, 2026. Both variants run on the same foundational intelligence, split by safety mitigations rather than m…
- 5.
Download Meta AI for Mac: Free Desktop App | AI at Meta AI at Meta
- 6.
AI 导读 · 谷歌密集发布Flash模型,暗示其轻量级AI战略提速,市场关注度陡增。
产品发布/更新6 条
- 7.
Agentic assistants have a structural problem: the context that makes them useful — deal documents, privileged files, client records — is exactly the context users cannot send to a…
- 8.
Speed and accuracy usually pull against each other in text-to-speech. Gradium AI's new default model reports both: an 81.0% human-rated pass rate on 500 hard sentences across five…
- 9.
面向抖音直播电商的 Windows 本地 AI Agent Studio,贯通主播发现、直播洞察、直播复盘与短视频内容编导的统一智能工作流。
- 10.
GLM-5.3-Flash × J-Space capability realization — benchmark presentation of the J-Space Cognition Suite
- 11.ZekunCheng/novoweave— GitHub
NovoWeave — a conceptual generative protein-design framework (non-functional pseudocode).
- 12.
SkyProduction(天工工作台)全新版本于8月31日正式上线
行业动态6 条
- 13.
AI 导读 · 产学研融合加速AI与量子计算落地,MIT与IBM合作模式成行业标杆。
- 14.Muse Spark 1.3— Hacker News
AI 导读 · 多模态推理新突破,Meta开源模型或改写行业格局。
- 15.
A new method, called CW-Net, translates the reasoning process of an autonomous vehicle’s AI system into understandable concepts that explain its behavior.
- 16.Gemini 3.8 Flash and 3.8 Flash Cyber— Hacker News
https://deepmind.google/models/model-cards/gemini-3-8-flash/
- 17.Fable 5.1 World Modeling— Hacker News
AI 导读 · 多模态世界模型新突破,预示AI从感知迈向认知推理的关键一步。
论文研究6 条
- 19.
LLM-based evaluators of natural language generation (NLG) quality are widely deployed as scoring tools and as automated training signals, yet the internal procedure by which they assign a rating remai…
- 20.
Evaluating software engineering agents on realistic benchmarks is costly, since each task may require multi-step code exploration, modification, and test execution. Existing efficient evaluation metho…
- 21.
The repository-level code generation task requires synthesizing code that satisfies task requirements while remaining consistent with the target repository context. Since real-world repositories often…
- 22.
Dynamic agent harnesses let language models change the software that shapes their own execution. This flexibility brings a new reasoning burden: a local plugin change can propagate through dependencie…
- 23.
Natural language is emerging as a primary feedback channel for improving language agents, capable of conveying intent, preferences, and causal structure in forms interpretable by both humans and moder…
- 24.
Real-world robotic assembly at sub-millimeter tolerances demands spatial precision, compliant interaction, and robustness to contact failures. We present Facet-0, a robotic foundation model that predi…
技巧与观点6 条
- 25.Real-Time Intelligence with IBM Time Series Models on Confluent— HuggingFace Blog
- 26.
AI 导读 · 让不同AI模型直接互通,绕开语言瓶颈,或开启多模型协作新范式。
- 27.
AI 导读 · 云厂商联手AI巨头,跨区域推理落地澳洲,标志多模型生态竞争加剧。
- 28.
AI 导读 · 用生成式AI将培训视频转为标准作业流程,并借RAG优化工单处理,展现云上客服自动化的落地范式。
- 29.
AI 导读 · 用AI检测仪表盘静默故障,填补运维盲区,是数据可靠性领域的新解法。
- 30.
AI 导读 · 自动生成架构图打通代码与文档,企业级效率革新值得关注。
快讯
- ·探索太阳风暴早期预警:AI 提前 9.24 小时捕捉太阳活动区信号— IT之家 · 9/3 08:08
- ·
- ·纽约时报诉 OpenAI 案升级:美国司法部首次就 AI 版权问题公开表态,主张 AI 训练属“合理使用”— IT之家 · 9/3 07:50
- ·
- ·Meta 发布最强 AI 模型 Muse Spark 1.3,编码能力超 GPT-5.6 Sol— IT之家 · 9/3 07:29
- ·欧洲最强 AI:西班牙公司推出 Quasar 438B 模型,1M 词元上下文— IT之家 · 9/3 07:16