2026-08-28 · AI 资讯日报
今日新收录 174 条公开资讯,按模型 / 产品 / 行业 / 论文 / 观点 自动归类汇编(非 AI 生成,点击可溯源原文)。
今日精选6 条
- 1.
Recent advances in inference-time scaling have significantly improved the reasoning performance of large language models (LLMs). However, these methods typically rely on repeated generation or externa…
- 2.
To improve large language models' ability to resolve real-world software issues, prior work has focused on constructing large-scale agent trajectory datasets and performing supervised fine-tuning (SFT…
- 3.
In real-world software development, code review typically involves iterative interactions between developers and reviewers to improve software quality, making the process costly and time-consuming. Al…
- 4.
LLM-based agents are increasingly deployed in product-level execution harnesses, where jailbreaks can trigger harmful tool use and persistent state changes, creating greater risks than unsafe text gen…
- 5.
Chemical reactions are fundamentally transformations in electron space, yet most machine learning approaches model them either through \textit{de novo} generation of product molecules or through heuri…
- 6.
Transduced language models (TLMs) compose a pretrained \emph{source} language model with a functional finite-state transducer to induce a language model over \emph{target} strings. Computing the proba…
模型发布/更新6 条
- 1.Beneficial Deployments— Anthropic
AI 导读 · Anthropic新部署方案,AI安全落地的关键一步。
- 2.Expanding our support for scientists— Anthropic
AI 导读 · AI实验室加大科研扶持,或催生更多前沿突破。
- 3.Gemini Omni 1.1 Flash lets you build with more control— Google DeepMind
- 4.3 new ways to plan and book travel in Search— Google AI
Book hotels and track airfares, plus view miles and rewards with AI Mode in Google Search.
- 5.
NVIDIA Vice President of Hyperscale and HPC Ian Buck hand-delivers Vera CPU systems across the AI ecosystem as Vera begins shipping at scale.
- 6.Piloting the world's first double-blind AI evaluations— Google DeepMind
Piloting the world's first double-blind AI evaluations
产品发布/更新6 条
- 7.deeplethe/utopia— GitHub
World's first open-source enterprise world model.
- 8.kulkarnirohit123/cra-agent— GitHub
Autonomous agentic AI for CRA (Cyber Resilience Act) compliance: scans repos, triages findings, opens Jira tickets, and auto-fixes vulnerabilities via PR.
- 9.calmrocks/ai-engineer-notebooks— GitHub
Hands-on, framework-free Colab notebooks for the AI Engineer / Forward Deployed Engineer (FDE) skill set — model APIs, structured output, tool calling, RAG, eva…
- 10.
AI 导读 · 突破实时重建瓶颈,开启万帧级3D场景流式处理新纪元。
- 11.
AI 导读 · 轻量低价直击痛点,智能眼镜或迎来真正大众化拐点。
- 12.
AI 导读 · 把AI助手塞进耳朵,重新定义耳机价值,值得关注。
行业动态6 条
- 13.Looking beyond natural sequences— MIT News
A new machine-learning framework aims to improve the success rate of computational protein design while moving away from results that reproduce sequences found in nature.
- 14.
AI 导读 · 警惕AI沦为简历注水工具,损害真实项目价值。
- 15.
Anthropic has struck a deal worth around 45 billion dollars with British cloud startup Nscale, Bloomberg reports. The article Anthropic locks in 45-billion-dollar compute deal with…
- 16.
Google's new Gemini 3.5 Transcribe recognizes over 85 languages, strips filler words, and corrects slips of the tongue in real time. It hits a 4.0 percent word error rate in stream…
- 17.
Anthropic is building its own browser into Claude Cowork, running inside the desktop app. The article Claude Cowork now runs its own browser inside the desktop app appeared first o…
- 18.
Letting an AI do your shopping might not get you the best deal. Researchers at the Wharton School show how erratic AI shopping agents really are: a single external source like Wire…
论文研究6 条
- 19.
Recent advances in inference-time scaling have significantly improved the reasoning performance of large language models (LLMs). However, these methods typically rely on repeated generation or externa…
- 20.
To improve large language models' ability to resolve real-world software issues, prior work has focused on constructing large-scale agent trajectory datasets and performing supervised fine-tuning (SFT…
- 21.
In real-world software development, code review typically involves iterative interactions between developers and reviewers to improve software quality, making the process costly and time-consuming. Al…
- 22.
LLM-based agents are increasingly deployed in product-level execution harnesses, where jailbreaks can trigger harmful tool use and persistent state changes, creating greater risks than unsafe text gen…
- 23.
Chemical reactions are fundamentally transformations in electron space, yet most machine learning approaches model them either through \textit{de novo} generation of product molecules or through heuri…
- 24.
Transduced language models (TLMs) compose a pretrained \emph{source} language model with a functional finite-state transducer to induce a language model over \emph{target} strings. Computing the proba…
技巧与观点6 条
- 25.
Claude Science AMA: How to accelerate scientific discovery Anthropic
- 26.
AI 导读 · 用Agent串联生成式AI工具,打通创作流程,效率提升看得见。
- 27.
Self-hosted speech AI carries an observability trade-off: the numbers that drive capacity planning and cost management stay locked inside the vendor container. Deepgram closes that…
- 28.
Serving automatic speech recognition (ASR) models at scale is costly when each request uses only a fraction of a GPU. Learn how NVIDIA CUDA Multi-Process Service (MPS) with NVIDIA…
- 29.Breaking Claude Code Opus 5 Auto Mode— Simon Willison
AI 导读 · 自动模式成防注入攻击新防线,Anthropic押注AI编码安全,行业风向标值得紧盯。
- 30.Who let the agents in— Ben's Bites
It’s you, and it’s getting easier
快讯
- ·Cloudflare 发布 Kitesurf,一款面向 AI 智能智能体的浏览器引擎— InfoQ · 8/28 22:00
- ·智谱认领“牛来”模型,实测:“牛马”友好— InfoQ · 8/28 19:32
- ·Cloudflare 将工程规范改造为 AI 强制执行的管控系统— InfoQ · 8/28 18:53
- ·曝英伟达129 亿美元收购Hugging Face!黄仁勋:我只后悔没有更早、更多投资OpenAI、Anthropic— InfoQ · 8/28 17:29
- ·M3 系列 Mac 原生 Linux 支持最后冲刺,Asahi 攻克 USB 3.0 高速传输等底层壁垒— IT之家 · 8/28 13:45
- ·Anthropic 推出 MHS 标准,首度重拳杀入物理 AI 领域— IT之家 · 8/28 12:46