We look at NVIDIA Personal AI Router (PAIR), an open source virtual inference router that spreads local AI requests across the machines already on a home network. We cover how PAIR…
AI 点评 · 开源路由打通多设备算力,让本地AI请求自动分流,家庭算力池化迎来新玩法。
共 167 条相关资讯 · 来自历史归档
We look at NVIDIA Personal AI Router (PAIR), an open source virtual inference router that spreads local AI requests across the machines already on a home network. We cover how PAIR…
AI 点评 · 开源路由打通多设备算力,让本地AI请求自动分流,家庭算力池化迎来新玩法。
AI 点评 · Transformer 生态剧变,关键人物去向牵动开源推理风向。

Building a Physical AI system takes a continuous pipeline, not a single training job. This post shows how to run that model factory (synthetic data generation, post-training, and c…
It’s officially the Ternus era at Apple. Tim Cook stepped down as CEO this week, handing the company to former hardware chief John Ternus, whose first memo promised a “huge launch…

Nvidia's PAIR (Personal AI Router) automatically spreads local AI requests across all available devices on a home network, cutting wait times for parallel agent tasks. The article…

IT之家 9 月 4 日消息,根据 NVIDIA(英伟达)官网更新,其面向 Windows on Arm PC 的 RTX Spark N1X 超级芯片处理器在笔记本电脑端提供 2 档配置,而 桌面主机上则仅有“满血”高配 。 可以看到,RTX Spark N1X 面向移动端提供了 5120 核 GPU + 18 核 CPU 的低配, GPU 规模较 614…
AI 点评 · 移动端低配曝光,凸显英伟达细分市场策略,Arm PC竞争再添变数。

Frontier intelligence is going local. At IFA 2026, NVIDIA, Microsoft and its partners are teaming up to provide faster inference and new tools that make agents easier to set up and…
Nvidia is announcing its new Personal AI Router (PAIR), a free tool that syncs up your home computers for tackling local AI inference tasks with tools like Ollama and LM Studio. Le…

Nvidia plans to acquire Hugging Face for about $12.9 billion, securing the central platform for open AI models. More than 18 million developers and 200,000 companies use the hub. C…

Nvidia says Hugging Face will stay open even as the chipmaker takes control of a key AI hub.

At IFA 2026, Nvidia and its partners showed off the first RTX Spark-powered laptops and mini PCs, designed to run AI models right on your computer.

The long-rumored deal will give the chip giant access to—and help it promote—a huge repository of open-source AI models and data sets.
Nvidia said Hugging Face hosts over 3 million models and is used by over 18 million developers.
Nvidia has agreed to buy Hugging Face for $12.93 billion, bringing one of the most popular hosting platforms for open-source AI models, datasets, and tools under the ownership of t…
https://blogs.nvidia.com/blog/nvidia-to-acquire-hugging-face...

Anthropic has signed a $35 billion cloud computing deal with Lambda, an Nvidia-backed cloud provider. The article Anthropic ramps up Claude infrastructure with $35 billion Lambda d…
On our new Real World AI stage, we’ll be focusing on the intersection between the digital and physical, and all the ways we’ll continue to see a blending of the two.
AI 点评 · 虚实融合成AI新战场,Nvidia与机器人同台,看点在于前沿技术如何落地现实场景。
NVIDIA has released Switchyard, an Apache-2.0 Rust proxy and library for LLM traffic. It decodes requests into provider-neutral types, routes them with passthrough, random, LLM-cla…
AI 点评 · 跨厂商LLM流量统一路由,开源Rust实现,填补多API兼容空白。
Stable 4-bit floating-point (FP4) pretraining is difficult because the E2M1 payload represents only a narrow range of magnitudes. NVIDIA's Transformer Engine \nv{} recipe addresses this with current-t…
Nvidia is officially launching DLSS 5 this week, following a divisive announcement in March where we likened the AI upscaling tech to a "real-time generative AI filter for video ga…

IT之家 9 月 1 日消息,英伟达与联发科 8 月 31 日宣布加深长期合作,打造涵盖 AI 基础设施、本地 AI 计算和汽车领域的下一代 AI 计算平台。 作为扩展合作的一部分,联发科将采用 NVIDIA NVLink Fusion 平台,为超大规模企业、云服务提供商和前沿模型开发者提供预验证的路径, 开发定制 XPU 并将其引入 NVIDIA NVLi…
AI 点评 · 芯片巨头跨界结盟,AI、PC、汽车三线并进,生态版图再扩张。
Nvidia invests $3.5 billion into Taiwanese chipmaker MediaTek. The deal shows how Nvidia plans to stay essential to AI infrastructure as Big Tech begins to build its own AI chips.
In this tutorial, we build an ensemble weather forecasting workflow with NVIDIA Earth2Studio. We install the required Earth2Studio components while preserving Colab’s existing CUDA…
AI 点评 · 开源工具链让AI气象预测门槛骤降,批量集成预报实战教程可复现,对研究者和工程团队极具参考价值。
The new generation of data center systems is increasing efficiency with smarter traffic control instead of just more processor cycles.
AI 点评 · 芯片巨头转向系统级优化,智能调度成新战场,算力效率比拼升级。
Neocloud Lambda has raised $1B in private debt to buy Nvidia AI chips and lease them to Microsoft. It's the latest in a string of loans, underscoring the high cost of the AI boom.
AI 点评 · AI算力军备竞赛白热化,千亿融资凸显芯片租赁商业模式的高杠杆与高风险。
AI 点评 · 英伟达自证造血能力,AI投资热潮的关键风向标。

Nvidia is nabbing critical infrastructure for open models as interest grows.
AI 点评 · 英伟达揽下开源模型枢纽,AI生态话语权再落一子。
On Nvidia's earnings call Wednesday, CEO Jensen Huang casually announced the company had "achieved AGI," one of the tech industry's ultimate goals some of its biggest players have…

Serving automatic speech recognition (ASR) models at scale is costly when each request uses only a fraction of a GPU. Learn how NVIDIA CUDA Multi-Process Service (MPS) with NVIDIA…

NVIDIA Vice President of Hyperscale and HPC Ian Buck hand-delivers Vera CPU systems across the AI ecosystem as Vera begins shipping at scale.

Z.ai releases GLM-5.3-Flash, an open-source model with 320 billion parameters that lands just three points behind the larger GLM-5.3 on Artificial Analysis's Intelligence Index, at…

Nvidia is buying open-source AI platform Hugging Face for $12.9 billion, about 80 times its $150 million annual revenue. The deal fits Nvidia's push to invest billions in open AI m…
Nvidia has reportedly agreed to buy Hugging Face, the popular open source AI hub, for $12.9 billion in a move that would let Nvidia both protect its chip empire and jump back into…

IT之家 8 月 27 日消息,韩媒 ZDNET Korea 当地时间 26 日援引消息人士报道称,NVIDIA(英伟达)已要求上游 HBM 内存供应商调整 HBM4 内存的供应规划, 提升 8 层堆叠 (8Hi) 产品占比 ,更高堆叠的 12Hi HBM4 比例相应减少。 全球内存市场当前处于供不应求的状态, 降低堆叠高度能以等量的 HBM4 DRAM D…
AI 点评 · 看点在于英伟达调整HBM4堆叠策略,或暗示成本与良率考量,影响内存供应链格局。

Open Source wins!
AI 点评 · 开源社区迎来巨变,英伟达重金收购HuggingFace,AI生态格局或将重塑。
Amazon is adding another 2 million Nvidia GPU chips to its data centers over the next two years. But this extended partnerships stretches beyond buying more chips.
AI 点评 · 云巨头加码英伟达芯片,凸显AI算力军备竞赛白热化,供应链格局或生变。
Nvidia's predicting it will pull in $108 billion in revenue within just a few months. It wouldn't be the first company to rake in over $100 billion in quarterly revenue - Amazon, A…
AI 点评 · AI巨头单季营收破千亿美元,标志算力需求爆发,或将重塑科技股估值体系。

The next wave of AI is placing new demands on infrastructure. As AI agents and trillion-parameter workloads become mainstream, the performance of AI infrastructure depends not only…
AI 点评 · AI算力瓶颈转向内存,NVIDIA自研高带宽存储或重塑硬件格局。
Perplexity releases Portable Computer, packaging local models, harness, sandbox, and connectors into one system running on NVIDIA DGX Spark. The post Perplexity Ships Portable Comp…

OpenAI showed off "Jalapeño," its first in-house inference chip, with benchmarks at the Hot Chips conference. According to SemiAnalysis tests, the chip beats Nvidia's Blackwell and…
AI 点评 · 自研芯片性能反超英伟达,OpenAI硬件话语权大增,AI算力格局或将生变。
https://www.bloomberg.com/news/articles/2026-08-25/openai-cl... , https://archive.ph/yCTrr
AI 点评 · 自研芯片性能对标英伟达,OpenAI硬件自主化提速,或重塑AI算力格局。

Nvidia is moving its Groq 3 LPX inference chip into full production and reports 3,400 tokens per second on Gemma 4 31B, four times faster than Cerebras. But the numbers don't tell…
AI 点评 · 数据亮眼但算法存疑,行业竞争焦点转向推理速度的真实性价比。
Morphological transforms are long-standing tools for shape and mask processing, but the de facto reference implementation in the Python ecosystem, i.e. scipy.ndimage, is CPU-only, single-array, and th…

Nvidia worker indicted after Jensen Huang scolded Supermicro for AI server smuggling.
AI 点评 · 英伟达高管卷入对华AI服务器走私案,凸显巨头涉地缘政治风险,合规警钟再响。
The next era of AI inference won’t be defined by a single breakthrough chip, network or system. It’ll be defined by how every layer of the AI factory works together. That’s why NVI…
AI 点评 · 英伟达量产新一代推理芯片,聚焦AI工厂全栈协同,标志智能体规模化部署进入新阶段。

According to OpenRouter data, agentic AI workloads consume 15x more tokens than a simple chat request. Why? Consider what happens when an AI agent researches a company for an inves…
AI 点评 · 能耗效率提升30倍,直击智能体高消耗痛点,重新定义AI推理性价比标杆。

Nvidia is negotiating an investment in Perplexity at a valuation above $30 billion, more than 50 percent higher than its last funding round, The Information reports. Perplexity's a…
**Agent harnesses** are becoming a key optimization focus, with NVIDIA research showing traditional skill checks poorly predict agent usefulness and proposing a new metric called *…

Nvidia servers with Vera Rubin and Grace Blackwell chips are set to cost about 15 percent more due to an ongoing DRAM shortage from Samsung, SK Hynix, and Micron, Bloomberg reports…
AI 点评 · AI服务器涨价15%,凸显存储芯片瓶颈正成为算力扩张的新关卡。
Nvidia continues to pour money into data center development — just as AI data centers bring lots of money into Nvidia.
AI 点评 · 英伟达重金押注数据中心,AI基建红利循环加速,产业链风向标意义显著。
Nvidia research shows that AI agents can perform well, and not go off the deep end, through fine-tuning, even if the AI model isn't that great at the task.
AI 点评 · 模型平庸也能靠微调出彩,AI应用重心正从算法转向工程控制。

Waymo built its own chip for its robotaxis, cutting its reliance on Nvidia. The article Waymo builds its own chip for its robotaxis, cutting its reliance on Nvidia appeared first o…

Nvidia is paying $6 billion for software that builds AI models from the startup Poolside, and it wants to bring on 109 employees. The article Nvidia is acquiring Poolside's "Model…

Yes, we’re confused too.

China is letting small batches of Nvidia's H200 chips onto the mainland to help domestic AI firms in the race with the US. The article China lets Nvidia's H200 chips trickle onto t…
April - 1805 Napoleon is master of Europe Only the British fleet stands before him Compute is now an asset class I see it is once again time to talk financial innovation. Apollo, B…
NVIDIA has released TensorRT Model Connect (TRTMC) in public preview, an Apache-2.0 project that takes a supported Hugging Face or local checkpoint to end-to-end TensorRT inference…
Take three frontier mixture-of-experts models (Alibaba, OpenAI, NVIDIA; 3.6-4.0B active parameters each) and fine-tune them to reason in a low-resource language. On accuracy benchmarks almost nothing…
NVIDIA teams use ChatGPT Work to reduce manual tasks, connect fast-moving signals, and scale successful workflows globally.

NVIDIA Nemotron 3.5 Lightning, an open model built for high-volume agentic workloads, is now available in Amazon SageMaker JumpStart. This post shows how to deploy the 30B Mixture-…
AI 点评 · 大模型上云提速,30B参数轻量部署,企业智能体落地门槛再降。
Groq raised $350 million at a $3.5 billion valuation as the former AI chipmaker pivots to a neocloud business and expands its Nvidia-powered data center footprint.
AI 点评 · 芯片商转型云服务,估值翻倍印证AI算力需求新风口。
Nvidia's investment in SoftBank's data center developer will guarantee its chips power an OpenAI data center.
AI 点评 · 英伟达押注软银数据中心,锁定OpenAI算力订单,跨界资本绑定凸显AI基建竞争白热化。

Nvidia wants you building your own model, not buying from Anthropic/OpenAI.
AI 点评 · 英伟达推动自建模型,挑战闭源巨头,生态之争再升级。

OpenAI has signed a 20-year lease for an 8-gigawatt data center in Ohio. Nvidia is guaranteeing up to $105 billion for the residual value of the facilities and becomes the exclusiv…
**OpenAI** is advancing its power-and-compute infrastructure with a **4+ GW NVIDIA** capacity commitment and an **8 GW Ohio campus** buildout through **2032**, emphasizing vertical…

Nvidia has cut its guarantee for OpenAI's planned data center in Ohio nearly in half, from $250 billion to just under $120 billion, after investors pushed back on the risk. Meanwhi…
AI 点评 · 英伟达收缩OpenAI押注,恰逢Anthropic数据亮眼,AI投资风向生变。
Tool: CORS Chat I built this today ( with GPT-5.6-Sol xhigh ) to help test Qwen 3.8 27B running in LM Studio on both my M5 MacBook Pro and an NVIDIA DGX Spark. It provides a web UI…

Indonesia is taking charge of its AI future. This week, the Ministry of Communication and Digital Affairs (Komdigi), Indosat Ooredoo Hutchison (Indosat or IOH), NVIDIA and Universi…
AI 点评 · 校企联手建国家级AI中心,东南亚人才竞争升级,中国需警惕生态外溢。
Multi-GPU acceleration for MiniMax H3 video generation on NVIDIA V100 (sm_70). Ulysses sequence parallelism as a drop-in ComfyUI custom node — ~19 min to ~7 min…
Nvidia has a plan to make sure its GPUs won't lose value. It wants to convince a new crop of financiers to keep lending for AI buildouts.

Nvidia is working on Nemotron 4, a new open-weight model designed to rival the world’s best freely available models. The article Nvidia's Nemotron 4 aims for one trillion parameter…
NVIDIA's open 30B MoE targets the agent execution layer, with Switchyard routing each step to the cheapest capable model. The post NVIDIA AI Releases Nemotron 3.5 Lightning: A 30B…

We announced partnerships with Apollo, BlackRock, Blackstone, Brookfield, Goldman Sachs and KKR to establish independent financing platforms designed to mobilize over $500 billion…
LTX-2.5 brings frontier video generation to local NVIDIA hardware: 6.8-second clips, native multishot, day-one ComfyUI, open weights. The post The Video Production Stack Now Fits o…
AI 点评 · 本地开源视频生成首次媲美前沿水平,多镜头与ComfyUI支持将大幅降低创作门槛。

Nvidia's Nemotron 3.5 Lightning is an open-weights model with just 3.6 billion active parameters that matches OpenAI's gpt-oss-120b on the Intelligence Index despite being four tim…

The open source ecosystem is making it easier for AI enthusiasts and developers to build, customize and run increasingly capable agents locally. Throughout August, NVIDIA is celebr…

As AI shifts from chatbots to autonomous agents, open models are serving market demands for full control over where AI runs and how it’s deployed and evolves. Today, NVIDIA is expa…

Nvidia is teaming up with Apollo, BlackRock, Blackstone, Brookfield, Goldman Sachs, and KKR to mobilize over $500 billion for AI infrastructure. To win over investors, the chipmake…
AI 点评 · 开源权重加全栈部署控制,让多语言语音代理延迟优化门槛大降。
Run full MiniMax-H3 FL2VA on one RTX A6000: 1344x768 video + stereo audio, 6.16x Turbo, and formal Sol-Attn N=10.
NVIDIA releases NemotronLabs VoiceChat 11B, an open full-duplex speech-to-speech model with 448 ms latency and live tool calling. The post NVIDIA Releases NemotronLabs VoiceChat 11…

The AI industry's hunger for power keeps growing. Nvidia is investing up to $3 billion in Lancium, a power infrastructure developer that already has four gigawatts under contract i…
AI 点评 · 算力军备竞赛延伸至电力赛道,巨头抢滩能源基建成新焦点。
In this tutorial, we build an advanced multimodal retrieval-augmented generation pipeline with NVIDIA NeMo Retriever. We begin by configuring a Python 3.12 environment, installing…
NVIDIA Labs has open-sourced NOOA (NVIDIA Object-Oriented Agents), a model-agnostic Python framework for building AI agents. Agent development today is split across prompt template…

Anthropic and OpenAI are racing to scale up while reducing dependence on Nvidia.
AI 点评 · AI巨头自研芯片,摆脱英伟达依赖,算力军备竞赛升级。

In July, NVIDIA joined more than 200 companies and organizations in signing “Open Weights and American AI Leadership,” an open letter arguing that AI leadership will be measured no…
Modern Greek is absent from NVIDIA's Nemotron retrieval models and from major multilingual retrieval benchmarks, despite being important for retrieval-augmented generation (RAG) in legal, energy, fina…

SpaceX plans to more than 5x its compute capacity by the end of 2027, betting exclusively on Nvidia's Vera Rubin platform. The expansion could require well over a million new GPUs.…
NVIDIA released Alpamayo 2 Super, a 34B vision-language-action model for autonomous driving, under OpenMDW-1.1 — a permissive license covering fine-tuning, derivatives and commerci…
CUDA又危了?
Modern Greek is absent from NVIDIA's Nemotron retrieval models and from major multilingual retrieval benchmarks, despite being important for retrieval-augmented generation (RAG) in legal, energy, fina…
The week-old Open Secure AI Alliance, spearheaded by Nvidia and grown to over 120 companies, already has proposals out for defending against AI agents.
AI 点评 · 英伟达主导百家企业联盟,一周即出AI安全方案,动作之快凸显行业对智能体威胁的紧迫共识。

NVIDIA is participating in the U.S. National Science Foundation’s (NSF) State and Regional Artificial Intelligence Infrastructure Hubs program, an effort launching today to expand…
**Alibaba** launched **Qwen3.8-Max**, enhancing multimodal capabilities and agent ecosystem integration. **NVIDIA** introduced **Alpamayo 2 Super** for autonomous vehicle reasoning…
NVIDIA Sol-Attn for ComfyUI / Triton kernel on SM89 - SM121, with zero-copy MiniMax H3 nodes: memory-efficient attention, scheduled tau with graph preview, and…
AI 点评 · 从芯片到电网全链路降本,直击AI算力成本痛点,行业风向标。

This week on Uncanny Valley, we discuss the open- vs. closed-source debate in AI, key players in White House AI policy, and how to stop your chatbot logs from showing up in search-…
AI 点评 · 英伟达开源联盟缺失OpenAI和Anthropic,揭示AI巨头间开源与闭源路线的深层博弈。

IT之家 7 月 30 日消息,英伟达昨日(7 月 29 日)发布公告,宣布太平洋时间 8 月 26 日星期三下午 2 点(IT之家注:北京时间 8 月 27 日早上 5 点)召开财报电话会议, 讨论 2027 财年第二财季(截至 2026 年 7 月 26 日)的财务业绩。 在直播方面,本次财报电话会议将通过 investor.nvidia.com 网站开…
AI 点评 · AI风向标,业绩关乎全球科技股走势,市场紧盯其增长能否持续。
Figure 1: CUDA-to-MLX optimization translation map. CUDA optimization knowledge can be translated into architecture-native MLX strategies rather than copied instruction-for-instruc…
As a discerning AI investor who values style and substance, Sarah Guo knows this season’s standout accessory isn’t the latest designer purse — but what’s inside it. In a recent vid…

Taiwan's prosecutors have detained an Nvidia employee in connection with the alleged illegal export of Super Micro AI servers to China, according to Bloomberg and Reuters. The arti…

Nvidia is pouring what it calls a "substantial" sum into Safe Superintelligence (SSI), the AI lab run by Ilya Sutskever, OpenAI's former chief scientist. The article Nvidia invests…
AI 点评 · 英伟达统一三大技术栈,加速工业级AI仿真与自主系统落地。
In this tutorial, we deploy the 1-bit Bonsai-27B language model using the PrismML fork of llama.cpp, which provides the specialized CUDA kernels required to decode the model’s Q1_0…

IT之家 7 月 28 日消息,联想 (Lenovo) 来酷 (Lecoo) 斗战者 (BELLATOR) 今日正式公布了战 7000P 锐龙版游戏笔记本电脑。 这一型号搭载 AMD 锐龙 9 8940HX 处理器和 NVIDIA GeForce RTX 5060 笔记本电脑 GPU,配备 16GB DDR5-4800 SO-DIMM 内存和 512GB P…
AI 点评 · 锐龙9配RTX 5060组合首次亮相,性能与性价比或成最大看点。

一句话,一个周末,Claude 直接把自家最前沿的模型跑在了一台全新的 AMD MI355X 机架上! 人类工程师全程没动手改过一行代码 。 谁能想到,英伟达花 20 年堆起的 CUDA 护城河,就这样被跨了过去。 一个周末,Claude 把 AMD 新 GPU 调通了 故事是这样的。 前段时间,AMD 给 Anthropic 送来一台搭载 MI355X 的…
After two years in stealth, Safe Superintelligence has announced a long-term partnership with Nvidia as it prepares to scale to its next phase.
Nvidia on Monday said it is joining forces with Microsoft, SpaceX, IBM, and other tech companies to build and share open-source AI security tools. The new Open Secure AI Alliance s…

IT之家 7 月 27 日消息,英伟达今日在官方博客发文,随着工程团队致力于开发越来越复杂的 CPU、GPU 和人工智能系统,现代芯片设计的复杂性也在不断增加。 为应对这一挑战,英伟达正与楷登电子和新思科技合作, 优化 NVIDIA Vera CPU 的关键电子设计自动化(EDA)应用 。 英伟达现正部署 Vera 应用于其下一代 CPU 和 GPU 开发的…

The complexity of modern chip design continues to grow as engineering teams work to develop increasingly sophisticated CPUs, GPUs and AI systems. To help meet that challenge, NVIDI…
AI 点评 · 用自家芯片加速设计下一代芯片,体现软硬件协同创新,AI时代效率提升关键。
Since Volta introduced Independent Thread Scheduling (ITS), NVIDIA GPUs have been widely assumed to handle warp divergence in a fixed manner. We test this assumption across Ampere, Hopper, and datacen…
英伟达宣布,SK集团与英伟达在AI工厂及下一代内存领域扩大战略合作。SK集团与英伟达今日宣布计划建立一项超过5000亿美元的综合合作伙伴关系,以建立满足全球计算需求激增的人工智能基础设施。双方签署了意向书,正式化协议,涵盖从人工智能工厂建设到人工智能内存供应。SK电信将建设2吉瓦的NVIDIA Vera Rubin DSX AI工厂,以满足全球计算需求。英伟…

Microsoft, along with Meta, Nvidia, and more than 20 other companies, is pushing for open-weight AI models in an open letter. The strategic logic is simple: the more models running…
AI companies, including Nvidia and Mistral, urge policymakers to avoid broad restrictions on open-weight AI models as Washington debates responses to Chinese AI and alleged model d…
Letter: https://images.nvidia.com/pdf/Open-Weights-and-American-AI-L... [pdf] https://x.com/JensenHuang/status/2080643682408321103 , https://xcancel.com/JensenHuang/status/20806436…

At this week’s AI Summit in San Francisco, South Korean President Jae Myung Lee and some of the country’s top business leaders and researchers are meeting with NVIDIA and ecosystem…
AI 点评 · 韩国总统亲自带队与英伟达合作,预示国家AI战略加速落地。
AMD is challenging its chipmaker rival with a new rack-scale system that will start shipping to customers later this year.
AI 点评 · AMD推出机架级AI系统,挑战英伟达霸主地位,产品年内交付值得关注。
If there's a place in the universe without GPUs, Nvidia is sending them there.
AI 点评 · 英伟达将GPU送上月球,展现其芯片在极端环境下的应用潜力。

NVIDIA founder and CEO Jensen Huang today visited the Naval Postgraduate School in Monterey, California, to commission an NVIDIA DGX GB300 system — bringing one of the world’s most…
AI 点评 · 军事与AI深度融合新标杆,顶级算力部署国防教育,预示AI军事应用加速落地。

Before a healthcare robot can be useful in the real world, it has to learn how the physical world pushes back. Anatomy varies. Instruments bend, press, slip and interact with tissu…
AI 点评 · 开源首个GPU加速医疗物理仿真,降低机器人精准医疗开发门槛。

IT之家 7 月 22 日消息,英伟达(NVIDIA)今天(7 月 22 日)发布博文,报道称纬创资通(Wistron)在美国得克萨斯州沃斯堡开设其首家美国制造工厂, 将生产英伟达 GB300、Vera Rubin 等 AI 超级芯片,总投资达 7 亿美元。 IT之家援引博文介绍,该工厂占地面积约为 32.4 万平方英尺(IT之家注:约 30101 平方米)…
AI 点评 · 台系代工厂首度切入美国本土生产英伟达最强AI芯片,供应链本土化加速。

The AI era runs on AI infrastructure. Many of these advanced systems are built and tested in Texas. Wistron opened its first U.S. manufacturing facility today in Fort Worth — a 324…
AI 点评 · 纬创在美建厂专产英伟达AI系统,标志AI硬件供应链加速向本土化转移。
Traditional agent development is split across prompt templates, tool schemas, callback code, and workflow graphs. We present NVIDIA Object-Oriented Agents (NOOA), a model-agnostic Python framework for…
AI 点评 · 打破传统AI代理开发碎片化,统一框架降低门槛,加速多模型应用落地。

NVIDIA Vera Rubin is here, and it’s going gigascale. Vera Rubin NVL72 production is ramping up with racks running at partners CoreWeave, Google Cloud, Microsoft Azure, Oracle Cloud…
AI 点评 · Vera Rubin NVL72量产,性能功耗比突破,为合作伙伴降低推理成本。

AI has entered the gigascale era. The world’s most advanced AI factories are bringing together hundreds of thousands of GPUs and CPUs to train frontier models, power agentic AI and…
AI 点评 · 为天文观测打造的超大规模AI工厂,标志AI基础设施迈入超大规模计算新时代。
.jpg)
Nvidia’s Vera Rubin platform combines CPUs and GPUs into a single system, reflecting the company’s growing ambition to power every layer of AI infrastructure.
AI 点评 · 英伟达从GPU霸主向数据中心全栈芯片布局,或重塑AI基础设施竞争格局。

NVIDIA and its partners are investing in American manufacturing, supply chains, energy grids and skilled workforces so the U.S. can produce the infrastructure needed for better hea…
AI 点评 · 英伟达联合伙伴布局美国本土制造链,强化AI基础设施自给能力。

In this post, we show how Amazon Quick can serve as the business-user front door for specialized agent workflows. We use the NVIDIA NeMo Agent Toolkit to build a supply-chain risk…
AI 点评 · 企业级AI智能体开发门槛降低,Quick与NVIDIA工具链结合实现供应链风险管理,展示行业落地新范
From open models to real-time simulation, AI and graphics breakthroughs are transforming media, content creation and robotics.

Erin Davis calls it the “SuperDuperPOD.” That’s two things in one name: pharmaceutical giant Bristol Myers Squibb (BMS) already runs one of the largest AI clusters in life sciences…
AI 点评 · NVIDIA与Hugging Face联手,实现视频图像模型大规模微调,降低AI应用门槛。

Lowest cost per token from extreme codesign maximizes intelligence per dollar for post-training in the agentic era.
AI 点评 · 极致软硬协同设计降低每token成本,让后训练阶段的智能性价比达到新高度,是智能体AI落地的关键指标
AI 点评 · 英伟达新嵌入模型登顶基准,加速智能体检索技术落地,AI搜索能力再升级。
Run NVIDIA's tri-modal Nemotron Omni (text+vision+audio) entirely on Apple Silicon — pure MLX runtime, parity-verified against NVIDIA's reference

IT之家 7 月 16 日消息,丰田与英伟达于当地时间 14 日声明称,双方将扩大合作,以汽车为起点,把物理 AI 应用拓展至机器人和智能设施等领域。 目前,丰田正在采用 NVIDIA DRIVE AGX 平台开发下一代汽车,并搭载通过安全认证的 NVIDIA DriveOS 操作系统,以实现先进的 L2++ 驾驶辅助功能。丰田与英伟达希望 打通汽车、基础设…
AI 点评 · 英伟达加持丰田L2++,汽车、机器人、设施全场景AI化,巨头跨界布局加速落地。

General-purpose robots and autonomous machines are moving from research labs to real-world mass-market deployment, creating demand for compact, power-efficient AI supercomputers ca…
AI 点评 · 英伟达推Jetson Thor,瞄准机器人量产与边缘AI,算力能效比成关键看点。
Home to leading manufacturers, robotics pioneers, infrastructure builders and iconic gaming companies, of course, Japan is one of the world’s centers of AI — building across the fu…

Enterprises have plenty of powerful models to choose from. The real test is whether the AI an enterprise builds uniquely addresses the needs of the business: improving workflows, t…
AI 点评 · 开放模型让企业和国家掌控AI定制权,破解信任与安全难题,这才是行业关键突破。
AI 点评 · 揭秘GPU热潮背后的资金循环游戏,看巨头如何用资本撬动算力产业链。

In this post, we explore what makes the Nemotron 3 architecture unique, walk through the fine-tuning techniques available, and show you step-by-step how to get started with serverl…
AI 点评 · 降低大模型微调门槛,NVIDIA与AWS合作让企业无需管理服务器即可定制AI模型。
The company is using the cash to open an office in the Bay Area and compete for talent there, "strengthening its position at the heart of the world's leading AI ecosystem."
AI 点评 · Nvidia领投,欧洲AI语音新星获巨额种子轮融资,正面挑战硅谷人才争夺战。
Having proven how valuable compute can be, the company finds itself at the center of a market everyone wants to be in — while simpler technologies and less interesting companies ge…

NVIDIA Nemotron 3 Ultra is offering leading performance at lower cost than top closed models with the largest and most widely adopted AI agent orchestration platform. LangChain tun…

It's early, but the plan is to reduce dependency on Nvidia and Huawei.

Max single-threaded CPUs at scale are a new category of CPUs built for the agentic AI era. Across the creation and deployment of an agentic system, the CPU is on the critical path…

Open source AI has shown how quickly developers can innovate when models, data and tools are shared. Robotics has the same opportunity, but advancements in physical AI development…
We introduce Nemotron-Labs-Diffusion, a tri-mode language model (LM) that unifies AR, diffusion, and self-speculation decoding within a single architecture. Trained with a joint AR-diffusion objective…
Audio intelligence involves understanding, reasoning about, and generating both audio and speech. In this work, we introduce Nemotron-Labs-Audex-30B-A3B (Audex), a unified audio-text LLM built on Nemo…

As AI moves from model development to production inference, compute demand is accelerating and shifting toward continuously operating AI factories that generate tokens at scale. Th…

We're excited to introduce US-based frontier open-weight models in AWS GovCloud (US). With this release, Amazon Bedrock now supports OpenAI’s open-weight GPT OSS models (120B and 2…
AI 点评 · 政府云首次集成前沿开源大模型,提升敏感数据场景的AI安全与合规能力。
Generate textured, segmented and rigged GLB assets entirely on your own GPU.
Why is LLM inference slow — and how do you make it fast? A hands-on, first-principles course: roofline → KV cache → quantization → parallelism → vLLM/SGLang, wi…
Hi everyone, I started working on nanoeuler after the ban of anthropic's fable because my ambition and dream is to work in the AI field in anthropic. The two interesting reasons th…
8-bit Adafactor Optimizer with Fused CUDA Kernels

NVIDIA and Microsoft birthed a new computer

Nvidia announced the Rubin CPX, a solution that is specifically designed to be optimized for the prefill phase, with the single-die Rubin CPX heavily emphasizing compute FLOPS over…