

Celebrating 25 years of visual search innovation
(翻译)庆祝视觉搜索创新25周年
Google Images is turning 25. Here’s a look back at some major milestones — and new ways to explore and create visual content.

NVIDIA Cosmos-H-Dreams: Bringing Real-Time Generative Simulation to Surgical Robotics
(翻译)英伟达 Cosmos-H-Dreams:为手术机器人带来实时生成式模拟
A Blog post by NVIDIA on Hugging Face

对话纬钛机器人CEO李瑞:具身智能仅靠视觉容易到天花板,触觉是未来必选项
作者丨 向 欣 编辑丨 高景辉 &nbs

Scale AV Perception Across Vehicle Platforms with NVIDIA Omniverse NuRec
(翻译)使用 NVIDIA Omniverse NuRec 扩展跨车辆平台的自动驾驶感知
A perception stack is shaped by the vehicle that carries it. Move the same software to a new carline—for example, from an SUV to a sedan or another vehicle…

李飞飞、吴佳俊再联手,打破世界模型和机器人动作的「巴别塔」
作者丨 幸丽娟 编辑丨岑 峰  

SCAFFOLD: A Large-Scale Structured Dataset of Computer Science Research Figures with Diagram QA and Chain-of-Thought Reasoning Traces
(翻译)SCAFFOLD:面向计算机科学研究图表的大规模结构化数据集,包含图表问答与思维链推理轨迹
Abstract page for arXiv paper 2609.00018: SCAFFOLD: A Large-Scale Structured Dataset of Computer Science Research Figures with Diagram QA and Chain-of-Thought Reasoning Traces
刚刚,李飞飞掀桌!全球首个多模态世界模型发布,几张照片省掉几百个机位
#欢迎关注爱范儿官方微信公众号:爱范儿(微信号:ifanr),更多精彩内容第一时间为您奉上。

DeepSeek V4 多模态开源,我们把它的视觉链路拆了一遍
作者丨 郑 佳 美 编辑丨 岑 峰 &n

打破黑盒猜想:大模型通往真正「空间智能」的破局之路
作者丨 张 璐 编辑丨 幸丽娟 &nbs

群核科技联手英伟达、英特尔、浙大,三篇 ECCV 论文给物理 AI 造基础设施
作者丨 张 璐 编辑丨 幸丽娟 &nbs

腾讯混元、清华、南洋理工联手,「以小博大」破解空间智能算力与记忆断裂难题 | ECCV 2026
作者丨 张 璐 编辑丨 幸丽娟 &nbs

GPT Image 2.5 对比实测:变化与局限
GPT Image 2.5 在复杂中文图文任务中表现如何?本文通过十一类任务、72张输出,对比两代模型在海报、安装说明、经营简报等场景下的能力。结果显示,2.5 在商品信息对应和数字复现上有所提升,但在工具接触点、零件数量和图表比例上仍会出现错误,书法风格差异明显。

