

Celebrating 25 years of visual search innovation
(翻译)庆祝视觉搜索创新25周年
Google Images is turning 25. Here’s a look back at some major milestones — and new ways to explore and create visual content.

NVIDIA Cosmos-H-Dreams: Bringing Real-Time Generative Simulation to Surgical Robotics
(翻译)英伟达 Cosmos-H-Dreams:为手术机器人带来实时生成式模拟
A Blog post by NVIDIA on Hugging Face

A survey detection channel overrides the pixels in an astronomical foundation model, and biases tomographic mean redshifts
(翻译)巡天检测通道覆盖天文学基础模型中的像素,并使层析平均红移产生偏置
arXiv:2608.23626v1 Announce Type: new Abstract: Foundation models for astronomy are trained on survey pixels together with the catalogue products derived from those pixels. Those catalogues are incomplete at a measurable rate, and a model trained on both inherits that incompleteness as a systematic.
高德发布首个无长程依赖的万帧级流式3D重建模型ABot-Recon,以12帧重建万帧3D场景
8月28日,阿里巴巴集团旗下高德正式发布首个无长程依赖的万帧级流式3D重建模型ABot-Recon。

对话纬钛机器人CEO李瑞:具身智能仅靠视觉容易到天花板,触觉是未来必选项
作者丨 向 欣 编辑丨 高景辉 &nbs

Scale AV Perception Across Vehicle Platforms with NVIDIA Omniverse NuRec
(翻译)使用 NVIDIA Omniverse NuRec 扩展跨车辆平台的自动驾驶感知
A perception stack is shaped by the vehicle that carries it. Move the same software to a new carline—for example, from an SUV to a sedan or another vehicle…

李飞飞、吴佳俊再联手,打破世界模型和机器人动作的「巴别塔」
作者丨 幸丽娟 编辑丨岑 峰  

SCAFFOLD: A Large-Scale Structured Dataset of Computer Science Research Figures with Diagram QA and Chain-of-Thought Reasoning Traces
(翻译)SCAFFOLD:面向计算机科学研究图表的大规模结构化数据集,包含图表问答与思维链推理轨迹
Abstract page for arXiv paper 2609.00018: SCAFFOLD: A Large-Scale Structured Dataset of Computer Science Research Figures with Diagram QA and Chain-of-Thought Reasoning Traces
刚刚,李飞飞掀桌!全球首个多模态世界模型发布,几张照片省掉几百个机位
#欢迎关注爱范儿官方微信公众号:爱范儿(微信号:ifanr),更多精彩内容第一时间为您奉上。

李飞飞的 World Labs,把赛博朋克的「超梦」做出来了
#欢迎关注爱范儿官方微信公众号:爱范儿(微信号:ifanr),更多精彩内容第一时间为您奉上。

DeepSeek V4 多模态开源,我们把它的视觉链路拆了一遍
作者丨 郑 佳 美 编辑丨 岑 峰 &n

