Google AI Blog发布于 07/15 00:00

Celebrating 25 years of visual search innovation

(翻译)庆祝视觉搜索创新25周年

Google Images is turning 25. Here’s a look back at some major milestones — and new ways to explore and create visual content.

查看原文
NVIDIA Technical Blog发布于 09/01 00:00

Scale AV Perception Across Vehicle Platforms with NVIDIA Omniverse NuRec

(翻译)使用 NVIDIA Omniverse NuRec 扩展跨车辆平台的自动驾驶感知

A perception stack is shaped by the vehicle that carries it. Move the same software to a new carline—for example, from an SUV to a sedan or another vehicle…

查看原文
arXiv cs.AI发布于 09/02 12:00

SCAFFOLD: A Large-Scale Structured Dataset of Computer Science Research Figures with Diagram QA and Chain-of-Thought Reasoning Traces

(翻译)SCAFFOLD:面向计算机科学研究图表的大规模结构化数据集,包含图表问答与思维链推理轨迹

Abstract page for arXiv paper 2609.00018: SCAFFOLD: A Large-Scale Structured Dataset of Computer Science Research Figures with Diagram QA and Chain-of-Thought Reasoning Traces

查看原文
雷峰网发布于 09/01 18:38

DeepSeek V4 多模态开源,我们把它的视觉链路拆了一遍

    作者丨 郑 佳 美     编辑丨 岑   峰                                                                       &n

查看原文
雷峰网发布于 09/03 18:07

打破黑盒猜想:大模型通往真正「空间智能」的破局之路

    作者丨 张   璐     编辑丨 幸丽娟                                                                       &nbs

查看原文
雷峰网发布于 09/09 10:26

腾讯混元、清华、南洋理工联手,「以小博大」破解空间智能算力与记忆断裂难题 | ECCV 2026

    作者丨 张   璐     编辑丨 幸丽娟                                                                       &nbs

查看原文
人人都是产品经理发布于 09/15 09:46

GPT Image 2.5 对比实测:变化与局限

GPT Image 2.5 在复杂中文图文任务中表现如何?本文通过十一类任务、72张输出,对比两代模型在海报、安装说明、经营简报等场景下的能力。结果显示,2.5 在商品信息对应和数字复现上有所提升,但在工具接触点、零件数量和图表比例上仍会出现错误,书法风格差异明显。

查看原文