ResearchRadar
ResearchRadar
工作原理来源使用场景每周雷达
ENESCADEZHJAHI登录免费开始
实时模板 · 只读无需账户

计算机视觉

从预印本到期刊,跟踪视觉学习、图像理解和感知方法。

搜索查询“computer vision”
过去 7 天4 个来源3 条发现

此雷达的来源

选择一个或多个来源来筛选结果。

3 条结果
arXiv
预印本arXiv · Computer Vision·2026年10月1日

VISTA: A Visual Harness for Reasoning in an Interactive World

We show that multimodal models possess strong reasoning abilities and that an appropriate harness can unlock their potential to solve tasks across diverse interactive environments. We introduce VISTA, a visual harness that gives a general-purpose multimodal model long-horizon vision. VISTA allows the model to directly perceive the environment through visual observations and maintains a lossless visual memory that pr…

DOI
10.48550/arXiv.2610.02200 ↗
作者
Qiushi Han, Keya Hu, Linlu Qiu, Cathy Wu, Kaiming He
期刊 / 发布平台
未提供
出版类型
预印本
状态
预印本
已发表版本
来源元数据中未链接
计算机视觉70% 相关阅读原文 ↗
arXiv
预印本arXiv · Computer Vision·2026年10月1日

GeoLatent: Geometry-Guided Latent Structuring with Routed Optimization for 3D Reasoning

Despite progress in vision-language models, 3D spatial reasoning from 2D images remains challenging. Text-based methods describe intermediate geometry with discrete tokens, limiting fidelity for continuous spatial relations. Continuous latents offer richer representations, but a single latent type does not explicitly separate the cues needed across spatial tasks. Decomposed spatial latents address this by representi…

DOI
10.48550/arXiv.2610.02091 ↗
作者
Yakun Zhu, Yi Bin, Yujuan Ding, Zheng Wang, Pengpeng Zeng, Duo Peng, Jingkuan Song, Heng Tao Shen
期刊 / 发布平台
未提供
出版类型
预印本
状态
预印本
已发表版本
来源元数据中未链接
计算机视觉70% 相关阅读原文 ↗
arXiv
预印本arXiv · Computer Vision·2026年10月1日

Learning from Failure: Leveraging Unreliable Predictions in Semi-Supervised Real-World Adverse Weather Removal

Adverse weather image restoration aims to recover images degraded by rain, haze, snow, and other weather-induced artifacts, thereby improving the robustness of outdoor vision systems. Existing unified restoration models exhibit limited generalization to real-world scenes due to their reliance on synthetic supervision and insufficient semantic constraints. In this paper, we propose a novel student--teacher semi-super…

DOI
10.48550/arXiv.2610.02051 ↗
作者
Cap Dang Xuan Kiet, Tat-Jen Cham
期刊 / 发布平台
未提供
出版类型
预印本
状态
预印本
已发表版本
来源元数据中未链接
计算机视觉70% 相关阅读原文 ↗
由 ResearchRadar 提供支持结果收集时间: 2026年10月2日 04:35 UTC大约每 15 分钟更新一次
真实的雷达,无需承诺。打开每篇原始出版物,浏览排序结果。只有保存发现、修改设置或接收每日摘要时才需要账户。

想把这个雷达变成自己的吗?

免费创建账户,主题、查询、时间范围和来源集已准备好,可随时编辑。

试用其他模板