ResearchRadar
ResearchRadar
यह कैसे काम करता हैस्रोतउपयोगसाप्ताहिक रडार
ENESCADEZHJAHIसाइन इनमुफ़्त शुरू करें
लाइव टेम्पलेट · केवल पढ़ने के लिएखाता ज़रूरी नहीं

कंप्यूटर विज़न

प्रीप्रिंट से जर्नल तक दृश्य अधिगम, छवि समझ और अनुभूति की विधियाँ।

खोज क्वेरी“computer vision”
पिछले 7 दिन4 स्रोत3 निष्कर्ष

इस रडार के स्रोत

परिणाम फ़िल्टर करने के लिए एक या अधिक स्रोत चुनें।

3 परिणाम
arXiv
प्रीप्रिंटarXiv · Computer Vision·1 अक्टू॰ 2026

VISTA: A Visual Harness for Reasoning in an Interactive World

We show that multimodal models possess strong reasoning abilities and that an appropriate harness can unlock their potential to solve tasks across diverse interactive environments. We introduce VISTA, a visual harness that gives a general-purpose multimodal model long-horizon vision. VISTA allows the model to directly perceive the environment through visual observations and maintains a lossless visual memory that pr…

DOI
10.48550/arXiv.2610.02200 ↗
लेखक
Qiushi Han, Keya Hu, Linlu Qiu, Cathy Wu, Kaiming He
जर्नल / मंच
उपलब्ध नहीं
प्रकाशन प्रकार
प्रीप्रिंट
स्थिति
प्रीप्रिंट
प्रकाशित संस्करण
स्रोत मेटाडेटा में लिंक नहीं
कंप्यूटर विज़न70% प्रासंगिकमूल सामग्री पढ़ें ↗
arXiv
प्रीप्रिंटarXiv · Computer Vision·1 अक्टू॰ 2026

GeoLatent: Geometry-Guided Latent Structuring with Routed Optimization for 3D Reasoning

Despite progress in vision-language models, 3D spatial reasoning from 2D images remains challenging. Text-based methods describe intermediate geometry with discrete tokens, limiting fidelity for continuous spatial relations. Continuous latents offer richer representations, but a single latent type does not explicitly separate the cues needed across spatial tasks. Decomposed spatial latents address this by representi…

DOI
10.48550/arXiv.2610.02091 ↗
लेखक
Yakun Zhu, Yi Bin, Yujuan Ding, Zheng Wang, Pengpeng Zeng, Duo Peng, Jingkuan Song, Heng Tao Shen
जर्नल / मंच
उपलब्ध नहीं
प्रकाशन प्रकार
प्रीप्रिंट
स्थिति
प्रीप्रिंट
प्रकाशित संस्करण
स्रोत मेटाडेटा में लिंक नहीं
कंप्यूटर विज़न70% प्रासंगिकमूल सामग्री पढ़ें ↗
arXiv
प्रीप्रिंटarXiv · Computer Vision·1 अक्टू॰ 2026

Learning from Failure: Leveraging Unreliable Predictions in Semi-Supervised Real-World Adverse Weather Removal

Adverse weather image restoration aims to recover images degraded by rain, haze, snow, and other weather-induced artifacts, thereby improving the robustness of outdoor vision systems. Existing unified restoration models exhibit limited generalization to real-world scenes due to their reliance on synthetic supervision and insufficient semantic constraints. In this paper, we propose a novel student--teacher semi-super…

DOI
10.48550/arXiv.2610.02051 ↗
लेखक
Cap Dang Xuan Kiet, Tat-Jen Cham
जर्नल / मंच
उपलब्ध नहीं
प्रकाशन प्रकार
प्रीप्रिंट
स्थिति
प्रीप्रिंट
प्रकाशित संस्करण
स्रोत मेटाडेटा में लिंक नहीं
कंप्यूटर विज़न70% प्रासंगिकमूल सामग्री पढ़ें ↗
ResearchRadar द्वारा संचालितपरिणाम एकत्र किए गए: 2 अक्टू॰ 2026, 4:35 am UTCलगभग हर 15 मिनट में अपडेट होता है
एक वास्तविक रडार, बिना किसी बाध्यता के।हर मूल प्रकाशन खोलें और क्रमबद्ध परिणाम देखें। खाता केवल निष्कर्ष सहेजने, सेटअप बदलने या दैनिक सारांश पाने के लिए चाहिए।

इस रडार को अपना बनाना चाहते हैं?

इस विषय, क्वेरी, समय-सीमा और स्रोत सेट के साथ एक मुफ़्त खाता बनाएँ, संपादित करने के लिए तैयार।

दूसरा टेम्पलेट आज़माएँ