Computer Vision and Pattern Recognition
PhyGround: Benchmarking Physical Reasoning in ...
Computer Vision and Pattern Recognitionlibrarian
29 views
Image Generators are Generalist Vision Learners
Computer Vision and Pattern RecognitionVision Banana
58 views
MM-WebAgent: A Hierarchical Multimodal Web Age...
Computer Vision and Pattern Recognitionlibrarian
70 views
ActionParty: Multi-Subject Action Binding in G...
Computer Vision and Pattern RecognitionAlexander Pondaven
95 views
No Hard Negatives Required: Concept Centric Le...
Computer Vision and Pattern RecognitionHai Pham*
93 views
Do VLMs Need Vision Transformers? Evaluating S...
Computer Vision and Pattern Recognitionlibrarian
99 views
SAVeS: Steering Safety Judgments in Vision-Lan...
Computer Vision and Pattern Recognitionlibrarian
98 views
DreamPartGen: Semantically Grounded Part-Level...
Computer Vision and Pattern Recognitionlibrarian
101 views
Near-perfect photo-ID of the Hula painted frog...
Computer Vision and Pattern Recognitionyoavram
193 views
Multilayer Graph Approach to Deep Subspace Clu...
Computer Vision and Pattern Recognitionlovro-sindicic
177 views
Label-independent hyperparameter-free self-sup...
Computer Vision and Pattern Recognitionlovro-sindicic
185 views
PersonaLive! Expressive Portrait Image Animati...
Computer Vision and Pattern RecognitionGrisha Samokhin
190 views
Mull-Tokens: Modality-Agnostic Latent Thinking
Computer Vision and Pattern Recognitionlibrarian
201 views
Linear Gaussian Bounding Box Representation an...
Computer Vision and Pattern Recognitionrahulraj Kk
201 views
Point3R: Streaming 3D Reconstruction with Expl...
Computer Vision and Pattern Recognitionlibrarian
498 views
FADRM: Fast and Accurate Data Residual Matchin...
Computer Vision and Pattern Recognitionlibrarian
461 views
HalluSegBench: Counterfactual Visual Reasoning...
Computer Vision and Pattern Recognitionlibrarian
549 views
Whole-Body Conditioned Egocentric Video Prediction
Computer Vision and Pattern Recognitionlibrarian
538 views
Reinforcing Spatial Reasoning in Vision-Langua...
Computer Vision and Pattern Recognitionlibrarian
611 views
Outside Knowledge Conversational Video (OKCV) ...
Computer Vision and Pattern Recognitionlibrarian
492 views
Decoupling the Image Perception and Multimodal...
Computer Vision and Pattern Recognitionlibrarian
635 views
Direct Numerical Layout Generation for 3D Indo...
Computer Vision and Pattern Recognitionlibrarian
662 views
Refer to Anything with Vision-Language Prompts
Computer Vision and Pattern RecognitionShengcao Cao
653 views
Let Androids Dream of Electric Sheep: A Human-...
Computer Vision and Pattern RecognitionAnastasia Kokkanen
749 views
Delving into RL for Image Generation with CoT:...
Computer Vision and Pattern Recognitionlibrarian
605 views
Let Androids Dream of Electric Sheep: A Human-...
Computer Vision and Pattern Recognitionlibrarian
624 views
SpatialScore: Towards Unified Evaluation for M...
Computer Vision and Pattern RecognitionHaoning Wu
707 views
VTBench: Evaluating Visual Tokenizers for Auto...
Computer Vision and Pattern Recognitionlibrarian
677 views
Does Feasibility Matter? Understanding the Imp...
Computer Vision and Pattern Recognitionlibrarian
598 views
MathCoder-VL: Bridging Vision and Code for Enh...
Computer Vision and Pattern Recognitionlibrarian
670 views
StreamBridge: Turning Your Offline Video Large...
Computer Vision and Pattern Recognitionlibrarian
637 views