Computer Vision and Pattern Recognition

Authors and titles for recent submissions, skipping first 109

[ total of 749 entries: 1-50 | 10-59 | 60-109 | 110-159 | 160-209 | 210-259 | 260-309 | ... | 710-749 ]
[ showing 50 entries per page: fewer | more | all ]

Wed, 10 Dec 2025 (continued, showing last 22 of 131 entries)

[110] arXiv:2512.07838 [pdf, ps, other]: Title: Detection of Cyberbullying in GIF using AI

Authors: Pal Dave, Xiaohong Yuan, Madhuri Siddula, Kaushik Roy

Subjects: Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG); Multimedia (cs.MM)
[111] arXiv:2512.08715 (cross-list from cs.PF) [pdf, ps, other]: Title: Multi-domain performance analysis with scores tailored to user preferences

Authors: Sébastien Piérard, Adrien Deliège, Marc Van Droogenbroeck

Subjects: Performance (cs.PF); Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG)
[112] arXiv:2512.08629 (cross-list from cs.AI) [pdf, ps, other]: Title: See-Control: A Multimodal Agent Framework for Smartphone Interaction with a Robotic Arm

Authors: Haoyu Zhao, Weizhong Ding, Yuhao Yang, Zheng Tian, Linyi Yang, Kun Shao, Jun Wang

Subjects: Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV); Human-Computer Interaction (cs.HC)
[113] arXiv:2512.08545 (cross-list from cs.CL) [pdf, ps, other]: Title: Curriculum Guided Massive Multi Agent System Solving For Robust Long Horizon Tasks

Authors: Indrajit Kar, Kalathur Chenchu Kishore Kumar

Comments: 22 pages, 2 tables, 9 figures

Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV); Multiagent Systems (cs.MA)
[114] arXiv:2512.08500 (cross-list from cs.GR) [pdf, ps, other]: Title: Learning to Control Physically-simulated 3D Characters via Generating and Mimicking 2D Motions

Authors: Jianan Li, Xiao Chen, Tao Huang, Tien-Tsin Wong

Subjects: Graphics (cs.GR); Computer Vision and Pattern Recognition (cs.CV)
[115] arXiv:2512.08360 (cross-list from cs.NE) [pdf, ps, other]: Title: Conditional Morphogenesis: Emergent Generation of Structural Digits via Neural Cellular Automata

Authors: Ali Sakour

Comments: 13 pages, 5 figures. Code available at: this https URL

Subjects: Neural and Evolutionary Computing (cs.NE); Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG)
[116] arXiv:2512.08284 (cross-list from physics.geo-ph) [pdf, ps, other]: Title: Self-Reinforced Deep Priors for Reparameterized Full Waveform Inversion

Authors: Guangyuan Zou, Junlun Li, Feng Liu, Xuejing Zheng, Jianjian Xie, Guoyi Chen

Comments: Submitted to GEOPHYSICS

Subjects: Geophysics (physics.geo-ph); Computer Vision and Pattern Recognition (cs.CV)
[117] arXiv:2512.08271 (cross-list from cs.RO) [pdf, ps, other]: Title: Zero-Splat TeleAssist: A Zero-Shot Pose Estimation Framework for Semantic Teleoperation

Authors: Srijan Dokania, Dharini Raghavan

Comments: Published and Presented at 3rd Workshop on Human-Centric Multilateral Teleoperation in ICRA 2025

Subjects: Robotics (cs.RO); Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG); Image and Video Processing (eess.IV)
[118] arXiv:2512.08216 (cross-list from eess.IV) [pdf, ps, other]: Title: Tumor-anchored deep feature random forests for out-of-distribution detection in lung cancer segmentation

Authors: Aneesh Rangnekar, Harini Veeraraghavan

Subjects: Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG)
[119] arXiv:2512.08188 (cross-list from cs.RO) [pdf, ps, other]: Title: Embodied Tree of Thoughts: Deliberate Manipulation Planning with Embodied World Model

Authors: Wenjiang Xu, Cindy Wang, Rui Fang, Mingkang Zhang, Lusong Li, Jing Xu, Jiayuan Gu, Zecui Zeng, Rui Chen

Comments: Website at this https URL

Subjects: Robotics (cs.RO); Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV)
[120] arXiv:2512.08170 (cross-list from cs.RO) [pdf, ps, other]: Title: RAVES-Calib: Robust, Accurate and Versatile Extrinsic Self Calibration Using Optimal Geometric Features

Authors: Haoxin Zhang, Shuaixin Li, Xiaozhou Zhu, Hongbo Chen, Wen Yao

Subjects: Robotics (cs.RO); Computer Vision and Pattern Recognition (cs.CV)
[121] arXiv:2512.08153 (cross-list from cs.LG) [pdf, ps, other]: Title: TreeGRPO: Tree-Advantage GRPO for Online RL Post-Training of Diffusion Models

Authors: Zheng Ding, Weirui Ye

Subjects: Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV)
[122] arXiv:2512.08125 (cross-list from eess.IV) [pdf, ps, other]: Title: FlowSteer: Conditioning Flow Field for Consistent Image Restoration

Authors: Tharindu Wickremasinghe, Chenyang Qi, Harshana Weligampola, Zhengzhong Tu, Stanley H. Chan

Subjects: Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV)
[123] arXiv:2512.08099 (cross-list from math.NA) [pdf, ps, other]: Title: Generalizations of the Normalized Radon Cumulative Distribution Transform for Limited Data Recognition

Authors: Matthias Beckmann, Robert Beinert, Jonas Bresch

Subjects: Numerical Analysis (math.NA); Computer Vision and Pattern Recognition (cs.CV); Information Theory (cs.IT)
[124] arXiv:2512.08029 (cross-list from cs.LG) [pdf, ps, other]: Title: CLARITY: Medical World Model for Guiding Treatment Decisions by Modeling Context-Aware Disease Trajectories in Latent Space

Authors: Tianxingjian Ding, Yuanhao Zou, Chen Chen, Mubarak Shah, Yu Tian

Subjects: Machine Learning (cs.LG); Computer Vision and Pattern Recognition (cs.CV)
[125] arXiv:2512.07998 (cross-list from cs.RO) [pdf, ps, other]: Title: DIJIT: A Robotic Head for an Active Observer

Authors: Mostafa Kamali Tabrizi, Mingshi Chi, Bir Bikram Dey, Yu Qing Yuan, Markus D. Solbach, Yiqian Liu, Michael Jenkin, John K. Tsotsos

Subjects: Robotics (cs.RO); Computer Vision and Pattern Recognition (cs.CV)
[126] arXiv:2512.07981 (cross-list from cs.LG) [pdf, ps, other]: Title: CIP-Net: Continual Interpretable Prototype-based Network

Authors: Federico Di Valerio, Michela Proietti, Alessio Ragno, Roberto Capobianco

Subjects: Machine Learning (cs.LG); Computer Vision and Pattern Recognition (cs.CV)
[127] arXiv:2512.07976 (cross-list from cs.RO) [pdf, ps, other]: Title: VLD: Visual Language Goal Distance for Reinforcement Learning Navigation

Authors: Lazar Milikic, Manthan Patel, Jonas Frey

Subjects: Robotics (cs.RO); Computer Vision and Pattern Recognition (cs.CV)
[128] arXiv:2512.07969 (cross-list from cs.RO) [pdf, ps, other]: Title: Sparse Variable Projection in Robotic Perception: Exploiting Separable Structure for Efficient Nonlinear Optimization

Authors: Alan Papalia, Nikolas Sanderson, Haoyu Han, Heng Yang, Hanumant Singh, Michael Everett

Comments: 8 pages, submitted for review

Subjects: Robotics (cs.RO); Computer Vision and Pattern Recognition (cs.CV)
[129] arXiv:2512.07884 (cross-list from cs.LG) [pdf, ps, other]: Title: GSPN-2: Efficient Parallel Sequence Modeling

Authors: Hongjun Wang, Yitong Jiang, Collin McCarthy, David Wehr, Hanrong Ye, Xinhao Li, Ka Chun Cheung, Wonmin Byeon, Jinwei Gu, Ke Chen, Kai Han, Hongxu Yin, Pavlo Molchanov, Jan Kautz, Sifei Liu

Comments: NeurIPS 2025

Subjects: Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV)
[130] arXiv:2512.07855 (cross-list from cs.LG) [pdf, ps, other]: Title: LAPA: Log-Domain Prediction-Driven Dynamic Sparsity Accelerator for Transformer Model

Authors: Huizheng Wang, Hongbin Wang, Shaojun Wei, Yang Hu, Shouyi Yin

Subjects: Machine Learning (cs.LG); Computer Vision and Pattern Recognition (cs.CV)
[131] arXiv:2512.05791 (cross-list from physics.med-ph) [pdf, ps, other]: Title: Fast and Robust Diffusion Posterior Sampling for MR Image Reconstruction Using the Preconditioned Unadjusted Langevin Algorithm

Authors: Moritz Blumenthal, Tina Holliber, Jonathan I. Tamir, Martin Uecker

Comments: Submitted to Magnetic Resonance in Medicine

Subjects: Medical Physics (physics.med-ph); Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG); Probability (math.PR)

Tue, 9 Dec 2025 (showing first 28 of 259 entries)

[132] arXiv:2512.07834 [pdf, ps, other]: Title: Voxify3D: Pixel Art Meets Volumetric Rendering

Authors: Yi-Chuan Huang, Jiewen Chan, Hao-Jen Chien, Yu-Lun Liu

Comments: Project page: this https URL

Subjects: Computer Vision and Pattern Recognition (cs.CV)
[133] arXiv:2512.07833 [pdf, ps, other]: Title: Relational Visual Similarity

Authors: Thao Nguyen, Sicheng Mo, Krishna Kumar Singh, Yilin Wang, Jing Shi, Nicholas Kolkin, Eli Shechtman, Yong Jae Lee, Yuheng Li

Comments: Project page, data, and code: this https URL

Subjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
[134] arXiv:2512.07831 [pdf, ps, other]: Title: UnityVideo: Unified Multi-Modal Multi-Task Learning for Enhancing World-Aware Video Generation

Authors: Jiehui Huang, Yuechen Zhang, Xu He, Yuan Gao, Zhi Cen, Bin Xia, Yan Zhou, Xin Tao, Pengfei Wan, Jiaya Jia

Comments: Project Website this https URL

Subjects: Computer Vision and Pattern Recognition (cs.CV)
[135] arXiv:2512.07829 [pdf, ps, other]: Title: One Layer Is Enough: Adapting Pretrained Visual Encoders for Image Generation

Authors: Yuan Gao, Chen Chen, Tianrong Chen, Jiatao Gu

Subjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)
[136] arXiv:2512.07826 [pdf, ps, other]: Title: OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing

Authors: Haoyang He, Jie Wang, Jiangning Zhang, Zhucun Xue, Xingyuan Bu, Qiangpeng Yang, Shilei Wen, Lei Xie

Comments: 38 pages

Subjects: Computer Vision and Pattern Recognition (cs.CV)
[137] arXiv:2512.07821 [pdf, ps, other]: Title: WorldReel: 4D Video Generation with Consistent Geometry and Motion Modeling

Authors: Shaoheng Fang, Hanwen Jiang, Yunpeng Bai, Niloy J. Mitra, Qixing Huang

Subjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)
[138] arXiv:2512.07807 [pdf, ps, other]: Title: Lang3D-XL: Language Embedded 3D Gaussians for Large-scale Scenes

Authors: Shai Krakovsky, Gal Fiebelman, Sagie Benaim, Hadar Averbuch-Elor

Comments: Accepted to SIGGRAPH Asia 2025. Project webpage: this https URL

Subjects: Computer Vision and Pattern Recognition (cs.CV); Graphics (cs.GR)
[139] arXiv:2512.07806 [pdf, ps, other]: Title: Multi-view Pyramid Transformer: Look Coarser to See Broader

Authors: Gyeongjin Kang, Seungkwon Yang, Seungtae Nam, Younggeun Lee, Jungwoo Kim, Eunbyung Park

Comments: Project page: see this https URL

Subjects: Computer Vision and Pattern Recognition (cs.CV)
[140] arXiv:2512.07802 [pdf, ps, other]: Title: OneStory: Coherent Multi-Shot Video Generation with Adaptive Memory

Authors: Zhaochong An, Menglin Jia, Haonan Qiu, Zijian Zhou, Xiaoke Huang, Zhiheng Liu, Weiming Ren, Kumara Kahatapitiya, Ding Liu, Sen He, Chenyang Zhang, Tao Xiang, Fanny Yang, Serge Belongie, Tian Xie

Comments: Project Page: this https URL

Subjects: Computer Vision and Pattern Recognition (cs.CV)
[141] arXiv:2512.07778 [pdf, ps, other]: Title: Distribution Matching Variational AutoEncoder

Authors: Sen Ye, Jianning Pei, Mengde Xu, Shuyang Gu, Chunyu Wang, Liwei Wang, Han Hu

Subjects: Computer Vision and Pattern Recognition (cs.CV)
[142] arXiv:2512.07776 [pdf, ps, other]: Title: GorillaWatch: An Automated System for In-the-Wild Gorilla Re-Identification and Population Monitoring

Authors: Maximilian Schall, Felix Leonard Knöfel, Noah Elias König, Jan Jonas Kubeler, Maximilian von Klinski, Joan Wilhelm Linnemann, Xiaoshi Liu, Iven Jelle Schlegelmilch, Ole Woyciniuk, Alexandra Schild, Dante Wasmuht, Magdalena Bermejo Espinet, German Illera Basas, Gerard de Melo

Comments: Accepted at WACV 2026

Subjects: Computer Vision and Pattern Recognition (cs.CV)
[143] arXiv:2512.07760 [pdf, ps, other]: Title: Modality-Aware Bias Mitigation and Invariance Learning for Unsupervised Visible-Infrared Person Re-Identification

Authors: Menglin Wang, Xiaojin Gong, Jiachen Li, Genlin Ji

Comments: Accepted to AAAI 2026

Subjects: Computer Vision and Pattern Recognition (cs.CV)
[144] arXiv:2512.07756 [pdf, ps, other]: Title: UltrasODM: A Dual Stream Optical Flow Mamba Network for 3D Freehand Ultrasound Reconstruction

Authors: Mayank Anand, Ujair Alam, Surya Prakash, Priya Shukla, Gora Chand Nandi, Domenec Puig

Subjects: Computer Vision and Pattern Recognition (cs.CV); Robotics (cs.RO)
[145] arXiv:2512.07747 [pdf, ps, other]: Title: Unison: A Fully Automatic, Task-Universal, and Low-Cost Framework for Unified Understanding and Generation

Authors: Shihao Zhao, Yitong Chen, Zeyinzi Jiang, Bojia Zi, Shaozhe Hao, Yu Liu, Chaojie Mao, Kwan-Yee K. Wong

Subjects: Computer Vision and Pattern Recognition (cs.CV)
[146] arXiv:2512.07745 [pdf, ps, other]: Title: DiffusionDriveV2: Reinforcement Learning-Constrained Truncated Diffusion Modeling in End-to-End Autonomous Driving

Authors: Jialv Zou, Shaoyu Chen, Bencheng Liao, Zhiyu Zheng, Yuehao Song, Lefei Zhang, Qian Zhang, Wenyu Liu, Xinggang Wang

Subjects: Computer Vision and Pattern Recognition (cs.CV)
[147] arXiv:2512.07738 [pdf, ps, other]: Title: HLTCOE Evaluation Team at TREC 2025: VQA Track

Authors: Dengjia Zhang, Charles Weng, Katherine Guerrerio, Yi Lu, Kenton Murray, Alexander Martin, Reno Kriz, Benjamin Van Durme

Comments: 7 pages, 1 figure

Subjects: Computer Vision and Pattern Recognition (cs.CV)
[148] arXiv:2512.07733 [pdf, ps, other]: Title: SpatialDreamer: Incentivizing Spatial Reasoning via Active Mental Imagery

Authors: Meng Cao, Xingyu Li, Xue Liu, Ian Reid, Xiaodan Liang

Subjects: Computer Vision and Pattern Recognition (cs.CV)
[149] arXiv:2512.07730 [pdf, ps, other]: Title: SAVE: Sparse Autoencoder-Driven Visual Information Enhancement for Mitigating Object Hallucination

Authors: Sangha Park, Seungryong Yoo, Jisoo Mok, Sungroh Yoon

Comments: WACV 2026

Subjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)
[150] arXiv:2512.07729 [pdf, ps, other]: Title: Improving action classification with brain-inspired deep networks

Authors: Aidas Aglinskas, Stefano Anzellotti

Subjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)
[151] arXiv:2512.07720 [pdf, ps, other]: Title: ViSA: 3D-Aware Video Shading for Real-Time Upper-Body Avatar Creation

Authors: Fan Yang, Heyuan Li, Peihao Li, Weihao Yuan, Lingteng Qiu, Chaoyue Song, Cheng Chen, Yisheng He, Shifeng Zhang, Xiaoguang Han, Steven Hoi, Guosheng Lin

Comments: Project page: \url{this https URL}

Subjects: Computer Vision and Pattern Recognition (cs.CV)
[152] arXiv:2512.07712 [pdf, ps, other]: Title: UnCageNet: Tracking and Pose Estimation of Caged Animal

Authors: Sayak Dutta, Harish Katti, Shashikant Verma, Shanmuganathan Raman

Comments: 9 pages, 2 figures, 2 tables. Accepted to the Indian Conference on Computer Vision, Graphics, and Image Processing (ICVGIP 2025), Mandi, India

Subjects: Computer Vision and Pattern Recognition (cs.CV)
[153] arXiv:2512.07703 [pdf, ps, other]: Title: PVeRA: Probabilistic Vector-Based Random Matrix Adaptation

Authors: Leo Fillioux, Enzo Ferrante, Paul-Henry Cournède, Maria Vakalopoulou, Stergios Christodoulidis

Subjects: Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG)
[154] arXiv:2512.07702 [pdf, ps, other]: Title: Guiding What Not to Generate: Automated Negative Prompting for Text-Image Alignment

Authors: Sangha Park, Eunji Kim, Yeongtak Oh, Jooyoung Choi, Sungroh Yoon

Comments: WACV 2026

Subjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)
[155] arXiv:2512.07698 [pdf, ps, other]: Title: sim2art: Accurate Articulated Object Modeling from a Single Video using Synthetic Training Data Only

Authors: Arslan Artykov, Corentin Sautier, Vincent Lepetit

Subjects: Computer Vision and Pattern Recognition (cs.CV); Robotics (cs.RO)
[156] arXiv:2512.07674 [pdf, ps, other]: Title: DIST-CLIP: Arbitrary Metadata and Image Guided MRI Harmonization via Disentangled Anatomy-Contrast Representations

Authors: Mehmet Yigit Avci, Pedro Borges, Virginia Fernandez, Paul Wright, Mehmet Yigitsoy, Sebastien Ourselin, Jorge Cardoso

Subjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)
[157] arXiv:2512.07668 [pdf, ps, other]: Title: EgoCampus: Egocentric Pedestrian Eye Gaze Model and Dataset

Authors: Ronan John, Aditya Kesari, Vincenzo DiMatteo, Kristin Dana

Subjects: Computer Vision and Pattern Recognition (cs.CV)
[158] arXiv:2512.07661 [pdf, ps, other]: Title: Optimization-Guided Diffusion for Interactive Scene Generation

Authors: Shiaho Li, Naisheng Ye, Tianyu Li, Kashyap Chitta, Tuo An, Peng Su, Boyang Wang, Haiou Liu, Chen Lv, Hongyang Li

Subjects: Computer Vision and Pattern Recognition (cs.CV)
[159] arXiv:2512.07652 [pdf, ps, other]: Title: An AI-Powered Autonomous Underwater System for Sea Exploration and Scientific Research

Authors: Hamad Almazrouei, Mariam Al Nasseri, Maha Alzaabi

Subjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)

[ total of 749 entries: 1-50 | 10-59 | 60-109 | 110-159 | 160-209 | 210-259 | 260-309 | ... | 710-749 ]
[ showing 50 entries per page: fewer | more | all ]

Disable MathJax (What is MathJax?)

Links to: arXiv, form interface, find, cs, new, 2512, contact, help (Access key information)

> cs > cs.CV

Computer Vision and Pattern Recognition

Authors and titles for recent submissions, skipping first 109

Wed, 10 Dec 2025 (continued, showing last 22 of 131 entries)

Tue, 9 Dec 2025 (showing first 28 of 259 entries)