Skip to main content
Cornell University
We gratefully acknowledge support from the Simons Foundation, member institutions, and all contributors. Donate
arxiv logo > cs.CV
arXiv logo
Cornell University Logo

quick links

  • Login
  • Help Pages
  • About

Computer Vision and Pattern Recognition

Authors and titles for recent submissions

  • Fri, 1 Aug 2025
  • Thu, 31 Jul 2025
  • Wed, 30 Jul 2025
  • Tue, 29 Jul 2025
  • Mon, 28 Jul 2025

See today's new changes

Total of 642 entries : 1-50 ... 451-500 501-550 551-600 601-642
Showing up to 50 entries per page: fewer | more | all

Mon, 28 Jul 2025 (continued, showing last 42 of 118 entries )

[601] arXiv:2507.18713 [pdf, html, other]
Title: SaLF: Sparse Local Fields for Multi-Sensor Rendering in Real-Time
Yun Chen, Matthew Haines, Jingkang Wang, Krzysztof Baron-Lis, Sivabalan Manivasagam, Ze Yang, Raquel Urtasun
Subjects: Computer Vision and Pattern Recognition (cs.CV); Robotics (cs.RO)
[602] arXiv:2507.18678 [pdf, html, other]
Title: Towards Scalable Spatial Intelligence via 2D-to-3D Data Lifting
Xingyu Miao, Haoran Duan, Quanhao Qian, Jiuniu Wang, Yang Long, Ling Shao, Deli Zhao, Ran Xu, Gongjie Zhang
Comments: ICCV 2025 (Highlight)
Subjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)
[603] arXiv:2507.18677 [pdf, html, other]
Title: HeartUnloadNet: A Weakly-Supervised Cycle-Consistent Graph Network for Predicting Unloaded Cardiac Geometry from Diastolic States
Siyu Mu, Wei Xuan Chan, Choon Hwai Yap
Comments: Codes are available at this https URL
Subjects: Computer Vision and Pattern Recognition (cs.CV); Medical Physics (physics.med-ph)
[604] arXiv:2507.18675 [pdf, other]
Title: Advancing Vision-based Human Action Recognition: Exploring Vision-Language CLIP Model for Generalisation in Domain-Independent Tasks
Utkarsh Shandilya, Marsha Mariya Kappan, Sanyam Jain, Vijeta Sharma
Subjects: Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG)
[605] arXiv:2507.18667 [pdf, html, other]
Title: Gen-AI Police Sketches with Stable Diffusion
Nicholas Fidalgo, Aaron Contreras, Katherine Harvey, Johnny Ni
Subjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)
[606] arXiv:2507.18661 [pdf, other]
Title: Eyes Will Shut: A Vision-Based Next GPS Location Prediction Model by Reinforcement Learning from Visual Map Feed Back
Ruixing Zhang, Yang Zhang, Tongyu Zhu, Leilei Sun, Weifeng Lv
Subjects: Computer Vision and Pattern Recognition (cs.CV); Robotics (cs.RO)
[607] arXiv:2507.18660 [pdf, html, other]
Title: Fuzzy Theory in Computer Vision: A Review
Adilet Yerkin, Ayan Igali, Elnara Kadyrgali, Maksat Shagyrov, Malika Ziyada, Muragul Muratbekova, Pakizar Shamoi
Comments: Submitted to Journal of Intelligent and Fuzzy Systems for consideration (8 pages, 6 figures, 1 table)
Subjects: Computer Vision and Pattern Recognition (cs.CV)
[608] arXiv:2507.18657 [pdf, other]
Title: VGS-ATD: Robust Distributed Learning for Multi-Label Medical Image Classification Under Heterogeneous and Imbalanced Conditions
Zehui Zhao, Laith Alzubaidi, Haider A.Alwzwazy, Jinglan Zhang, Yuantong Gu
Comments: The idea is still underdeveloped, not yet enough to be published
Subjects: Computer Vision and Pattern Recognition (cs.CV); Cryptography and Security (cs.CR)
[609] arXiv:2507.18656 [pdf, html, other]
Title: ShrinkBox: Backdoor Attack on Object Detection to Disrupt Collision Avoidance in Machine Learning-based Advanced Driver Assistance Systems
Muhammad Zaeem Shahzad, Muhammad Abdullah Hanif, Bassem Ouni, Muhammad Shafique
Comments: 8 pages, 8 figures, 1 table
Subjects: Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG)
[610] arXiv:2507.18655 [pdf, html, other]
Title: Part Segmentation of Human Meshes via Multi-View Human Parsing
James Dickens, Kamyar Hamad
Subjects: Computer Vision and Pattern Recognition (cs.CV); Image and Video Processing (eess.IV)
[611] arXiv:2507.18653 [pdf, html, other]
Title: Adapt, But Don't Forget: Fine-Tuning and Contrastive Routing for Lane Detection under Distribution Shift
Mohammed Abdul Hafeez Khan, Parth Ganeriwala, Sarah M. Lehman, Siddhartha Bhattacharyya, Amy Alvarez, Natasha Neogi
Comments: Accepted to ICCV 2025, 2COOOL Workshop. Total 14 pages, 5 tables, and 4 figures
Subjects: Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG)
[612] arXiv:2507.18650 [pdf, other]
Title: Features extraction for image identification using computer vision
Venant Niyonkuru, Sylla Sekou, Jimmy Jackson Sinzinkayo
Journal-ref: World Journal of Advanced Research and Reviews, 2025, 27(01), 1341-1351
Subjects: Computer Vision and Pattern Recognition (cs.CV)
[613] arXiv:2507.18649 [pdf, html, other]
Title: Livatar-1: Real-Time Talking Heads Generation with Tailored Flow Matching
Haiyang Liu, Xiaolin Hong, Xuancheng Yang, Yudi Ruan, Xiang Lian, Michael Lingelbach, Hongwei Yi, Wei Li
Comments: Technical Report
Subjects: Computer Vision and Pattern Recognition (cs.CV)
[614] arXiv:2507.18645 [pdf, html, other]
Title: Quantum-Cognitive Tunnelling Neural Networks for Military-Civilian Vehicle Classification and Sentiment Analysis
Milan Maksimovic, Anna Bohdanets, Immaculate Motsi-Omoijiade, Guido Governatori, Ivan S. Maksymov
Subjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)
[615] arXiv:2507.19328 (cross-list from eess.IV) [pdf, html, other]
Title: NerT-CA: Efficient Dynamic Reconstruction from Sparse-view X-ray Coronary Angiography
Kirsten W.H. Maas, Danny Ruijters, Nicola Pezzotti, Anna Vilanova
Subjects: Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV)
[616] arXiv:2507.19284 (cross-list from cs.CG) [pdf, html, other]
Title: Relaxed Total Generalized Variation Regularized Piecewise Smooth Mumford-Shah Model for Triangulated Surface Segmentation
Huayan Zhang, Shanqiang Wang, Xiaochao Wang
Subjects: Computational Geometry (cs.CG); Computer Vision and Pattern Recognition (cs.CV)
[617] arXiv:2507.19282 (cross-list from eess.IV) [pdf, other]
Title: SAM2-Aug: Prior knowledge-based Augmentation for Target Volume Auto-Segmentation in Adaptive Radiation Therapy Using Segment Anything Model 2
Guoping Xu, Yan Dai, Hengrui Zhao, Ying Zhang, Jie Deng, Weiguo Lu, You Zhang
Comments: 26 pages, 10 figures
Subjects: Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV); Medical Physics (physics.med-ph)
[618] arXiv:2507.19230 (cross-list from eess.IV) [pdf, html, other]
Title: Unstable Prompts, Unreliable Segmentations: A Challenge for Longitudinal Lesion Analysis
Niels Rocholl, Ewoud Smit, Mathias Prokop, Alessa Hering
Subjects: Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV)
[619] arXiv:2507.19225 (cross-list from cs.SD) [pdf, html, other]
Title: Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation
Fang Kang, Yin Cao, Haoyu Chen
Subjects: Sound (cs.SD); Computer Vision and Pattern Recognition (cs.CV); Multimedia (cs.MM); Audio and Speech Processing (eess.AS)
[620] arXiv:2507.19201 (cross-list from eess.IV) [pdf, html, other]
Title: Joint Holistic and Lesion Controllable Mammogram Synthesis via Gated Conditional Diffusion Model
Xin Li, Kaixiang Yang, Qiang Li, Zhiwei Wang
Comments: Accepted, ACM Multimedia 2025, 10 pages, 5 figures
Subjects: Image and Video Processing (eess.IV); Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV)
[621] arXiv:2507.19199 (cross-list from eess.IV) [pdf, html, other]
Title: Enhancing Diabetic Retinopathy Classification Accuracy through Dual Attention Mechanism in Deep Learning
Abdul Hannan, Zahid Mahmood, Rizwan Qureshi, Hazrat Ali
Comments: submitted to Computer Methods in Biomechanics and Biomedical Engineering: Imaging & Visualization
Subjects: Image and Video Processing (eess.IV); Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV)
[622] arXiv:2507.19197 (cross-list from cs.LG) [pdf, html, other]
Title: WACA-UNet: Weakness-Aware Channel Attention for Static IR Drop Prediction in Integrated Circuit Design
Youngmin Seo, Yunhyeong Kwon, Younghun Park, HwiRyong Kim, Seungho Eum, Jinha Kim, Taigon Song, Juho Kim, Unsang Park
Comments: 9 pages, 5 figures
Subjects: Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV)
[623] arXiv:2507.19186 (cross-list from eess.IV) [pdf, other]
Title: Reconstruct or Generate: Exploring the Spectrum of Generative Modeling for Cardiac MRI
Niklas Bubeck, Yundi Zhang, Suprosanna Shit, Daniel Rueckert, Jiazhen Pan
Subjects: Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV)
[624] arXiv:2507.19172 (cross-list from cs.AI) [pdf, html, other]
Title: PhysDrive: A Multimodal Remote Physiological Measurement Dataset for In-vehicle Driver Monitoring
Jiyao Wang, Xiao Yang, Qingyong Hu, Jiankai Tang, Can Liu, Dengbo He, Yuntao Wang, Yingcong Chen, Kaishun Wu
Comments: It is the initial version, not the final version
Subjects: Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV)
[625] arXiv:2507.19165 (cross-list from eess.IV) [pdf, html, other]
Title: Extreme Cardiac MRI Analysis under Respiratory Motion: Results of the CMRxMotion Challenge
Kang Wang, Chen Qin, Zhang Shi, Haoran Wang, Xiwen Zhang, Chen Chen, Cheng Ouyang, Chengliang Dai, Yuanhan Mo, Chenchen Dai, Xutong Kuang, Ruizhe Li, Xin Chen, Xiuzheng Yue, Song Tian, Alejandro Mora-Rubio, Kumaradevan Punithakumar, Shizhan Gong, Qi Dou, Sina Amirrajab, Yasmina Al Khalil, Cian M. Scannell, Lexiaozi Fan, Huili Yang, Xiaowu Sun, Rob van der Geest, Tewodros Weldebirhan Arega, Fabrice Meriaudeau, Caner Özer, Amin Ranem, John Kalkhof, İlkay Öksüz, Anirban Mukhopadhyay, Abdul Qayyum, Moona Mazher, Steven A Niederer, Carles Garcia-Cabrera, Eric Arazo, Michal K. Grzeszczyk, Szymon Płotka, Wanqin Ma, Xiaomeng Li, Rongjun Ge, Yongqing Kou, Xinrong Chen, He Wang, Chengyan Wang, Wenjia Bai, Shuo Wang
Subjects: Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV)
[626] arXiv:2507.19138 (cross-list from eess.IV) [pdf, html, other]
Title: RealisVSR: Detail-enhanced Diffusion for Real-World 4K Video Super-Resolution
Weisong Zhao, Jingkai Zhou, Xiangyu Zhu, Weihua Chen, Xiao-Yu Zhang, Zhen Lei, Fan Wang
Subjects: Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV)
[627] arXiv:2507.19132 (cross-list from cs.AI) [pdf, html, other]
Title: OS-MAP: How Far Can Computer-Using Agents Go in Breadth and Depth?
Xuetian Chen, Yinghao Chen, Xinfeng Yuan, Zhuo Peng, Lu Chen, Yuekeng Li, Zhoujia Zhang, Yingqian Huang, Leyan Huang, Jiaqing Liang, Tianbao Xie, Zhiyong Wu, Qiushi Sun, Biqing Qi, Bowen Zhou
Comments: Work in progress
Subjects: Artificial Intelligence (cs.AI); Computation and Language (cs.CL); Computer Vision and Pattern Recognition (cs.CV); Human-Computer Interaction (cs.HC)
[628] arXiv:2507.19125 (cross-list from eess.IV) [pdf, html, other]
Title: Learned Image Compression with Hierarchical Progressive Context Modeling
Yuqi Li, Haotian Zhang, Li Li, Dong Liu
Comments: 17 pages, ICCV 2025
Subjects: Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV); Multimedia (cs.MM)
[629] arXiv:2507.19089 (cross-list from cs.AI) [pdf, html, other]
Title: Fine-Grained Traffic Inference from Road to Lane via Spatio-Temporal Graph Node Generation
Shuhao Li, Weidong Yang, Yue Cui, Xiaoxing Liu, Lingkai Meng, Lipeng Ma, Fan Zhang
Subjects: Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV)
[630] arXiv:2507.19074 (cross-list from eess.IV) [pdf, other]
Title: A Self-training Framework for Semi-supervised Pulmonary Vessel Segmentation and Its Application in COPD
Shuiqing Zhao, Meihuan Wang, Jiaxuan Xu, Jie Feng, Wei Qian, Rongchang Chen, Zhenyu Liang, Shouliang Qi, Yanan Wu
Subjects: Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV)
[631] arXiv:2507.19041 (cross-list from quant-ph) [pdf, html, other]
Title: PGKET: A Photonic Gaussian Kernel Enhanced Transformer
Ren-Xin Zhao
Subjects: Quantum Physics (quant-ph); Computer Vision and Pattern Recognition (cs.CV)
[632] arXiv:2507.19035 (cross-list from eess.IV) [pdf, html, other]
Title: Dual Path Learning -- learning from noise and context for medical image denoising
Jitindra Fartiyal, Pedro Freire, Yasmeen Whayeb, James S. Wolffsohn, Sergei K. Turitsyn, Sergei G. Sokolov
Comments: 10 pages, 7 figures
Subjects: Image and Video Processing (eess.IV); Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV)
[633] arXiv:2507.18915 (cross-list from cs.CL) [pdf, html, other]
Title: Mining Contextualized Visual Associations from Images for Creativity Understanding
Ananya Sahu, Amith Ananthram, Kathleen McKeown
Subjects: Computation and Language (cs.CL); Computer Vision and Pattern Recognition (cs.CV)
[634] arXiv:2507.18895 (cross-list from eess.IV) [pdf, html, other]
Title: Dealing with Segmentation Errors in Needle Reconstruction for MRI-Guided Brachytherapy
Vangelis Kostoulas, Arthur Guijt, Ellen M. Kerkhof, Bradley R. Pieters, Peter A.N. Bosman, Tanja Alderliesten
Comments: Published in: Proc. SPIE Medical Imaging 2025, Vol. 13408, 1340826
Journal-ref: Proc. SPIE Medical Imaging 2025: Image-Guided Procedures, Robotic Interventions, and Modeling, Vol. 13408, 1340826 (2025)
Subjects: Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV)
[635] arXiv:2507.18830 (cross-list from eess.IV) [pdf, html, other]
Title: RealDeal: Enhancing Realism and Details in Brain Image Generation via Image-to-Image Diffusion Models
Shen Zhu, Yinzhu Jin, Tyler Spears, Ifrah Zawar, P. Thomas Fletcher
Comments: 19 pages, 10 figures
Subjects: Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG)
[636] arXiv:2507.18808 (cross-list from cs.RO) [pdf, html, other]
Title: Perpetua: Multi-Hypothesis Persistence Modeling for Semi-Static Environments
Miguel Saavedra-Ruiz, Samer B. Nashed, Charlie Gauthier, Liam Paull
Comments: Accepted to the IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS 2025) Code available at this https URL. Webpage and additional videos at this https URL
Subjects: Robotics (cs.RO); Computer Vision and Pattern Recognition (cs.CV)
[637] arXiv:2507.18740 (cross-list from eess.IV) [pdf, html, other]
Title: Learned Single-Pixel Fluorescence Microscopy
Serban C. Tudosie, Valerio Gandolfi, Shivaprasad Varakkoth, Andrea Farina, Cosimo D'Andrea, Simon Arridge
Comments: 10 pages, 6 figures, 1 table
Subjects: Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG); Optics (physics.optics)
[638] arXiv:2507.18681 (cross-list from cs.LG) [pdf, other]
Title: Concept Probing: Where to Find Human-Defined Concepts (Extended Version)
Manuel de Sousa Ribeiro, Afonso Leote, João Leite
Comments: Extended version of the paper published in Proceedings of the International Conference on Neurosymbolic Learning and Reasoning (NeSy 2025)
Subjects: Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV); Neural and Evolutionary Computing (cs.NE)
[639] arXiv:2507.18664 (cross-list from cs.GR) [pdf, html, other]
Title: Generating real-time detailed ground visualisations from sparse aerial point clouds
Aidan Murray, Eddie Waite, Caleb Ross, Scarlet Mitchell, Alexander Bradley, Joanna Jamrozy, Kenny Mitchell
Comments: CVMP Short Paper. 1 page, 3 figures, CVMP 2022: The 19th ACM SIGGRAPH European Conference on Visual Media Production, London. This work was supported by the European Union's Horizon 2020 research and innovation programme under Grant 101017779
Subjects: Graphics (cs.GR); Computer Vision and Pattern Recognition (cs.CV)
[640] arXiv:2507.18654 (cross-list from cs.LG) [pdf, html, other]
Title: Diffusion Models for Solving Inverse Problems via Posterior Sampling with Piecewise Guidance
Saeed Mohseni-Sehdeh, Walid Saad, Kei Sakaguchi, Tao Yu
Subjects: Machine Learning (cs.LG); Computer Vision and Pattern Recognition (cs.CV)
[641] arXiv:2507.18647 (cross-list from eess.IV) [pdf, html, other]
Title: XAI-Guided Analysis of Residual Networks for Interpretable Pneumonia Detection in Paediatric Chest X-rays
Rayyan Ridwan
Comments: 13 pages, 14 figures
Subjects: Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV)
[642] arXiv:2507.18640 (cross-list from cs.HC) [pdf, html, other]
Title: How good are humans at detecting AI-generated images? Learnings from an experiment
Thomas Roca, Anthony Cintron Roman, Jehú Torres Vega, Marcelo Duarte, Pengce Wang, Kevin White, Amit Misra, Juan Lavista Ferres
Subjects: Human-Computer Interaction (cs.HC); Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV)
Total of 642 entries : 1-50 ... 451-500 501-550 551-600 601-642
Showing up to 50 entries per page: fewer | more | all
  • About
  • Help
  • Click here to contact arXiv Contact
  • Click here to subscribe Subscribe
  • Copyright
  • Privacy Policy
  • Web Accessibility Assistance
  • arXiv Operational Status
    Get status notifications via email or slack