Computer Vision and Pattern Recognition

Authors and titles for recent submissions

See today's new changes

Total of 642 entries : 1-50 ... 451-500 501-550 551-600 601-642

Showing up to 50 entries per page: fewer | more | all

[601] arXiv:2507.18713 [pdf, html, other]: Title: SaLF: Sparse Local Fields for Multi-Sensor Rendering in Real-Time

Yun Chen, Matthew Haines, Jingkang Wang, Krzysztof Baron-Lis, Sivabalan Manivasagam, Ze Yang, Raquel Urtasun

Subjects: Computer Vision and Pattern Recognition (cs.CV); Robotics (cs.RO)
[602] arXiv:2507.18678 [pdf, html, other]: Title: Towards Scalable Spatial Intelligence via 2D-to-3D Data Lifting

Xingyu Miao, Haoran Duan, Quanhao Qian, Jiuniu Wang, Yang Long, Ling Shao, Deli Zhao, Ran Xu, Gongjie Zhang

Comments: ICCV 2025 (Highlight)

Subjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)
[603] arXiv:2507.18677 [pdf, html, other]: Title: HeartUnloadNet: A Weakly-Supervised Cycle-Consistent Graph Network for Predicting Unloaded Cardiac Geometry from Diastolic States

Siyu Mu, Wei Xuan Chan, Choon Hwai Yap

Comments: Codes are available at this https URL

Subjects: Computer Vision and Pattern Recognition (cs.CV); Medical Physics (physics.med-ph)
[604] arXiv:2507.18675 [pdf, other]: Title: Advancing Vision-based Human Action Recognition: Exploring Vision-Language CLIP Model for Generalisation in Domain-Independent Tasks

Utkarsh Shandilya, Marsha Mariya Kappan, Sanyam Jain, Vijeta Sharma

Subjects: Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG)
[605] arXiv:2507.18667 [pdf, html, other]: Title: Gen-AI Police Sketches with Stable Diffusion

Nicholas Fidalgo, Aaron Contreras, Katherine Harvey, Johnny Ni

Subjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)
[606] arXiv:2507.18661 [pdf, other]: Title: Eyes Will Shut: A Vision-Based Next GPS Location Prediction Model by Reinforcement Learning from Visual Map Feed Back

Ruixing Zhang, Yang Zhang, Tongyu Zhu, Leilei Sun, Weifeng Lv

Subjects: Computer Vision and Pattern Recognition (cs.CV); Robotics (cs.RO)
[607] arXiv:2507.18660 [pdf, html, other]: Title: Fuzzy Theory in Computer Vision: A Review

Adilet Yerkin, Ayan Igali, Elnara Kadyrgali, Maksat Shagyrov, Malika Ziyada, Muragul Muratbekova, Pakizar Shamoi

Comments: Submitted to Journal of Intelligent and Fuzzy Systems for consideration (8 pages, 6 figures, 1 table)

Subjects: Computer Vision and Pattern Recognition (cs.CV)
[608] arXiv:2507.18657 [pdf, other]: Title: VGS-ATD: Robust Distributed Learning for Multi-Label Medical Image Classification Under Heterogeneous and Imbalanced Conditions

Zehui Zhao, Laith Alzubaidi, Haider A.Alwzwazy, Jinglan Zhang, Yuantong Gu

Comments: The idea is still underdeveloped, not yet enough to be published

Subjects: Computer Vision and Pattern Recognition (cs.CV); Cryptography and Security (cs.CR)
[609] arXiv:2507.18656 [pdf, html, other]: Title: ShrinkBox: Backdoor Attack on Object Detection to Disrupt Collision Avoidance in Machine Learning-based Advanced Driver Assistance Systems

Muhammad Zaeem Shahzad, Muhammad Abdullah Hanif, Bassem Ouni, Muhammad Shafique

Comments: 8 pages, 8 figures, 1 table

Subjects: Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG)
[610] arXiv:2507.18655 [pdf, html, other]: Title: Part Segmentation of Human Meshes via Multi-View Human Parsing

James Dickens, Kamyar Hamad

Subjects: Computer Vision and Pattern Recognition (cs.CV); Image and Video Processing (eess.IV)
[611] arXiv:2507.18653 [pdf, html, other]: Title: Adapt, But Don't Forget: Fine-Tuning and Contrastive Routing for Lane Detection under Distribution Shift

Mohammed Abdul Hafeez Khan, Parth Ganeriwala, Sarah M. Lehman, Siddhartha Bhattacharyya, Amy Alvarez, Natasha Neogi

Comments: Accepted to ICCV 2025, 2COOOL Workshop. Total 14 pages, 5 tables, and 4 figures

Subjects: Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG)
[612] arXiv:2507.18650 [pdf, other]: Title: Features extraction for image identification using computer vision

Venant Niyonkuru, Sylla Sekou, Jimmy Jackson Sinzinkayo

Journal-ref: World Journal of Advanced Research and Reviews, 2025, 27(01), 1341-1351

Subjects: Computer Vision and Pattern Recognition (cs.CV)
[613] arXiv:2507.18649 [pdf, html, other]: Title: Livatar-1: Real-Time Talking Heads Generation with Tailored Flow Matching

Haiyang Liu, Xiaolin Hong, Xuancheng Yang, Yudi Ruan, Xiang Lian, Michael Lingelbach, Hongwei Yi, Wei Li

Comments: Technical Report

Subjects: Computer Vision and Pattern Recognition (cs.CV)
[614] arXiv:2507.18645 [pdf, html, other]: Title: Quantum-Cognitive Tunnelling Neural Networks for Military-Civilian Vehicle Classification and Sentiment Analysis

Milan Maksimovic, Anna Bohdanets, Immaculate Motsi-Omoijiade, Guido Governatori, Ivan S. Maksymov

Subjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)
[615] arXiv:2507.19328 (cross-list from eess.IV) [pdf, html, other]: Title: NerT-CA: Efficient Dynamic Reconstruction from Sparse-view X-ray Coronary Angiography

Kirsten W.H. Maas, Danny Ruijters, Nicola Pezzotti, Anna Vilanova

Subjects: Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV)
[616] arXiv:2507.19284 (cross-list from cs.CG) [pdf, html, other]: Title: Relaxed Total Generalized Variation Regularized Piecewise Smooth Mumford-Shah Model for Triangulated Surface Segmentation

Huayan Zhang, Shanqiang Wang, Xiaochao Wang

Subjects: Computational Geometry (cs.CG); Computer Vision and Pattern Recognition (cs.CV)
[617] arXiv:2507.19282 (cross-list from eess.IV) [pdf, other]: Title: SAM2-Aug: Prior knowledge-based Augmentation for Target Volume Auto-Segmentation in Adaptive Radiation Therapy Using Segment Anything Model 2

Guoping Xu, Yan Dai, Hengrui Zhao, Ying Zhang, Jie Deng, Weiguo Lu, You Zhang

Comments: 26 pages, 10 figures

Subjects: Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV); Medical Physics (physics.med-ph)
[618] arXiv:2507.19230 (cross-list from eess.IV) [pdf, html, other]: Title: Unstable Prompts, Unreliable Segmentations: A Challenge for Longitudinal Lesion Analysis

Niels Rocholl, Ewoud Smit, Mathias Prokop, Alessa Hering

Subjects: Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV)
[619] arXiv:2507.19225 (cross-list from cs.SD) [pdf, html, other]: Title: Face2VoiceSync: Lightweight Face-Voice Consistency for Text-Driven Talking Face Generation

Fang Kang, Yin Cao, Haoyu Chen

Subjects: Sound (cs.SD); Computer Vision and Pattern Recognition (cs.CV); Multimedia (cs.MM); Audio and Speech Processing (eess.AS)
[620] arXiv:2507.19201 (cross-list from eess.IV) [pdf, html, other]: Title: Joint Holistic and Lesion Controllable Mammogram Synthesis via Gated Conditional Diffusion Model

Xin Li, Kaixiang Yang, Qiang Li, Zhiwei Wang

Comments: Accepted, ACM Multimedia 2025, 10 pages, 5 figures

Subjects: Image and Video Processing (eess.IV); Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV)
[621] arXiv:2507.19199 (cross-list from eess.IV) [pdf, html, other]: Title: Enhancing Diabetic Retinopathy Classification Accuracy through Dual Attention Mechanism in Deep Learning

Abdul Hannan, Zahid Mahmood, Rizwan Qureshi, Hazrat Ali

Comments: submitted to Computer Methods in Biomechanics and Biomedical Engineering: Imaging & Visualization

Subjects: Image and Video Processing (eess.IV); Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV)
[622] arXiv:2507.19197 (cross-list from cs.LG) [pdf, html, other]: Title: WACA-UNet: Weakness-Aware Channel Attention for Static IR Drop Prediction in Integrated Circuit Design

Youngmin Seo, Yunhyeong Kwon, Younghun Park, HwiRyong Kim, Seungho Eum, Jinha Kim, Taigon Song, Juho Kim, Unsang Park

Comments: 9 pages, 5 figures

Subjects: Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV)
[623] arXiv:2507.19186 (cross-list from eess.IV) [pdf, other]: Title: Reconstruct or Generate: Exploring the Spectrum of Generative Modeling for Cardiac MRI

Niklas Bubeck, Yundi Zhang, Suprosanna Shit, Daniel Rueckert, Jiazhen Pan

Subjects: Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV)
[624] arXiv:2507.19172 (cross-list from cs.AI) [pdf, html, other]: Title: PhysDrive: A Multimodal Remote Physiological Measurement Dataset for In-vehicle Driver Monitoring

Jiyao Wang, Xiao Yang, Qingyong Hu, Jiankai Tang, Can Liu, Dengbo He, Yuntao Wang, Yingcong Chen, Kaishun Wu

Comments: It is the initial version, not the final version

Subjects: Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV)
[625] arXiv:2507.19165 (cross-list from eess.IV) [pdf, html, other]: Title: Extreme Cardiac MRI Analysis under Respiratory Motion: Results of the CMRxMotion Challenge

Kang Wang, Chen Qin, Zhang Shi, Haoran Wang, Xiwen Zhang, Chen Chen, Cheng Ouyang, Chengliang Dai, Yuanhan Mo, Chenchen Dai, Xutong Kuang, Ruizhe Li, Xin Chen, Xiuzheng Yue, Song Tian, Alejandro Mora-Rubio, Kumaradevan Punithakumar, Shizhan Gong, Qi Dou, Sina Amirrajab, Yasmina Al Khalil, Cian M. Scannell, Lexiaozi Fan, Huili Yang, Xiaowu Sun, Rob van der Geest, Tewodros Weldebirhan Arega, Fabrice Meriaudeau, Caner Özer, Amin Ranem, John Kalkhof, İlkay Öksüz, Anirban Mukhopadhyay, Abdul Qayyum, Moona Mazher, Steven A Niederer, Carles Garcia-Cabrera, Eric Arazo, Michal K. Grzeszczyk, Szymon Płotka, Wanqin Ma, Xiaomeng Li, Rongjun Ge, Yongqing Kou, Xinrong Chen, He Wang, Chengyan Wang, Wenjia Bai, Shuo Wang

Subjects: Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV)
[626] arXiv:2507.19138 (cross-list from eess.IV) [pdf, html, other]: Title: RealisVSR: Detail-enhanced Diffusion for Real-World 4K Video Super-Resolution

Weisong Zhao, Jingkai Zhou, Xiangyu Zhu, Weihua Chen, Xiao-Yu Zhang, Zhen Lei, Fan Wang

Subjects: Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV)
[627] arXiv:2507.19132 (cross-list from cs.AI) [pdf, html, other]: Title: OS-MAP: How Far Can Computer-Using Agents Go in Breadth and Depth?

Xuetian Chen, Yinghao Chen, Xinfeng Yuan, Zhuo Peng, Lu Chen, Yuekeng Li, Zhoujia Zhang, Yingqian Huang, Leyan Huang, Jiaqing Liang, Tianbao Xie, Zhiyong Wu, Qiushi Sun, Biqing Qi, Bowen Zhou

Comments: Work in progress

Subjects: Artificial Intelligence (cs.AI); Computation and Language (cs.CL); Computer Vision and Pattern Recognition (cs.CV); Human-Computer Interaction (cs.HC)
[628] arXiv:2507.19125 (cross-list from eess.IV) [pdf, html, other]: Title: Learned Image Compression with Hierarchical Progressive Context Modeling

Yuqi Li, Haotian Zhang, Li Li, Dong Liu

Comments: 17 pages, ICCV 2025

Subjects: Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV); Multimedia (cs.MM)
[629] arXiv:2507.19089 (cross-list from cs.AI) [pdf, html, other]: Title: Fine-Grained Traffic Inference from Road to Lane via Spatio-Temporal Graph Node Generation

Shuhao Li, Weidong Yang, Yue Cui, Xiaoxing Liu, Lingkai Meng, Lipeng Ma, Fan Zhang

Subjects: Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV)
[630] arXiv:2507.19074 (cross-list from eess.IV) [pdf, other]: Title: A Self-training Framework for Semi-supervised Pulmonary Vessel Segmentation and Its Application in COPD

Shuiqing Zhao, Meihuan Wang, Jiaxuan Xu, Jie Feng, Wei Qian, Rongchang Chen, Zhenyu Liang, Shouliang Qi, Yanan Wu

Subjects: Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV)
[631] arXiv:2507.19041 (cross-list from quant-ph) [pdf, html, other]: Title: PGKET: A Photonic Gaussian Kernel Enhanced Transformer

Ren-Xin Zhao

Subjects: Quantum Physics (quant-ph); Computer Vision and Pattern Recognition (cs.CV)
[632] arXiv:2507.19035 (cross-list from eess.IV) [pdf, html, other]: Title: Dual Path Learning -- learning from noise and context for medical image denoising

Jitindra Fartiyal, Pedro Freire, Yasmeen Whayeb, James S. Wolffsohn, Sergei K. Turitsyn, Sergei G. Sokolov

Comments: 10 pages, 7 figures

Subjects: Image and Video Processing (eess.IV); Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV)
[633] arXiv:2507.18915 (cross-list from cs.CL) [pdf, html, other]: Title: Mining Contextualized Visual Associations from Images for Creativity Understanding

Ananya Sahu, Amith Ananthram, Kathleen McKeown

Subjects: Computation and Language (cs.CL); Computer Vision and Pattern Recognition (cs.CV)
[634] arXiv:2507.18895 (cross-list from eess.IV) [pdf, html, other]: Title: Dealing with Segmentation Errors in Needle Reconstruction for MRI-Guided Brachytherapy

Vangelis Kostoulas, Arthur Guijt, Ellen M. Kerkhof, Bradley R. Pieters, Peter A.N. Bosman, Tanja Alderliesten

Comments: Published in: Proc. SPIE Medical Imaging 2025, Vol. 13408, 1340826

Journal-ref: Proc. SPIE Medical Imaging 2025: Image-Guided Procedures, Robotic Interventions, and Modeling, Vol. 13408, 1340826 (2025)

Subjects: Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV)
[635] arXiv:2507.18830 (cross-list from eess.IV) [pdf, html, other]: Title: RealDeal: Enhancing Realism and Details in Brain Image Generation via Image-to-Image Diffusion Models

Shen Zhu, Yinzhu Jin, Tyler Spears, Ifrah Zawar, P. Thomas Fletcher

Comments: 19 pages, 10 figures

Subjects: Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG)
[636] arXiv:2507.18808 (cross-list from cs.RO) [pdf, html, other]: Title: Perpetua: Multi-Hypothesis Persistence Modeling for Semi-Static Environments

Miguel Saavedra-Ruiz, Samer B. Nashed, Charlie Gauthier, Liam Paull

Comments: Accepted to the IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS 2025) Code available at this https URL. Webpage and additional videos at this https URL

Subjects: Robotics (cs.RO); Computer Vision and Pattern Recognition (cs.CV)
[637] arXiv:2507.18740 (cross-list from eess.IV) [pdf, html, other]: Title: Learned Single-Pixel Fluorescence Microscopy

Serban C. Tudosie, Valerio Gandolfi, Shivaprasad Varakkoth, Andrea Farina, Cosimo D'Andrea, Simon Arridge

Comments: 10 pages, 6 figures, 1 table

Subjects: Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG); Optics (physics.optics)
[638] arXiv:2507.18681 (cross-list from cs.LG) [pdf, other]: Title: Concept Probing: Where to Find Human-Defined Concepts (Extended Version)

Manuel de Sousa Ribeiro, Afonso Leote, João Leite

Comments: Extended version of the paper published in Proceedings of the International Conference on Neurosymbolic Learning and Reasoning (NeSy 2025)

Subjects: Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV); Neural and Evolutionary Computing (cs.NE)
[639] arXiv:2507.18664 (cross-list from cs.GR) [pdf, html, other]: Title: Generating real-time detailed ground visualisations from sparse aerial point clouds

Aidan Murray, Eddie Waite, Caleb Ross, Scarlet Mitchell, Alexander Bradley, Joanna Jamrozy, Kenny Mitchell

Comments: CVMP Short Paper. 1 page, 3 figures, CVMP 2022: The 19th ACM SIGGRAPH European Conference on Visual Media Production, London. This work was supported by the European Union's Horizon 2020 research and innovation programme under Grant 101017779

Subjects: Graphics (cs.GR); Computer Vision and Pattern Recognition (cs.CV)
[640] arXiv:2507.18654 (cross-list from cs.LG) [pdf, html, other]: Title: Diffusion Models for Solving Inverse Problems via Posterior Sampling with Piecewise Guidance

Saeed Mohseni-Sehdeh, Walid Saad, Kei Sakaguchi, Tao Yu

Subjects: Machine Learning (cs.LG); Computer Vision and Pattern Recognition (cs.CV)
[641] arXiv:2507.18647 (cross-list from eess.IV) [pdf, html, other]: Title: XAI-Guided Analysis of Residual Networks for Interpretable Pneumonia Detection in Paediatric Chest X-rays

Rayyan Ridwan

Comments: 13 pages, 14 figures

Subjects: Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV)
[642] arXiv:2507.18640 (cross-list from cs.HC) [pdf, html, other]: Title: How good are humans at detecting AI-generated images? Learnings from an experiment

Thomas Roca, Anthony Cintron Roman, Jehú Torres Vega, Marcelo Duarte, Pengce Wang, Kevin White, Amit Misra, Juan Lavista Ferres

Subjects: Human-Computer Interaction (cs.HC); Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV)

Total of 642 entries : 1-50 ... 451-500 501-550 551-600 601-642

Showing up to 50 entries per page: fewer | more | all

Computer Vision and Pattern Recognition

Authors and titles for recent submissions

Mon, 28 Jul 2025 (continued, showing last 42 of 118 entries )