


default search action
15th ICVGIP 2025: Mandi Himachal Pradesh, India
- Ragini Verma, Anand Mishra, Arnav Bhavsar:

Proceedings of the Sixteen Indian Conference on Computer Vision, Graphics and Image Processing, ICVGIP 2025, Mandi Himachal Pradesh, India, December 17-20, 2025. ACM 2025, ISBN 979-8-4007-1930-1 - Suresh Nehra

, Aupendu Kar, Prabir Kumar Biswas, Jayanta Mukhopadhyay:
Light-Field Dataset for Disparity Based Depth Estimation. 1:1-1:10 - Samarth Garg, Debanjan Sadhya, Sunil Kumar:

ReviewGuard: Context-Aware Detection of AI-Generated Text in Academic Peer Reviews. 2:1-2:8 - Aditi Palit, Pranay Varna Chinthapatla

, Kalidas Yeturu
:
Mask What Matters: Action-Specific Privacy via Object-Layer Transformation. 3:1-3:9 - Debapriya Roy

, Angshul Majumdar
:
Smoothness-Preserving Deep Semi Non-negative Matrix Factorization for Clustering. 4:1-4:9 - Ankita Raj

, Kaashika Prajaapat, Tapan Kumar Gandhi
, Chetan Arora:
Mimicking Human Visual Development for Learning Robust Image Representations. 5:1-5:10 - Mrinmoy Sen, Akshay Bankar, Ravikiran Kundapur Subraya, Roopa Sheshadri, Kyuwon Kim:

Multi-modal Image Colorization with Instance-Aware Transformer network. 6:1-6:9 - Dillip Sathiyamoorthy

, Prithwijit Guha
:
VAD-FedHSM: Video Anomaly Detection in Federated Learning Framework with Local Hard Sample Mining. 7:1-7:10 - Abhay Kumar Dwivedi, Shanu Saklani, Soumya Dutta

:
Compressive Modeling and Visualization of Multivariate Scientific Data using Implicit Neural Representation. 8:1-8:9 - Gyanendra Chaubey

, Azad Singh
, Aiman Farooq
, Deepak Mishra
:
RibCageImp: A Deep Learning Framework for 3D Ribcage Implant Generation. 9:1-9:9 - Ketan P. Detroja, Soham Chatterjee:

Lightweight and Accurate Deepfake Detection using Gradient-Aware SE-Enhanced EfficientNetV2-S. 10:1-10:8 - Akash Manna

, Radha Krishna Deshpande
, Ajoy Mondal, C. V. Jawahar
:
Label-Free Adaptation of Indic Printed OCR. 11:1-11:9 - Anup Kushwaha

, Banseedhar Gondaliya, Chandramouli Sanchi, Surya Kumar, Apil Thapa, Edula Sri Ranga Srihit Reddy:
Screen2UI: Pixel-Only On-Device UI Hierarchy Generation. 12:1-12:9 - Siva Manohar Reddy Kesu, Neelam Sinha

, Hariharan Ramasangu
, Thomas Gregor Issac:
Alzheimer's Disease Classification using Retinal OCT: TransNetOCT and Swin Transformer Models. 13:1-13:9 - Bavneet Kaur, Gurpreet Singh Josan

:
Evaluating Illumination Effects on AU Intensity in Bone-Driven Facial Animation of 3D Avatar. 14:1-14:9 - Sunayana Samavedam

, Venkat Adithya Amula, Saurabh Saini
, Avani Gupta
, P. J. Narayanan
:
CAV Styler- Making Neural Style Transfer Interpretable and controllable. 15:1-15:9 - Pratyush Jena

, Amal Joseph, Arnav Sharma
, Ravi Kiran Sarvadevabhatla
:
Unveiling Text in Challenging Stone Inscriptions: A Character-Context-Aware Patching Strategy for Binarization. 16:1-16:9 - Prabhat Kumar

, Himanshu Singh
, Badri Narayan Subudhi
, Geatanon Di Caterina
, Vinit Jakhetiya, T. Veerakumar:
Enhancing Spiking Neural Networks with Evidential Deep Learning for Object Classification on Event Based Dataset. 17:1-17:9 - Chiranjeev Bhaya

, Mithil Shail, Chandra Mohan Velpula
, Dhrubajyoti Mandal
, Sourav Kumar Parida
, Subhadeep Bhadra
, Sujoy Datta:
A Pixel-Mapping Based Fast Image Rotation Algorithm. 18:1-18:9 - Sujoy Nath

, Arkaprabha Basu
, Sharanya Dasgupta, Swagatam Das
:
HalluShift++: Bridging Language and Vision through Internal Representation Shifts for Hierarchical Hallucinations in MLLMs. 19:1-19:9 - S. Jeba Berlin, Antara Roy Choudhury, Arita Noelle Prasad

, Sowmya Murli, Dinesh Babu Jayagopi
, Vijaya Raman
, Shyam Sundar Rajagopalan
:
Early Diagnosis of Autism through Behaviors from Semi-Structured Triadic Play Interaction Videos. 20:1-20:10 - Dara Deepthi, P. Radha Krishna

:
Classification of Brain Tumours Using Multimodal Driven Deep Learning Techniques. 21:1-21:8 - Vikas Sharma, Jagannath Prasad Sahoo

, Amit Shukla, Durgesh Ameta
, Jay Shorey, Sindhuja Reddy, Pranav Prashant Shewale, Sushant Sharma:
TongueSight: AI-Powered Tongue Analysis for Digitizing Traditional Medicine Diagnostics. 22:1-22:7 - Shehroz S. Khan

, Petar Przulj
, Ahmed Ashraf
, Ali Abedi
:
ChestGPT: Integrating Large Language Models and Vision Transformers for Disease Detection and Localization in Chest X-Rays. 23:1-23:8 - Aupendu Kar, Krishnendu Ghosh

, Prabir Kumar Biswas:
Sharing the Learned Knowledge-base to Estimate Convolutional Filter Parameters for Continual Image Restoration. 24:1-24:9 - V. S. Sukesh Babu

, Pawan Kumar Bamne
, Rahul Raman
:
CamPedV2: A Comprehensive Dataset for Advancing Pedestrian Detection Models. 25:1-25:8 - Shashank Krishna Vempati

, Gaurav Talebailkar
, Sai Prabhath Bogam
, Bhaskar Arun
, Kayalvizhi Ganesan, Chetan Arora:
From Words to Paragraphs: Hierarchical Dense Text Detection with SAM-Adaptive Backbone. 26:1-26:10 - Arghya Pratihar, Swagato Das, Swagatam Das:

Hyperbolic Fuzzy C-Means with Adaptive Weight-based Filtering for Efficient Clustering. 27:1-27:9 - Somarowthu Gani Lakshmi

, Makkapati Prasanna Lakshmi, Bandaru Ganesh Sai Manideep
, Katikela Manaswani
, Kalapatapu V. S. K. R. Shiva Kumar
:
A CBAM-Enhanced EfficientNet-DCGAN Hybrid CNN for Class-Imbalanced Skin Lesion Classification on HAM10000. 28:1-28:7 - Chirayu Agrawal, Snehasis Mukherjee

:
SCoDA: Self-supervised Continual Domain Adaptation. 29:1-29:9 - Swapnil Thatte, Promita Banerjee, Mayank Vatsa

, Richa Singh, Bhanu Duggal, Angshuman Paul:
MFA-U-Net: U-Net with Multi-level Feature Aggregation for Scale-free Medical Image Segmentation. 30:1-30:8 - Surbhi Madan

, Shreya Ghosh
, Ramanathan Subramanian
, Tom Gedeon, Abhinav Dhall
:
CSGaze: Context-aware Social Gaze Prediction. 31:1-31:9 - Shaon Bhattacharyya, Ajoy Mondal, C. V. Jawahar

:
FormLens: From Ink to Insight with Adapting Vision-Language Models for Handwritten Form Digitization. 32:1-32:9 - Advait Amit Kisar, Sumeet Shekhar, Gopi Raju Matta

, Kaushik Mitra
:
A Hybrid Dataset and Depth-Guided Transformer for Nighttime Flare Removal. 33:1-33:9 - Kriti Khare, Parimala Kancharla:

Quality Aware Text-to-Image Synthesis with distortion maps as structural guides. 34:1-34:8 - Roshin Roy, Konda Reddy Mopuri

:
Metric-based Regularization for Training GANs on Long-Tailed Datasets. 35:1-35:10 - Jagadish Kashinath Kamble, Jayanta Mukhopadhyay, Debaditya Roy

, Partha Pratim Das:
Generating Key Postures of Bharatanatyam Adavus with Pose Estimation. 36:1-36:10 - Sayan Kahali, Ajoy Dey

, Aalekhya Mukhopadhyay, Sounak Dey:
SPIKECODE: Spiking Autoencoder FrameWork for Lossy Image Compression on Constrained Low-Power Edge Devices. 37:1-37:8 - Rutvik Patel, Anureet Chhabra, Ankita Raj

, Chetan Arora:
Rethinking Detection Heads: Enhancing YOLO for Drone Image Object Detection. 38:1-38:10 - Harichandana B. S. S

, Sumit Kumar
, Deepak Chaurasia:
Adaptive Binary Wallpaper Abstraction for Power-Conscious Always-On Displays. 39:1-39:9 - Anuran Basu

, Kotha Kartheek, Lingamaneni Gnanesh Chowdary, Snehasis Mukherjee
:
A Unified model for the Adverse Weather Removal using Squeeze-and-Excitation Attention. 40:1-40:9 - Deepika Kamboj, Gaurav Harit

:
Towards Structured Multimodal Understanding: Scientific Diagram Captioning with Modern VLMs. 41:1-41:9 - Kunal Dargan, Nikhil Kumar Jangamreddy, Chetan Arora:

CP-SIS: Learnable Class Prototypes for Efficient Surgical Instrument Segmentation. 42:1-42:10 - S. Balasubramanian

, M. Sai Subramaniam
, Sai Sriram Talasu
, Yedu Krishna P
, M. Pranav Phanindra Sai
, Darshan Gera
, Ravi Mukkamala
:
Contextual Memory Recall: A Novel Metric for Class Incremental Learning. 43:1-43:8 - Rajeev Ranjan Dwivedi

, Mohammedkaif Rafiq Kalagond, Monika Sharma, Amit Sangroya:
Equitable Dermatology: Adversarial and Spectral Techniques for Fair Skin Lesion Classification. 44:1-44:9 - Debmani Saha, Kumari Rina, Deepshikha Ray

, Kaushik Mukhopadhyay
, Ayoleena Roy
, Oishila Bandyopadhyay
:
Severity Grading of Autism Spectrum Disorder from Eye Gaze Scanpath Trajectory using Deep Learning. 45:1-45:8 - Meghana Shankar, Akanxit Upadhyay, Anmol Namdev

, Green Rosh K. S, Pawan Prasad Bh
:
PHAF: Personalised Hand Avatar in a Flash. 46:1-46:9 - Darshan Gera

, Vignesh Sai Sankalp Sham
, Ritu Raj Pradhan, S. Balasubramanian
, Ravi Mukkamala
, Sreepada V. Satya N. Sarma:
Zero-Shot Autism Behaviour Recognition Leveraging Multimodal Representations and Cross-Dataset Evaluation. 47:1-47:9 - Udayraj Mangal

, Anchal Pandey:
B2R2Net: Robust and Lightweight Bounding Box Refining & Refactoring Network for Geometry-Agnostic Screenshots. 48:1-48:7 - Thomas Apotheker, Siraj T. M

, Raji Susan Mathew, Navchetan Awasthi:
Context-QSMnet: Enhancing Generalization in Quantitative Susceptibility Mapping with Context-Aware Deep Neural Networks. 49:1-49:8 - Shayon Dasgupta, Avijit Dasgupta

, C. V. Jawahar
:
Are We There Yet? Exploring the Capabilities of MLLMs in Assistive AI Applications. 50:1-50:9 - Aparna Agrawal

, Seshadri Mazumder
, C. V. Jawahar
, Vinay P. Namboodiri
:
Towards Scalable Sign Production: Leveraging Co-Articulated Gloss Dictionary for Fluid Sign Synthesis. 51:1-51:10 - Om Rajendra Kathalkar

, Nitin Nilesh, Sachin Chaudhari, Anoop Namboodiri:
AQIFormer: A Transformer-Based Multi-View Architecture for Cross-City Air Quality Classification. 52:1-52:9 - Franklin Burhagohain

, Shovan Barma
:
T2M-CycleGAN: A Knowledge-Distilled CycleGAN Framework for High-Fidelity 3T-to-7T MRI Translation. 54:1-54:8 - Kanika Arora, Nazreen Shah, Ashish Sethi

, Ranjitha Prasad
, Mithun Uliyar, Krishna A G:
Personalized Federated Autoencoders: Unsupervised Representation Learning at the Edge. 55:1-55:9 - Devansh Garg

:
Strongly Connected Components Are All You Need: Graph-Theoretic Interpretability and Optimization for Vision Transformers. 56:1-56:9 - Kajal Sanklecha

, Siddharth Mangipudi, Prayushi Mathur
, P. J. Narayanan
:
Mesh Object Retrieval using Self-Supervised Mesh Convolution. 57:1-57:10 - Gowthamaan Palani

, Ganapathy Krishnamurthi:
MDFN: Efficient Image Super-Resolution through Multi-Domain Feature Fusion. 58:1-58:9 - Seema Kumari

, Mitkumar Patel, Ankeshwar Ruthesha, Shanmuganathan Raman:
PointxLSTM: Integrating xLSTM for 3D Point Cloud Completion and Analysis. 59:1-59:10 - Suruchi Kumari, Aryan Das

, Swalpa Kumar Roy
, Indu Joshi, Pravendra Singh
:
Leveraging Task-Specific Knowledge from LLM for Semi-Supervised 3D Medical Image Segmentation. 60:1-60:10 - Namrata Das, Samit Biswas

:
Writer Identification from Multilingual Handwritten Indic Text Document Images. 61:1-61:9 - Chiranjeev Bhaya

, Mithil Shail, Chandra Mohan Velpula
, Sarbojit Ganguly, Praveen Bangre Prabhakar Rao:
A Fast Lossless Compression Algorithm for Color Filter Arrays. 62:1-62:9 - Devansh Shrestha, Varun Dutt

:
Behavioral Evaluation of Prospect-Theoretic Instance-Based Learning: Modeling Risk and Alternation in Experience-Based Decisions. 63:1-63:8 - Kashish Jain, Narayan Ji Mishra

, Vireshwar Kumar, Ashok Kumar Bhateja
:
Chaff-Free Fuzzy Vault using Multimodal Biometrics for Secure Key Access. 64:1-64:9 - Samik Some, Vinay Namboodiri

:
Can Unsupervised Segmentation Reduce Annotation Costs for Video Semantic Segmentation? 65:1-65:9 - Devika R. G.

, Sivakumar Ramachandran, Linu Shine
, Jiji C. V.:
UNet-Mamba Fusion for Automated Glaucoma Screening via Cup-to-Disc Ratio Estimation. 66:1-66:9 - Akhilaraj D, Joseph Zacharias:

IDANet: Identity-Driven Attention Network for Unsupervised Synthetic Noise Removal. 67:1-67:9 - Sumesh Prasad

, Ojaswini Sharma, Saket Anand:
Lightweight Models for Classification and Detection via Knowledge Distillation and Distance Correlation. 68:1-68:9 - Yash Arora

, Aditya Arun
, C. V. Jawahar
:
What is there in an Indian Thali? 69:1-69:9 - Aditi Palit, Kalidas Yeturu

:
Pooling Diverse Voices: Logarithmic Binary Fusion for Server-Side Pseudo-Labeling in Federated Learning. 70:1-70:7 - Farzana S

, C. V. Rishi, Shubham Goel
, Aditya Arun
, C. V. Jawahar
:
How Does India Cook Biryani? 71:1-71:10 - Jayasree Saha

, vinay P. Namboodiri
, C. V. Jawahar
:
Low SNR Speech Perception with HuBERT: A Discussion on Visual-Audio Fusion and Domain-Specific Modeling. 72:1-72:9 - Subramanyam Sahoo, Rajesh Kesavalalji

, Joel Jojo
, Sonal Singh
, Bharat Kumar Bolla
:
The Green Mile: A Multi-Layered Quest to Reveal, Measure, and Slash Carbon Emissions in AI Training & Inference. 73:1-73:8 - Hanvitha Saraswathi Mukkamala, Shankar Gangisetty

, Ananya Kulkarni
, Veera Ganesh Yalla
, C. V. Jawahar
:
Enhancing Driving Visibility via Semantic-Guided Knowledge Distillation Framework for Adverse Weather Removal. 74:1-74:9 - Sanjay Bhargav Dharavath

, Hanvitha Saraswathi Mukkamala:
Quantformers: Bits to Qubits. 75:1-75:8 - Arkapal Panda, Aditya Shankar Pal, Utpal Garain

:
Theoretical Perspective on Histogram Binning with Extension to Multi-label Medical Image Classification. 76:1-76:8 - Dhivya S. D, P. L. Chithra:

Spatial - Aware Hierarchical Graph Networks for Volumetric Medical Image Segmentation. 77:1-77:9 - Raghav Mittal

, Lokender Tiwari
, Shivam Ashok Shukla
, Mritunjoy Halder
, Brojeshwar Bhowmick:
GeMR : Multi-modal Interactive 3D Scene Composition In XR. 78:1-78:10 - Vasudha Joshi

, Adrish Adhikari, Pabitra Mitra
, Supratik Bose
:
Medical Semantic-Aware Image-Text Alignment loss for Medical Visual Question Answering. 79:1-79:9 - Evani Lalitha, Ajoy Mondal, C. V. Jawahar

:
SemiHastakshar: Generalizable Indic Handwritten OCR through Semi-Supervised Learning. 80:1-80:9 - Harinandan Shukla

, Prankur Shukla, Umarani Jayaraman
:
Self-Supervised Learning for Annotation-Efficient Endoscopic Instrument Segmentation. 81:1-81:6 - Himanshu Patil, Geo Jolly, Ramana Raja Buddala

, Ganesh Ramakrishnan:
Improving Video Question Answering through query-based frame selection. 82:1-82:8 - Padmasri P, B. Sathya Bama

, Mohamed Mansoor Roomi Sindha
:
Milk or mimic? MilkNet: Spectral-Spatial Fusion Net for Brand Authentication in Commercial Milk Products Using Hyperspectral Imaging. 83:1-83:9 - Anju J. S, Pradeep R, Linu Shine

, Sreeni K. G:
HGRLiteNet: A Lightweight Hand Gesture Recognition Network. 84:1-84:6 - PushapDeep Singh, Jyoti Nigam, Medicherla Vamsi Krishna, Arnav Bhavsar

, Aditya Nigam:
On the Transferability of LaBraM: Evaluating Representations across Visual and Clinical Domains with Convolutional Adapters. 85:1-85:7 - Susant Kumar Panigrahi

, Subhoshri Pal, Pradipta Sasmal, Debdoot Sheet
:
TransUNet-Recon: A Transformer-Augmented UNet Architecture for Accelerated MRI Reconstruction. 86:1-86:8 - Raj Bahadur Singh

, Aloke Datta:
RMA-Net: Residual Multiscale Attention CNN for Hyperspectral Brain Tissue Classification. 87:1-87:9 - Supreet Kaur, Ajay Parakh, Vishakha Pareek, Rajendra Nagar, Santanu Chaudhury:

Knowledge-Based Metaverse for Crafts. 88:1-88:8 - Madan Sharma

, Aditya Singh, Ashank Kunwar
, Nirbhay Kumar Tagore, Sachin Kumar
:
Retrieval Augmented Continuous Person Tracking and Re-Identification. 89:1-89:11 - Priya Kannapiran, Mohamed Mansoor Roomi Sindha

, Mathu sudhanan Sreenivasan
, Shaan Sindha M, Uma Maheswari Pandyan:
A Lightweight Dual-Stream Framework for Vision-Based Musth Detection in Elephants. 90:1-90:9 - Moumita Dholey

, Rosina Ahmed, Sanjoy Chatterjee
, Jayanta Mukhopadhyay:
DWT-Mamba: A Spatially-Aware State Space Model for Medical Image Classification. 91:1-91:9 - Sayak Dutta, Harish Katti

, Shashikant Verma
, Shanmuganathan Raman:
UnCageNet: Tracking and Pose Estimation of Caged Animals. 92:1-92:9 - Sukanya Das

, Debasis Samanta
, Monalisa Sarma:
Statistically Validated Hybrid Fusion and Ensemble Feature Selection for Static Sign Language Recognition. 93:1-93:11 - Parshiv Kapoor

, Yash Bansal, Deychen Myes
, Sherrin Jacob
, Venkateswaran K. Iyer
, Lavleen Singh, Aparajita Khan, Partha Pratim Roy
:
Weighted Color‑Morphology Feature Fusion for Tuberculosis Bacilli Detection from Cytopathology Images. 94:1-94:9 - Mirothali Chand

, Varun Dutt
, Kv Uday:
A Deep Learning Integrated Stacked Ensemble Framework for Futuristic Landslide Prediction using Multimodal Sensor Data. 95:1-95:9 - Prasenjit Betal

, Pabitra Mitra
, Arindam Dasgupta, Chittaranjan Mandal, Prateek Gupta
:
Topological Analysis of Plant Structure Using Persistent Homology for Genotype Classification. 96:1-96:8 - Jyoti Nigam, Utsav Jain

, Piyush Kumar, Aryan Raj, Gopesh Sharma, Aaditya Singh, Asmit Kumar, Aditya Nigam, Arnav Bhavsar
:
EEG-Guided Image Synthesis via Hybrid Embedding and Diffusion Framework. 97:1-97:8 - Priyobrata Mondal, Soumi Pal

, Swagatam Das
:
TranSMOTE: Transformer Based Synthetic Minority Oversampling Technique. 98:1-98:11 - Ayushman Datta

, Priyobrata Mondal, Swagatam Das
:
Margin Adaptation: Enhanced Label-Distribution-Aware Margins with Strategic Reweighting and Representation Optimization. 99:1-99:12 - Navdha Bhardwaj, Yashit Verma, Arnav Bhavsar

:
A New Diverse Dataset for rPPG Estimation, and Benchmarking with Standard Frameworks. 100:1-100:7

manage site settings
To protect your privacy, all features that rely on external API calls from your browser are turned off by default. You need to opt-in for them to become active. All settings here will be stored as cookies with your web browser. For more information see our F.A.Q.


Google
Google Scholar
Semantic Scholar
Internet Archive Scholar
CiteSeerX
ORCID













