Latest Face Recognition Research Papers
The newest Face Recognition papers from across the field — arXiv, NeurIPS, CVPR, Nature, and more — refreshed daily and ranked by relevance. Distill AI tracks Face Recognition so you don’t have to: get the standout work delivered to your inbox every morning, with 2-sentence summaries and the option to chat with any paper.
Get the latest Face Recognition papers in your inbox — free →Recent papers
- A Lightweight Browser-Based Facial Recognition Framework for Preventing Impersonation in Computer-Based ExaminationsMuyideen Abdulraheem · Zenodo (CERN European Organ... · Nov 11, 2026
The rise of Computer-Based Testing (CBT) in academic institutions has streamlined assessment processes but introduced challenges in preventing impersonation, threatening exam integrity. This study proposes a real-time facial recognition sys…
- A Lightweight Browser-Based Facial Recognition Framework for Preventing Impersonation in Computer-Based ExaminationsMuyideen Abdulraheem · Zenodo (CERN European Organ... · Nov 11, 2026
The rise of Computer-Based Testing (CBT) in academic institutions has streamlined assessment processes but introduced challenges in preventing impersonation, threatening exam integrity. This study proposes a real-time facial recognition sys…
- SenseNova-U1.5: Towards Native Unified Visual IntelligenceHaiwen Diao, Jiahao Wang, Chenjing Ding, Hanming Deng et al. · arXiv · Sep 10, 2026
We launch SenseNova-U1.5, an 8B-MoT native unified multimodal model that understands, reasons about, and generates visual content within an encoder-free and VAE-free architecture. We strengthen its visual interface through spatially coheren…
- Guided Super-Resolution of Digital Elevation Models with Diffusion-Based Image GeneratorsArmand Mihai Nicolicioiu, Dominik Narnhofer, Nando Metzger, Daniel Panangian et al. · arXiv · Sep 10, 2026
High-resolution digital surface models (DSMs) play an important role in urban analysis, 3D building reconstruction, and infrastructure monitoring, yet their availability remains limited due to the high cost and complexity of data acquisitio…
- Language-Augmented Semantic Priors for B-Spline Surface FittingYunzhong Lou, Yusheng Luo, Jiahao Li, Yu Song et al. · arXiv · Sep 10, 2026
The use of B-splines and Non-Uniform Rational B-Splines surfaces constitutes the mathematical foundation of contemporary computer-aided design (CAD) systems. Despite long-term progress, geometric kernels in traditional CAD still rely heavil…
- Single-Stream Multi-Feature Fusion with Temporal Robustness for Gait Emotion RecognitionShirong Lyu, Silu Quan, Yixuan Ding, Chengpeng Wang · arXiv · Sep 10, 2026
3D skeleton-based gait emotion recognition faces high annotation costs, data scarcity, and poor generalization on heterogeneous data. This paper proposes SV-GCN, a single-stream multi-feature fusion framework with temporal invariance. We in…
- Learn the Solid, Not the File: Canonical Inputs for Neural Networks on CAD Boundary RepresentationsHeinrich Jiang, Hager Yasser Mohamed, Alexander Hitt, Valeriia Lomakina et al. · arXiv · Sep 10, 2026
Boundary representation (B-rep) is the standard format used by modern CAD systems for parametric 3D models. It turns out, the exact same solid can be represented by different B-reps: for example, two engineers using different operations, a …
- Harnessing Intrinsic Subject-Aware Attention for Controllable Multi-Subject Video GenerationNiange Yu, Ye Tian, Biaolong Chen, Miao Lu et al. · arXiv · Sep 10, 2026
Multi-subject video generation faces two key challenges: uncontrollable fidelity strength and potential semantic drift. We address these by analyzing the internal mechanisms of Diffusion Transformers (DiTs). We found that certain attention …
- FreeFlow: A Bias-free Hierarchical Transformer for Optical Flow EstimationVladislav Bargatin, Alexander Yakovenko, Khaled Abud, Dmitriy Vatolin · arXiv · Sep 10, 2026
Optical flow methods typically rely on task-specific inductive biases, such as correlation volumes, feature warping, and iterative refinement, among others, to reach high accuracy. While effective, such biases constrain the model to predefi…
- Show-Harness: Just a VLM Agent Can Play RobotsYanzhe Chen, Zechen Bai, Zhijun Cao, Wenzheng Zeng et al. · arXiv · Sep 9, 2026
Foundation vision-language models (VLMs) exhibit broad intelligence about the world, yet translating this intelligence into robot control remains challenging. We present Show-Harness, an Embodied Harness that enables VLMs to "play" robots t…
- SynThermFace: Amplifying Limited Paired Data for Visible-Thermal Face Recognition via Synthetic Data GenerationAnjith George, Adam Unal, Sebastien Marcel · arXiv · Sep 9, 2026
Face recognition (FR) is a widely used modality for biometric authentication, but conventional models rely on visible-spectrum imagery and degrade when high-quality RGB images cannot be captured. Cross-spectral face recognition addresses th…
- CoSA: Correlation-Guided Change A ttention with Learnable Residual Gating for Remote Sensing Change DetectionAbdirashid Omar, Jonghyuk Park · arXiv · Sep 8, 2026
Pixel-level annotation of fixed traffic-camera imagery is expensive, while crosswalk models trained from street-level imagery face a substantial viewpoint and appearance shift when applied to elevated CCTV. We investigate a data-efficient t…
- FRAME: Factored Retrieval via Attribute Readouts for Object-Centric Scene MemoryWoosang Jeon, Sanghyeok Choi, Minwoo Kim, Taehyun Jung et al. · arXiv · Sep 8, 2026
Language-guided robots need persistent scene memories to follow instructions, revisit objects, and resolve references to objects encountered over time. While much of language-guided scene-memory retrieval has emphasized spatial or relationa…
- SeGDeP: Semantic- and Geometric-Aware Decoupled Prompts for Reasoning SegmentationLinnan Zhao, Xu Liu, Lingling Li, Licheng Jiao et al. · arXiv · Sep 8, 2026
Reasoning segmentation converts an implicit linguistic conclusion into a precise mask, requiring both semantic identification and spatial grounding. Existing MLLM-segmenter interfaces either use a special trigger or compress both signals in…
- DifFoundMAD: Foundation Models meet Differential Morphing Attack DetectionLázaro J. González‐Soler, A. Dörsch, Christian Rathgeb, Christoph Busch · Zenodo (CERN European Organ... · Sep 4, 2026
In this work, we introduce DifFoundMAD, a parameter-efficient D-MAD framework that exploits the generalisation capabilities of vision foundation models (FM) to capture discrepancies between suspected morphs and live capture images. In contr…
- DifFoundMAD: Foundation Models meet Differential Morphing Attack DetectionLázaro J. González‐Soler, A. Dörsch, Christian Rathgeb, Christoph Busch · Zenodo (CERN European Organ... · Sep 4, 2026
In this work, we introduce DifFoundMAD, a parameter-efficient D-MAD framework that exploits the generalisation capabilities of vision foundation models (FM) to capture discrepancies between suspected morphs and live capture images. In contr…
- Catalogue Photography as a Cold Start: Toward Deployable Carbide Burr RecognitionAbilash Philip Madavath, Chandra Yuvesh Aubeeluck, Augustin Raju, Nicolas Pyschny et al. · arXiv · Sep 3, 2026
Verifying that manufactured batches of milling tools or carbide rotary burrs conform to production order sheets remains a largely manual and error-prone quality assurance task. Automating this process with computer vision faces a critical c…
- Thinking in Pictures: A Systematic Benchmark for Reasoning-driven Image GenerationYutong Liu, Nan Huang, Xu Cao, James M. Rehg · arXiv · Sep 2, 2026
Recent advancements in unified generative models (UGMs) and world simulators have achieved unprecedented results in visual perception and synthesis. However, these models primarily rely on surface-level event alignment, leaving the capacity…
- Video-Based Palm-Vein Authentication under Challenging ConditionsXiaofeng Yan, Kechen Liu, Abhilash Venkatesh, Cathy Zhang et al. · arXiv · Sep 2, 2026
Palm-vein biometrics are increasingly used for secure, contactless authentication. Yet real-world deployment exposes them to surface noise (sweat, dirt), illumination and motion variation, and temperature-driven changes in vascular visibili…
- Multi-Tool Image Editing Attribution in Facial ForgerySheng Liu, Qiang Sheng, Danding Wang, Yu Li et al. · arXiv · Sep 2, 2026
As generative AI tools become increasingly powerful and easy to use, people can easily edit portrait images with a prompt, necessitating the task of image editing attribution, which predicts the involved editing tools from the given image. …
- H3-World: Turning Language Understanding into World ControlDanze Chen, Zeqing Wang, Ziyue Lin, Xingyi Yang et al. · arXiv · Sep 1, 2026
We present H3-World, an efficient framework that turns the 33B MiniMax-H3 video generator into an interactive world model. Our key finding is that, as large video generators become more capable, language is emerging as a natural interface f…
- A Sensor-Adaptive Incremental Learning Framework for Artifact Detection in Satellite Precipitation DataAndres F. Monsalve, Hernan A. Moreno, Christian D. Kummerow · arXiv · Sep 1, 2026
Historically, retrieving rainfall data from satellite imagery has been the domain of space agencies. However, in recent years, the development of cheaper, more compact satellites (SmallSats) capable of detecting rainfall proxies has led to …
- Benchmarking Spatial, Spectral, and Self-Supervised Cues for Face Forgery Detection under Realistic DegradationLucas Cunha, Lucas Sotomaior, Lucas Gasperin, Beatriz Caldas et al. · arXiv · Sep 1, 2026
Face forgery detectors often achieve strong results on controlled benchmarks, but their reliability under realistic image degradations remains limited. This paper presents a standardized benchmark for face forgery detection using the Multi-…
- CameraEditor: Camera-Controlled Image Editing via Video-Prior Sequential ModelingXin Shen, Chengyou Jia, Keshuo Xing, Zifeng Zhu et al. · arXiv · Sep 1, 2026
Beyond semantic content, camera parameters play a pivotal role in dictating the geometric perspective and appearance of any given image. While recent image editing models excel at semantic and stylistic manipulation, they struggle with expl…
- Longstanding mental representations of familiar facesSarah Laurence, A. Mike Burton, Camilla Düring, Jen Pink et al. · White Rose Research Online ... · Sep 1, 2026
- BRF-GS: Hyperspectral Bidirectional Reflectance Factor Modeling and Image Generation Based on 3D Gaussian SplattingYiling Yao, Wenjuan Zhang, Bowen Wang, Bocheng Li et al. · arXiv · Aug 31, 2026
The bidirectional reflectance factor (BRF) characterizes the directional radiative properties of terrestrial surfaces. However, existing three-dimensional (3D) radiative transfer models require complex scene construction and computationally…
- Robust retinal biometrics for patient identity verification and retrieval across age and imaging devicesJose D. Vargas-Quiros, Dennis Bontempi, Jeroen Vermeulen, Bart Liefers et al. · arXiv · Aug 31, 2026
Patient identity errors can compromise longitudinal medical records, research databases, and downstream clinical decisions. We present a retinal biometric system for verifying claimed identities and retrieving the correct identity from colo…
- Identity-Conditioned Latent Consistency Distillation for Face SynthesisTiago Kienen Chaves, Bernardo Biesseck, David Menotti · arXiv · Aug 31, 2026
Diffusion models have achieved strong results in high-fidelity image synthesis, but their iterative sampling process makes large-scale generation computationally expensive. This limitation is especially relevant when generating synthetic fa…
- FaceSnap: Real-Time Personalized Lightstage Facial Performance CaptureRukhshanda Hussain, Noé Artru, Emeline Got, Luiz Gustavo Hafemann et al. · arXiv · Aug 31, 2026
Lightstage facial capture produces production-quality digital humans, but it is resource and labor-intensive. Multi-camera setups, hours of computation, and massive data storage create bottlenecks that hinder iterative workflows. This paper…
- DARP: A Calibrated Dual-Arm RGB-D-IR Dataset for Multi-View Robotic PerceptionManish Kansana, Mohammed Yusuf Mujawar, Sudip Mittal, Shahram Rahimi et al. · arXiv · Aug 31, 2026
Robotic perception from a single viewpoint is often limited by self-occlusion and incomplete surface visibility. This paper presents DARP(Dual-Arm Robotic Perception) https://doi.org/10.21227/rmv3-be47, a calibrated dual-arm RGB-D-IR datase…