接前文:CVPR 2026 目标检测(object detection)方向上接收论文总结1
其他
以下论文标题中出现 detect/detection 等字样,但不完全属于目标检测主线(如异常检测、伪造/生成内容检测、OOD 检测、变化检测、动作检测等)。按子方向分组列出。
异常检测
- A Semantically Disentangled Unified Model for Multi-category 3D Anomaly Detection
Team: SuYeon Kim,Wongyu Lee,MyeongAh Cho
- ADSeeker: A Knowledge-Grounded Reasoning Framework for Industry Anomaly Detection and Reasoning
Team: Kai Zhang,Zekai Zhang,Xihe Sun 等
- Alert-CLIP: Abnormality-aware Latent-Enhanced Representation Tuning of CLIP for Video Anomaly Detection
Team: Yiyan Zhu,Menghao Zhang,Haifeng Sun 等
- Anomaly-Related Residual Fields for Cross-domain Anomaly Detection
Team: Kewei Gao,Jiayi Xie,Zhengda Shen 等
- AnomalyVFM -- Transforming Vision Foundation Models into Zero-Shot Anomaly Detectors
Team: Matic Fučka,Vitjan Zavrtanik,Danijel Skočaj
- Back to Point: Exploring Point-Language Models for Zero-Shot 3D Anomaly Detection
Team: Kaiqiang Li,Gang Li,Mingle Zhou 等
- Bidirectional Multimodal Prompt Learning with Scale-Aware Training for Few-Shot Multi-Class Anomaly Detection
Team: Yujin Lee,Sewon Kim,Daeun Moon 等
- CHAL: Causal-guided Hierarchical Anomaly-aware Learning for Moving Infrared Small Target Detection
Team: Weiwei Duan,Luping Ji,Shipeng Lei 等
- Complementary Prototype Mapping for Efficient Multimodal Anomaly Detection
Team: Yuan Zhao,Xiaoqin Zhang,Huchuan Lu,Lihe Zhang
- Defect Cue-Preserved Structural Feature Refinement for Few-Shot Anomaly Detection
Team: Le Jiang,Yan Huang,Zhen Xu 等
- DLVP-CLIP: Enhancing Fine-Grained Zero-Shot Anomaly Detection via Dynamic Local Visual Prompting
Team: Gaowei Zhang,Lihe Zhang
- Dual-Prototype-Guided Multi-task Learning for Unsupervised Anomaly Detection and Classification
Team: Qianhao Luo,Jiajia Mi,Mingtao Yan 等
- FastRef: Fast Prototype Refinement for Few-shot Industrial Anomaly Detection
Team: Yufei Li,Long Tian,Yuyang Dai 等
- FB-CLIP: Fine-Grained Zero-Shot Anomaly Detection with Foreground-Background Disentanglement
Team: Ming Hu,Yongsheng Huo,Mingyu Dou 等
- Fine-VAD: Towards Fine-Grained Video Anomaly Detection via Progressive Cross-Granularity Learning
Team: Menghao Zhang,Yiyan Zhu,Pengfei Ren 等
- From Attraction to Equilibrium: Physics-Inspired Semantic Gravitons for Zero-Shot Anomaly Detection
Team: Yuwen Pan,Yuan Wang,Shaohui Li 等
- Geometry-Aligned and Anomaly-Aware Reconstruction for 3D Anomaly Detection
Team: Linchun Wu,Qin Zou,Yuanhao Yue,Zhongyuan Wang
- GPFlow: Gaussian Prototype Probability Flow for Unsupervised Multi-Modal Anomaly Detection
Team: Yiting Li,Xulei Yang,Jingyi Liao 等
- GS-CLIP: Zero-shot 3D Anomaly Detection by Geometry-Aware Prompt and Synergistic View Representation Learning
Team: Zehao Deng,An Liu,Yan Wang
- Hierarchical Point-Patch Fusion with Adaptive Patch Codebook for 3D Shape Anomaly Detection
Team: Xueyang Kang,Zizhao Li,Tian Lan 等
- Hunting Normality from Query Sample via Residual Learning for Generalist Anomaly Detection
Team: Xiaolei Wang,Yuexin Wang,Tianhong Dai 等
- InvAD: Inversion-based Reconstruction-Free Anomaly Detection with Diffusion Models
Team: Shunsuke Sakai,Xiangteng He,Chunzhi Gu 等
- Joint Learning of General and Diverse Patterns with Mixture of Memory Experts for Weakly-Supervised Video Anomaly Detection
Team: Bo Sun,Junxi Chen,Zhe Wu 等
- LayoutAD: Exploring Semantic-Geometric Misalignment Reasoning for Scene Layout Anomaly Detection
Team: Zhichao Zeng,Jiasheng Zhang,Jiyun Sun 等
- Learning from Noisy Supervision: A Denoising-Debiasing Framework for Weakly Supervised Video Anomaly Detection
Team: Yaxin Zhao,Yang Wang,Wenya Guo 等
- MMR-AD: A Large-Scale Multimodal Dataset for Benchmarking General Anomaly Detection with Multimodal Large Language Models
Team: Xincheng Yao,Zefeng Qian,Chao Shi 等
- MoECLIP: Patch-Specialized Experts for Zero-shot Anomaly Detection
Team: Jun Yeong Park,JunYoung Seo,Minji Kang,Yu Rang Park
- Multi-Prototype Compactness and Boundary-Aware Synthesis for Unsupervised Anomaly Detection
Team: Kailun Liao,Jianfeng Yang,Tao Tao 等
- No Need For Real Anomaly: MLLM Empowered Zero-Shot Video Anomaly Detection
Team: Zunkai Dai,Ke Li,Jiajia Liu 等
- Omni-AD: A Large-scale and Versatile Benchmark for Industrial Anomaly Detection
Team: Dahu Shi,Chengshen He,Shaochen Zhang 等
- PDD: Manifold-Prior Diverse Distillation for Medical Anomaly Detection
Team: Xijun Lu,Hongying Liu,Fanhua Shang 等
- RAID: Retrieval-Augmented Anomaly Detection
Team: Mingxiu Cai,Zhe Zhang,Gaochang Wu 等
- RC-NF: Robot-Conditioned Normalizing Flow for Real-Time Anomaly Detection in Robotic Manipulation
Team: Shijie Zhou,Bin Zhu,Jiarui Yang 等
- Reasoning-Driven Anomaly Detection and Localization with Image-Level Supervision
Team: Yizhou Jin,Yuezhu Feng,Jinjin Zhang 等
- SubspaceAD: Training-Free Few-Shot Anomaly Detection via Subspace Modeling
Team: Camile Lendering,Erkut Akdag,Egor Bondarau
- The Road Less Seen: Segment Exploration for Weakly Supervised Video Anomaly Detection
Team: Anusha Acharya,Hitesh Sapkota,Qi Yu,Xumin Liu
- TLMA: Mitigating the Impact of Weakly Labeled Information for Video Anomaly Detection
Team: Rong Xu,Runqi Wang,Yingjun Zhang 等
- Towards an Incremental Unified Multimodal Anomaly Detection: Augmenting Multimodal Denoising From an Information Bottleneck Perspective
Team: Kaifang Long,Lianbo Ma,Jiaqi Liu 等
- UniMMAD: Unified Multi-Modal and Multi-Class Anomaly Detection via MoE-Driven Feature Decompression
Team: Yuan Zhao,Youwei Pang,Lihe Zhang 等
- VisualAD: Language-Free Zero-Shot Anomaly Detection via Vision Transformer
Team: Yanning Hou,Peiyuan Li,Zirui Liu 等
- Wavelet-Driven 3D Anomaly Detection under Pose-Agnostic and Sparse-View
Team: Mingwen Shao,Qiao Zhang,Xinyuan Chen 等
- Weakly Supervised Video Anomaly Detection with Anomaly-Connected Components and Intention Reasoning
Team: Yu Wang,Shengjie Zhao
伪造/生成内容检测
- A Debiased Reconstruction-based Framework for Training-Free Detection of AI-Generated Images
Team: Sungik Choi,Hankook Lee,Jaehoon Lee 等
- A Difference-in-Difference Approach to Detecting AI-Generated Images
Team: Xinyi Qi,Kai Ye,Chengchun Shi 等
- A Sanity Check for Multi-In-Domain Face Forgery Detection in the Real World
Team: Jikang Cheng,Renye Yan,Zhiyuan Yan 等
- Agent4FaceForgery: Multi-Agent LLM Framework for Realistic Face Forgery Detection
Team: Yingxin Lai,Zitong YU,Jun Wang 等
- All in One: Unifying Deepfake Detection, Tampering Localization, and Source Tracing with a Robust Landmark-Identity Watermark
Team: Junjiang Wu,Liejun Wang,Zhiqing Guo
- AVFakeBench: A Comprehensive Audio-Video Forgery Detection Benchmark for AV-LMMs
Team: Shuhan Xia,Peipei Li,Xuannan Liu 等
- Beyond [CLS] Token: Query-Driven Token-Level Forgery Purification for Generalizable Deepfake Detection
Team: Changshuo Wang,Jiangming Wang,Ke-Yue Zhang 等
- CoCoVideo: The High-Quality Commercial-Model-Based Contrastive Benchmark for AI-Generated Video Detection
Team: Huidong Feng,Wentao Chen,Jie Chen 等
- Cross-modal Representation Learning for Diffusion-generated Image Detection
Team: Tao Gong,Dayong Wang,Qi Chu 等
- Decoupling Bias, Aligning Distributions: Synergistic Fairness Optimization for Deepfake Detection
Team: Feng Ding,Wenhui Yi,Yunpeng Zhou 等
- DeepfakeImpact: A Two-Stage Benchmark with Real-World Impact in Deepfake Detection
Team: Chaoyu Gong,Han Zhang,Siqiang Luo
- Detecting AI-Generated Forgeries via Iterative Manifold Deviation Amplification
Team: Jiangling Zhang,Shuxuan Gao,Bofan Liu 等
- Detecting Compressed AI-Generated Images via Phase Spectrum Robustness
Team: Kai Li,Wenqi Ren,Wei Wang,Xiaochun Cao
- DFD-HR: Generalizable Deepfake Detection via Hierarchical Routing Learning
Team: Jiamu Sun,Zhiyuan Yan,Ke-Yue Zhang 等
- DiffusionFF: A Diffusion-based Framework for Joint Face Forgery Detection and Fine-Grained Artifact Localization
Team: Siran Peng,Haoyuan Zhang,Li Gao 等
- Diversity over Uniformity: Rethinking Representation in Generated Image Detection
Team: Qinghui He,Haifeng Zhang,Qiao Qin 等
- Enabling Supervised Learning of Generative Signatures for Generalized AI-Generated Images Detection
Team: Jianwei Fei,Yunshu Dai,Xiaoyu Zhou 等
- FVBench: Benchmarking Deepfake Video Detection Capability of Large Multimodal Models
Team: Jiarui Wang,Huiyu Duan,Juntong Wang,Xiongkuo Min
- Investigating Self-Supervised Representations for Audio-Visual Deepfake Detection
Team: Dragos-Alexandru Boldisor,Stefan Smeu,Dan Oneata,Elisabeta Oneata
- Layer Consistency Matters: Elegant Latent Transition Discrepancy for Generalizable Synthetic Image Detection
Team: Yawen Yang,Feng Li,Shuqi Kong 等
- Locate-Then-Examine: Grounded Region Reasoning Improves Detection of AI-Generated Images
Team: Yikun Ji,Yan Hong,Bowen Deng 等
- Omni-Fake: Benchmarking Unified Multimodal Social Media Deepfake Detection
Team: Tianxiao Li,Zhenglin Huang,Haiquan Wen 等
- Pixels Don't Lie (But Your Detector Might): Bootstrapping MLLM-as-a-Judge for Trustworthy Deepfake Detection and Reasoning Supervision
Team: Kartik Kuckreja,Parul Gupta,Muhammad Haris Khan,Abhinav Dhall
- PPM-CLIP: Probabilistic Prompt Modeling for Generalizable AI-Generated Image Detection
Team: Xinyuan Wang,Yingxin Lai,Zhiming Luo,Zhihui Liu
- ReAlign: Generalizable Image Forgery Detection via Reasoning-Aligned Representation
Team: Qing Huang,Zhipei Xu,Xuanyu Zhang 等
- SAIDO: Generalizable Detection of AI-Generated Images via Scene-Aware and Importance-Guided Dynamic Optimization in Continual Learning
Team: Yongkang Hu,Yu Cheng,Yushuo Zhang 等
- Scaling Up AI-Generated Image Detection with Generator-Aware Prototypes
Team: Ziheng Qin,Yuheng Ji,Renshuai Tao 等
- Skyra: AI-Generated Video Detection via Grounded Artifact Reasoning
Team: Yifei Li,Wenzhao Zheng,Yanran Zhang 等
- Towards Generalizable AI-Generated Image Detection via Image-Adaptive Prompt Learning
Team: Yiheng Li,Zichang Tan,Guoqing Xu 等
- TriDF: Evaluating Perception, Detection, and Hallucination for Interpretable DeepFake Detection
Team: Jian-Yu Jiang-Lin,Kang-Yang Huang,Ling Zou 等
- Tutor-Student Reinforcement Learning: A Dynamic Curriculum for Robust Deepfake Detection
Team: Zhanhe Lei,Zhongyuan Wang,Jikang Cheng 等
- UniGenDet: A Unified Generative-Discriminative Framework for Co-Evolutionary Image Generation and Generated Image Detection
Team: Yanran Zhang,Wenzhao Zheng,Yifei Li 等
- Unleashing Vision-Language Semantics for Deepfake Video Detection
Team: Jiawen Zhu,Yunqi Miao,Xueyi Zhang 等
- VMD-FACT: A New Video Dataset and MLLM-based method for Detecting Realistic AI-Generated Video Misinformation
Team: Yongkang Zhang,Dongyu She,Baiyu Ji 等
- X-AVDT: Audio-Visual Cross-Attention for Robust Deepfake Detection
Team: Youngseo Kim,Kwan Yun,Seokhyeon Hong 等
- Your One-Stop Solution for AI-Generated Video Detection
Team: Long Ma,Zihao Xue,Yan Wang 等
- Zero-shot Detection of AI-Generated Image via RAW-RGB Alignment
Team: Haiwei Wu,Fengpeng Li,Zhilin Tu 等
OOD检测
- Activation Matters: Test-time Activated Negative Labels for OOD Detection with Vision-Language Models
Team: Yabin Zhang,Maya Varma,Yunhe Gao 等
- ANTS: Adaptive Negative Textual Space Shaping for OOD Detection via Test-Time MLLM Understanding and Reasoning
Team: Wenjie Zhu,Yabin Zhang,Xin Jin 等
- Bypassing the Transport Plan: Dynamic Reweighting for Out-of-Distribution Detection with Optimal Transport
Team: Yang Xiao,Weiming Liu,Jun Dan 等
- Enhancing Out-of-Distribution Detection with Extended Logit Normalization
Team: Yifan Ding,Xixi Liu,Jonas Unger,Gabriel Eilertsen
- Learning Latent Concepts for Detecting Out-of-Distribution Objects
Team: Ting Peng,Junhao Dong,Yew-Soon Ong
- Mind the Way You Select Negative Texts: Pursuing the Distance Consistency in OOD Detection with VLMs
Team: Zhikang Xu,Qianqian Xu,Zitai Wang 等
- Mitigating Simplicity Bias in OOD Detection through Object Co-occurrence Analysis
Team: Boyang Dai,Chaoqi Chen,Yizhou Yu
- Neural Distribution Prior for LiDAR Out-of-Distribution Detection
Team: Zizhao Li,Zhengkang Xiang,Jiayang Ao 等
- RankOOD - Class Ranking-based Out-of-Distribution Detection
Team: Dishanika Denipitiyage,Naveen Karunanayake,Suranga Seneviratne,Sanjay Chawla
- Sparsity as a Key: Unlocking New Insights from Latent Structures for Out-of-Distribution Detection
Team: Ahyoung Oh,Wonseok Shin,Songkuk Kim
- The Invisible Gorilla Effect in Out-of-distribution Detection
Team: Harry Anthony,Ziyun Liang,Hermione Warr,Konstantinos Kamnitsas
- TTL: Test-time Textual Learning for OOD Detection with Pretrained Vision-Language Models
Team: Jinlun Ye,Jiang Liao,Runhe Lai 等
- UNI-OOD: Unified Object- and Image-level Out-of-Distribution Detection via Cross-Context Attentive Vision-Language Modeling
Team: Yuchuan Li,Azadeh Motamedi,Hyock Ju Kwon 等
变化检测
- Changes in Real Time: Online Scene Change Detection with Multi-View Fusion
Team: Chamuditha Jayanga Galappaththige,Jason Lai,Lloyd Windrim 等
- OpenDPR: Open-Vocabulary Change Detection via Vision-Centric Diffusion-Guided Prototype Retrieval for Remote Sensing Imagery
Team: Qi Guo,Jue Wang,Yinhe Liu,Yanfei Zhong
- RDF-MIG: A Robust Diffusion Framework for Masked Image Generation to Augment Semantic Segmentation and Change Detection
Team: Zian Cao,Wei Wei,Qingshan Gao,Yuanyuan Fu
- SRGCD: Stability-Driven Region Growth Framework for 3D Change Detection
Team: Yue Wu,Tao Peng,Yongzhe Yuan 等
- UniChange: Unifying Change Detection with Multimodal Large Language Model
Team: Xu Zhang,Danyang Li,Xiaohang Dong 等
动作检测
- Decompose and Transfer: CoT-Prompting Enhanced Alignment for Open-Vocabulary Temporal Action Detection
Team: Sa Zhu,Wanqian Zhang,Lin Wang 等
- Mining Instance-Centric Vision-Language Contexts for Human-Object Interaction Detection
Team: Soo Won Seo,KyungChae Lee,Hyungchan Cho 等
- MoVie: Broaden Your Views with Human Motion for Action Detection
Team: Di Yang,Mahmoud Ali,Xuanlong Yu 等
- RegFormer: Transferable Relational Grounding for Efficient Weakly-Supervised Human-Object Interaction Detection
Team: Jihwan Park,Chanhyeong Yang,Jinyoung Park 等
- Streamlined Open-Vocabulary Human-Object Interaction Detection
Team: Chang Sun,Dongliang Liao,Changxing Ding
- TF-CADE: Foreground-Concentrated Text-Video Alignment for Zero-Shot Temporal Action Detection
Team: Yearang Lee,Ho-Joong Kim,Seong-Whan Lee
关键点/地标检测
- BEV-SLD: Self-Supervised Scene Landmark Detection for Global Localization with LiDAR Bird's-Eye View Images
Team: David Skuddis,Vincent Ress,Wei Zhang 等
- EV-CGNet: Co-visible Focused 3D-guided 2D Event Keypoint Detection Network
Team: Yuan Gao,Tianle Ding,Yuqing Zhu,Tianzhu Zhang
- From Pairs to Sequences: Track-Aware Policy Gradients for Keypoint Detection
Team: Yepeng Liu,Hao Li,Liwen Yang 等
幻觉检测
- Beyond the Global Scores: Fine-Grained Token Grounding as a Robust Detector of LVLM Hallucinations
Team: Tuan Dung Nguyen,Minh Khoi Ho,Qi Chen 等
- Lyapunov Probes for Hallucination Detection in Large Foundation Models
Team: Bozhi Luan,Gen Li,Yalan Qin 等
- PAS: Prelim Attention Score for Detecting Object Hallucinations in Large Vision-Language Models
Team: Nhat Hoang,Minh Vu,My T. Thai,Manish Bhattarai
- Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination
Team: Zichuan Wang,Songlin Yang,Bo Peng 等
- ZINA: Multimodal Fine-grained Hallucination Detection and Editing
Team: Yuiga Wada,Kazuki Matsuda,Komei Sugiura,Graham Neubig
讽刺/语义检测
- MMSD3.0: A Multi-Image Benchmark for Real-World Multimodal Sarcasm Detection
Team: Haochen Zhao,Yuyao Kong,Yongxiu Xu 等
跟踪相关
- From Detection to Association: Learning Discriminative Object Embeddings for Multi-Object Tracking
Team: Yuqing Shao,Yuchen Yang,Rui Yu 等
其他检测相关
- Adaptive Confidence Regularization for Multimodal Failure Detection
Team: Moru Liu,Hao Dong,Olga Fink,Mario Trapp
- ArchSym: Detecting 3D-Grounded Architectural Symmetries in the Wild
Team: Hanyu Chen,Ruojin Cai,Steve Marschner,Noah Snavely
- AutoDebias: An Automated Framework for Detecting and Mitigating Backdoor Biases in Text-to-Image Models
Team: Hongyi Cai,Mohammad Mahdinur Rahman,MingKang Dong 等
- AXG-Reasoner: Error Detection and Explanation in Long Task Videos with Vision-Language Models
Team: Shih-Po Lee,Ehsan Elhamifar
- BlackMirror: Black-Box Backdoor Detection for Text-to-Image Models via Instruction-Response Deviation
Team: Feiran Li,Qianqian Xu,Shilong Bao 等
- Breaking Spurious Correlations: Uncertainty-Driven Causal Transformers for AU Detection
Team: Yuru Wang,Yue Zhou
- Bulk RNA-seq Guided Multi-modal Detection of Anomalous Regions in Human Cancer via Spatial Transcriptomics
Team: Hang Shi,Ruocheng Yang,Wenjie You 等
- BUSSARD: Normalizing Flows for Bijective Universal Scene-Specific Anomalous Relationship Detection
Team: Melissa Schween,Mathis Kruse,Bodo Rosenhahn
- Conan: Progressive Learning to Reason Like a Detective over Multi-Scale Visual Evidence
Team: Kun Ouyang,Yuanxin Liu,Linli Yao 等
- COPYLENS: Towards Copyrighted Characters Infringement Detection via Copyright-Aware Prompt Learning
Team: Yaoyu Jin,Xiaochun Yang,Hong Liu 等
- CrossVL: Complexity-Aware Feature Routing and Paired Curriculum for Cross-View Vision-Language Detection
Team: Zhipeng Liu,Chunbo Luo
- Data Leakage Detection and De-duplication in Large Scale Geospatial Image Datasets
Team: Yeshwanth Kumar Adimoolam,Charalambos Poullis,Melinos Averkiou
- DetAny4D: Detect Anything 4D Temporally in a Streaming RGB Video
Team: Jiawei Hou,Shenghao Zhang,Can Wang 等
- Detect Any AI-Counterfeited Text Image
Team: Chenfan Qu,Yiwu Zhong,Xuekang Zhu 等
- Detect Anything via Next Point Prediction
Team: Qing Jiang,Junan Huo,Xingyu Chen 等
- DetectSCI: Toward Object-Guided ROI Reconstruction for High-Resolution Video Snapshot Compressive Imaging
Team: Xingjian Jiang,Lishun Wang,Ping Wang,Xin Yuan
- EReCu: Pseudo-label Evolution Fusion and Refinement with Multi-Cue Learning for Unsupervised Camouflage Detection
Team: Shuo Jiang,Gaojia Zhang,Min Tan 等
- FedSDR: Federated Graph Learning with Structural Noise Detection and Reconstruction
Team: Jiaqi Liu,Zihan Tan,Guancheng Wan 等
- Geometry-driven OOD Detectors Are Class-Incremental Learners
Team: Wangwang Jia,Zijian Gao,Tianjiao Wan 等
- Ghost-FWL: A Large-Scale Full-Waveform LiDAR Dataset for Ghost Detection and Removal
Team: Kazuma Ikeda,Ryosei Hara,Rokuto Nagata 等
- GuardTrace-VL: Detecting Unsafe Multimodel Reasoning via Iterative Safety Supervision
Team: Yuxiao Xiang,Junchi Chen,Zhenchao Jin 等
- Homaloidal parametrization for detecting critical two-view configurations
Team: Rakshith Madhavan,Matteo Forlivesi,Marina Bertolini 等
- KLIP: Localized Distribution Shift Detection via KL-Divergence with Diffusion Priors in Inverse Problems
Team: Alireza Kheirandish,Jihoon Hong,Sara Fridovich-Keil
- Learnability-Driven Submodular Optimization for Active Roadside 3D Detection
Team: Ruiyu Mao,Baoming Zhang,Nicholas Ruozzi,Yunhui Guo
- Learning to Diversify and Focus: A Reinforcement Framework for Open-Vocabulary HOI Detection
Team: Yongchao Xu,Jiawei Liu,Junfeng Wang 等
- LocateAnything3D: Vision-Language 3D Detection with Chain-of-Sight
Team: Yunze Man,Shihao Wang,Guowen Zhang 等
- Look Before You Fuse: 2D-Guided Cross-Modal Alignment for Robust 3D Detection
Team: Xiang Li,Zhangchi Hu,Xu Xiao,Bin Kong
- MatchED: Crisp Edge Detection Using End-to-End, Matching-based Supervision
Team: Bedrettin Cetinkaya,Sinan Kalkan,Emre Akbas
- MEMO: Human-like Crisp Edge Detection Using Masked Edge Prediction
Team: Jiaxin Cheng,Yue Wu,Yicong Zhou
- MRD: Multi-resolution Retrieval-Detection Fusion for High-Resolution Image Understanding
Team: Fan Yang,Xingping Dong,Xin Yu 等
- Neural Field-Based 3D Surface Reconstruction of Microstructures from Multi-Detector Signals in Scanning Electron Microscopy
Team: Shuo Chen,Yijin Li,Xi Zheng,Guofeng Zhang
- Off The Grid: Detection of Primitives for Feed-Forward 3D Gaussian Splatting
Team: Arthur Moreau,Richard Shaw,Michal Nazarczuk 等
- OpenFS: Multi-Hand-Capable Fingerspelling Recognition with Implicit Signing-Hand Detection and Frame-Wise Letter-Conditioned Synthesis
Team: Junuk Cha,Jihyeon Kim,Han-Mu Park
- OVOD-Agent: A Markov-Bandit Framework for Proactive Visual Reasoning and Self-Evolving Detection
Team: Chujie Wang,Jianyu Lu,Zhiyuan Luo 等
- Physical Adversarial Clothing Evades Visible-Thermal Detectors via Non-Overlapping RGB-T Pattern
Team: Xiaopei Zhu,Guanning Zeng,Zhanhao Hu 等
- Probabilistic Concept Graph Reasoning for Multimodal Misinformation Detection
Team: Ruichao Yang,Wei Gao,Xiaobin Zhu 等
- Real-Time Multimodal Fingertip Contact Detection via Depth and Motion Fusion for Vision-Based Human-Computer Interaction
Team: Mukhiddin Toshpulatov,Wookey Lee,Suan Lee,Geehyuk Lee
- ReManNet: A Riemannian Manifold Network for Monocular 3D Lane Detection
Team: Chengzhi Hong,Bijun Li
- RPGFusion: 4D Radar Prior-Guided Multi-Modal Fusion for 3D Detection
Team: Xin Qiu,Wenjie Liu
- SAVA-X: Ego-to-Exo Imitation Error Detection via Scene-Adaptive View Alignment and Bidirectional Cross View Fusion
Team: Xiang Li,Heqian Qiu,Lanxiao Wang 等
- Scene Reconstruction as Mapping Priors for 3D Detection
Team: Yang Fu,Yuliang Zou,Hao Xiang 等
- Seeing Through the Noise: Improving Infrared Small Target Detection and Segmentation from Noise Suppression Perspective
Team: Maoxun Yuan,Duanni Meng,Ziteng Xi 等
- SFR-Net: Steering-Fusion-Refining Network in Multi-label Zero-Shot Sewer Defect Detection
Team: Zhao-Min Chen,Xinjian Huang,Yisu Ge,Yu Li
- Similarity-Consistent Likelihood Diffusion enables Hidden Person Detection from Wall Reflections
Team: Zhiwen Zheng,Hao Zhou,Huiyu Qi 等
- SimLBR: Learning to Detect Fake Images by Learning to Detect Real Images
Team: Aayush Dhakal,Subash Khanal,Srikumar Sastry 等
- Synergistic Bleeding Region and Point Detection in Laparoscopic Surgical Videos
Team: Jialun Pei,Zhangjun Zhou,Diandian Guo 等
- Target-Aware Invertible Encoder with Reconstruction Guidance for Infrared Small Target Detection
Team: Shule Yan,Zetian Zhang,Xiao Ma,Zexuan Ji
- Towards Stealthy and Effective Backdoor Attacks on Lane Detection: A Naturalistic Data Poisoning Approach
Team: Yifan Liao,Yuxin Cao,Yedi Zhang 等
- Training-free Detection of Generated Videos via Spatial-Temporal Likelihoods
Team: Omer Ben Hayun,Roy Betser,Meir Yossef Levi 等
- TTP: Test-Time Padding for Adversarial Detection and Robust Adaptation on Vision-Language Models
Team: Zhiwei Li,Yitian Pang,Weining Wang 等
- TVHighlights: LLM-Guided Human-Free Collaborative Training for Video Highlight Detection in Movies and TV Dramas
Team: Qi Qiu,Xuan Wu,Jiawei Peng 等
- UAV-CB: A Complex-Background RGB-T Dataset and Local Frequency Bridge Network for UAV Detection
Team: Shenghui Huang,Menghao Hu,Longkun Zou 等
- Unlearning without Forgetting: Securely Removing Targeted Concepts from Large-Scale Vision-Language Open-Vocabulary Detectors
Team: Zhongze Wu,Xiu Su,Feng Yang 等
- Wavelet-based Frame Selection by Detecting Semantic Boundary for Long Video Understanding
Team: Wang Chen,Yuhui Zeng,Yongdong Luo 等
总结
从本届接收论文来看,CVPR 2026 目标检测方向呈现以下趋势:
3D 目标检测体量最大:单目、多视角、BEV、LiDAR/Radar 融合与室内外统一检测持续活跃;雷达-相机融合、Gaussian Splatting 先验、token 压缩与不确定性估计是常见技术点。
开放词汇/开放世界检测成为主线之一:Open-Vocabulary Detection、Open-World Detection、未知类别发现、检索式检测(如 WeDetect)与热成像开放词汇检测等方向快速增长。
数据高效学习受重视:少样本、跨域少样本、增量检测、主动学习、在线数据筛选与弱监督设定显著增多,反映标注成本与持续部署需求。
实时高效架构回潮:YOLO 体系、Mamba/SSM 混合结构、轻量化模型与训练策略优化重新成为焦点。
场景专用化加深:遥感/旋转框、UAV、小目标、伪装/显著性、水下、X-ray 安检、Person Search 等方法继续细分。
检测概念外延明显:异常检测、深度伪造/生成内容检测、OOD 检测、变化检测等“泛检测”任务数量可观,但与经典目标检测主线有所区分。
总体而言,CVPR 2026 目标检测研究在通用检测框架演进之外,更强调开放词汇泛化、三维感知、数据高效学习与真实场景鲁棒落地。
参考资料
(注:文档部分内容由 AI 生成;Code/Blog/单位信息以公开网页检索为准,如有遗漏欢迎补充指正。)