Conceptio
›
computer-vision-and-pattern-recognition
Topic
computer-vision-and-pattern-recognition
Knowledge-graph topic
· documents ABOUT computer-vision-and-pattern-recognition across the archive
4,969
Documents about computer-vision-and-pattern-recognition
Documents about computer-vision-and-pattern-recognition
Time Blindness: Why Video-Language Models Can't See What Humans Can?
#179178
arXiv (OAI)
Learning Locally, Revising Globally: Global Reviser for Federated Learning with Noisy Labels
#189226
arXiv (OAI)
DeCo-MIL: Debiased Counterfactual Reasoning for Long-Tailed Whole Slide Image Analysis
#609152
arXiv (OAI Expanded)
Towards Airborne Object Detection: A Deep Learning Analysis
#609860
arXiv (All)
MotionGS-SLAM: Event-Modulated Gaussian Splatting for Motion-Blur Robust SLAM
#610978
arXiv (OAI Expanded)
CETalk: Continuous Valence-Arousal Control for Audio-Driven 3D Talking Head Generation
#611060
arXiv (OAI Expanded)
Two by Two: Learning Multi-Task Pairwise Objects Assembly for Generalizable Robot Manipulation
#611296
arXiv (OAI Expanded)
MotionGS-SLAM: Event-Modulated Gaussian Splatting for Motion-Blur Robust SLAM
#612455
arXiv (All)
CETalk: Continuous Valence-Arousal Control for Audio-Driven 3D Talking Head Generation
#612537
arXiv (All)
Two by Two: Learning Multi-Task Pairwise Objects Assembly for Generalizable Robot Manipulation
#612773
arXiv (All)
Omni-LiveAvatar: Minute-Level Real-Time Streaming Joint Audio-Video Avatar Generation
#614246
arXiv (OAI Expanded)
Shape Operator PCA: Curvature-Aware Projections for Geometric Machine Learning
#614318
arXiv (OAI Expanded)
TEA: Text Encoder Alignment for Robust Concept Erasure in Text-to-Image Models
#614344
arXiv (OAI Expanded)
Omni-LiveAvatar: Minute-Level Real-Time Streaming Joint Audio-Video Avatar Generation
#615213
arXiv (All)
Shape Operator PCA: Curvature-Aware Projections for Geometric Machine Learning
#615285
arXiv (All)
TEA: Text Encoder Alignment for Robust Concept Erasure in Text-to-Image Models
#615311
arXiv (All)
RRFC: Recursive Refinement via Feedback Conditioning for Iterative Image-to-Image Generation
#616339
arXiv (OAI Expanded)
PWLR: Pairwise Witness Local Rejection for Boundary-Aware Out-of-Distribution Detection
#616440
arXiv (OAI Expanded)
UI-Mate: Advancing Open-Weight Foundation GUI Agents with In-Context Demonstrations
#616551
arXiv (OAI Expanded)
RRFC: Recursive Refinement via Feedback Conditioning for Iterative Image-to-Image Generation
#616839
arXiv (All)
PWLR: Pairwise Witness Local Rejection for Boundary-Aware Out-of-Distribution Detection
#616940
arXiv (All)
UI-Mate: Advancing Open-Weight Foundation GUI Agents with In-Context Demonstrations
#617051
arXiv (All)
Breaking the Compression Barrier: Cross-Architecture Compression Boundary Learning via Reverse Regrowth
#617184
arXiv (OAI Expanded)
A Highly Efficient Diversity-based Input Selection for DNN Improvement Using VLMs
#617434
arXiv (OAI Expanded)
SemVideo: Reconstructs What You Watch from Brain Activity via Hierarchical Semantic Guidance
#617469
arXiv (OAI Expanded)
Less Data, Faster Convergence: Goal-Driven Data Optimization for Multimodal Instruction Tuning
#617488
arXiv (OAI Expanded)
Human Pose Estimation in Trampoline Gymnastics: How to Improve Performance on Extreme Poses
#617511
arXiv (OAI Expanded)
Breaking the Compression Barrier: Cross-Architecture Compression Boundary Learning via Reverse Regrowth
#617684
arXiv (All)
A Highly Efficient Diversity-based Input Selection for DNN Improvement Using VLMs
#617934
arXiv (All)
SemVideo: Reconstructs What You Watch from Brain Activity via Hierarchical Semantic Guidance
#617969
arXiv (All)
Less Data, Faster Convergence: Goal-Driven Data Optimization for Multimodal Instruction Tuning
#617988
arXiv (All)
Human Pose Estimation in Trampoline Gymnastics: How to Improve Performance on Extreme Poses
#618011
arXiv (All)
Visko Orbis 1.0: A Live Model for Real-Time Interactive Long Video Generation
#618714
arXiv (OAI Expanded)
Defake-o3: From Speculative Rationales to Verifiable Evidence for Explainable AIGI Detection
#618985
arXiv (OAI Expanded)
StreamOPD: A Post-Training Recipe with Spatio-Temporal Cue Gating for Streaming Video Understanding
#619041
arXiv (OAI Expanded)
SIGMA-Lane: Scale-pyramId Gated MAmba for Temporally Consistent Video Lane Detection
#619059
arXiv (OAI Expanded)
Visko Orbis 1.0: A Live Model for Real-Time Interactive Long Video Generation
#620393
arXiv (All)
Defake-o3: From Speculative Rationales to Verifiable Evidence for Explainable AIGI Detection
#620664
arXiv (All)
StreamOPD: A Post-Training Recipe with Spatio-Temporal Cue Gating for Streaming Video Understanding
#620720
arXiv (All)
SIGMA-Lane: Scale-pyramId Gated MAmba for Temporally Consistent Video Lane Detection
#620738
arXiv (All)
Automatic Cephalometric Landmark Localization on CBCT-Derived Digitally Reconstructed Radiographs for Skeletal Malocclusion Classification
#622846
arXiv (OAI Expanded)
GeoPose: Patient-agnostic CTA-to-DSA registration through projection-space calibration
#622903
arXiv (OAI Expanded)
Training-Free Reconstruction-Based AI-Generated Image Detectors Are Inherently Vulnerable to Adversarial Examples
#622943
arXiv (OAI Expanded)
Flattery in Motion: Benchmarking and Analyzing Sycophancy in Video-LLMs
#183258
arXiv (OAI)
ViFiCon: Vision and Wireless Association Via Self-Supervised Contrastive Learning
#163009
arXiv (OAI)
ViFiCon: Vision and Wireless Association Via Self-Supervised Contrastive Learning
#163968
arXiv (OAI Expanded)
FZ-VLM: A Two Stage Florence-Zephyr Vision Language Model Framework for Pulmonary Nodule Characterization and Clinical Decision Making
#610961
arXiv (OAI Expanded)
FZ-VLM: A Two Stage Florence-Zephyr Vision Language Model Framework for Pulmonary Nodule Characterization and Clinical Decision Making
#612438
arXiv (All)
Music Audio-Visual Question Answering Requires Specialized Multimodal Designs
#131330
arXiv (OAI)
LaMI: Augmenting Large Language Models via Late Multi-Image Fusion
#139076
arXiv (OAI)
← Previous
Page 13 of 20
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.