Conceptio
›
computer-vision-and-pattern-recognition
Topic
computer-vision-and-pattern-recognition
Knowledge-graph topic
· documents ABOUT computer-vision-and-pattern-recognition across the archive
4,969
Documents about computer-vision-and-pattern-recognition
Documents about computer-vision-and-pattern-recognition
Tracking the Truth: Object-Centric Spatio-Temporal Monitoring for Video Large Language Models
#609935
arXiv (All)
EchoMask: Speech-Queried Attention-based Mask Modeling for Holistic Co-Speech Motion Generation
#611297
arXiv (OAI Expanded)
EchoMask: Speech-Queried Attention-based Mask Modeling for Holistic Co-Speech Motion Generation
#612774
arXiv (All)
Thinking with Gaze: Sequential Eye-Tracking as Visual Reasoning Supervision for Medical VLMs
#613981
arXiv (OAI Expanded)
Thinking with Gaze: Sequential Eye-Tracking as Visual Reasoning Supervision for Medical VLMs
#614948
arXiv (All)
EgoGazeLite: On-Device Egocentric Gaze Prediction for Token-Efficient Multimodal LLM Video Input
#616266
arXiv (OAI Expanded)
EgoGazeLite: On-Device Egocentric Gaze Prediction for Token-Efficient Multimodal LLM Video Input
#616766
arXiv (All)
The Many Senses of Visual Similarity: A Text-Prompted Image Perceptual Metric
#618701
arXiv (OAI Expanded)
GRNEdit: Efficient General Video Editing from a New Binary-Evidence Perspective in Generative Refinement Networks
#619049
arXiv (OAI Expanded)
The Many Senses of Visual Similarity: A Text-Prompted Image Perceptual Metric
#620380
arXiv (All)
GRNEdit: Efficient General Video Editing from a New Binary-Evidence Perspective in Generative Refinement Networks
#620728
arXiv (All)
HarmTrace: Anchor-Calibrated Decoupled Optimization for Fine-Grained Target Identification in Harmful Memes
#622922
arXiv (OAI Expanded)
Unsupervised Anomaly Detection for Image Dataset Quality Assurance in Multi-Center Breast MRI
#623012
arXiv (OAI Expanded)
Text-to-Image Models and Their Representation of People from Different Nationalities Engaging in Activities
#139147
arXiv (OAI)
LoREnc: Low-Rank Encryption for Securing Foundation Models and LoRA Adapters
#617569
arXiv (OAI Expanded)
LoREnc: Low-Rank Encryption for Securing Foundation Models and LoRA Adapters
#618069
arXiv (All)
Language Guided Adversarial Purification
#128266
arXiv (OAI)
Horticultural Temporal Fruit Monitoring via 3D Instance Segmentation and Re-Identification using Colored Point Clouds
#128320
arXiv (OAI)
Gen-n-Val: Agentic Image Data Generation and Validation
#131334
arXiv (OAI)
When 'YES' Meets 'BUT': Can Large Models Comprehend Contradictory Humor Through Comparative Reasoning?
#144270
arXiv (OAI)
Federated Breast Cancer Detection Enhanced by Synthetic Ultrasound Image Augmentation
#147120
arXiv (OAI)
GCDance: Genre-Controlled Music-Driven 3D Full Body Dance Generation
#153031
arXiv (OAI)
AdaMMS: Model Merging for Heterogeneous Multimodal Large Language Models with Unsupervised Coefficient Optimization
#158686
arXiv (OAI)
Cross-Distribution Diffusion Priors-Driven Iterative Reconstruction for Sparse-View CT
#160958
arXiv (OAI)
Culture-inspired Multi-modal Color Palette Generation and Colorization: A Chinese Youth Subculture Case
#182956
arXiv
Culture-inspired Multi-modal Color Palette Generation and Colorization: A Chinese Youth Subculture Case
#183006
arXiv Biology
Culture-inspired Multi-modal Color Palette Generation and Colorization: A Chinese Youth Subculture Case
#183056
arXiv (OAI Expanded)
Culture-inspired Multi-modal Color Palette Generation and Colorization: A Chinese Youth Subculture Case
#183106
arXiv (OAI)
Event-based Civil Infrastructure Visual Defect Detection: ev-CIVIL Dataset and Benchmark
#189254
arXiv (OAI)
Path2ST: Hierarchical Cell-Tissue Grounded Cross-Modal Translation for Spatial Transcriptomics
#609143
arXiv (OAI Expanded)
IP Protection in the Era of Visual Generative AI: A Survey
#609163
arXiv (OAI Expanded)
Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models
#609821
arXiv (All)
Mitigating Error Accumulation in Co-Speech Motion Generation via Global Rotation Diffusion and Multi-Level Constraints
#611395
arXiv (OAI Expanded)
Mitigating Error Accumulation in Co-Speech Motion Generation via Global Rotation Diffusion and Multi-Level Constraints
#612872
arXiv (All)
Synthetic Image Detection with CLIP: Understanding and Assessing Predictive Cues
#613946
arXiv (OAI Expanded)
FusionBERT: Multi-View Image--3D Retrieval via Cross-Attention Visual Fusion and Normal-Aware 3D Encoder
#614008
arXiv (OAI Expanded)
Lost in Adaptation: Layer-Selective Recovery of Temporal Reasoning in Video-Language Models
#614019
arXiv (OAI Expanded)
Efficient Table QA via TableGrid Navigation and Progressive Inference Prompting
#614052
arXiv (OAI Expanded)
Synthetic Image Detection with CLIP: Understanding and Assessing Predictive Cues
#614913
arXiv (All)
FusionBERT: Multi-View Image--3D Retrieval via Cross-Attention Visual Fusion and Normal-Aware 3D Encoder
#614975
arXiv (All)
Lost in Adaptation: Layer-Selective Recovery of Temporal Reasoning in Video-Language Models
#614986
arXiv (All)
Efficient Table QA via TableGrid Navigation and Progressive Inference Prompting
#615019
arXiv (All)
Catching Hallucinated Citations in Video-LLM Question Answering: A Self-Verification Pipeline and Verifier Ablation Study
#616230
arXiv (OAI Expanded)
Catching Hallucinated Citations in Video-LLM Question Answering: A Self-Verification Pipeline and Verifier Ablation Study
#616730
arXiv (All)
TIMA: Text-Image Mutual Awareness for Balancing Zero-Shot Adversarial Robustness and Generalization Ability
#617254
arXiv (OAI Expanded)
VesselBridge3D: A Foundation Model Adaptation Framework for Label-Efficient 3D Vessel Segmentation
#617473
arXiv (OAI Expanded)
Understanding Sources of Demographic Predictability in Brain MRI via Disentangling Anatomy and Contrast
#617479
arXiv (OAI Expanded)
TIMA: Text-Image Mutual Awareness for Balancing Zero-Shot Adversarial Robustness and Generalization Ability
#617754
arXiv (All)
VesselBridge3D: A Foundation Model Adaptation Framework for Label-Efficient 3D Vessel Segmentation
#617973
arXiv (All)
Understanding Sources of Demographic Predictability in Brain MRI via Disentangling Anatomy and Contrast
#617979
arXiv (All)
← Previous
Page 15 of 20
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.