Conceptio
›
computer-vision-and-pattern-recognition
Topic
computer-vision-and-pattern-recognition
Knowledge-graph topic
· documents ABOUT computer-vision-and-pattern-recognition across the archive
4,969
Documents about computer-vision-and-pattern-recognition
Documents about computer-vision-and-pattern-recognition
Matched Outcomes, Divergent Gaze: How Foveated MLLMs Search Compared to Humans
#622826
arXiv (OAI Expanded)
Autonomous vision-based UAV system for oil spill sampling and precision recovery
#394898
Digital Commons @ Michigan Tech
T2I-FactualBench: Benchmarking the Factuality of Text-to-Image Models with Knowledge-Intensive Concepts
#149203
arXiv (OAI)
Improving Dynamic Object Interactions in Text-to-Video Generation with AI Feedback
#153015
arXiv (OAI)
Are Explanations Helpful? A Comparative Analysis of Explainability Methods in Skin Lesion Classifiers
#153016
arXiv (OAI)
AIM 2025 Rip Current Segmentation (RipSeg) Challenge Report
#153116
arXiv (OAI)
BEVal: A Cross-dataset Evaluation Study of BEV Segmentation Models for Autonomous Driving
#176998
arXiv (OAI)
A Deep Learning-Based CCTV System for Automatic Smoking Detection in Fire Exit Zones
#189323
arXiv (OAI)
PE-CSNet: An equivariant network architecture with learnable patch-based sparse representation
#609141
arXiv (OAI Expanded)
Emergence of Transfer Learning towards Specific Identification of Alzheimer's Disease A Prospective Approach
#609164
arXiv (OAI Expanded)
MegaParts: Scaling Part-Aware 3D Object Generation to 300 Parts via Token-Efficient Autoregressive Modeling
#609209
arXiv (OAI Expanded)
Rethinking Token-wise Feature Caching: Accelerating Diffusion Transformers with Dual Feature Caching
#609763
arXiv (All)
MetaReason: Precise Interleaved Multimodal Reasoning via Editing Meta Information for Solving Geometry Problems
#610963
arXiv (OAI Expanded)
MetaReason: Precise Interleaved Multimodal Reasoning via Editing Meta Information for Solving Geometry Problems
#612440
arXiv (All)
UAV Video Deblurring via Motion-Aware Diffusion: A Path to Robust Target Detection
#614270
arXiv (OAI Expanded)
UAV Video Deblurring via Motion-Aware Diffusion: A Path to Robust Target Detection
#615237
arXiv (All)
What You Ask is What You Ground: Bridging Question Intent to Temporal Evidence for Grounded VideoQA
#616352
arXiv (OAI Expanded)
What You Ask is What You Ground: Bridging Question Intent to Temporal Evidence for Grounded VideoQA
#616852
arXiv (All)
Camera-Agnostic Pruning of 3D Gaussian Splats via Descriptor-Based Beta Evidence
#617502
arXiv (OAI Expanded)
Camera-Agnostic Pruning of 3D Gaussian Splats via Descriptor-Based Beta Evidence
#618002
arXiv (All)
When Do Cheap Probes Predict Expensive Training? Probing 3D-CT Encoders for Text Generation
#618707
arXiv (OAI Expanded)
Synthetic Data Augmentation for Satellite-Based Analysis of Battle-Damaged Agricultural Fields in Ukraine
#619099
arXiv (OAI Expanded)
When Do Cheap Probes Predict Expensive Training? Probing 3D-CT Encoders for Text Generation
#620386
arXiv (All)
Synthetic Data Augmentation for Satellite-Based Analysis of Battle-Damaged Agricultural Fields in Ukraine
#620778
arXiv (All)
JoyVASA: Portrait and Animal Image Animation with Diffusion-Based Audio-Driven Facial Dynamics and Head Motion Generation
#147033
arXiv (OAI)
Towards a General-Purpose Zero-Shot Synthetic Low-Light Image and Video Pipeline
#158692
arXiv (OAI)
ARQ: A Mixed-Precision Quantization Framework for Accurate and Certifiably Robust DNNs
#177011
arXiv (OAI)
MINOS: A Multimodal Evaluation Model for Bidirectional Generation Between Image and Text
#179179
arXiv (OAI)
Chain-of-Zoom: Extreme Super-Resolution via Scale Autoregression and Preference Alignment
#131329
arXiv (OAI)
OctoTools: An Agentic Framework with Extensible Tools for Complex Reasoning
#141699
arXiv (OAI)
What to Say and When to Say it: Live Fitness Coaching as a Testbed for Situated Interaction
#144212
arXiv (OAI)
Weakly Supervised Attention-based Models Using Activation Maps for Citrus Mite and Insect Pest Classification
#152949
arXiv (OAI)
RA-RRG: Multimodal Retrieval-Augmented Radiology Report Generation with Key Phrase Extraction
#153048
arXiv (OAI)
TextTIGER: Text-based Intelligent Generation with Entity Prompt Refinement for Text-to-Image Generation
#153051
arXiv (OAI)
A Generalist Model for Diverse Text-Guided Medical Image Synthesis
#155482
arXiv (OAI)
Sharpness-Aware Minimization with Z-Score Gradient Filtering
#160876
arXiv (OAI)
LLMPhy: Parameter-Identifiable Physical Reasoning Combining Large Language Models and Physics Engines
#163070
arXiv (OAI)
GOSPA and T-GOSPA quasi-metrics for evaluation of multi-object tracking algorithms
#163172
arXiv (OAI)
Test-Time Instance Selection for Improved Whole Slide Image Analysis
#609188
arXiv (OAI Expanded)
MiDAS: A Multimodal Data Acquisition System and Dataset for Robot-Assisted Minimally Invasive Surgery
#613947
arXiv (OAI Expanded)
Where did the ambiguity go? Examining how multimodal models interpret polysemous words
#614167
arXiv (OAI Expanded)
MiDAS: A Multimodal Data Acquisition System and Dataset for Robot-Assisted Minimally Invasive Surgery
#614914
arXiv (All)
Where did the ambiguity go? Examining how multimodal models interpret polysemous words
#615134
arXiv (All)
The role of object-centric representations, guided attention, and external memory on generalizing visual relations
#617233
arXiv (OAI Expanded)
The role of object-centric representations, guided attention, and external memory on generalizing visual relations
#617733
arXiv (All)
GaussianDWM++: Language-Grounded 3D Gaussian Driving World Model for Unified Scene Understanding, Editing, and Multi-Modal Generation
#618961
arXiv (OAI Expanded)
GaussianDWM++: Language-Grounded 3D Gaussian Driving World Model for Unified Scene Understanding, Editing, and Multi-Modal Generation
#620640
arXiv (All)
Bridging the Gap between Labeled and Unlabeled Data via Unified Flow with Feature Memory Bank
#622971
arXiv (OAI Expanded)
A Unified Backbone--Expert Framework with Relation-Token and Residual--Classifier Interfaces for Automatic Modulation Recognition
#611108
arXiv (OAI Expanded)
A Unified Backbone--Expert Framework with Relation-Token and Residual--Classifier Interfaces for Automatic Modulation Recognition
#612585
arXiv (All)
← Previous
Page 16 of 20
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.