Conceptio
›
pattern-recognition
Topic
pattern-recognition
Knowledge-graph topic
· documents ABOUT pattern-recognition across the archive
89
Documents about pattern-recognition
Documents about pattern-recognition
Tracking the Truth: Object-Centric Spatio-Temporal Monitoring for Video Large Language Models
#609935
arXiv (All)
Auditing Frame-Level AUC in Weakly Supervised Video Anomaly Detection: Granularity, Resolution, and Scene Bias
#610048
arXiv (All)
UC-VLM: Consistency-Driven Learning for AI-Generated Image Detection with Vision-Language Large Models
#611180
arXiv (OAI Expanded)
EchoMask: Speech-Queried Attention-based Mask Modeling for Holistic Co-Speech Motion Generation
#611297
arXiv (OAI Expanded)
UC-VLM: Consistency-Driven Learning for AI-Generated Image Detection with Vision-Language Large Models
#612657
arXiv (All)
EchoMask: Speech-Queried Attention-based Mask Modeling for Holistic Co-Speech Motion Generation
#612774
arXiv (All)
Thinking with Gaze: Sequential Eye-Tracking as Visual Reasoning Supervision for Medical VLMs
#613981
arXiv (OAI Expanded)
VGGT-Align: Bridging Local Reconstruction and Global Consistency for Long-Sequence 3D Reconstruction
#614271
arXiv (OAI Expanded)
Thinking with Gaze: Sequential Eye-Tracking as Visual Reasoning Supervision for Medical VLMs
#614948
arXiv (All)
VGGT-Align: Bridging Local Reconstruction and Global Consistency for Long-Sequence 3D Reconstruction
#615238
arXiv (All)
Dual-Branch State-Displacement Network for Sea Surface Temperature Super-Resolution
#616090
arXiv (OAI Expanded)
NumerosityVLM: A Cognitively Inspired Benchmark for Interpreting Numerosity Representations in Vision-Language Models
#616092
arXiv (OAI Expanded)
EgoGazeLite: On-Device Egocentric Gaze Prediction for Token-Efficient Multimodal LLM Video Input
#616266
arXiv (OAI Expanded)
Beyond Single Object: Learning 3D Relations with Large Language Models
#616354
arXiv (OAI Expanded)
Dual-Branch State-Displacement Network for Sea Surface Temperature Super-Resolution
#616590
arXiv (All)
NumerosityVLM: A Cognitively Inspired Benchmark for Interpreting Numerosity Representations in Vision-Language Models
#616592
arXiv (All)
EgoGazeLite: On-Device Egocentric Gaze Prediction for Token-Efficient Multimodal LLM Video Input
#616766
arXiv (All)
Beyond Single Object: Learning 3D Relations with Large Language Models
#616854
arXiv (All)
RingMo-Agent: A Unified Remote Sensing Foundation Model for Multi-Platform and Multi-Modal Reasoning
#617343
arXiv (OAI Expanded)
RingMo-Agent: A Unified Remote Sensing Foundation Model for Multi-Platform and Multi-Modal Reasoning
#617843
arXiv (All)
The Multimodal Brain Tumor Image Segmentation Benchmark (BRATS)
#7411
OpenAlex
Language Guided Adversarial Purification
#128266
arXiv (OAI)
Horticultural Temporal Fruit Monitoring via 3D Instance Segmentation and Re-Identification using Colored Point Clouds
#128320
arXiv (OAI)
Gen-n-Val: Agentic Image Data Generation and Validation
#131334
arXiv (OAI)
Text-to-Image Models and Their Representation of People from Different Nationalities Engaging in Activities
#139147
arXiv (OAI)
When 'YES' Meets 'BUT': Can Large Models Comprehend Contradictory Humor Through Comparative Reasoning?
#144270
arXiv (OAI)
Federated Breast Cancer Detection Enhanced by Synthetic Ultrasound Image Augmentation
#147120
arXiv (OAI)
T2I-FactualBench: Benchmarking the Factuality of Text-to-Image Models with Knowledge-Intensive Concepts
#149203
arXiv (OAI)
Improving Dynamic Object Interactions in Text-to-Video Generation with AI Feedback
#153015
arXiv (OAI)
Are Explanations Helpful? A Comparative Analysis of Explainability Methods in Skin Lesion Classifiers
#153016
arXiv (OAI)
GCDance: Genre-Controlled Music-Driven 3D Full Body Dance Generation
#153031
arXiv (OAI)
AIM 2025 Rip Current Segmentation (RipSeg) Challenge Report
#153116
arXiv (OAI)
AdaMMS: Model Merging for Heterogeneous Multimodal Large Language Models with Unsupervised Coefficient Optimization
#158686
arXiv (OAI)
Cross-Distribution Diffusion Priors-Driven Iterative Reconstruction for Sparse-View CT
#160958
arXiv (OAI)
V-MAGE: A Game Evaluation Framework for Assessing Vision-Centric Capabilities in Multimodal Large Language Models
#163115
arXiv (OAI)
BEVal: A Cross-dataset Evaluation Study of BEV Segmentation Models for Autonomous Driving
#176998
arXiv (OAI)
Culture-inspired Multi-modal Color Palette Generation and Colorization: A Chinese Youth Subculture Case
#182956
arXiv
Culture-inspired Multi-modal Color Palette Generation and Colorization: A Chinese Youth Subculture Case
#183006
arXiv Biology
Culture-inspired Multi-modal Color Palette Generation and Colorization: A Chinese Youth Subculture Case
#183056
arXiv (OAI Expanded)
Culture-inspired Multi-modal Color Palette Generation and Colorization: A Chinese Youth Subculture Case
#183106
arXiv (OAI)
Event-based Civil Infrastructure Visual Defect Detection: ev-CIVIL Dataset and Benchmark
#189254
arXiv (OAI)
Revisiting CroPA: A Reproducibility Study and Enhancements for Cross-Prompt Adversarial Transferability in Vision-Language Models
#189293
arXiv (OAI)
A Deep Learning-Based CCTV System for Automatic Smoking Detection in Fire Exit Zones
#189323
arXiv (OAI)
A Pattern-Recognition Method for Highway Construction Project Expenditure Cash Flows Using Clustering-Based K-Means Approach
#283826
DigitalCommons@University of Nebraska
Energy-Efficient Plant Monitoring via Knowledge Distillation
#333847
HAL (France)
Probabilistic models for multi-classifier biometric authentication using quality measures
#456708
DataCite
PE-CSNet: An equivariant network architecture with learnable patch-based sparse representation
#609141
arXiv (OAI Expanded)
Path2ST: Hierarchical Cell-Tissue Grounded Cross-Modal Translation for Spatial Transcriptomics
#609143
arXiv (OAI Expanded)
IP Protection in the Era of Visual Generative AI: A Survey
#609163
arXiv (OAI Expanded)
Emergence of Transfer Learning towards Specific Identification of Alzheimer's Disease A Prospective Approach
#609164
arXiv (OAI Expanded)
← Previous
Page 13 of 20
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.