Conceptio
›
Archive
›
arXiv (All)
arXiv (All)
open access
COMET: Contrastive Motion-Enhanced Temporal Reasoning for Video Multimodal Large Language Models
Zhu, Chenghua et al.
arXiv (All) · Papers · License: Open Access
Open Source ↗
Direct PDF ↓
computation-and-language
computer-vision-and-pattern-recognition
computer vision and pattern recognition, computation and language, machine learning
This document is indexed with metadata only — full text is not available in the archive for this record.
Open the official source ↗
Related documents
Are Human-Aligned Models Models of Humans? A Turing-Test Gap in Preference Alignment
#1052296
RPMem: Learning Long-Term Recurrent Parametric Memory Across Sessions for LLM Agents
#1052293
Mira-Scene: Pixel-Aligned Layouts for Generative 3D Scene Reconstruction
#1052299
Density-Ratio Rescoring for Imbalanced Classification Using Raking Duals and Classifier Scores
#1052300
Record
· ID 662209
Retrieved via
Conceptio
— every document is proof-bundled with source, license, and retrieval metadata.