video-text-to-texttransformersapache-2.0

OpenMOSS-Team/MOSS-VL-Instruct-0408

huggingface.co/OpenMOSS-Team/MOSS-VL-Instruct-0408

107Likes
671Downloads
2026-09-11Updated
transformerssafetensorsmoss_vlfeature-extractionSFTVideo-UnderstandingImage-UnderstandingMOSS-VLOpenMOSSmultimodalvideovision-languagevideo-text-to-textcustom_codeenzharxiv:2608.15045arxiv:2606.07639base_model:OpenMOSS-Team/MOSS-VL-Base-0408base_model:finetune:OpenMOSS-Team/MOSS-VL-Base-0408license:apache-2.0region:us
The README has not been fetched yet (metadata-first ingest). The model page and the metadata below are live.

Model card

Mirrored from the Hugging Face Hub and served from the Conceptio Open Knowledge Archive. Read the original card at https://huggingface.co/OpenMOSS-Team/MOSS-VL-Instruct-0408.