image-text-to-texttransformersapache-2.0

HuggingFaceTB/SmolVLM2-2.2B-Instruct

huggingface.co/HuggingFaceTB/SmolVLM2-2.2B-Instruct

333Likes
127,138Downloads
2025-04-08Updated
transformerssafetensorssmolvlmimage-text-to-textvideo-text-to-textconversationalendataset:HuggingFaceM4/the_cauldrondataset:HuggingFaceM4/Docmatixdataset:lmms-lab/LLaVA-OneVision-Datadataset:lmms-lab/M4-Instruct-Datadataset:HuggingFaceFV/finevideodataset:MAmmoTH-VL/MAmmoTH-VL-Instruct-12Mdataset:lmms-lab/LLaVA-Video-178Kdataset:orrzohar/Video-STaRdataset:Mutonix/Vriptdataset:TIGER-Lab/VISTA-400Kdataset:Enxin/MovieChat-1K_traindataset:ShareGPT4Video/ShareGPT4Videoarxiv:2504.05299base_model:HuggingFaceTB/SmolVLM-Instructbase_model:finetune:HuggingFaceTB/SmolVLM-Instructlicense:apache-2.0endpoints_compatible
The README has not been fetched yet (metadata-first ingest). The model page and the metadata below are live.

Model card

Mirrored from the Hugging Face Hub and served from the Conceptio Open Knowledge Archive. Read the original card at https://huggingface.co/HuggingFaceTB/SmolVLM2-2.2B-Instruct.