image-text-to-textapache-2.0

inference-net/ClipTagger-12b

huggingface.co/inference-net/ClipTagger-12b

59Likes
63Downloads
2025-08-14Updated
safetensorsgemma3VLMvideo-understandingimage-captioninggemmajson-modestructured-outputvideo-analysisimage-text-to-textconversationalenlicense:apache-2.0model-indexcompressed-tensorsregion:us
The README has not been fetched yet (metadata-first ingest). The model page and the metadata below are live.

Model card

Mirrored from the Hugging Face Hub and served from the Conceptio Open Knowledge Archive. Read the original card at https://huggingface.co/inference-net/ClipTagger-12b.