Home/Models/SmolVLM2 500M Video Instruct

SmolVLM2 500M Video Instruct

by Hugging Face TB · Released Feb 2025

Open Source
Compare
Context
N/A
Input $/1M
n/a
Output $/1M
n/a
Type
image-text-to-text
License
Open Source
Benchmarks
0 tested
Data as of
About

HuggingFaceTB image text to text model. 294K downloads on HuggingFace.

No benchmark data available yet.

Links
Documentation
BenchGecko API
huggingfacetb-smolvlm2-500m-video-instruct
Specifications
  • Typeimage-text-to-text
  • ContextN/A
  • ReleasedFeb 2025
  • LicenseOpen Source
  • StatusActive
Available On
Hugging Face TBn/a
Share & Export
Tweet
SmolVLM2 500M Video Instruct is an open-source image-text-to-text AI model by Hugging Face TB, released in February 2025.

Key facts · as of 2026-04-09

  • SmolVLM2 500M Video Instruct by Hugging Face TB. Not enough public benchmark scores to rank yet.
  • List price n/a input · n/a output per 1M tokens (as of 2026-04-09).

How to cite · data as of 2026-04-09

SmolVLM2 500M Video Instruct · benchmarks, pricing and providers. BenchGecko, data as of 2026-04-09. https://benchgecko.ai/model/huggingfacetb-smolvlm2-500m-video-instruct

Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP