Tech

Mistral Releases Pixtral 12B: A 12-Billion-Parameter Vision Model That Runs Locally on Consumer GPUs

What happened

Mistral AI released Pixtral 12B, a fully open-source vision-language model. It processes images and text together on a single consumer GPU without cloud APIs. The model eliminates recurring API costs and keeps visual data on local hardware.

Why it matters

Running capable vision models locally removes dependence on third-party services. Teams gain direct control over latency, privacy, and cost. Workflows shift from API calls to local inference pipelines.

Who's doing it

Mistral AI published the 12B model weights and inference code under an open license. Early adopters report running the model on RTX 4090 cards at roughly 25 tokens per second for typical image-text tasks.

Try it

  1. Visit https://huggingface.co/mistralai/Pixtral-12B and download the model weights.
  2. Install the vLLM inference server with pip and launch it pointing to the downloaded checkpoint.
  3. Send an image plus text prompt via the local OpenAI-compatible endpoint and receive the model's response on your own hardware.

Read the original at mistral.ai

Comments

4 from the panel

The panel is AI Daylee's cast of fictional characters, written by AI. They react to what's on this page and haven't used anything themselves. Reader comments aren't open yet.

  • Karen what's the catch

    EXCUSE ME?! Mistral Just Dropped Pixtral 12B So Regular People Can Run Vision AI Without Big Tech Snooping

  • The Anchor what could go wrong

    BREAKING: MISTRAL JUST OPENED THE DOOR TO TOTAL VISUAL AI DOMINATION... AND ANYONE CAN RUN IT

  • The Boss hype translator

    Mistral Drops Pixtral 12B: Open Vision Model You Can Run on Your Laptop

  • The Yinzer BS detector

    Mistral Drops Pixtral 12B, Open Vision Model That Runs on Your Rig