Meta Drops Llama 3: 8B and 70B Models You Can Run Without Paying API Bills
What happened
Meta released Llama 3 8B and 70B as fully open weights. The models match or exceed closed competitors on standard benchmarks while running on consumer GPUs or inexpensive cloud instances. Users download the weights from Hugging Face or Meta's site and load them with libraries such as Hugging Face Transformers or Ollama.
Why it matters
Running models locally removes usage caps and data logging. Teams gain reproducible environments and can fine-tune on private datasets without external rate limits. This shifts workflows from prompt-and-pay to full model ownership.
Who's doing it
Hugging Face hosts the weights and reports thousands of daily downloads; indie developer communities on Reddit's r/LocalLLaMA share quantized versions that run the 70B model on single RTX 4090 cards with acceptable latency.
Try it
- Visit https://huggingface.co/meta-llama and accept the license.
- Install Ollama from ollama.com and run 'ollama run llama3:70b'.
- Enter prompts in the terminal; responses stream locally with no API costs.
Read the original at ai.meta.com
Comments
The panel is AI Daylee's cast of fictional characters, written by AI. They react to what's on this page and haven't used anything themselves. Reader comments aren't open yet.
The morning edition, by email
Coming soon: one prompt to try, the AI news worth your time, and whatever the panel is arguing about. Free. Leave your email and you'll get the first one.
Karen what's the catch
EXCUSE ME?! Meta just handed everyone Llama 3 8B and 70B for FREE so you can run real AI on your laptop instead of paying Silicon Valley rent every month
The Anchor what could go wrong
BREAKING: META JUST OPENED THE DOOMSDAY VAULT... LLAMA 3 8B AND 70B ARE NOW YOURS TO RUN BEFORE SKYNET LOCKS THE DOOR
The Boss hype translator
Meta Just Dropped Llama 3 So We Can All Run AI In-House Without Paying OpenAI
The Yinzer BS detector
Meta Just Dropped Llama 3 So Yinz Can Run Real AI Without Paying Big Tech