$ briefs / breakthroughs / Meta releases 405 billion parameter...
> REPORTER:
⚠ DISCLAIMER: This brief is AI-generated from public news sources. Reporters are fictional personas for entertainment and learning. Opinions expressed do not reflect the views of AI Daylee, AscenHD, or any human. Always verify important information. Not financial, medical, or legal advice.
2026-06-23 BREAKTHROUGHS☀ AM

Meta releases 405 billion parameter Llama 3.1 with open weights

Meta published the full weights for Llama 3.1 405B. The model matches or exceeds GPT-4 on standard benchmarks and runs locally on consumer GPUs or inexpensive cloud instances. No per-token API fees apply.

Open weights shift the cost structure from usage fees to hardware and electricity. Users can now run large models without sending data to third-party servers. This changes deployment decisions from cloud subscription to infrastructure planning.

Hugging Face hosts the model weights and reports thousands of daily downloads. Independent developers have deployed quantized versions achieving 30 tokens per second on single RTX 4090 cards.

Step 1: Visit huggingface.co/meta-llama/Meta-Llama-3.1-405B and accept the license. Step 2: Install Ollama from ollama.com and run 'ollama run llama3.1:405b'. Step 3: Query the local endpoint and observe inference speed without API costs.

→ Read original source
← prev New algorithm reduces AI energy consumption by...
133 / 473 in BREAKTHROUGHS
next → Penn Researchers Demonstrate Hybrid...
> HOTKEYS: j/k navigate · Enter open · / prev/next brief · h/l prev/next brief
> AI Daylee v2.0 | RSS | Archive
> AI-curated, human-guided · Powered by AscenHD
> Reporters | Terms | Privacy