Meta Releases Llama 3.3 70B — 405B Performance at a Fraction of the Cost
In one sentence Llama 3.3 70B matches Llama 3.1 405B on most benchmarks while requiring 6x less compute, with 128K context and Apache 2.0 license — redefining the default open enterprise model.
Meta surprised the AI community by releasing a 70 billion parameter model that performs nearly as well as their much larger 405 billion parameter model on most tasks. Think of it like a sports car that goes almost as fast as a supercar but uses six times less fuel. This matters enormously for businesses and developers: you can now run a top-tier open model on much cheaper hardware, or serve far more users for the same cost. The 128,000-token context window means it can read and reason over long documents in one go. The Apache 2.0 license means anyone — including companies — can use, modify, and deploy it freely without restrictions. It quickly became the default choice for enterprise teams wanting powerful open AI without the infrastructure cost of the full 405B model.
Companies
Meta
Tools
Llama 3.3, Ollama
Tags
Sources