Small Language Models Are the New Rage, Researchers Say

The original version of this story appeared in Quanta Magazine.

Large language models work well because they’re so large. The latest models from OpenAI, Meta, and DeepSeek use hundreds of billions of “parameters”—the adjustable knobs that determine connections among data and get tweaked during the training process. With more parameters, the models are better able to identify patterns and connections, which in turn makes them more powerful and accurate.

But this power comes at a cost. Training a model with hundreds of billions of parameters takes huge computational resources. To train its Gemini 1.0 Ultra model, for example, Google reportedly spent $191 million. Large language models (LLMs) also require considerable computational

Related News

Prediction Markets Let You Bet on Whether a Wildfire Will Burn Down Your Town

What Are Fish Oil Supplements Good For? Here’s Your Crash Course

Workers claim unsafe conditions at a restaurant owned by the South Park creators. They have Brooke Shields on their side

Trump Accounts are now live. Here’s what you need to know

How I Went From Side Hustle to 7 Figures in 12 Months Using 4 AI Tools (No Employees, No Investors)

AI Can Do a Lot — But Most Companies Don’t Want It Talking to Their Clients. Here’s Why.