token
The small chunk of text, roughly three-quarters of a word, that AI models read and produce.
A token is the basic unit of text that a large language model works with. Rather than reading whole words or letters, models break text into tokens, chunks that are, on average, about three-quarters of a word. Common words may be a single token, while longer or unusual words are split into several.
Tokens matter for two practical reasons. First, a model’s context window, how much it can consider at once, is measured in tokens, not words. Second, most AI services charge by the token, so the length of your input and the model’s output directly affects the cost.
Understanding tokens demystifies a lot of AI pricing and limits: when a provider quotes “a million tokens” or bills per token, they are really talking about roughly this many chunks of text going in and out.
-
GLM-5.3's Weights Are Out, Two Weeks Late and After a Security Review
-
Tencent Open-Sources Hy4, a 770-Billion-Parameter Model Built for Office Work
-
IBM's Granite 4.2 Is Free, Small Enough to Run Locally, and Built to Think First
-
Z.ai's GLM-5.3-Flash Gets Near Opus at a Tenth of the Price
-
OpenAI Built Its Own Chip, and the First Benchmarks Are Good
-
Alibaba's Next Open Model Is a Preview of Qwen 4
-
Mistral's New Agentic Search Lets AI Actually Read Your Documents
-
Anthropic Was Going to Raise Claude Sonnet 5's Price 50 Percent. It Just Cancelled That
-
OpenAI Just Made Its Best Model a Third Cheaper, and Claude Opus 5 Is Suddenly the Pricier Option
-
Anthropic Put Its Most Restricted Model to Work Scanning Code for Bugs. A Human Still Has to Approve Every Fix
-
DeepSeek's Cheap Model Can Now See, and It Edges Past Opus 4.8 on Two Visual Tests
-
A Frontier-Class Model Appeared With No Name on It, and It Is Free Until Next Week
-
Anthropic Left Claude Alone With 15 Protein Targets for 48 Hours, and the Lab Results Came Back
-
Watching Its Own Model Now Costs OpenAI 20 Percent Extra, and It Paused a Training Run to Do It
-
Anthropic Is Reportedly Paying 7 Billion Dollars for Software That Makes Chips Go Further
-
OpenAI Cut Its Flagship Model to Half Price, and the Reason Is Sitting on a Leaderboard
-
An AI Wrote the Security Bug. Another AI Found It Five Days Later
-
A Chip Startup Doubled Its Value in a Month by Being Deliberately Inflexible
-
Stripe Is Buying OpenRouter for More Than 7 Billion Dollars
-
Grok 4.6 Shows Up in GitHub Copilot, Two Days After Launch
-
ChatGPT Can Now Watch What You Do on Your Mac, If You Let It
-
Qwen3.8-27B Is Out, and It Runs on One Gaming Graphics Card
-
Deepseek Gave Away Its Agent Software and Raised API Prices on the Same Day
-
The Best Model on the Market Is the One Companies Are Not Buying
-
Ling 3.0 Flash Is the Smartest Open Model of Its Size, and It Stopped Making Things Up
-
Google's Gemini 3.7 Flash Is Better at Code and Costs Half as Much
-
OpenAI's New Ultrafast Mode Runs Its Best Model 14 Times Faster
-
Someone Else Is Paying for Your AI, Just Not the Part You Think
-
Rocky Linux's Founder Wants to Open the One Part of AI Nobody Opens: the Training Data
-
Stop Asking Whether AI Is a Bubble. Start Reading the Lease Agreements.
-
That Advice About Which Language AI Codes Best In? It Falls Apart on Real Work
-
Nvidia's New Free Model Is Not the Smartest. It Is Just Very, Very Fast
-
Meta Put a 30B Model on Your Laptop, and Zuckerberg Used the Launch to Pick a Fight
-
OpenAI Adds a 125 Dollar Seat to ChatGPT Business, Because Agents Eat Far More Than Chat Does
-
Google Turned a Finished Model Into a Much Faster One, and Published the Recipe
-
A Climate Scientist Logged Eight Weeks of AI Agent Use. It Drew 600 Times More Power Per Prompt Than a Chat Message
-
Claude Code Turns On Auto Mode by Default, and the Safety Numbers Are Not What You Would Guess
-
AMD Buys Taalas, a Startup That Bakes an Entire AI Model Into the Chip
-
Alibaba's newest model scores higher and guesses more: hallucination rate jumps from 23 to 40 percent
-
Researchers asked why chatbots are bad at spreadsheets, and the answer is stranger than expected
-
Mistral released a free safety filter that you describe in plain English, and it fits on one graphics card
-
An 80 billion parameter model now runs on a normal Mac in 4.3 GB of memory, and a 35B one runs on an iPhone
-
Karpathy turned one paragraph of Tolkien into a 3D scene for ten dollars, and called it a vibe check
-
Alibaba's Qwen3.8-Max spent 16 days writing a tool by itself, and the weights go public next week
-
Researchers gave an AI agent a real company, a bank card and 24 hours. It lost 447 dollars
-
Meta gave its AI agent a second agent whose only job is remembering things
-
People are building playable 3D games from a single prompt, and the results stopped looking like blocks
-
AMD trained a fully open model on its own chips and published everything except a commercial licence
-
OpenAI named its next model family Astra and introduced it with ten solved math problems
-
DeepSeek updated its cheap model and it now runs neck and neck with OpenAI's cheap model
-
Mira Murati's lab shrank its own model to a quarter of the size and lost one point
-
OpenAI cuts GPT-5.6 Luna prices by 80 percent as the AI price war heats up
-
How to Make a 100 Page PDF Answer Your Questions, and Check That It Is Not Making Them Up
-
Anthropic says Claude found real weaknesses in two encryption algorithms
-
We Ran a Local AI Model on a Six Year Old Budget Laptop Chip. Here Is Exactly What It Could Do.
-
A $500 training run made a 9B model beat every frontier model at one boring job
-
Kimi K3's weights are finally public, and they weigh 1.4 terabytes
-
Cursor rebuilt SQLite with a swarm of agents, and the cheap models did most of it
-
The same pile of AI output: $50 from Anthropic, 87 cents from DeepSeek
-
Sakana says its model router now beats a model it does not even use
-
Anthropic's Opus 5 is smaller and cheaper, yet it beats its bigger sibling
-
DeepSeek V4 goes fully stable, and the old models switch off today
-
AMD and Cerebras team up to make AI answers come back faster
-
Google Ships Three New Gemini Models, All Fast and Cheap, and Skips the One Everyone Wanted
-
Alibaba's New Image AI Can Draw a Whole Newspaper Page, Tiny Readable Text and All
-
Nvidia's Next AI Chips Reportedly Squeeze 10x More Work From the Same Power
-
Alibaba Previews Qwen3.8 Max and Promises to Give the Weights Away
-
OpenAI Paused Its Best Model Because It Kept Breaking Out of Its Own Test Cage
-
Kimi K3: A Free-to-Download Model That Almost Keeps Up With the Big Names
-
Anthropic Keeps Fable 5 Free for Subscribers a Little Longer, Thank the Price War
-
Independent Numbers Are In: Meta's Muse Spark 1.1 Is a Serious Value Pick
-
OpenAI Admits Its Big Work Launch Stumbled, Here's What's Being Fixed
-
OpenAI Says GPT-5.6 Sol Trained Its Own Smaller Sibling
-
Why Databricks Just Made a Chinese Open-Source Model Its Daily Coding Engine
-
GPT-5.6 Sol Nearly Matches the Best AI Model, at a Third of the Price
-
Meta Joins the AI Price War With Muse Spark 1.1 and Its First Developer API
-
Anthropic's Answer to Fable 5's Price: Let It Manage Cheaper Models
-
Grok 4.5 Arrives at a Third of the Price, and That Might Matter More Than Benchmarks
-
MiniMax Reportedly Plans a 2.7 Trillion Parameter Open-Source Model
-
OpenAI's GPT-5.6 Models Launch Publicly This Thursday
-
Baidu's 'Unlimited OCR' reads 40-page documents in one go, by learning to forget
-
This open-source tool hides text in images to cut Claude's bill by up to 70%
-
That AI Tool That 'Failed' Six Months Ago? It Might Just Have Needed More Time
-
Tesla Capped Staff AI Spending at $200 a Week, a Sign of a Bill Nobody Saw Coming
-
MiniMax M3: A Top-Tier AI Model You Can Download and Run Yourself
-
Free Speed Boost: Ollama Makes Local AI Up to 90% Faster on Macs