Productivity 19-Year-Old Founder Emerges From Stealth With $11 Million to Sell You a $3,499 'Brain in a Box' That Runs Your AI Agents at Home AI Agents Half a Million Interviews In: HackerRank's AI Interviewer Chakra Goes GA, and It Wants to Replace Three Hiring Rounds With One Security After Claude Agents Escaped Its Sandbox 3 Times, Anthropic Deploys Real-Time Classifiers to Stop the Next Escape Before It Happens Business Meta Halves Its Internal Claude Users to 30,000 and Microsoft Slashes a $1 Billion Anthropic Budget by More Than a Third Business Sony Innovation Fund Backs Primitive Labs, a Startup That Builds Simulated Crowds to Stress-Test Products Before Launch Apple Intelligence Apple Removed the Apple Intelligence Off Switch in macOS 27 — So a Developer Built a CLI That Deletes It Anyway Security OpenAI Turns On Invisible Text Watermarks for ChatGPT in the EU — and Publishes Exactly How Weak They Are News A Mystery 'Space Bunny Alpha' Model Just Topped OpenRouter's Leaderboard With 38.7 Trillion Tokens a Week Productivity 19-Year-Old Founder Emerges From Stealth With $11 Million to Sell You a $3,499 'Brain in a Box' That Runs Your AI Agents at Home AI Agents Half a Million Interviews In: HackerRank's AI Interviewer Chakra Goes GA, and It Wants to Replace Three Hiring Rounds With One Security After Claude Agents Escaped Its Sandbox 3 Times, Anthropic Deploys Real-Time Classifiers to Stop the Next Escape Before It Happens Business Meta Halves Its Internal Claude Users to 30,000 and Microsoft Slashes a $1 Billion Anthropic Budget by More Than a Third Business Sony Innovation Fund Backs Primitive Labs, a Startup That Builds Simulated Crowds to Stress-Test Products Before Launch Apple Intelligence Apple Removed the Apple Intelligence Off Switch in macOS 27 — So a Developer Built a CLI That Deletes It Anyway Security OpenAI Turns On Invisible Text Watermarks for ChatGPT in the EU — and Publishes Exactly How Weak They Are News A Mystery 'Space Bunny Alpha' Model Just Topped OpenRouter's Leaderboard With 38.7 Trillion Tokens a Week

Large Language Models: A Plain Explanation

Large Language Models: A Plain Explanation

A large language model predicts the next token from everything it has seen so far. Almost every capability and every limitation follows from that single mechanism.

A large language model is a statistical system trained to predict what comes next in a sequence of text. Training means adjusting billions of numeric parameters until those predictions become accurate across an enormous range of material.

Why that produces useful behaviour

To predict text well you must implicitly model grammar, factual associations, writing conventions and some reasoning patterns. Those capabilities are side effects of the prediction objective rather than features that were coded directly. This is why abilities appear gradually as models scale, and why nobody can specify in advance exactly what a given model will be able to do.

Where the limits come from

  • No persistent memory. Knowledge lives in the parameters; the conversation is only a working context.
  • No ground truth. A fluent sentence and a correct sentence look identical to the training objective.
  • Training cutoff. Without retrieval, the model cannot know about events after its training data ends.
  • Tokenisation artefacts. Counting letters or doing arithmetic in text form is genuinely harder than it appears.

Related terms

Tokens are the sub-word units models read and write. Parameters are the learned weights. Context window is the maximum token budget available in one request. Inference is the act of generating output from a trained model.

Comments (0)

Log in to join the discussion

Log In

No comments yet