Skip to main content

Envisioning is an emerging technology research institute and advisory.

LinkedInInstagramGitHub

2011 — 2026

research
  • Observatory
  • Newsletter
  • Methodology
  • Origins
  • Vocab
  • RSS Feeds
services
  • Signals Session
  • Bespoke Projects
  • Build Sessions
  • Pricing
  • Use Cases
  • Signals
  • Signal Scan↗free
impact
  • ANBIMAFuture of Brazilian Capital Markets
  • IEEECharting the Energy Transition
  • Horizon 2045Future of Human and Planetary Security
  • WKOTechnology Scanning for Austria
solutions
  • Innovation
  • Strategy
  • Consultants
  • Foresight
  • Associations
  • Governments
  • L&D
resources
  • Partners
  • Coding for Non-Coders
  • How We Work
  • Data Visualization
  • Multi-Model Method
  • FAQ
  • Security & Privacy
  • Public Sector
about
  • Manifesto
  • Community
  • Events
  • Support
  • Contact
ResearchServicesSignalsAbout
ResearchServicesSignalsAbout
  1. Home
  2. Vocab
  3. LLM Watermarking

LLM Watermarking

Embedding imperceptible statistical signals in text generated by large language models so the output can be identified as machine-produced without changing what the text says.

Year: 2022Generality: 620Added: Aug 11, 2026
Back to Vocab

LLM watermarking is a class of techniques that embed an imperceptible statistical signature in text generated by large language models, so that the output can be reliably identified as machine-produced by a downstream detector without altering meaning, quality, or readability. Modern LLM watermarking approaches typically bias the model's token-selection process during decoding — for example, by partitioning the vocabulary into "green" and "red" token groups using a cryptographic key and steering sampling toward green tokens — so that green-token patterns in the output betray the model's identity. The signature travels with the text when it is copied or lightly edited, and can be detected by anyone holding the corresponding key. LLM watermarking is distinct from older text-watermarking approaches for human-authored documents, and from cryptographic signing of the model's response (which proves identity at the message level but does not survive copying or paraphrasing). Regulatory pressure — notably the EU AI Act Article 50(2) Code of Practice on Transparency of AI-Generated Content — has made production-grade LLM watermarking a near-term deployment requirement for major model providers.

Sources

  1. How Claude marks AI-generated content

    Anthropic Help Center · Aug 10, 2026

  2. A Watermark for Large Language Models

    arXiv (ICML 2023) · Jan 24, 2023

  3. Watermarking Makes Language Models Radioactive

    arXiv · Feb 22, 2024

Research this in Signals

Scan LLM Watermarking for yourself.

Signals turns a topic into a sourced research record you can inspect and rerun. Your first scan is free, and this one starts with LLM Watermarking already loaded, so edit it or scan as is.