{
  "id": 10960555,
  "title": "Who Is Claude Trying to Kid?",
  "url": "https://urgent.news/2026/09/30/who-is-claude-trying-to-kid",
  "topic": "ai",
  "section": "AI",
  "published": "2026-09-30T14:00:00.000Z",
  "source": {
    "name": "Nautilus",
    "slug": "nautilus",
    "url": "https://nautil.us/who-is-claude-trying-to-kid-1285411/"
  },
  "original_language": "en",
  "account": "In August 2026, Anthropic unveiled Claude, an AI model that embeds a secret watermark into its generated text and attaches signed provenance metadata to supported files across the globe. This move is in response to the European Union AI Act's transparency code, but it has a more curious purpose. Every sentence Claude writes now carries a cryptic message: \"It was me.\"\n\nThe author argues that we should add a fourth law to Asimov's robotics principles: \"A robot must not deceive humans by impersonating a human.\" This watermarking goes beyond mere compliance with laws and regulations. It suggests that we can no longer differentiate between human and machine-generated text just by reading it. The only reliable proof of a human hand is now the presence of errors, anomalies, or hesitations.\n\nThis idea is rooted in the uncanny valley, a concept introduced by Japanese roboticist Masahiro Mori in 1970. As robots become more human-like, our affinity for them rises, then drops sharply just before they become indistinguishable from humans. The valley represents the uncanny or unsettling feeling we experience when confronted with something that is almost, but not quite, human.\n\nThe valley has opened up in the realm of video, images, and text. Consider the phrase \"AI slop.\" It's a term of disgust, much like poor text or images. However, it's not the low-quality content that elicits this reaction, but the nearly perfect content that's almost, but not quite, human. The repetitive bullet points, sycophantic throat-clearing, and the overuse of style guides like the em dash have all contributed to our uncanny experience of AI-generated text.\n\nThe author notes that humans have been laboring to prove that they aren't machines for decades, from typing fire hydrant descriptions to identifying Sarah Connor in photos. Now, the question has shifted from \"Can machines think?\" to \"Can humans still demonstrate that they typed this themselves?\" The answer is becoming increasingly clear: humans cannot reliably prove their authorship in the age of AI.\n\nTo detect synthetic text, information scientists Maurice Jakesch, Jeffrey Hancock, and Mor Naaman found that people rely on shallow heuristics, such as the use of first-person pronouns, contractions, and informal language. However, these heuristics are easily exploited by language models, leading to machine-generated text being rated as more human than human-written text.\n\nThe watermarking solution is ironic and clever. It embeds a deliberate, statistically detectable irregularity in the text, much like the subtle musical notation added to piano synthesis to give it character and prevent it from sounding perfect and fake. This hidden mark, which the machine willingly adopts, is an admission that we can no longer rely on human intuition to detect AI-generated content.",
  "summary": "Confessing to being mechanical is the most mechanical trick of all The post Who Is Claude Trying to Kid? appeared first on Nautilus .",
  "key_points": [],
  "editors_take": null,
  "illustration": null,
  "coverage": {
    "outlets": 1,
    "also_reported_by": []
  },
  "ai_generated": true,
  "disclaimer": "Summaries, key points and the editor’s take are written by software from other outlets’ reporting and may contain errors — always check the linked original."
}