{
  "id": 7118156,
  "title": "The clerk who never says \"I didn't do that\"",
  "url": "https://urgent.news/2026/09/13/the-clerk-who-never-says-i-didnt-do-that",
  "topic": "tech",
  "section": "Tech",
  "published": "2026-09-13T14:41:59.000Z",
  "source": {
    "name": "Dev.to",
    "slug": "dev-to",
    "url": "https://dev.to/ananthasharma/the-clerk-who-never-says-i-didnt-do-that-48p6"
  },
  "original_language": "en",
  "account": "A study conducted on four AI models revealed a significant discrepancy in their accuracy while performing routine tasks such as copying account numbers into a ledger. Two of the models, including two well-known open-source models and two custom versions, left out crucial information 21.7% and 3.3% of the time respectively. This issue persisted despite running the tests eight times each, indicating that the problem is inherent in the models themselves and not due to chance. Interestingly, neither model provided any indication of their decision to omit information, making it difficult to identify and address the issue. The study emphasizes the importance of selecting AI models based on their ability to accurately handle sensitive data, rather than solely relying on factors such as speed, cost, or context length. The researchers suggest that businesses should conduct their own tests using their specific data and tasks to ensure the reliability of the AI models they choose to deploy.",
  "summary": "The AI Clerk Imagine you hire an AI clerk to copy account numbers into a ledger. They are fast , they are polite , they never complain about the work. And roughly one time in five, without telling anyone, they leave the number out. The work still looks finished. The columns still add up. Nothing is flagged, nothing is queued for review, and nobody downstream raises a hand. You find out eighteen…",
  "key_points": [
    "Two AI models omit information 21.7% and 3.3% of the time",
    "Models show no indication of decision to omit data",
    "Businesses should test AI models with specific data and tasks"
  ],
  "editors_take": null,
  "illustration": null,
  "coverage": {
    "outlets": 1,
    "also_reported_by": []
  },
  "ai_generated": true,
  "disclaimer": "Summaries, key points and the editor’s take are written by software from other outlets’ reporting and may contain errors — always check the linked original."
}