{
  "id": 2787525,
  "title": "Calibration Is Bet Sizing",
  "url": "https://urgent.news/2026/08/23/calibration-is-bet-sizing",
  "topic": "ai",
  "section": "AI",
  "published": "2026-08-23T12:19:16.000Z",
  "source": {
    "name": "Dev.to",
    "slug": "dev-to",
    "url": "https://dev.to/romiteld/calibration-is-bet-sizing-9cm"
  },
  "original_language": "en",
  "account": "The final section of the audit covers the reason behind selective rollout of the calibration system. Six of the seven assets improved calibration, while one asset, LINK-USD, did not see any improvement and even underperformed in terms of calibration error and log-loss. The audit concludes that the calibration system should be selectively rolled out, with LINK remaining on raw softmax until further investigation. The mechanism for implementing this selective rollout is described as boring and straightforward, involving setting up an allow-list of assets to use the calibrated model and excluding LINK if necessary.",
  "summary": "The last post was about making a number trustworthy. Leakage geometry, purge widths, de-overlap, a baseline that could not cheat. It ended with a minute-scale ceiling that held at 52% across seven configurations and a model family swap. This one is about what happens after you trust the number. Because a probability you are going to bet on is a different object from a probability you are going to…",
  "key_points": [
    "Six assets improved calibration with new system",
    "LINK-USD did not see improvement, underperformed",
    "Selective rollout involves allow-list of assets"
  ],
  "editors_take": null,
  "illustration": null,
  "coverage": {
    "outlets": 1,
    "also_reported_by": []
  },
  "ai_generated": true,
  "disclaimer": "Summaries, key points and the editor’s take are written by software from other outlets’ reporting and may contain errors — always check the linked original."
}