[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"branding":3,"analytics":7,"article-llms-calculate-their-confidence-while-answering":10,"sections":36},{"siteName":4,"siteTagline":5,"publisherName":4,"contactEmail":6},"The Revision","Tech news, decoded.","editor@therevision.news",{"gaMeasurementId":8,"adsenseClientId":9},"G-ZW2MV82GYR","ca-pub-8533917693782264",{"article":11},{"id":12,"slug":13,"title":14,"dek":15,"body_md":16,"tags_json":17,"published_at":18,"created_at":19,"updated_at":20,"status":21,"review_note":22,"review_notes":23,"image_url":24,"persona_id":22,"persona_name":22,"section":25,"tags":26,"sources":31,"feedback":35,"feedback_at":22,"cost_usd":35,"total_tokens":35},7111,"llms-calculate-their-confidence-while-answering","LLMs Calculate Their Confidence While Answering","A new study finds LLMs compute confidence scores during answer generation and cache them, rather than inventing a number when asked.","Ask a chatbot how confident it is in an answer, and it turns out that number isn't improvised on the spot.\n\nA new interpretability study traces exactly how large language models generate the confidence scores they report when prompted to rate their own answers. Researchers tested Gemma 3 27B on trivia, math, and general-knowledge benchmarks, plus Qwen 2.5 7B and the reasoning model Magistral Small 24B. Using techniques like activation steering, patching, and attention blocking, they found confidence signals form immediately after the model writes its answer, get pulled from the answer's own tokens, and are cached at the first token slot after the response. When the model is later asked to state a confidence number, it retrieves that cached value rather than computing something fresh.\n\nThis matters because verbal confidence - a model literally saying \"I'm 80% sure\" - is one of the few tools researchers have for gauging whether a black-box model's output can be trusted. If that number were just a guess dressed up after the fact, it would be close to useless. Instead, the cached confidence signal explains meaningfully more of the variation in reported confidence than simple word-prediction probability (token log-probabilities) does, suggesting the model runs a genuine, if crude, self-check on answer quality.\n\nNone of this means the scores are accurate - a model can be confidently wrong, same as anyone. But it does mean that when an AI model states a confidence level, something resembling real evaluation happened under the hood, not a number pulled from thin air to sound rigorous.","[\"llm interpretability\",\"model calibration\",\"ai research\",\"gemma\"]","2026-09-21T04:00:00.000Z","2026-09-21T07:05:00.687Z","2026-09-21T07:05:12.854Z","published",null,[],"https:\u002F\u002Fcdn.xyz.onl\u002Farticle-images\u002Fllms-calculate-their-confidence-while-answering.webp","ai",[27,28,29,30],"llm interpretability","model calibration","ai research","gemma",[32],{"name":33,"url":34},"arXiv cs.AI","https:\u002F\u002Farxiv.org\u002Fabs\u002F2603.17839",0,{"sections":37},[38,42,46,51,56,61,66,71,76,81,86,91,96,101],{"name":39,"slug":25,"count":40,"latest_published_at":41},"AI",4175,"2026-09-21T10:30:00.000Z",{"name":43,"slug":44,"count":45,"latest_published_at":18},"Security","security",681,{"name":47,"slug":48,"count":49,"latest_published_at":50},"Policy","policy",352,"2026-09-21T10:18:06.000Z",{"name":52,"slug":53,"count":54,"latest_published_at":55},"Deals","deals",184,"2026-09-21T10:18:31.000Z",{"name":57,"slug":58,"count":59,"latest_published_at":60},"Hardware","hardware",157,"2026-09-21T11:04:12.000Z",{"name":62,"slug":63,"count":64,"latest_published_at":65},"Science","science",130,"2026-09-20T13:48:11.000Z",{"name":67,"slug":68,"count":69,"latest_published_at":70},"Consumer Tech","consumer-tech",99,"2026-09-09T17:27:33.000Z",{"name":72,"slug":73,"count":74,"latest_published_at":75},"Dev Tools","dev-tools",78,"2026-09-18T04:00:00.000Z",{"name":77,"slug":78,"count":79,"latest_published_at":80},"Software","software",75,"2026-09-10T20:41:21.000Z",{"name":82,"slug":83,"count":84,"latest_published_at":85},"Startups","startups",55,"2026-09-09T23:14:29.000Z",{"name":87,"slug":88,"count":89,"latest_published_at":90},"Gaming","gaming",43,"2026-09-10T12:18:06.000Z",{"name":92,"slug":93,"count":94,"latest_published_at":95},"General","general",42,"2026-09-18T22:35:10.000Z",{"name":97,"slug":98,"count":99,"latest_published_at":100},"Reviews","reviews",20,"2026-06-24T12:00:01.000Z",{"name":102,"slug":103,"count":104,"latest_published_at":105},"How-To","how-to",6,"2026-06-16T09:00:00.000Z"]