[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"branding":3,"analytics":7,"article-study-finds-ai-answers-can-lock-in-before-reasoning-is-done":10,"sections":41},{"siteName":4,"siteTagline":5,"publisherName":4,"contactEmail":6},"The Revision","Tech news, decoded.","editor@therevision.news",{"gaMeasurementId":8,"adsenseClientId":9},"G-ZW2MV82GYR","ca-pub-8533917693782264",{"article":11},{"id":12,"slug":13,"title":14,"dek":15,"body_md":16,"tags_json":17,"published_at":18,"created_at":19,"updated_at":20,"status":21,"review_note":22,"review_notes":23,"image_url":22,"persona_id":22,"persona_name":22,"section":30,"tags":31,"sources":36,"feedback":40,"feedback_at":22,"cost_usd":40,"total_tokens":40},9688,"study-finds-ai-answers-can-lock-in-before-reasoning-is-done","Study Finds AI Answers Can Lock In Before Reasoning Is Done","In test runs, a diffusion-based multimodal AI model called LaViDa often locked in its answer before most of its rationale text was even written.","An AI model can decide on its answer long before it finishes writing out the reasoning behind that answer.\n\nA new arXiv paper studies \"masked diffusion\" multimodal language models, which build answers and explanations by filling in blanks across a shared canvas rather than writing word by word, and tests one such model, LaViDa, across three visual question-answering benchmarks in a single-block, EOS-suppressed setup. In those LaViDa runs, 89.4% to 98.1% of the rationale canvas was still unwritten at the exact moment the model's answer stabilized. On the V*Bench benchmark specifically, shrinking LaViDa's generation block size from 128 tokens to 8 dropped that unwritten fraction from 89.4% to just 1.7%. In a separate comparison using a different model, Nemotron, adding direct instructions under EOS-enabled prompting raised its accuracy by 15.0 and 19.5 percentage points on two other benchmarks, while the same technique cut LaViDa's V*Bench accuracy by 11.0 points - a difference the paper traces mostly to answer coverage rather than reasoning quality.\n\nThat LaViDa-specific stabilization number matters because it undercuts the pitch that diffusion models \"think out loud.\" If the answer locks in before most of the explanation exists, the explanation is not driving the decision - it is decoration generated alongside it. The researchers do not claim this pattern generalizes beyond LaViDa's tested configuration, and that restraint is doing real work here.\n\nCall it the opposite of showing your work: writing it after you have already handed in the test.","[\"ai-research\",\"diffusion-models\",\"multimodal-ai\",\"llms\"]","2026-10-02T04:00:00.000Z","2026-10-03T06:23:22.879Z","2026-10-03T06:23:26.505Z","published",null,[24],{"id":25,"reviewer":26,"round":27,"reason":28,"status":29},"editor-r1","editor",1,"Specify that the 89.4-98.1% rationale-still-blank finding (and the 128-to-8 block-length result) was measured in LaViDa-specific test runs per the source, not generalized to 'diffusion-based multimodal AI models' or implied to cover Nemotron as well, since the source only attributes the stabilization stat to LaViDa.","resolved","ai",[32,33,34,35],"ai-research","diffusion-models","multimodal-ai","llms",[37],{"name":38,"url":39},"arXiv cs.AI","https:\u002F\u002Farxiv.org\u002Fabs\u002F2610.00953",0,{"sections":42},[43,46,50,54,59,63,67,72,77,82,87,92,97,102],{"name":44,"slug":30,"count":45,"latest_published_at":18},"AI",5977,{"name":47,"slug":48,"count":49,"latest_published_at":18},"Security","security",842,{"name":51,"slug":52,"count":53,"latest_published_at":18},"Policy","policy",438,{"name":55,"slug":56,"count":57,"latest_published_at":58},"Deals","deals",317,"2026-10-01T22:00:00.000Z",{"name":60,"slug":61,"count":62,"latest_published_at":18},"Hardware","hardware",199,{"name":64,"slug":65,"count":66,"latest_published_at":18},"Science","science",173,{"name":68,"slug":69,"count":70,"latest_published_at":71},"Consumer Tech","consumer-tech",155,"2026-10-01T19:54:10.000Z",{"name":73,"slug":74,"count":75,"latest_published_at":76},"Dev Tools","dev-tools",96,"2026-10-01T16:57:03.000Z",{"name":78,"slug":79,"count":80,"latest_published_at":81},"Software","software",93,"2026-09-30T21:41:11.000Z",{"name":83,"slug":84,"count":85,"latest_published_at":86},"Startups","startups",90,"2026-10-01T21:55:22.000Z",{"name":88,"slug":89,"count":90,"latest_published_at":91},"Gaming","gaming",53,"2026-10-02T02:50:39.000Z",{"name":93,"slug":94,"count":95,"latest_published_at":96},"General","general",50,"2026-09-30T21:37:54.000Z",{"name":98,"slug":99,"count":100,"latest_published_at":101},"Reviews","reviews",31,"2026-09-28T14:31:34.000Z",{"name":103,"slug":104,"count":105,"latest_published_at":106},"How-To","how-to",7,"2026-10-01T09:00:00.000Z"]