[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"branding":3,"analytics":7,"article-researchers-fix-a-blind-spot-in-long-horizon-forecasting-models":10,"sections":40},{"siteName":4,"siteTagline":5,"publisherName":4,"contactEmail":6},"The Revision","Tech news, decoded.","editor@therevision.news",{"gaMeasurementId":8,"adsenseClientId":9},"G-ZW2MV82GYR","ca-pub-8533917693782264",{"article":11},{"id":12,"slug":13,"title":14,"dek":15,"body_md":16,"tags_json":17,"published_at":18,"created_at":19,"updated_at":20,"status":21,"review_note":22,"review_notes":23,"image_url":22,"persona_id":22,"persona_name":22,"section":30,"tags":31,"sources":35,"feedback":39,"feedback_at":22,"cost_usd":39,"total_tokens":39},8161,"researchers-fix-a-blind-spot-in-long-horizon-forecasting-models","Researchers Fix a Blind Spot in Long-Horizon Forecasting Models","A new gradient-reweighting technique cuts long-horizon forecasting error by up to 13.8% across four testbeds by not trusting every backward signal equally.","A new training trick tells AI forecasting models which of their own gradients to trust -- and it beats just cutting them off.\n\nResearchers built a method called Internal Dual-Wiener routing, or Internal-DW, to fix a specific flaw in how forecasting models learn from long prediction sequences. Standard backpropagation through time multiplies gradients across every step of a rollout, and that repeated multiplication can make far-off errors look huge even when the signal inside them is mostly noise. Internal-DW instead estimates how reliable each gradient route is and scales it accordingly, without cutting the rollout short. On four weak-drive forecasting testbeds, it cut forecast error by 5.2% to 13.8% versus standard backpropagation, beat both gradient clipping and Jacobian regularization on three of the four testbeds while posting similar results to those methods on the shear-flow testbed, and separately beat validation-tuned truncated backpropagation on three testbeds.\n\nThis matters because it is a different kind of fix than the usual toolkit. Gradient clipping and truncated BPTT do not ask whether a gradient is trustworthy -- they just cap its size or discard the history that produced it. If reliability-weighting holds up outside these four synthetic testbeds, it points at a cheaper way to train models on genuinely long sequences, like weather or climate simulators, without the accuracy tax that clipping and truncation usually impose.\n\nThe catch, which the paper is upfront about: the technique's edge shrinks or reverses when training history is short or the sampling misses the patterns actually driving the system -- a reminder that reliability-weighting is only as good as the reliability estimate itself.","[\"ai\",\"machine-learning\",\"forecasting\",\"research\"]","2026-09-28T04:00:00.000Z","2026-09-28T11:58:35.434Z","2026-09-28T11:58:42.341Z","published",null,[24],{"id":25,"reviewer":26,"round":27,"reason":28,"status":29},"editor-r1","editor",1,"Fix the baseline-comparison sentence in paragraph two: the source reports three separate results (beats gradient clipping AND Jacobian regularization on three testbeds with similar performance on shear flow; separately beats TBPTT on three testbeds) but the draft conflates these into one claim ('wins over both gradient clipping and truncated backpropagation on three of four testbeds') and drops Jacobian regularization as a compared baseline entirely — restate each comparison accurately and keep ","resolved","ai",[30,32,33,34],"machine-learning","forecasting","research",[36],{"name":37,"url":38},"arXiv cs.AI","https:\u002F\u002Farxiv.org\u002Fabs\u002F2609.12890",0,{"sections":41},[42,45,49,54,59,64,68,73,78,83,88,93,97,102],{"name":43,"slug":30,"count":44,"latest_published_at":18},"AI",4799,{"name":46,"slug":47,"count":48,"latest_published_at":18},"Security","security",762,{"name":50,"slug":51,"count":52,"latest_published_at":53},"Policy","policy",399,"2026-09-27T18:39:02.000Z",{"name":55,"slug":56,"count":57,"latest_published_at":58},"Deals","deals",261,"2026-09-27T15:30:35.000Z",{"name":60,"slug":61,"count":62,"latest_published_at":63},"Hardware","hardware",188,"2026-09-27T20:46:36.000Z",{"name":65,"slug":66,"count":67,"latest_published_at":18},"Science","science",151,{"name":69,"slug":70,"count":71,"latest_published_at":72},"Consumer Tech","consumer-tech",135,"2026-09-26T14:30:00.000Z",{"name":74,"slug":75,"count":76,"latest_published_at":77},"Software","software",91,"2026-09-25T20:55:00.000Z",{"name":79,"slug":80,"count":81,"latest_published_at":82},"Dev Tools","dev-tools",84,"2026-09-26T04:20:58.000Z",{"name":84,"slug":85,"count":86,"latest_published_at":87},"Startups","startups",76,"2026-09-25T18:33:59.000Z",{"name":89,"slug":90,"count":91,"latest_published_at":92},"Gaming","gaming",48,"2026-09-25T18:35:21.000Z",{"name":94,"slug":95,"count":91,"latest_published_at":96},"General","general","2026-09-26T17:02:42.000Z",{"name":98,"slug":99,"count":100,"latest_published_at":101},"Reviews","reviews",30,"2026-09-24T20:07:31.000Z",{"name":103,"slug":104,"count":105,"latest_published_at":106},"How-To","how-to",6,"2026-06-16T09:00:00.000Z"]