[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"branding":3,"analytics":7,"article-ai-model-merging-method-nearly-matches-specialist-models":10,"sections":44},{"siteName":4,"siteTagline":5,"publisherName":4,"contactEmail":6},"The Revision","Tech news, decoded.","editor@therevision.news",{"gaMeasurementId":8,"adsenseClientId":9},"G-ZW2MV82GYR","ca-pub-8533917693782264",{"article":11},{"id":12,"slug":13,"title":14,"dek":15,"body_md":16,"tags_json":17,"published_at":18,"created_at":19,"updated_at":20,"status":21,"review_note":22,"review_notes":23,"image_url":22,"persona_id":22,"persona_name":22,"section":34,"tags":35,"sources":39,"feedback":43,"feedback_at":22,"cost_usd":43,"total_tokens":43},10115,"ai-model-merging-method-nearly-matches-specialist-models","AI model merging method nearly matches specialist models","A new method called ReForge merges several specialist AI vision models into one with almost no accuracy loss, though language merging still lags behind.","A new technique called ReForge closes most of the gap between merged AI models and the specialist models they are built from, without extra training.\n\nResearchers describe ReForge as a bilevel optimization framework that treats merging as Bayesian linear regression anchored to a strong prior model, with an inner step producing a closed-form estimate from unlabeled calibration data and an outer step tuning regularization and scaling via Bayesian optimization on a validation set. A data-free variant swaps the calibration activations for task-vector statistics, so it needs no extra data at all. Tested on merges of up to 20 tasks in vision and 5 tasks in language, ReForge beat baseline methods including TA, WUDI-Merging, and TSV across the board. On a 20-task ViT-B\u002F32 vision benchmark, it pushed the previous best method from 77.6% to 82.8% accuracy with calibration data, and to 81.5% without it.\n\nOn an eight-task ViT-L\u002F14 vision benchmark, the calibration-assisted version hit 95.1% mean accuracy, within a percentage point of the 95.8% you get from keeping each task's specialist model separate. That is a meaningful result: merging several specialist vision models into one no longer costs much accuracy, which matters for anyone trying to avoid training and serving a pile of separate models. Language merging, by contrast, only went up to 5 tasks in this paper, suggesting the technique is further along for vision than for text.\n\nThe code has not shipped yet, so nobody outside the authors' own benchmarks can confirm the gains hold up.","[\"ai\",\"model-merging\",\"machine-learning\",\"computer-vision\"]","2026-10-05T04:00:00.000Z","2026-10-05T23:52:42.465Z","2026-10-05T23:52:48.312Z","published",null,[24,30],{"id":25,"reviewer":26,"round":27,"reason":28,"status":29},"editor-r1","editor",1,"The 77.6%-to-82.8% figure is explicitly attributed in the source to beating the ISO-CTS baseline specifically, not TA\u002FWUDI-Merging\u002FTSV as the draft implies — fix the attribution (name ISO-CTS as the baseline for that stat, or rephrase so the named baselines aren't credited with a number that's actually about a different, undefined baseline).","resolved",{"id":31,"reviewer":26,"round":32,"reason":33,"status":29},"editor-r2",2,"The eight-task 95.1%-vs-95.8% comparison is sourced to ViT-L\u002F14, a vision model, not a 'language setup' as the draft states — and the source caps language merging at 5 tasks — so fix the domain attribution to vision.","ai",[34,36,37,38],"model-merging","machine-learning","computer-vision",[40],{"name":41,"url":42},"arXiv cs.AI","https:\u002F\u002Farxiv.org\u002Fabs\u002F2605.12843",0,{"sections":45},[46,50,54,59,64,69,73,78,83,88,93,98,103,108],{"name":47,"slug":34,"count":48,"latest_published_at":49},"AI",6317,"2026-10-05T09:51:57.000Z",{"name":51,"slug":52,"count":53,"latest_published_at":18},"Security","security",871,{"name":55,"slug":56,"count":57,"latest_published_at":58},"Policy","policy",446,"2026-10-05T10:25:00.000Z",{"name":60,"slug":61,"count":62,"latest_published_at":63},"Deals","deals",340,"2026-10-05T09:18:03.000Z",{"name":65,"slug":66,"count":67,"latest_published_at":68},"Hardware","hardware",205,"2026-10-05T10:58:22.000Z",{"name":70,"slug":71,"count":72,"latest_published_at":18},"Science","science",179,{"name":74,"slug":75,"count":76,"latest_published_at":77},"Consumer Tech","consumer-tech",160,"2026-10-05T10:23:15.000Z",{"name":79,"slug":80,"count":81,"latest_published_at":82},"Dev Tools","dev-tools",99,"2026-10-05T10:47:06.000Z",{"name":84,"slug":85,"count":86,"latest_published_at":87},"Software","software",97,"2026-10-04T10:00:00.000Z",{"name":89,"slug":90,"count":91,"latest_published_at":92},"Startups","startups",93,"2026-10-05T11:13:51.000Z",{"name":94,"slug":95,"count":96,"latest_published_at":97},"Gaming","gaming",53,"2026-10-02T02:50:39.000Z",{"name":99,"slug":100,"count":101,"latest_published_at":102},"General","general",51,"2026-10-05T02:35:01.000Z",{"name":104,"slug":105,"count":106,"latest_published_at":107},"Reviews","reviews",32,"2026-10-02T18:00:00.000Z",{"name":109,"slug":110,"count":111,"latest_published_at":112},"How-To","how-to",8,"2026-10-05T09:00:00.000Z"]