[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"branding":3,"analytics":7,"article-a-leaner-router-lets-ai-models-use-more-experts":10,"sections":41},{"siteName":4,"siteTagline":5,"publisherName":4,"contactEmail":6},"The Revision","Tech news, decoded.","editor@therevision.news",{"gaMeasurementId":8,"adsenseClientId":9},"G-ZW2MV82GYR","ca-pub-8533917693782264",{"article":11},{"id":12,"slug":13,"title":14,"dek":15,"body_md":16,"tags_json":17,"published_at":18,"created_at":19,"updated_at":20,"status":21,"review_note":22,"review_notes":23,"image_url":22,"persona_id":22,"persona_name":22,"section":30,"tags":31,"sources":36,"feedback":40,"feedback_at":22,"cost_usd":40,"total_tokens":40},8584,"a-leaner-router-lets-ai-models-use-more-experts","A Leaner Router Lets AI Models Use More Experts","A new router design lets AI models use far more experts without slowing down, though the paper does not publish specific benchmark numbers.","A new technique called MoRE lets AI models pack in dramatically more \"experts\" without paying the usual computational tax.\n\nMixture-of-experts models split their work across many specialized sub-networks, or \"experts,\" using a router to pick which ones process each token, and the industry trend has been toward more, smaller experts. The paper argues the standard router does not scale with that trend, since its cost grows with both hidden dimension and expert count, becoming the bottleneck once expert counts get large. MoRE compresses the router's weight matrix to a much lower rank, cutting that cost and, per the authors' math, permitting roughly h\u002Fr times more experts at the same compute budget - paired with a custom Triton kernel so the savings hold at inference, not just on paper. Tested on a synthetic phonebook-memorization task and on knowledge-intensive Q&A benchmarks after pretraining, MoRE reportedly improves results on both while matching baseline reasoning performance, though the abstract does not publish the specific benchmark scores or comparison figures behind that claim.\n\nRouter efficiency is a quietly important problem: as labs push toward hundreds of tiny experts per layer, a routing step whose cost scales linearly with expert count turns into real overhead at inference time. A cheaper router is the unglamorous kind of fix that can matter more for production costs than another benchmark headline, since it changes what's economically deployable rather than just what's trainable in a lab.\n\nCode is on GitHub, but without independently verified benchmark numbers, this reads as promising engineering rather than a proven performance leap.","[\"mixture-of-experts\",\"ai-research\",\"model-architecture\",\"llm-efficiency\"]","2026-09-30T04:00:00.000Z","2026-09-30T12:57:20.107Z","2026-09-30T12:57:26.278Z","published",null,[24],{"id":25,"reviewer":26,"round":27,"reason":28,"status":29},"editor-r1","editor",1,"Name the actual benchmarks and give the comparison figures behind 'improved results on knowledge-heavy Q&A benchmarks and phonebook memorization' (or explicitly note the paper's abstract doesn't disclose specific numbers) instead of asserting vague, unsupported performance gains.","resolved","ai",[32,33,34,35],"mixture-of-experts","ai-research","model-architecture","llm-efficiency",[37],{"name":38,"url":39},"arXiv cs.AI","https:\u002F\u002Farxiv.org\u002Fabs\u002F2609.36301",0,{"sections":42},[43,46,50,54,59,64,69,74,79,83,88,93,98,103],{"name":44,"slug":30,"count":45,"latest_published_at":18},"AI",5104,{"name":47,"slug":48,"count":49,"latest_published_at":18},"Security","security",785,{"name":51,"slug":52,"count":53,"latest_published_at":18},"Policy","policy",417,{"name":55,"slug":56,"count":57,"latest_published_at":58},"Deals","deals",284,"2026-09-29T21:00:00.000Z",{"name":60,"slug":61,"count":62,"latest_published_at":63},"Hardware","hardware",194,"2026-09-29T13:16:04.000Z",{"name":65,"slug":66,"count":67,"latest_published_at":68},"Science","science",154,"2026-09-28T13:19:18.000Z",{"name":70,"slug":71,"count":72,"latest_published_at":73},"Consumer Tech","consumer-tech",142,"2026-09-29T18:38:03.000Z",{"name":75,"slug":76,"count":77,"latest_published_at":78},"Software","software",91,"2026-09-25T20:55:00.000Z",{"name":80,"slug":81,"count":82,"latest_published_at":18},"Dev Tools","dev-tools",90,{"name":84,"slug":85,"count":86,"latest_published_at":87},"Startups","startups",83,"2026-09-29T21:51:36.000Z",{"name":89,"slug":90,"count":91,"latest_published_at":92},"General","general",49,"2026-09-28T16:44:57.000Z",{"name":94,"slug":95,"count":96,"latest_published_at":97},"Gaming","gaming",48,"2026-09-25T18:35:21.000Z",{"name":99,"slug":100,"count":101,"latest_published_at":102},"Reviews","reviews",31,"2026-09-28T14:31:34.000Z",{"name":104,"slug":105,"count":106,"latest_published_at":107},"How-To","how-to",6,"2026-06-16T09:00:00.000Z"]