[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"branding":3,"analytics":7,"article-new-training-trick-speeds-up-ai-agent-memory-compaction":10,"sections":40},{"siteName":4,"siteTagline":5,"publisherName":4,"contactEmail":6},"The Revision","Tech news, decoded.","editor@therevision.news",{"gaMeasurementId":8,"adsenseClientId":9},"G-ZW2MV82GYR","ca-pub-8533917693782264",{"article":11},{"id":12,"slug":13,"title":14,"dek":15,"body_md":16,"tags_json":17,"published_at":18,"created_at":19,"updated_at":20,"status":21,"review_note":22,"review_notes":23,"image_url":22,"persona_id":22,"persona_name":22,"section":30,"tags":31,"sources":35,"feedback":39,"feedback_at":22,"cost_usd":39,"total_tokens":39},8738,"new-training-trick-speeds-up-ai-agent-memory-compaction","New training trick speeds up AI agent memory compaction","A technique called KV-streams lets AI agents compact long context histories 2.6 to 5x faster during training without hurting performance.","Training AI agents to handle long tasks just got cheaper, thanks to a fix for a bottleneck nobody outside ML research labs has heard of.\n\nThe problem: agentic AI models that run for a long time build up huge context histories, which eat GPU memory. The common fix, called context compaction, summarizes or trims that history to keep memory use flat. But compaction usually forces the model to reprocess its entire context from scratch every time it compacts, which slows training to a crawl. A new paper proposes KV-streams, a method that streams the model's cached internal state forward instead of flushing and rebuilding it after each compaction. Tested across three different compaction strategies, it delivered a 2.6 to 5x wall-clock speedup during training with no measured drop in performance.\n\nThe more interesting finding is a side effect, not the speed gain itself. The researchers found the streamed cache can act like a recurrent memory, letting the model recall information that has already scrolled out of its visible context window. That's the kind of long-horizon memory researchers have spent years trying to bolt onto transformers with separate memory modules. Here it shows up for free from reinforcement learning alone, no extra architecture required.\n\nIf that holds up outside a controlled test setup, it's a bigger deal than the throughput number. Plenty of papers promise faster training; fewer show a model developing memory-like behavior as a side effect of an efficiency tweak.","[\"ai\",\"llm-training\",\"reinforcement-learning\",\"arxiv\"]","2026-09-30T04:00:00.000Z","2026-09-30T22:49:46.609Z","2026-09-30T22:49:53.218Z","published",null,[24],{"id":25,"reviewer":26,"round":27,"reason":28,"status":29},"editor-r1","editor",1,"The dek rounds the paper's 2.6x-5x speedup up to 'three to five times faster,' which contradicts the precise figure the body itself correctly cites — fix the dek to say '2.6 to 5x' (or 'roughly 3 to 5x' only if hedged) so it matches the body.","resolved","ai",[30,32,33,34],"llm-training","reinforcement-learning","arxiv",[36],{"name":37,"url":38},"arXiv cs.AI","https:\u002F\u002Farxiv.org\u002Fabs\u002F2609.35750",0,{"sections":41},[42,46,51,56,61,65,69,73,78,82,87,92,97,102],{"name":43,"slug":30,"count":44,"latest_published_at":45},"AI",5214,"2026-09-30T13:00:00.000Z",{"name":47,"slug":48,"count":49,"latest_published_at":50},"Security","security",793,"2026-09-30T12:55:00.000Z",{"name":52,"slug":53,"count":54,"latest_published_at":55},"Policy","policy",419,"2026-09-30T12:24:32.000Z",{"name":57,"slug":58,"count":59,"latest_published_at":60},"Deals","deals",292,"2026-09-30T14:15:18.000Z",{"name":62,"slug":63,"count":64,"latest_published_at":45},"Hardware","hardware",196,{"name":66,"slug":67,"count":68,"latest_published_at":18},"Science","science",155,{"name":70,"slug":71,"count":72,"latest_published_at":45},"Consumer Tech","consumer-tech",144,{"name":74,"slug":75,"count":76,"latest_published_at":77},"Dev Tools","dev-tools",91,"2026-09-30T12:58:00.000Z",{"name":79,"slug":80,"count":76,"latest_published_at":81},"Software","software","2026-09-25T20:55:00.000Z",{"name":83,"slug":84,"count":85,"latest_published_at":86},"Startups","startups",83,"2026-09-29T21:51:36.000Z",{"name":88,"slug":89,"count":90,"latest_published_at":91},"General","general",49,"2026-09-28T16:44:57.000Z",{"name":93,"slug":94,"count":95,"latest_published_at":96},"Gaming","gaming",48,"2026-09-25T18:35:21.000Z",{"name":98,"slug":99,"count":100,"latest_published_at":101},"Reviews","reviews",31,"2026-09-28T14:31:34.000Z",{"name":103,"slug":104,"count":105,"latest_published_at":106},"How-To","how-to",6,"2026-06-16T09:00:00.000Z"]