[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"branding":3,"analytics":7,"article-strata-teaches-game-ai-to-learn-from-its-own-matches":10,"sections":40},{"siteName":4,"siteTagline":5,"publisherName":4,"contactEmail":6},"The Revision","Tech news, decoded.","editor@therevision.news",{"gaMeasurementId":8,"adsenseClientId":9},"G-ZW2MV82GYR","ca-pub-8533917693782264",{"article":11},{"id":12,"slug":13,"title":14,"dek":15,"body_md":16,"tags_json":17,"published_at":18,"created_at":19,"updated_at":20,"status":21,"review_note":22,"review_notes":23,"image_url":22,"persona_id":22,"persona_name":22,"section":30,"tags":31,"sources":35,"feedback":39,"feedback_at":22,"cost_usd":39,"total_tokens":39},8905,"strata-teaches-game-ai-to-learn-from-its-own-matches","STRATA Teaches Game AI to Learn From Its Own Matches","A new AI system splits real-time strategy decisions across separate agents, then learns from its own matches to raise win rates.","A new AI system taught itself to win a classic real-time strategy (RTS) game by studying its own past matches.\n\nResearchers built STRATA to play Red Alert by splitting command decisions across three specialized agents: one handles high-level strategy, one handles logistics, and one handles fast tactical calls. After each match, a fourth Review Agent mines the game trace for useful lessons, checks those lessons against evidence from later matches, and compresses the validated ones into short \"experience cards\" the strategy agent can pull from next time. In one fixed test scenario, using the learned experience cards lifted the system's win rate from 30% to 100%. Run sequentially against AI opponents with different play styles, STRATA also developed distinct long-term strategies tailored to each.\n\nEarlier LLM-based RTS systems leaned on manually written, experience-based prompts and could miss fast-moving tactical events while waiting on slow model inference. STRATA's three-agent split handles time-sensitive combat separately from slower strategic planning, and its review-and-compress loop replaces hand-tuned prompts with evidence pulled from the system's own games.\n\nIt's still a lab result in a decades-old game, not a battlefield-ready product, but the pattern of specializing, reviewing, compressing, and repeating looks like an early sketch of AI agents that keep improving after deployment without constant human prompt-tuning.","[\"ai\",\"gaming\",\"llm agents\",\"research\"]","2026-10-01T04:00:00.000Z","2026-10-01T10:12:09.773Z","2026-10-01T10:12:15.635Z","published",null,[24],{"id":25,"reviewer":26,"round":27,"reason":28,"status":29},"editor-r1","editor",1,"The acronym 'RTS' is used in the body ('large language models for RTS command') without ever being defined — introduce it as 'real-time strategy (RTS)' on first use, as the source does, before using the bare acronym.","resolved","ai",[30,32,33,34],"gaming","llm agents","research",[36],{"name":37,"url":38},"arXiv cs.AI","https:\u002F\u002Farxiv.org\u002Fabs\u002F2609.38881",0,{"sections":41},[42,45,50,55,60,65,70,75,80,84,89,93,98,103],{"name":43,"slug":30,"count":44,"latest_published_at":18},"AI",5351,{"name":46,"slug":47,"count":48,"latest_published_at":49},"Security","security",801,"2026-09-30T22:18:23.000Z",{"name":51,"slug":52,"count":53,"latest_published_at":54},"Policy","policy",429,"2026-10-01T02:26:17.000Z",{"name":56,"slug":57,"count":58,"latest_published_at":59},"Deals","deals",298,"2026-09-30T21:00:26.000Z",{"name":61,"slug":62,"count":63,"latest_published_at":64},"Hardware","hardware",196,"2026-09-30T13:00:00.000Z",{"name":66,"slug":67,"count":68,"latest_published_at":69},"Science","science",157,"2026-09-30T15:00:56.000Z",{"name":71,"slug":72,"count":73,"latest_published_at":74},"Consumer Tech","consumer-tech",149,"2026-09-30T22:57:11.000Z",{"name":76,"slug":77,"count":78,"latest_published_at":79},"Dev Tools","dev-tools",93,"2026-10-01T02:30:48.000Z",{"name":81,"slug":82,"count":78,"latest_published_at":83},"Software","software","2026-09-30T21:41:11.000Z",{"name":85,"slug":86,"count":87,"latest_published_at":88},"Startups","startups",84,"2026-09-30T20:39:09.000Z",{"name":90,"slug":32,"count":91,"latest_published_at":92},"Gaming",51,"2026-09-30T16:24:30.000Z",{"name":94,"slug":95,"count":96,"latest_published_at":97},"General","general",50,"2026-09-30T21:37:54.000Z",{"name":99,"slug":100,"count":101,"latest_published_at":102},"Reviews","reviews",31,"2026-09-28T14:31:34.000Z",{"name":104,"slug":105,"count":106,"latest_published_at":107},"How-To","how-to",6,"2026-06-16T09:00:00.000Z"]