<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"><channel><title>Composite Research</title><description>Benchmarks, evaluations, and methods from Composite, the team building AI for professional computer work.</description><link>https://research.composite.com</link><item><title>Introducing Composite-Bench: the strongest open-weights model isn&apos;t Kimi K3</title><link>https://research.composite.com/composite-bench</link><guid isPermaLink="true">https://research.composite.com/composite-bench</guid><description>The first long-horizon, browser-based computer-use benchmark with certified optimality and verified compute. In our evaluation, GLM-5.2 lands 11 points behind a two-way Claude tie and beats every closed model we measured.</description><pubDate>Fri, 24 Jul 2026 00:00:00 GMT</pubDate></item></channel></rss>