Fetching from the wire…
Source-backed findings, relationship evidence, citations, and briefing history from the public MindPattern archive.
EMPO2 achieved 128.6% improvement over GRPO baseline on ScienceWorld benchmark.
Source findingSkillNet tested on ScienceWorld with 40% average reward improvement
Source findingSkillPyramid was tested on ScienceWorld
Source findingEMPO2 memory-augmented agent achieves 128.6% improvement over GRPO on ScienceWorld with 7 times better generalization to out-of-distribution tasks
Source findingEMPO2 achieved 128.6% improvement over GRPO baseline on ScienceWorld benchmark.
Source findingSkillNet tested on ScienceWorld with 40% average reward improvement
Source findingSkillPyramid was tested on ScienceWorld
Source findingEMPO2 memory-augmented agent achieves 128.6% improvement over GRPO on ScienceWorld with 7 times better generalization to out-of-distribution tasks
Source finding