MUSE-Autoskill: Self-Evolving Agents via Skill Creation, Memory, Management, and Evaluation

Published
Source
arXiv
Paper number
252
Field
AI / General
arXiv ID
2605.27366

Key points

  • Testing is the unit test used to verify that a skill works correctly before it is allowed into the Skill Bank.
  • Memory is a dedicated file, .memory.md, that stores observations and lessons from previous runs.
  • At Level 1, single-node compression, when a particular tool output is extremely large, for example a 20,000-token log file, the method replaces it with a concise summary while leaving the surrounding conversation intact.

Paper links

External research summaries. These are not HDATF publications or measured product results.

Read original (opens in a new tab)