Project
PIDM: Predictive Inverse Dynamic Models
Imitation learning enables agents to learn complex behaviour from demonstrations, but in practice it often requires large datasets that are costly or impractical to collect. Our project studies how to make imitation learning significantly more…
Publication
Teaming Up with AI: Coordination and Cooperation
Microsoft Research Blog
SkillOpt: Agent skills as trainable parameters
AI agents often fail because their instructions, or skills, are manually modified with no guarantee of improvement. Learn how SkillOpt turns skill editing into a training process, making agent behavior more reliable without changing model…