プロジェクト
PIDM: Predictive Inverse Dynamic Models
Imitation learning enables agents to learn complex behaviour from demonstrations, but in practice it often requires large datasets that are costly or impractical to collect. Our project studies how to make imitation learning significantly more…
Microsoft Research ブログ
SkillOpt: Agent skills as trainable parameters
AI agents often fail because their instructions, or skills, are manually modified with no guarantee of improvement. Learn how SkillOpt turns skill editing into a training process, making agent behavior more reliable without changing model…