Meta Superintelligence Labs
Muse Product Post-Training
- Co-developed steering mechanisms for agent rollouts, including reflection, plan-and-execute, and adaptive replanning; migrated experimentation to a controlled white-box sandbox for reproducible iteration.
- Contributed personalized-research SFT data to Muse Spark post-training; the trained model improved by 3pp across three internal agent benchmarks.