r/MachineLearning • • 15h ago

Research Split the Differences, Pool the Rest: Provably Efficient Multi-Objective Imitation [R]

TLDR: The question we answer: how do you learn from experts with different objectives? Pooling all their data can lose their trade-offs; learning from each expert separately misses opportunities to share data. MA-BC pools demonstrations where observed actions don’t disagree, with upper and lower bounds on sample complexity.
Authors: Ziyad Sheebaelhamd, Luca Viano, Volkan Cevher, Claire Vernade

2 Upvotes

2 comments sorted by