Multi-Expert Distillation for Few-Shot Coordination (Student Abstract)

Yujian Zhu,Hao Ding,Zongzhang Zhang

doi:10.1609/aaai.v38i21.30539

Multi-Expert Distillation for Few-Shot Coordination (Student Abstract)

Yujian Zhu, Hao Ding + Show 1 more

Open Access

https://doi.org/10.1609/aaai.v38i21.30539

Copy DOI

Journal: Proceedings of the AAAI Conference on Artificial Intelligence

Publication Date: Mar 24, 2024

#Population-based Training #Stable Training + Show 6 more

Abstract
Full-Text PDF
Similar Papers

Abstract

Ad hoc teamwork is a crucial challenge that aims to design an agent capable of effective collaboration with teammates employing diverse strategies without prior coordination. However, current Population-Based Training (PBT) approaches train the ad hoc agent through interaction with diverse teammates from scratch, which suffer from low efficiency. We introduce Multi-Expert Distillation (MED), a novel approach that directly distills diverse strategies through modeling across-episodic sequences. Experiments show that our algorithm achieves more efficient and stable training and has the ability to improve its behavior using historical contexts. Our code is available at https://github.com/LAMDA-RL/MED.

Full Text

Paper version not known

Open DOI Link

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Similar Papers

Paper Title

Journal

Date

Author

View more papers

More From: Proceedings of the AAAI Conference on Artificial Intelligence

Paper Title

Journal

Date

Author

View more papers

Disclaimer: All third-party content on this website/platform is and will remain the property of their respective owners and is provided on "as is" basis without any warranties, express or implied. Use of third-party content does not indicate any affiliation, sponsorship with or endorsement by them. Any references to third-party content is to identify the corresponding services and shall be considered fair use under The CopyrightLaw.