Self-Distillation for Further Pre-training of Transformers - researchr publication related

researchr

You are not signed in
Sign in
Sign up

Seanie Lee, Minki Kang, Juho Lee 0001, Sung Ju Hwang, Kenji Kawaguchi. Self-Distillation for Further Pre-training of Transformers. In The Eleventh International Conference on Learning Representations, ICLR 2023, Kigali, Rwanda, May 1-5, 2023. OpenReview.net, 2023. [doi]

The following publications are possibly variants of this publication:

runs on WebDSL