Zihang Dai
Record snapshot: Sep 3, 2026. Coverage and affiliations may be incomplete or historical.
Selected work
5- Transformer-XL: Attentive Language Models Beyond a Fixed-Length ContextSaved credit: Contributor · 2019Explore people connected to this work →
- XLNet: Generalized Autoregressive Pretraining for Language UnderstandingSaved credit: Contributor · 2019Explore people connected to this work →
- Funnel-Transformer: Filtering out Sequential Redundancy for Efficient Language ProcessingSaved credit: Contributor · 2020Explore people connected to this work →
- CoAtNet: Marrying Convolution and Attention for All Data SizesSaved credit: Contributor · 2021Explore people connected to this work →
- Gemini: A Family of Highly Capable Multimodal ModelsSaved credit: Contributor · 2023Explore people connected to this work →