Yuchen Hao
Source-listed: Meta
Record snapshot: Sep 3, 2026. Coverage and affiliations may be incomplete or historical.
Selected work
5- PyTorch FSDP: Experiences on Scaling Fully Sharded Data ParallelSaved credit: Contributor · 2023Explore people connected to this work →
- Software-Hardware Co-design for Fast and Scalable Training of Deep Learning Recommendation ModelsSaved credit: Contributor · 2022Explore people connected to this work →
- Scaling Llama 3 Training with Efficient Parallelism StrategiesSaved credit: Contributor · 2025Explore people connected to this work →
- LoKA: Low-precision Kernel Applications for Recommendation Models At ScaleSaved credit: Contributor · 2026Explore people connected to this work →
- ROCS: Request-Oriented Compute Sharing for Efficient Large-Scale RecommendationSaved credit: Contributor · 2026Explore people connected to this work →