Yu (Sid) Wang
Record snapshot: Sep 3, 2026. Coverage and affiliations may be incomplete or historical.
Selected work
5- The Llama 3 Herd of ModelsSaved credit: Contributor · 2024Explore people connected to this work →
- LlamaRL: A Distributed Asynchronous Reinforcement Learning Framework for Efficient Large-scale LLM TrainingSaved credit: Contributor · 2025Explore people connected to this work →
- Beyond Verifiable Rewards: Scaling Reinforcement Learning in Language Models to Unverifiable DataSaved credit: Contributor · 2025Explore people connected to this work →
- Positional Encoding via Token-Aware Phase AttentionSaved credit: First author · 2025Explore people connected to this work →
- TreePiece: Faster Semantic Parsing via Tree TokenizationSaved credit: First author · 2023Explore people connected to this work →