← People

Yu (Sid) Wang

Record snapshot: Sep 3, 2026. Coverage and affiliations may be incomplete or historical.

Selected work

5
  1. The Llama 3 Herd of ModelsSaved credit: Contributor · 2024Explore people connected to this work →
  2. LlamaRL: A Distributed Asynchronous Reinforcement Learning Framework for Efficient Large-scale LLM TrainingSaved credit: Contributor · 2025Explore people connected to this work →
  3. Beyond Verifiable Rewards: Scaling Reinforcement Learning in Language Models to Unverifiable DataSaved credit: Contributor · 2025Explore people connected to this work →
  4. Positional Encoding via Token-Aware Phase AttentionSaved credit: First author · 2025Explore people connected to this work →
  5. TreePiece: Faster Semantic Parsing via Tree TokenizationSaved credit: First author · 2023Explore people connected to this work →