Guanting Dong
Source-listed: Research Intern on RL for General Agent (Top Seed Program) · ByteDance
Record snapshot: Sep 3, 2026. Coverage and affiliations may be incomplete or historical.
Selected work
5- Agent-World: Scaling Real-World Environment Synthesis for Evolving General Agent IntelligenceSaved credit: Contributor · 2026Explore people connected to this work →
- Agentic Reinforced Policy OptimizationSaved credit: First author · 2025Explore people connected to this work →
- WebThinker: Empowering Large Reasoning Models with Deep Research CapabilitySaved credit: Contributor · 2025Explore people connected to this work →
- Self-play with Execution Feedback: Improving Instruction-following Capabilities of Large Language ModelsSaved credit: First author · 2024Explore people connected to this work →
- Qwen2 Technical ReportSaved credit: Contributor · 2024Explore people connected to this work →