Peng Xia
Source-listed: Student researcher; Ph.D. student · Google; University of North Carolina at Chapel Hill
- Multimodal AI
- Language models
- Reinforcement learning
Selected work
5- MMed-RAG: Versatile Multimodal RAG System for Medical Vision Language ModelsFirst author · 2024
- MMIE: Massive Multimodal Interleaved Comprehension Benchmark for Large Vision-Language ModelsContributor · 2024
- MMedAgent-RL: Optimizing Multi-Agent Collaboration for Multimodal Medical ReasoningFirst author · 2025
- Agent0: Unleashing Self-Evolving Agents from Zero Data via Tool-Integrated ReasoningFirst author · 2025
- SkillRL: Evolving Agents via Recursive Skill-Augmented Reinforcement LearningFirst author · 2026