Xiangyu Zhao
Source-listed: The Hong Kong Polytechnic University and Shanghai Artificial Intelligence Laboratory
- Multimodal AI
- Language models
- Information retrieval
Selected work
5- UniFashion: A Unified Vision-Language Model for Multimodal Fashion Retrieval and GenerationFirst author · 2024
- EasyGen: Easing Multimodal Generation with BiDiffuser and LLMsFirst author · 2024
- WeatherGFM: Learning a Weather Generalist Foundation Model via In-context LearningFirst author · 2024
- MSEarth: A Multimodal Benchmark for Earth Science Phenomenon Discovery with MLLMsFirst author · 2026
- SCI-PRM: A Tool Aware Process Reward Model for Scientific Reasoning VerificationFirst author · 2026