← PeopleZhixin ZhangSource-listed: Beijing Institute of TechnologyWebsite ↗GitHub ↗Project ↗Publications ↗Institution ↗Multimodal AILanguage modelsInformation retrievalSelected work3Vision Search Assistant: Empower Vision-Language Models as Multimodal Search EnginesFirst author · 2024↗Online Vectorized HD Map Construction using GeometryFirst author · 2024↗InteractiveVideo: User-Centric Controllable Video Generation with Synergistic Multimodal InstructionsContributor · 2024↗