← People

Zhixin Zhang

Source-listed: Beijing Institute of Technology

  • Multimodal AI
  • Language models
  • Information retrieval

Selected work

3
  1. Vision Search Assistant: Empower Vision-Language Models as Multimodal Search EnginesFirst author · 2024
  2. Online Vectorized HD Map Construction using GeometryFirst author · 2024
  3. InteractiveVideo: User-Centric Controllable Video Generation with Synergistic Multimodal InstructionsContributor · 2024