November 28, 2025 Vidi2.5: Large Multimodal Models for Video Understanding and Creation 支持时空定位、时序检索和视频问答的大规模多模态模型,通过统一编码架构实现视频理解与创作 multimodal vision-language video-understanding spatio-temporal-grounding video-qa