VITA: Open Source Multimodal Large Language Model for Real-Time Interaction between Vision and Speech
General Introduction VITA is a leading open source interactive multimodal large language modeling project, pioneering the ability to achieve true full multimodal interaction. The project launched VITA-1.0 in August 2024, pioneering the first open source interactive fully-modal large language model.2024...































































































