Introduction Beijing Academy of Artificial Intelligence (BAAI) We collect, organize and open source the large scale multimodal instruction dataset, Infinity MM , consisting of tens of millions of samples. Through quality filtering and deduplication, the dataset has high quality and diversity. We propose a synthetic data generation method based on open source models and labeling system, using detailed image annotations and diverse question generation. Based on Infinity MM, we have successfully trained a 2 billion parameter VLM model, Aquila VL 2B , achieving SOTA performance among models of the same scale. News 2024/11/19 We have released Aquila VL 2B and all intermediate checkpoints obtained during different stages of training. Please feel free to use these models for analysis and experimentation. 2024/11/05 The data in stage2/7M 0712 math plus system release 0802 was incomplete. We have now updated it, and the new data is placed in stage2/7M 0712 math plus system release. Please replace the previous data with this updated version. 2024/10/28 All the data has been uploaded. 2024/10/24 The data of stage 2, stage 3 and stage 4 has been transferred. And the data of stage 1 will comple…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy