We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy
Dataset Card for Conceptual Captions 12M (CC12M) Dataset Summary Conceptual 12M (CC12M) is a dataset with 12 million image-text pairs specifically meant to be used for visionand-language pre-training. Its data collection pipeline is a relaxed version of the one used in Conceptual Captions 3M (CC3M). Usage This instance of Conceptual Captions is in webdataset .tar format. It can be used with webdataset library or upcoming releases of Hugging Face… See the full description on the dataset page: https://huggingface.co/datasets/Salmonnn/cc12.
No dataset card provided yet.
Mirrored from an external registry.
Last synced 6/19/2026
Preview not yet available for this dataset