4KLSDB: A Large Scale Dataset for 4K Image Restoration and Generation DataCV @ CVPR 2026 · Accepted 🎉 4KLSDB is a native 4K image dataset with 129,484 train / 2,000 val / 1,984 test images, spanning nature, urban scenes, people, food, artwork, CGI, animals, and architecture. It supports both image restoration (super resolution) and 4K text to image generation. Quick links · 🌐 Project page · 💻 Code (GitHub) · 📄 Paper (arXiv) · 🤗 Dataset · 🧱 Checkpoints What's in this repo 📝 Captions metadata.jsonl is the authoritative caption source from May 2026 onwards. It contains 129,484 entries produced by Qwen2.5 VL 7B Instruct prompted for detailed scene descriptions. Each line is a JSON object: The caption column inside the existing data/ .parquet shards reflects the original caption source (LAION style short captions / CogVLM); we have not rewritten those in place. Use metadata.jsonl for any task that wants the newest, longest, scene grounded captions (most importantly 4K T2I fine tuning ). Or via the Hub directly: 🧱 Pre trained checkpoints ( ckpts/ ) Every model used in the paper is released under ckpts/ / . They are 4KLSDB fine tuned variants of the upstream architectures: Folder…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy