Q Judger A fine tuned judge model for evaluating text to image (T2I) generation quality. Built on top of Qwen3.6 27B, it scores generated images across 5 hierarchical dimensions using structured checklists and outputs JSON formatted evaluation results. Links Resource Link 📑 Paper http://arxiv.org/abs/2605.28091 📊 Benchmark Dataset (HuggingFace) https://huggingface.co/datasets/Qwen/Qwen Image Bench 📊 Benchmark Dataset (ModelScope) https://www.modelscope.cn/datasets/Qwen/Qwen Image Bench 💻 GitHub https://github.com/QwenLM/Qwen Image Bench 🧑⚖️ Q Judger Model https://huggingface.co/Qwen/Qwen Image Bench 🧑⚖️ Q Judger Model https://modelscope.cn/models/Qwen/Qwen Image Bench Model Description Q Judger is a vision language model fine tuned specifically for automated evaluation of text to image generated images. Given a text prompt and a generated image, the model evaluates the image on fine grained quality criteria organized in a 3 level hierarchy and outputs structured JSON scores. Base Model : Qwen3.6 27B Task : Image quality evaluation / judging Input : Text prompt + generated image Output : Structured JSON with per dimension scores (0 = Fail, 1 = Pass, 2 = Excel, N/A) Thinking…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy