Overview HyperCLOVA X SEED 32B Think is an updated vision language thinking model that advances the SEED Think 14B line beyond simple scaling, pairing a unified vision language Transformer backbone with a reasoning centric training recipe. SEED 32B Think processes text tokens and visual patches within a shared embedding space, supports long context multimodal understanding up to 128K tokens, and provides an optional “thinking mode” for deep, controllable reasoning. Building on the earlier 14B model, SEED 32B Think further strengthens Korean centric reasoning and agentic capabilities, improving practical reasoning quality and reliability in real world use. Technical Report HyperCLOVAX SEED Think 32B Tech Report (PDF) Basic Information Architecture : Transformer based vision language model (VLM) architecture (Dense Model) Parameters : 32B Input Format : Text/Image/Video Output Format : Text Context Length : 128K Knowledge Cutoff : May 2025 Benchmarks General Knowledge (Korean Text) : KoBalt, CLIcK, HAERAE Bench 1.0 Vision Understanding : ChartVQA, TextVQA, K MMBench, K DTCBench Agentic Tasks : Tau^2 Airline, Tau^2 Retail, Tau^2 Telecom Examples Solving 2026 Korean CSAT Math Problem U…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy