Maestro1 9B A 9B multimodal reasoning model — math, code, and deep thinking that can see. Vection Labs Weights · Benchmarks · Quickstart · Limitations Abstract Maestro1 9B is a dense, 9 billion parameter vision language model built for hard problems : multi step mathematical proof, competitive programming grade code synthesis, and visual reasoning over images and video — within a single model and a single context window of up to 1M tokens . It is designed for users who care less about chat pleasantries and more about whether the model can actually solve the thing : derive the bound, find the bug, read the diagram, finish the proof. Maestro1 9B pairs an explicit step by step reasoning mode with native multimodal perception, so the same chain of thought that solves a math olympiad problem can also reason about a chart, a UI screenshot, or a short clip. Highlights Reasoning first. Produces structured, inspectable chains of thought for math, logic, and code. Genuinely multimodal. Images and video are first class inputs, not bolted on captioning. Long context. Up to 1M tokens via interleaved multimodal RoPE — whole codebases, long papers, or long videos in a single prompt. Open weights.…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy