APEX MTP Vision Apache 2.0 Agents A1 MTP APEX English 📖 中文文档 35B agentic MoE that reaches trillion parameter performance · APEX quantized GGUFs + BF16 + mmproj 🤖 About Agents A1 Agents A1 is a 35B parameter Mixture of Experts agentic model from InternScience , post trained on top of Qwen3.5 35B A3B via a three stage paradigm: full domain SFT → domain level teacher training → multi teacher multi domain on policy distillation. Despite operating in the ~35B model class, Agents A1 delivers highly competitive performance against frontier scale systems such as GPT 5.5, DeepSeek V4 pro, and Kimi K2.6 — achieving SOTA on Seal 0 (56.4), HiPhO (46.4), FrontierScience Olympiad (79.0), IFBench (80.6), IFEval (94.8), and best among comparable on BrowseComp (75.5), XBench DS 2510 (86.0), GAIA (96.0), SciCode (44.3), HLE (47.6), and MolBench bind (56.8). This GGUF package includes the mmproj F16.gguf vision projector for multimodal (image + text) capabilities with llama.cpp. MTP layers are extracted from Qwen3.5 35B A3B and injected into Agents A1's safetensors (see MTP Extraction & Injection section). License: Apache 2.0. 🧠 Model Details Architecture Qwen3.5 MoE (Mixture of Experts) Param…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy