Intern S1 Pro 💻Github Repo • 🤗Model Collections • 📜Technical Report • 💬Online Chat 👋 join us on Discord and WeChat Introduction We introduce Intern S1 Pro , a trillion scale MoE multimodal scientific reasoning model. Intern S1 Pro scales to 1T total parameters with 512 experts, activating 8 experts per token (22B activated parameters). The model delivers top tier performance on advanced reasoning benchmarks and achieves leading results across key AI4Science domains (chemistry, materials, life science, earth, etc.), while maintaining strong general multimodal and text capabilities. Features State of the art scientific reasoning , competitive with leading closed source models across AI4Science tasks. Strong general multimodal performance on various benchmarks. Trillion scale MoE training efficiency with STE routing (dense gradient for router training) and grouped routing for stable convergence and balanced expert parallelism. Fourier Position Encoding (FoPE) + upgraded time series modeling for better physical signal representation; supports long, heterogeneous time series (10^0–10^6 points). Performance We evaluate the Intern S1 Pro on various benchmarks, including genera…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy