Shanghai AI Lab Releases Intern-S2-Mobius and 397B MoE Architecture with Memory-Decoupled Attention

By Dr. Sarah Chen Published on 2026-09-10 4 min read Source: Shanghai AI Laboratory
Shanghai AI Lab details Intern-S2-Mobius, combining 397B MoE parameters with Intern-MemDec-4B memory decoders and InternLumina-U2 for complex scientific reasoning and mathematics.

The Shanghai AI Laboratory has published the **Intern-S2-Mobius** family, spearheaded by the 397B MoE preview model and the innovative **Intern-MemDec-4B** memory-decoupling system.

Memory-Decoupled Attention (MemDec) Intern-S2 separates working memory from historical associative retrieval. By utilizing a dedicated 4B auxiliary network to compress conversational context into compact memory vectors, the primary 397B MoE focuses solely on active mathematical deduction and theorem proving.

In tandem with **InternLumina-U2** multimodal diffusion checkpoints, Intern-S2 sets high-water marks on OlympiadBench mathematics and formal Isabelle theorem proof generation.

Advertisement
High-Throughput AI API & GPU Cloud Hosting Sponsor

Source & Fact Check

This technical dispatch was verified against primary documentation released by Shanghai AI Laboratory.

Read Original Announcement on Shanghai AI Laboratory →