MiniMax-M1

MiniMax-M1 is presented as the world’s first open-weight, large-scale hybrid-attention reasoning model, designed to push the frontier of long-context, tool-using, and deeply “thinking” language models. It is built on the MiniMax-Text-01 foundation and keeps the same massive parameter budget, but reworks the attention and training setup for better reasoning and test-time compute scaling. Architecturally, it combines Mixture-of-Experts layers with lightning attention, enabling the model to support a native context length of 1 million tokens while using far fewer FLOPs than comparable reasoning models for very long generations. The team emphasizes efficient scaling of test-time compute: at 100K-token generation lengths, M1 reportedly uses only about 25 percent of the FLOPs of some competing models, making extended “think step” traces more feasible. M1 is further trained with large-scale reinforcement learning over diverse tasks.

Features

Open-weight hybrid-attention reasoning model built atop the MiniMax-Text-01 architecture
Mixture-of-Experts plus lightning attention for 1M-token native context with efficient FLOP usage
Large-scale reinforcement learning training spanning math, coding, and sandboxed real-world tasks
CISPO RL algorithm that clips importance-sampling weights, designed for stable large-scale RL
Multiple variants with different “thinking budgets” such as 40K and 80K tokens for extended reasoning traces
Strong benchmark performance on software engineering, tool-use, and long-context reasoning compared to other open models

Project Samples

Project Activity

See All Activity >

License

Apache License V2.0

Follow MiniMax-M1

MiniMax-M1 Web Site

Other Useful Business Software

Custom VMs From 1 to 96 vCPUs With 99.95% Uptime

General-purpose, compute-optimized, or GPU/TPU-accelerated. Built to your exact specs.

Live migration and automatic failover keep workloads online through maintenance. One free e2-micro VM every month.

Try Free

Rate This Project

User Reviews

Be the first to post a review of MiniMax-M1!

Additional Project Details

Programming Language

Python

Related Categories

Python Large Language Models (LLM), Python AI Models

Registered

2025-12-01

Similar Business Software

Vertex AI

Build, deploy, and scale machine learning (ML) models faster, with fully managed ML tools for any use case. Through Vertex AI Workbench, Vertex AI is natively integrated with BigQuery, Dataproc, and Spark. You can use BigQuery ML to create and execute machine learning models in BigQuery...

See Software
LM-Kit.NET

LM-Kit.NET is a cutting-edge, high-level inference SDK designed specifically to bring the advanced capabilities of Large Language Models (LLM) into the C# ecosystem. Tailored for developers working within .NET, LM-Kit.NET provides a comprehensive suite of powerful Generative AI tools, making...

See Software
MiniMax M1

MiniMax‑M1 is a large‑scale hybrid‑attention reasoning model released by MiniMax AI under the Apache 2.0 license. It supports an unprecedented 1 million‑token context window and up to 80,000-token outputs, enabling extended reasoning across long documents. Trained using large‑scale reinforcement...

See Software
Google AI Studio

Google AI Studio is a unified development platform that helps teams explore, build, and deploy applications using Google’s most advanced AI models, including Gemini 3. It brings text, image, audio, and video models together in one interactive playground. With vibe coding, developers can use...

See Software
MiMo-V2-Flash

MiMo-V2-Flash is an open weight large language model developed by Xiaomi based on a Mixture-of-Experts (MoE) architecture that blends high performance with inference efficiency. It has 309 billion total parameters but activates only 15 billion active parameters per inference, letting it balance...

See Software
Step 3.5 Flash

Step 3.5 Flash is an advanced open source foundation language model engineered for frontier reasoning and agentic capabilities with exceptional efficiency, built on a sparse Mixture of Experts (MoE) architecture that selectively activates only about 11 billion of its ~196 billion parameters per...

See Software

Report inappropriate content

MiniMax-M1

Open-weight, large-scale hybrid-attention reasoning model

Get an email when there's a new version of MiniMax-M1

Features

Project Samples

Project Activity

Categories

License

Follow MiniMax-M1

User Reviews

Additional Project Details

Programming Language

Related Categories

Registered