MiniMax unveiled its M3 model with a sparse attention architecture that dramatically reduces compute costs while handling up to one million tokens.
Democratizing information: Free world-class API and RSS feeds for every business.
MiniMax unveiled its M3 model with a sparse attention architecture that dramatically reduces compute costs while handling up to one million tokens.