Live data from Hacker News

MiniMax M2.1: Built for Real-World Complex Tasks, Multi-Language Programming

minimaxi.com

31–40 of 86 posts

Re: MiniMax M2.1: Built for Real-World Complex Tasks, Multi-Language Programming

#31
post #18

> It exhibits consistent and stable results in tools such as Claude Code, Droid (Factory AI), Cline, Kilo Code, Roo Code, and BlackBox, while providing reliable support for Context Management mechanisms including Skill.md, Claude.md/agent.md/cursorrule, and Slash Commands. One of the demos shows them using Claude Code, which is interesting. And the next sections are titled 'Digital Employee' and 'End-to-End Office Au…

they are going IPO in HKEX in a few weeks. some hype up are necessary, not too far fetched imo, pretty much same as anthropic playbook.

anthropic playbook does include the false claim publicly made by its CEO that "in six months AI would be writing 90 percent of code". he made that claim 10 months ago. it is a criminal offence for intentionally misleading investors in many countries.

MiniMax is like 100x more honest.

Re: MiniMax M2.1: Built for Real-World Complex Tasks, Multi-Language Programming

#32
post #23

Would it kill them to use the words "AI coding agent" somewhere prominent? "MiniMax M2.1: Significantly Enhanced Multi-Language Programming, Built for Real-World Complex Tasks" could be an IDE, a UI framework, a performance library, or, or...

its main Chinese competitor GLM is like making 50 cents USD each in the past 6 months from its 40 million "developer users", calling your flagship model "AI coding agent" is like telling investors "we are doing this for fun, not for money".

Re: MiniMax M2.1: Built for Real-World Complex Tasks, Multi-Language Programming

#33
post #29

Earlier quoted context omitted.

If I use a software I need to trust it.

a model is not software, it is a bunch of weights. you are more than welcomed to pick whatever model or software you choose to trust, that is totally fine. However, that is vastly different from bad mouthing a model or software just because its release note contains a single sentence you don't like.

The API is software. You don't get the weights.

Re: MiniMax M2.1: Built for Real-World Complex Tasks, Multi-Language Programming

#34
post #23

Would it kill them to use the words "AI coding agent" somewhere prominent? "MiniMax M2.1: Significantly Enhanced Multi-Language Programming, Built for Real-World Complex Tasks" could be an IDE, a UI framework, a performance library, or, or...

It's not an AI coding agent. It's an LLM that can be used for whatever you'd like, including powering coding agents.

HA! I almost added a disclaimer to the original message that I wasn't certain in my identification, hence the request/complaint that they didn't make it clear. But I figured the message would be more effective if I "confidently got it wrong" rather than asking, so I went with it.

Re: MiniMax M2.1: Built for Real-World Complex Tasks, Multi-Language Programming

#35
post #23

Would it kill them to use the words "AI coding agent" somewhere prominent? "MiniMax M2.1: Significantly Enhanced Multi-Language Programming, Built for Real-World Complex Tasks" could be an IDE, a UI framework, a performance library, or, or...

It's not an AI coding agent. It's an LLM that can be used for whatever you'd like, including powering coding agents.

That reinforces OP’s point that it isn’t clear from their wording. I initially thought it was a speech model, then I saw Python, etc., and it took me a bit more reading to understand what it actually is

Re: MiniMax M2.1: Built for Real-World Complex Tasks, Multi-Language Programming

#39
post #27

I’ve spent a little bit of time testing Minimax M2. It’s quite good given the small size but it did make some odd mistakes and struggle with precise instructions.

This is an announcement for M2.1 not M2. It got a decent bump in agent capabilities.

Re: MiniMax M2.1: Built for Real-World Complex Tasks, Multi-Language Programming

#40
I think people should stop comparing to sonnet, but to opus instead since it's so far ahead on producing code I would actually want to use (gemini 3 pro tends to be lacking in generalization and wants things to be using it's own style rather than adapting).

Whatever benchmark opus is ahead in should be treated as a very important metric of proper generalization in models.

Post reply on HN