553 points by krackers 3 days ago | 155 comments | View on ycombinator
joelwallis 3 days ago |
dr_dshiv 3 days ago |
ricardobeat 3 days ago |
Fable scores 70%, Kimi K3 69%, Astra 74% (all on max effort).
passive 3 days ago |
I had used 2.5-pro for a hefty chunk of development, and found it to work like a somewhat forgetful senior engineer who was new to my project. Very capable, would almost always choose a reasonable option, if not always the best one for the project, and not great at multi-tasking. Generally, made me comfortable not scrutinizing the code line-by-line, but still needed a bit of steering once projects got to a reasonable size.
The next model is a clear step up in the multi-tasking capability at least, with me very rarely having to steer the implementation of a well-defined issue. In terms of code, I found MiMo-V.2.5-pro to be extremely conservative, implementing minimal solutions. The next model seems a little bit more ambitious, in positive ways, making good guesses about gaps/next steps. It also seems to be a fair bit better at design, at least for the little bit I've done, it was good at translating my concepts to practical elements on screen, and cleaned things up nicely as I made suggestions.
krm01 3 days ago |
liuliu 3 days ago |
ProfessorLayton 3 days ago |
For some reason I thought training took much, much longer than what the progress bar suggests.
This is really neat, I'm currently using mimo 2.5 pro, and it's decent (or great given the price). Hopefully their next one is multimodal.
fzysingularity 3 days ago |
thehamkercat 3 days ago |
ttul 3 days ago |
speedgoose 3 days ago |
esafak 3 days ago |
ssn2000 3 days ago |
rao-v 3 days ago |
If you are going to develop a near frontier model, and you don’t think you have special sauce up your sleeve, why not making training runs and RL environment scores etc. visible to the world?
I’m genuinely learning quite a bit just from the dashboard
ernsheong 3 days ago |
kkotak 3 days ago |
wolttam 3 days ago |
jstummbillig 3 days ago |
rozab 3 days ago |
wg0 3 days ago |
Google had this GPT long go and a wise man within Google noted:
"We don't have any maot neither does anyone else."
The AI bubble burst is guaranteed and is only delayed by IPOs.
dr_kiszonka 3 days ago |
thenews 3 days ago |
monneyboi 3 days ago |
Alifatisk 3 days ago |
singularity2001 3 days ago |
alescalaios 3 days ago |
tcbbd 3 days ago |
Toslink 3 days ago |
heronbank 3 days ago |
pppkin 3 days ago |
sinuhe69 3 days ago |
impulser_ 3 days ago |
Where is the cool shit from the US labs?
hsbalanxvxjsmab 3 days ago |
dude250711 3 days ago |
levocardia 3 days ago |
The cost is unbelievably low, and the quality of intelligence I get is equivalent to when I was working mostly with Anthropic models (late last year/early this year). I'm fully invested in MiMo and I'm very happy with it.
-- PS: I also check almost daily to see if other models are capable of doing such great work. And they do – DS4F is powerful and DS41 is impressive, GLM 5.3 Flash gets a job done well, etc. – but when I add cost of M-token in the ROI math, Jeez! MiMo is an order of magnitude better.