139 points by nateb2022 2 days ago | 33 comments | View on ycombinator
BoppreH 1 day ago |
nh43215rgb 1 day ago |
> Built on a sparse Mixture-of-Experts architecture, Step 5 Preview has 600B total parameters, with 27B active per token, and supports a 1M-token context window and vision input.
> Step 5 Preview scores 44 on the Artificial Analysis Intelligence Index.
> The model will be released with open weights on October 15.
I guess being Chinese company they decided to skip version 4, while also giving impression to be on the similar iteration with leading companies (claude opus 5).
I wonder if other Chinese labs like Kimi/Moonshot will follow suit.bethekind 1 day ago |
Finally FireRed is being used as a benchmark again! I believe Astra can beat it in 18 hours. Not sure how that compares.
garo-pro 1 day ago |
ghoshbishakh 1 day ago |
segmondy 1 day ago |
InsideOutSanta 1 day ago |
If this performs similarly in the real world, we're approaching a level of capability where for most devs, it only makes sense to pay for Anthropic or OpenAI subscriptions if they are heavily subsidized and actually cheaper than these alternative options.
* Oddly, because I perceived Devin as being kind of a joke before trying SWE-2.
Jacques2Marais 1 day ago |
wrs 1 day ago |
“the question it raises matters more than the answer” ???
“So ‘the river drifts from cool to warm’ is not a figure of speech.” Oh really?
conception 1 day ago |
dofm 1 day ago |
just60sec about 6 hours ago |
cboyardee 1 day ago |
derliebej 1 day ago |
> Interesting! It turns out there's already an existing project here [...] The project is fully built [...]
I'm always astounded how little effort is put into checking the AI answers displayed in these announcements. Back when I paid more attention, I remember OpenAI's and Google's demos constantly showed their AIs giving wrong answers.