I think there's such a thing as throwing too many marketing and sales people and too much polishing and "refinement" at something. I see in new product announcements from Microsoft as well. It's like seeing someone try too hard to impress you.
Very nice progress. Also I respect putting Kimi on those charts. Regardless of if they are beating the Pareto frontier (not now), model diversity is a good thing for humanity — I’m hopeful for the team to keep increasing their gains.
Excited to try this. The low costs v. benchmarks alone here are worth a serious test. K3 has been my daily driver for a month or two now and it's dramatically reduced token spend (while not having much of a negative impact on productivity).
This was the era of the AI race I was waiting for.
Bar charts should start at zero. If they don't start at zero, there should be a clear visual indicator that the chart has been trimmed without having to read the axis labels. I hate that this has to be repeated so often that it has become a cliché.
A bit disappointing to see it still lagging behind Chinese open models. Those Chinese models are pushing proprietary models to raise the bar, but we need equally strong non-Chinese open models to challenge the Chinese ones in turn.
For example, there's something about Anthropic's picked design and their little Claude avatars that's unsettling to me.
https://bulbapedia.bulbagarden.net/wiki/Lechonk_(Pok%C3%A9mo...
This was the era of the AI race I was waiting for.
https://news.ycombinator.com/item?id=49977979 (200 comments now)
Give them more compute!
Also 1T-A49B. Weights currently closed but promise to open source them by the end of the month.
Great release movie.
OpenAI's therapist: Le Chaton Fat isn't real and cannot hurt you
Le Chaton Fat: