I think there's such a thing as throwing too many marketing and sales people and too much polishing and "refinement" at something. I see in new product announcements from Microsoft as well. It's like seeing someone try too hard to impress you.
Very nice progress. Also I respect putting Kimi on those charts. Regardless of if they are beating the Pareto frontier (not now), model diversity is a good thing for humanity — I’m hopeful for the team to keep increasing their gains.
Excited to try this. The low costs v. benchmarks alone here are worth a serious test. K3 has been my daily driver for a month or two now and it's dramatically reduced token spend (while not having much of a negative impact on productivity).
This was the era of the AI race I was waiting for.
Bar charts should start at zero. If they don't start at zero, there should be a clear visual indicator that the chart has been trimmed without having to read the axis labels. I hate that this has to be repeated so often that it has become a cliché.
A bit disappointing to see it still lagging behind Chinese open models. Those Chinese models are pushing proprietary models to raise the bar, but we need equally strong non-Chinese open models to challenge the Chinese ones in turn.
I don't know why, but I personally find Mistral's marketing strategy much more appealing than that of other companies.
For example, there's something about Anthropic's picked design and their little Claude avatars that's unsettling to me.
[delayed]
The entire Anthropic branding is religious kitsch - deeply off putting, but apparently quite reflective of their reality.
I think there's such a thing as throwing too many marketing and sales people and too much polishing and "refinement" at something. I see in new product announcements from Microsoft as well. It's like seeing someone try too hard to impress you.
And their cookie banner. Never thought I would like a cookie banner
Very nice progress. Also I respect putting Kimi on those charts. Regardless of if they are beating the Pareto frontier (not now), model diversity is a good thing for humanity — I’m hopeful for the team to keep increasing their gains.
This is awesome, one of the coolest Pokemon ever too for those that don't follow that universe :)
https://bulbapedia.bulbagarden.net/wiki/Lechonk_(Pok%C3%A9mo...
Open weight, European, competes with GLM-5.3 on cybersecurity. What's not to like?
Excited to try this. The low costs v. benchmarks alone here are worth a serious test. K3 has been my daily driver for a month or two now and it's dramatically reduced token spend (while not having much of a negative impact on productivity).
This was the era of the AI race I was waiting for.
Curious now that SOL is cheaper than k3 - is k3 still your primary workhorse?
> Trained from scratch
How are they training without pirating the Z library corpus and all that?
Dupe.
https://news.ycombinator.com/item?id=49977979 (200 comments now)
How could I resist switching to a model named after my cat!?
Stats be damned irrelevant. The naming is good with this one!
Excited to see this! Nice that they are saying this is just a first step.
Give them more compute!
Did they release it again? https://news.ycombinator.com/item?id=49977979
Of course not. That's clearly a different URL.
I'm glad they're keeping at it!
Don't believe Mistral. They're wrong. It's really called "Le chaton fat".
Also 1T-A49B. Weights currently closed but promise to open source them by the end of the month.
Great release movie.
>Don't believe Mistral. They're wrong. It's really called "Le chaton fat".
OpenAI's therapist: Le Chaton Fat isn't real and cannot hurt you
Le Chaton Fat:
better than K3 and DS4, cool
Some more discussion: https://news.ycombinator.com/item?id=49977979
sorting the charts like that gives off weird vibes
Had the same thought - feels chart crime adjacent
Bar charts should start at zero. If they don't start at zero, there should be a clear visual indicator that the chart has been trimmed without having to read the axis labels. I hate that this has to be repeated so often that it has become a cliché.
Awh I was half expecting a zombie pirate..
We are so back
i subbmitted a partnership proposal in your contact.
A bit disappointing to see it still lagging behind Chinese open models. Those Chinese models are pushing proprietary models to raise the bar, but we need equally strong non-Chinese open models to challenge the Chinese ones in turn.
YES finally
Now THAT'S how you name a model. Take note, others.