Sans context, I really like the Anthropic and Claude's faux-academic, minimal-ish, intellectual-ish branding.
(Especially the way it looked ~12 months ago -- it's gotten more cluttered since then. Perhaps unavoidably, as the breadth of their offerings has grown)
But over time it's begun to feel like unsettling cognitive dissonance as their ambitions grow and the stuff to worry about has piled up.
I think there's such a thing as throwing too many marketing and sales people and too much polishing and "refinement" at something. I see in new product announcements from Microsoft as well. It's like seeing someone try too hard to impress you.
Very nice progress. Also I respect putting Kimi on those charts. Regardless of if they are beating the Pareto frontier (not now), model diversity is a good thing for humanity — I’m hopeful for the team to keep increasing their gains.
Excited to try this. The low costs v. benchmarks alone here are worth a serious test. K3 has been my daily driver for a month or two now and it's dramatically reduced token spend (while not having much of a negative impact on productivity).
This was the era of the AI race I was waiting for.
I think it's actually a wink and a nod towards the social media meme of "Le Chaton Fat," a fictional model that is jokingly attributed to Mistral, usually with century-defining benchmarks and unfathomable size.
Bar charts should start at zero. If they don't start at zero, there should be a clear visual indicator that the chart has been trimmed without having to read the axis labels. I hate that this has to be repeated so often that it has become a cliché.
A bit disappointing to see it still lagging behind Chinese open models. Those Chinese models are pushing proprietary models to raise the bar, but we need equally strong non-Chinese open models to challenge the Chinese ones in turn.
I don't know why, but I personally find Mistral's marketing strategy much more appealing than that of other companies.
For example, there's something about Anthropic's picked design and their little Claude avatars that's unsettling to me.
Sans context, I really like the Anthropic and Claude's faux-academic, minimal-ish, intellectual-ish branding.
(Especially the way it looked ~12 months ago -- it's gotten more cluttered since then. Perhaps unavoidably, as the breadth of their offerings has grown)
But over time it's begun to feel like unsettling cognitive dissonance as their ambitions grow and the stuff to worry about has piled up.
The entire Anthropic branding is religious kitsch - deeply off putting, but apparently quite reflective of their reality.
The logo isnt a stylized butthole. That sure helps endear me to them.
I think there's such a thing as throwing too many marketing and sales people and too much polishing and "refinement" at something. I see in new product announcements from Microsoft as well. It's like seeing someone try too hard to impress you.
They've polished all personality out of their companies.
Yes, sets off the spidey senses.
And their cookie banner. Never thought I would like a cookie banner
Very nice progress. Also I respect putting Kimi on those charts. Regardless of if they are beating the Pareto frontier (not now), model diversity is a good thing for humanity — I’m hopeful for the team to keep increasing their gains.
This is awesome, one of the coolest Pokemon ever too for those that don't follow that universe :)
https://bulbapedia.bulbagarden.net/wiki/Lechonk_(Pok%C3%A9mo...
Was wondering what Le Chonk meant (not French).
Open weight, European, competes with GLM-5.3 on cybersecurity. What's not to like?
It's French, duh. They'll probably surrender half of the company to the Germans after receiving a strongly (compound-)worded letter.
One step closer to Le Chaton Fat.
Excited to try this. The low costs v. benchmarks alone here are worth a serious test. K3 has been my daily driver for a month or two now and it's dramatically reduced token spend (while not having much of a negative impact on productivity).
This was the era of the AI race I was waiting for.
Curious now that SOL is cheaper than k3 - is k3 still your primary workhorse?
Is that a reference to LeChuck in Monkey Island? Love that game!
No, it's a play on the Twitter/Reddit memeing about "Le Chaton Fat", see: https://www.reddit.com/r/MistralAI/comments/1u6f0dm/what_is_...
You'd be surprised how much of the AI world is fueled by memes.
I think it's actually a wink and a nod towards the social media meme of "Le Chaton Fat," a fictional model that is jokingly attributed to Mistral, usually with century-defining benchmarks and unfathomable size.
https://x.com/i/trending/2066326562623127678
How could I resist switching to a model named after my cat!?
Stats be damned irrelevant. The naming is good with this one!
you and 10k other redditors
Dupe.
https://news.ycombinator.com/item?id=49977979 (200 comments now)
Excited to see this! Nice that they are saying this is just a first step.
Give them more compute!
> Trained from scratch
How are they training without pirating the Z library corpus and all that?
I interpret "scratch" to mean brand new weights. Not that they aren't training on a corpus of human text
Did they release it again? https://news.ycombinator.com/item?id=49977979
Of course not. That's clearly a different URL.
I'm glad they're keeping at it!
Don't believe Mistral. They're wrong. It's really called "Le chaton fat".
Also 1T-A49B. Weights currently closed but promise to open source them by the end of the month.
Great release movie.
>Don't believe Mistral. They're wrong. It's really called "Le chaton fat".
OpenAI's therapist: Le Chaton Fat isn't real and cannot hurt you
Le Chaton Fat:
sorting the charts like that gives off weird vibes
Had the same thought - feels chart crime adjacent
Bar charts should start at zero. If they don't start at zero, there should be a clear visual indicator that the chart has been trimmed without having to read the axis labels. I hate that this has to be repeated so often that it has become a cliché.
better than K3 and DS4, cool
Some more discussion: https://news.ycombinator.com/item?id=49977979
Awh I was half expecting a zombie pirate..
A bit disappointing to see it still lagging behind Chinese open models. Those Chinese models are pushing proprietary models to raise the bar, but we need equally strong non-Chinese open models to challenge the Chinese ones in turn.
We are so back
YES finally
Now THAT'S how you name a model. Take note, others.
[flagged]
[dead]
i subbmitted a partnership proposal in your contact.