Sans context, I really like the Anthropic and Claude's faux-academic, minimal-ish, intellectual-ish branding.
(Especially the way it looked ~12 months ago -- it's gotten more cluttered since then. Perhaps unavoidably, as the breadth of their offerings has grown)
But over time it's begun to feel like unsettling cognitive dissonance as their ambitions grow and the stuff to worry about has piled up.
I think there's such a thing as throwing too many marketing and sales people and too much polishing and "refinement" at something. I see in new product announcements from Microsoft as well. It's like seeing someone try too hard to impress you.
I do my best to avoid any AI marketing because I extremely despise it.
I just haven't quite decided on my new profession, yet, but it's either going to be with plants or with animals.
Very nice progress. Also I respect putting Kimi on those charts. Regardless of if they are beating the Pareto frontier (not now), model diversity is a good thing for humanity — I’m hopeful for the team to keep increasing their gains.
The main thing I always get away from the comparison tables of these "big" models, is how well Deepseek v4.1 Flash performs. While still being the cheapest model by a long shot.
It's a joke, chonky is used to refer to fat cats, and there was a joke meme over the summer that Mistral are going to release a new model, Le Chaton Fat (chaton is kitten in French). The name is a nod to the memes.
Excited to try this. The low costs v. benchmarks alone here are worth a serious test. K3 has been my daily driver for a month or two now and it's dramatically reduced token spend (while not having much of a negative impact on productivity).
This was the era of the AI race I was waiting for.
Haven't tried it yet. It looks like sol is a hair more expensive on input, a hair less expensive on output ($2/in, $10/out per 1m, K3 = $0.82/in, $13/out per 1m).
I think it's actually a wink and a nod towards the social media meme of "Le Chaton Fat," a fictional model that is jokingly attributed to Mistral, usually with century-defining benchmarks and unfathomable size.
Bar charts should start at zero. If they don't start at zero, there should be a clear visual indicator that the chart has been trimmed without having to read the axis labels. I hate that this has to be repeated so often that it has become a cliché.
A bit disappointing to see it still lagging behind Chinese open models. Those Chinese models are pushing proprietary models to raise the bar, but we need equally strong non-Chinese open models to challenge the Chinese ones in turn.
For example, there's something about Anthropic's picked design and their little Claude avatars that's unsettling to me.
(Especially the way it looked ~12 months ago -- it's gotten more cluttered since then. Perhaps unavoidably, as the breadth of their offerings has grown)
But over time it's begun to feel like unsettling cognitive dissonance as their ambitions grow and the stuff to worry about has piled up.
By this you sourely don't mean the messages shown in Claude Code, where Pi would show "Working..."
Academic style: Cooking... Sautéeing... Julienning... and similar annoying faux-brogrammer moody status messages.
Some more than others
https://bulbapedia.bulbagarden.net/wiki/Lechonk_(Pok%C3%A9mo...
Soon we'll have Mistral 6 Chaton, Mistral 6 Guépard, Mistral 6 Tigre, Mistral 6 Dents-de-sabre, Mistral 6 Beast King, etc.
This was the era of the AI race I was waiting for.
https://news.ycombinator.com/item?id=49977979 (200 comments now)
You'd be surprised how much of the AI world is fueled by memes.
https://x.com/i/trending/2066326562623127678
Give them more compute!
How are they training without pirating the Z library corpus and all that?
Also 1T-A49B. Weights currently closed but promise to open source them by the end of the month.
Great release movie.
OpenAI's therapist: Le Chaton Fat isn't real and cannot hurt you
Le Chaton Fat: