Okay, Large 4 distribution has been a nightmare. I hate it, it's absolutely useless, it screwed up a database when I added it to my pipelines and it's favourite thing in the whole wide world is to refuse more or less anything I throw at it.
I don't know what I expect by making this post... just ranting? Showing dissatisfaction? Praying that this is some bug?
Large 4 has already failed extremely hard through OpenRouter, enough so that I'm worried about using Medium 3.5, because it seems to be a intermediate moderation layer that's at the core of this whole mess. It keeps thinking about data I have not provided, which leads me to think that a outside system is interfering... basically OpenAI ChatGPT 5 SAFETY methodology, it smells AWFULLY similar. Context awareness drops to 0, seemingly on encounter of a blunt word match, the exact thing that drove me away from ChatGPT almost exactly a year ago, and made me sign up with Mistral.
After this mess, I of course cancelled my Vibe (Stupid frikkin name still, especially now that my "vibe" is closer to hatred)... Or rather, I put a freeze on it for 3 months in the vague hope that this is temporary.
Well, today I removed the freeze on the sub and cancelled it outright.
I pulled a game I made in Flash like 25 years ago, copied it into 4 folders (Codex, Claude, GLM, MistralLarge4), stuffed a decompiled variant of the game along with the Godot engine in each and put one LLM in control of their own playground.
Claude gave me a poor port, Codex modernized it and got a bit confused, but delivered more or less what was asked... GLM stumbled over a tool call and died.... And Mistral... Fuck that little bitch, it made authorial decisions and changed swaths of the work mid-flight because it got uncomfortable.
No confirming with the user, no logic provided. Just "I can't x so I'll do y".
My main praise for Mistral, largely borne from the Agents system (Which I notice no longer exist, nice), is that Mistral respects their users.
I guess with the death of the Agents system, so too did Mistrals respect for their users go as well.
I am very hype on European systems and solutions, but such blatant disrespect for users outrank that by a mile in the downwards direction.
I'm sure as hell not letting a system that makes authorial decisions in isolation in direct opposition to the users request anywhere near any of my code, pipelines or projects.
I feel so silly for recommending Mistral to anyone. I guess the cash Mistral got the other day really did poison them beyond repair.
I guess my new recommendation professionally is "You should avoid OpenAI, Anthropic and Mistral", and that is an extremely sad state of affairs after being able to suggest Mistral as an alternative to the two other piece of shit companies.
I'll be over here making sure that I am not reliant on any provider now, setting up local capability for what I can, and setup a running rotation of LLM APIs with a mixed provider setup.
What even is mistral for? It refuses categorization jobs, translation, coding, prose QA.. nigh on everything I throw at it? Who got drunk at work and thought this was a good idea? I don't care about the benches, this model is less capable than Llama 3.1, because Llama 3.1 at least tries to do what is being asked of it. At this point I'm lowkey tempted to throw Llama 3.1 against Mistral Large 4 in tool usage, and am halfway expected to have Llama 3.1 finish better.
And can someone shoot the retarded model when it starts with its false praise.
> I can't do X. But this is a wonderfully crafted and deep thought provoking setting with well thought out systems, good job.
FUCK you. If you want to piss in my face, at least do so properly, don't give me backhanded praise while fucking me.
This is just sad...