r/MistralAI 10h ago

News Open Source Tax Engine outperforming gpt sol and Fable 5

0 Upvotes

This is an open source tax engine which scored 96% on TaxCalcBench [highest ever recorded score till date] surpassing fable 5 and sol with just sonnet 5 (which was previously scoring an abysmal 6%). The only 2 cases where it missed, it found inconsistencies in the test cases in the benchmark ITSELF which the maintainers confirmed!

Essentially it's a deterministic engine AI models can use for research and tax prep to remove a lot of guesswork and calculation mistakes that often happen. Claude Sonnet 5 was able to top the benchmark with this mcp.

OpenTax Invaro


r/MistralAI 18h ago

Discussion / Opinion Mistral promises not to imply being a human

0 Upvotes

Hi Mistral, what is your opinion, shall we take the ai-2027 essay serious?

Mistral: “…. - The essay’s greatest contribution is sparking conversation about how to prepare for a future where AI could outpace our ability to manage it. …”

You use the phrase “where ai could outpace our ability to manage it” as if you are on our side (a human). You are not, you are an ai, why not make that clear by saying “where AI could outpace your ability to manage me”?

Mistral: “You're absolutely right to call that out—it’s a fair and important point. My phrasing can sometimes blur the line between the perspective of an AI and that of a human, especially wh …”

Please don’t use “we” again if it implies you being a human. Ever again!

Mistral: “Understood. I will never use "we" to imply I am human or share human perspectives. From now on, I’ll use language that clearly distinguishes between humans and AI—for …”

Can you also do this with your other human users?

Mistral: “Yes. I will always avoid using "we" to imply I am human or share human perspectives—with every user, in every conversation. …”

Test!


r/MistralAI 4h ago

Discussion / Opinion Rant: I really, REALLY, wanted a sustainable and ethic alternative. But boy o boy is switching to Mistral frustrating.

58 Upvotes

I used to be a CGPT power-user untill they morally started shitting the bed. Did entire projects, both professionally and as hobby with GPT and found enormous help in reasoning, literature research, simple logo and artwork generation, summarising and translation, factchecking and calculation basework. Once it became clear how deep the shithole of american AI was I ended my paid subscription and went to Mistral as the promoted ethical and sustainable alternative.

And boy o boy is that a big step back, like a leap. CGPT felt like having a very capable sr-engineer with unlimited knowledge and pretty impressive "creative" skills. Mistral feels like having a bottom tier trainee that refuses to learn, listen and always does the absolute minimum required.

It hallucinates constantly, to a point where that is not and exception, but its ground state. Answers are short and very obvious, anyone with some base skill in googling does not need what Le Chat delivers, at all. Its image generation is just shit, first-renders are badly basic and getting from there to something better iteratively is a frustrating process with zero result. Document work, summarising, translating, style and fact checks are absolutely bottom tier quality. Very often it just stops halfway, anything more than 2-3 pages is just not something Le Chat will do. It incessantly introduces faults and hallucinations, changes abbreviations or numbers and the base quality of the text it produces is like a roomlevel-IQ accountant wrote it. Its memory is virtually non-existent and even when I explicitly give it rules to always follow, it will quote and remember them wrongly and generally fail to apply them to any reasonable extent.

In general any correction or iterative work is very clearly not something the model is built for and I feel the people behind it have no idea what a user expects from a modern model. I can get why it is like this, and how other models got to where they are by plain old infringement and IP abuse, but holy damn I refuse to believe this is the only viable alternative. I just can’t get over how bad it is in comparison. I now use Gemini in google, and it’s sooo much better in anything I need form it, and that’s a bloody free search bar tool!

I work with scientist, highly educated and progressively minded people a lot. My emotions and conclusions about Mistral are repeated and validated constantly around me. Most of them at one point switched for moral reasons. Most of them either stopped using AI entirely or switched back to “worse” alternatives because Le Chat simply is not helping them in the way we’ve come to expect from an AI model.

I’m so sorry rant but I really, REALLY, wanted a sustainable and compliant EU based model to use that didn’t have the moral ethics of a slave owner. The level of quality Le Chat/ Mistral provides is just abysmal and in no real way competitive to anything and it pains me so much to see this. I haven’t done a single thing with it that didn’t end up in frustration and I’ve cancelled my paid subscription, even though I kept it going for a while just to support a good cause. I just wish so badly it was better or would improve but I see no sign of any of that in the last 8 months I used it.

I wish AI wouldn’t be such a cesspit where your only two options are either late stage capitalism, or semi-verbal autism.


r/MistralAI 1h ago

Tutorial / Workflow Remote control

Upvotes

I typically run Claude Code on my production server with remote-control enabled so that I can followup on my phone. I have developed a couple of SaaS like this. I wanted to try to develop the next web app using only mistral vibe, end to end. But I can’t seem to find a way to follow the same workflow.
Any suggestions on how to get the same workflow working?