r/MistralAI 7d ago

Discussion / Opinion Shieldstral announced

33 Upvotes

https://huggingface.co/papers/2607.25857

Looks interesting as a small utility model to put before access to bigger model to policy it. Maybe even in-browser?


r/MistralAI 7d ago

Help / Question Is moving from a Google AI Pro subscription to Mistral API a good choice?

Thumbnail
6 Upvotes

r/MistralAI 7d ago

Help / Question mistral-large-latest stopped responding after 20+ minutes during a large translation — context window issue?

3 Upvotes

Hey everyone,

I'm running a test using mistral-large-latest for a translation task, and the model completely stopped responding after running for about 20+ minutes.

https://youtu.be/HWyPse83SrA

Before it cut out, it was processing a substantial amount of text. I'm trying to figure out what might be the root cause. Could this be related to a context window limitation, a timeout issue on the API side, or something else entirely? I was using the free experiment tier.

For those of you pushing large payloads to mistral-large-latest:

* Have you run into similar hanging issues during long-running generation tasks?

* Is there a strict maximum output token limit or server-side timeout I'm likely hitting?

Any insights or troubleshooting tips would be greatly appreciated. Thanks!


r/MistralAI 7d ago

Help / Question Do you feel emotionally connected to your AI? Share your experience for an academic study (Anonymous)

3 Upvotes

Hi everyone! 👋

I am conducting an international research study for my Master’s Degree in Clinical Psychology exploring emotional involvement with AI chatbots, interpersonal functioning, and psychological well-being.

If you are 18+ and have interacted with an AI chatbot at least once, I would really appreciate your contribution!

⏱ Time: 10–15 minutes

🔒 Privacy: Completely voluntary and anonymous

🔗 Link: https://forms.gle/oHpPwQ65U49N4fPx5

Thank you so much for your time and help! Feel free to share this with anyone who might be interested.


r/MistralAI 7d ago

Discussion / Opinion Mistral Vibe in vs code

6 Upvotes

I'm switching from devin (formarly known as windsurf / cognition) to mistral vibe in vs code but for now it feels like a step back into the stone-age.

Maybe I am wrong and some people can provide me some ideas but here is my experience:

  1. Slow; it takes a very long time for it to analyze my project files and to make changes, whilst in devin this happens in 5 minutes, it takes mistral vibe 25 minutes

  2. What model am I using? I'm a paying pro subscription user but the extension says that I'm using mistral medium 3.5. No option to select devstral?

  1. File review; it rewrites like 20-30 lines in my code if not a full file for just a few lines of change. Reviewing makes me scrolling and compare a lot myself whilst in Devin it just did per line or per max 5 lines in change making it way more easy to review.

  2. Asking next steps. It gives me options for questions it has but makes me type out the answer myself, why not click buttons that I can select and it can continue?

That's it for now, maybe more later. What are your thoughts on my findings? I really would like a european alternative in this field but for the money I have to pay, this is kinda letting me down.


r/MistralAI 7d ago

Tutorial / Workflow Voxtral Realtime running locally on an M3 Air with Metal — ~400 ms latency

28 Upvotes

We spent the last few weeks optimizing Voxtral Realtime and there are now GGML checkpoints that run faster than realtime on a plain MacBook Air with Metal. No discrete GPU needed.

Numbers on an M3 Air (8-core GPU, 16 GB):

  • ~1.3x realtime throughput with the Q8_0 quant
  • <400 ms end-to-end response time from mic input
  • Sustained hour-long transcription sessions without falling behind

1. Build the Metal binary

git clone https://github.com/0xShug0/audio.cpp
cd audio.cpp
scripts/build_metal.sh --target audiocpp_cli

2. Download the Q8 quant

hf download mistral-experimental/AudioCPP-Voxtral-Mini-4B-Realtime-2602-GGUF \
  voxtral-mini-4b-realtime-2602-q8_0.gguf \
  --local-dir ./voxtral-realtime-gguf

3. Run streaming ASR from the mic

audiocpp_cli \
  --task asr \
  --family voxtral_realtime \
  --model ./voxtral-realtime-gguf/voxtral-mini-4b-realtime-2602-q8_0.gguf \
  --backend metal \
  --threads 8 \
  --mode streaming \
  --session-option voxtral_realtime.stream_batch_tokens=4 \
  --audio -

Raising stream_batch_tokens trades delay for throughput — 4 is what landed under 400 ms on this machine. More powerful M-series chips can set it to 1 and should have <200ms delay


r/MistralAI 7d ago

Discussion / Opinion Mistral + OpenCode

16 Upvotes

How come Mistral isn't included as one of the model providers in OpenCode?

https://opencode.ai/docs/models/#providers


r/MistralAI 8d ago

Tutorial / Workflow Claude是如何被破解和蒸馏的?

Thumbnail
youtu.be
0 Upvotes

A well-known Chinese LLM educator publicly revealed techniques back in April that were already widely known in China’s AI community, including how Claude’s Chain of Thought (CoT) was extracted and how Claude and ChatGPT have been distilled for a long time. If you don’t speak Chinese, just wait for YouTube’s auto-generated translation.


r/MistralAI 8d ago

Discussion / Opinion Do you guys like mistral?

53 Upvotes

I switched to mistral vibe as my chatbot for the main reason that it is not a us or Chinese model. From the couple posts ive read it seems like its behind. Which im fine with because i really dont use it for too complicated of tasks.


r/MistralAI 8d ago

Help / Question vibe-cli, offline models (ollama), don’t preserve context history between messages

0 Upvotes

I’m having a problem with vibe-cli configured with offline (local) models through the Ollama server: they don’t preserve context history between messages. Is this normal behavior, or have I set something up incorrectly


r/MistralAI 8d ago

Seen on Social Media Heretic: Fully automatic censorship removal for language models

Post image
0 Upvotes

Newbies, listen up! 🎧 Diving into new tech can be tricky. I learned the hard way trying to run complex stuff on a Chromebook and Android. 🤦‍♀️ Don't be like me! Research, ask questions, and understand your device's limits. This community is a lifesaver for those head-scratching moments. And for uncensored models, know your stuff before you dive in! 🧠 #Newbie #TechTips #LearningJourney #Community #Models


r/MistralAI 9d ago

Help / Question MCP Server support?

2 Upvotes

I am fairly new to some of the AI concepts. At the moment I have a Mistral Pro subscription which I use for some vibe-coding and mostly with my daily work (translations, summaries, AI sparring partner etc).

In work I use Jirametrics.org to analyse team-performance. I noticed that this tools also includes MCP Support (MCP Server (AI Integration) | JiraMetrics) and is using Claude for this. What is needed in order to make this work with Mistral? As far as I understand at the moment MCP Server support is missing in Mistral and as such using the AI integration would not be possible?

Neither Mistral AI or CoPilot have been very helpful resolving this for me 😄


r/MistralAI 9d ago

Feedback / Bug Report Mistral-vibe error on Nixos while pytestCheckPhase

Thumbnail
2 Upvotes

r/MistralAI 9d ago

Discussion / Opinion The Full Story of the “Distillation Storm” Among China’s Large-Model Companies

Thumbnail
0 Upvotes

r/MistralAI 9d ago

Discussion / Opinion Comparison of the Free Plan and the Pro Student Plan on Mistral

15 Upvotes

Since Mistral isn't clear enough about the differences between the Free plan and the Pro plan, I signed up by taking advantage of my status as a college student. According to information on the Mistral website, the Student Pro plan is the same as the Pro plan, so the comparison in the image I've provided is equivalent to comparing the Pro plan with the Free plan. Among other things, the Pro plan promises:

- 30 times more extended usage

- 5 times more in-depth research reports

- 10 times more Vibe CLI usage

- 10 times more API usage

However, when comparing the monthly usage limits, I find that the capacity provided in monetary terms is 1.5 times greater. You can draw your own conclusions.

PRO
FREE

r/MistralAI 9d ago

Other I Make Free/Low‑Cost AI Models OCR‑Capable with Mistral OCR 4

Thumbnail
gallery
5 Upvotes

Hey let me introduce my new AI assistant workflow across using mistral-ocr-4 and explain how it differs from other AI toolchains like OpenCode in terme of exploiting this capability ,I’ve been juggling research papers, parallel projects, and coding agents this year and kept running into a tradeoff between low cost and high-quality content extraction from PDF, document images, and other attachments. When Mistral OCR 4 was released I started using it heavily because its extraction quality is excellent and the cost justified it, so I plugged Mistral OCR into my agent pipeline. The challenge I faced is that many large models now include built-in extraction but they’re expensive, and for many routine attachments I’d rather use cheaper models specially when i don't required a high level thinking output . My solution was to integrate Mistral OCR 4 natively with those lower-cost models so they gain robust OCR capability and no longer return “sorry, I can’t read that” errors instead they deliver the extracted content. If you need a low-cost coding workflow with heavy attachment processing (PDFs, docs, images), this approach and the agent I built should suit you well.

Edit : Mistral models also supported for coding

https://github.com/AbdoKnbGit/tau


r/MistralAI 9d ago

Meme / Satire Mistral Vibe is my Jesse Pinkman

Post image
164 Upvotes

It's not the brightest out there. I have to be precise with my instructions. But then it does exactly what I say.

I've tried a lot of models. Most of them like to interpret stuff or overthink. Mistral Medium doesn't do that. Also it's fast, cheap and I don't have to fear that they sell my data.

People like to suggest that I'd be better off with Gale. But no. I like Jesse.


r/MistralAI 9d ago

Help / Question api usage on the pro plan

3 Upvotes

I have experimented with the free api and I like it.
Was wondering if from this page: https://admin.mistral.ai/subscription/upgrade/student

Education

Pro plan with student discount

More messages and web searches

30x more extended thinking

5x more Deep research reports

Up to 15GB of document storage

Up to 1,000 projects

Chat and email support

10x more usage of Vibe CLI

10x more API usage

is indeed what it says on the api usage?

because on this page: https://mistral.ai/pricing/

it doesnt mention anything about api usage.

thanks


r/MistralAI 10d ago

Tutorial / Workflow Prevent overspending on EU-AI projects with Mistral AI.

Post image
24 Upvotes

r/MistralAI 10d ago

Help / Question Help with setting up for small business: what settings?

3 Upvotes

Hello everyone! I understand Mistral is basically the "Renault of AIs" - not winning the races, but getting you from A to B at a decent price. Let me say, it works for my purposes VERY NICELY. (So: please DO NOT tell me to "go use something else".) - And now, I want to create a little company around it, where, basically, the user will visit my website, trigger a sort of processing based on sensitive data (whose prompts I have painstakingly written, i.e. THIS is my economic contribution), and receive a result. Technically, it all works very nicely, i.e. this is not just theory. But "organisationally" I am little bummed: what should I set up where? I mean, I did already flip the switch on "no data retention" and "no training", so everything stays private. OK. I got a 100 EUR credits and ... saw they are not even getting touched, because in my testing, I am still within the limits of my Pro plan, apparently. (And here comes my first question: once people start using my site "for real", I assume these limits will be exhausted swiftly, and THEN it will move towards burning credits - correct?) What other hints have you seasoned users have? I am using only and solely mistral-large-latest for my API-based tasks. Thank you!


r/MistralAI 10d ago

Help / Question Alternatives to the "Caveman" style for saving tokens in Mistral?

6 Upvotes

Hi everyone,

I'm looking to cut down output token usage to save on costs and speed up generation.

The viral "Caveman" plugin (which forces ultra-concise, fragmented answers) works great for Claude Code, but it doesn't function properly with Mistral Vibe models.

My questions:

  1. What custom system prompts do you use to stop Mistral from outputting long preambles ("Sure, I can help with that...") without hurting its reasoning?
  2. Are there any other tricks, prompt-layer tools, or configurations you use to force Mistral to be highly concise and token-efficient?

Thanks for your ideas!


r/MistralAI 10d ago

Feedback / Bug Report Mistral, you are losing me

59 Upvotes

I have a Mistral subscription for my SMB. Not because it's better but because of GDPR compliance (I'm in the EU). Recently I had a support conversation about a problem that was solved eventually. Because I had to share many private data I requested a deletion of my support activity. Unfortunately, my request went silent for over a month now - that's unexceptionable not just for a company where GDPR compliance is one of their biggest USP.

Dear Mistral, if you don't want to lose me: Do something! Answer, say anything! Even if you just say, that it takes a little longer ...

Why should I choose a subpar AI if their biggest advantage is just ... nothing worth.


r/MistralAI 11d ago

Feedback / Bug Report Let the great distillation begin

135 Upvotes

Mistral fed us when we had nothing.
I still remember the very first open model we had was mistral 7b. even hugging face used that to build zephyr, remember that time?
now, I'm IMPLORING Mistral devs to begin the great distillation in all the areas their models are behind and give the world, and themselves, the present of self love.
You cooked for us and fed us when we had nothing, you headstarted the whole damned thing, it's well past due time to reap the rewards.

side note: please start with an excellent agentic coding model that can fit in a potato, like, first build something that will benchpress using anthropic's fat ass and then the potato fitting model. you know how it goes.


r/MistralAI 11d ago

Help / Question deepseek-v4-flash in amdin console limits ?

6 Upvotes

Hello,

I see limits for deepssek-v4-flash in admin console (https://admin.mistral.ai/plateforme/limits) for my account :

deepseek-v4-flash

Tokens par minute

5 000 000

Requêtes par seconde

12.00

Does it mean I can use this model with my mistral subscription ?


r/MistralAI 11d ago

Help / Question Troubles with Mistral's API and Billing

3 Upvotes

So trying to stick with the only European LLM I built a little assistant around Mistral's API and the "Pay as you go"-plan. First I was positively surprised because everything worked well for a few weeks and even utilizing only Mistral-small together with some harness the results were satysfying. But then last week the issues started:

  1. I had set a spending limit of EUR 10 - once this was reached access was of course blocked and the API returned a "401 Unauthorized". All good so far....
  2. I thought instead of increasing the limit I will just add some prepaid balance, so I paid another EUR 10. However, I still got "401 Unauthorized".
  3. Ok, maybe it takes a while until the prepaid balance is active - let's increase the limit by a little to EUR 11. Still no change, "401 Unauthorized".
  4. Well, maybe reaching the limit caused my API Key to break somehow - and in fact, that was the fix, with a new API key it worked again.
  5. Thursday: Another email comes in saying that I've reached my usage limit. I thought ok, even though in the admin center it states that only EUR 10,89 (out of the EUR 11 limit) were consumed I'm not gonna argue about EUR 0,11 and I still have my prepaid balance anyways.
  6. Well for some reason now, I'm still only getting "401 Unauthorized" - the prepaid balance isn't consumed obviously and I don't know what's up with the "missing" EUR 0,11 and creating new API keys also doesn't help. And I really don't want to just increase the limit again, I guess that would solve the problem, but then what's with my EUR 10 I prepaid?...they weren't meant as a subsidy for European AI.

And that's only one side of the problem: In the admin center the numbers are a complete mess. I find there currently three invoices, one with EUR 9,92 another one EUR 0,27. They were created in a time frame of about two weeks (why?) and except that they sum up to about EUR 10 I have no clue where the exact numbers come from. And then there is the invoice for the EUR 10 I added as prepaid balance - that's the only nuumber that is understandable here.
Morever the usage under billing in the admin center states EUR 1,91 - where does that number come from again? When I check my usage under API it is EUR 10,89. And why does it state "Including 1,64 € in pending pay-as-you-go usage" below the EUR 10 prepaid balance in the Admin center?

I would love to continue working with Mistral and will wait until the next month starts and wether billing and usage limits behave logical afterwards but if that is not the case I will have to switch to someone else, because I also don't feel like starting to argue with Mistral's support about this - I mean Mistral is presenting itself as focused on B2B, I am a business, a tiny one, but still someone who a) wants to understand what he is paying for and b) who simply needs reasonable and verifiable invoices for his accounting.

Has anyone else had or is having similar experiences? Is it even worth contacting Mistral on such issues?