← Back to context

Comment by teekert

17 hours ago

Context: After careful research our organization preferred a European partner with good central privacy controls. We landed on Mistral, after being disappointed that the Pro tier was opt-in to training on prompts by default we switched up to the Team tier which provides an organization dashboard with some relevant settings. As we did that Mistral changed these options and the Team tier was now also opt-in by default and at the same time seemed to have lost the ability to centrally disable training on prompts for your entire organization. This even caused some of our (testing) prompts to be used for training (which Mistral removed after we expressed our disappointment).

For some time these pages conflicted with what our users reported (they said that in contrast to what I stated to our management they found they were opted into training on prompts by default as per their own privacy page). Mistral just now corrected their docs. I'm not sure how long the conflicting situation has lasted, but at least for several days.

For contrast: Claude disables training on prompts for organizations starting from the 18 euro tier [0]. As a European I'm disappointed.

[0] https://claude.com/pricing#team-&-enterprise

> and at the same time seemed to have lost the ability to centrally disable training on prompts for your entire organization

There is a toggle on https://admin.mistral.ai that allows you to disable training for both Vibe and Console/API for your entire organisation. And I'm not on the enterprise plan. I've disabled training the first time I created an account, and it has remained that way.

You story is also very confusing due to the wording around "opt-in by default" and "disappointed [about] opt-in to training" (most people would be disappointed about an opt-out) and probably conveys the wrong message to most people.

  • I don't have that toggle (but could indeed have sworn I saw it earlier).

    Sorry, I have always thought that "opting in" is, "opting for the presented option" and opting out is "opting out of it", so opting out [of sharing prompts for training] is choosing to not share, but apparently I was wrong my whole life. I'm not a native speaker, and I think most people here (in my country) would interpret this the way I do? Weird but TIL.

    • FWIW I don’t find what you wrote confusing, but I do think it is following a sort of… bad convention that some companies have been pushing.

      Opt-in and opt-out describe the nature of the choice that you make. “Opt” means to choose (apparently it is a French word we stole). Opt-in means you have to proactively choose to be in. Opt-out means you have to proactively choose to be out. “Opt-in by default” is an overly verbose way of saying “opt-out.”

      In either case it describes the choice that you need to proactively make to override the default behavior.

      Edit: I should also say that it is a “known point of contention” where pro-privacy people have been pushing back on this phrasing. So, you have probably accidentally stumbled into an ongoing discussion, which is why some of the comments might be unexpectedly prickly.

      3 replies →

    • If you opt in it means that by default you're not in and you chose that option. So if you're now included in training by default, then it's an opt-out feature as in you can opt out of it.

      2 replies →

    • I think there is a simple, unambiguous way to describe this, which is to use the passive voice or an explicit subject with the past tense to clarify that you didn't do the opting in.

      Saying "we were disappointed to be opted in to training by default" clarifies the point you're trying to make, which is that the toggle (whatever it may be called) was set to training by someone else, not you.

      Actually in your case you'd say "...after being disappointed that the Pro tier opted us in to training on prompts by default", which gives an explicit subject ("the [Mistral] Pro tier"). No one will confuse that with the opposite "the Pro tier opted us out of training on prompts by default".

  • My native-US-English speaker read:

    “disappointed that the Pro tier was opted-in to training on prompts by default” [and required manually opting out]

    “the Team tier was now also opted-in by default” [and required manually opting out]

    In context of each sentence and the larger comment, read smoothly here.

    Also - have seen more than one lively discussion on these phrases, since defaults can stick 95% of the time and Big Tech has done their best to be abusive about what they automatically enable for users by default for some time.

    • It's understandable but it is technically inconsistent and dilutes the meaning of opt-in. Opt implies an active choice, so you never are opting for the default.

      1 reply →

"For contrast: Claude disables training on prompts for organizations starting from the 18 euro tier "

In theory also for individuals?

At least I have that toggle to deactivate that. But how would I ever know if they actually respect that?

  • This post is only about the defaults, and expectations therefore for Team and Enterprise plans and their organizational controls.

  • If you don’t trust the company to keep their promise, don’t use their products.

    • How can you trust any company?

      The only company you could think of trusting is one where an external , independent auditor is doing its work.

    • > If you don’t trust the company to keep their promise, don’t use their products.

      Your comment is perplexing. No company on earth meets your requirement. What are you expected to do? Move to a hut in the woods?

      2 replies →

    • That is ridiculous. Vote to ensure products have to be legally private and make it a crime to share or reuse your data for anything. This should be the norm that people vote for. The idea that we have to give up privacy so some nerdy pervert can be a billionaire off the backs of people who do the work is absurd.

      Criminalize failure to keep private data private. Arrest CEOs and executives. Put them in jail when it happens.

  • How do you know an LLM provider does not kill puppies every time you send a prompt longer than 14 words?

    • Because they have no incentive as a company to do that? But do have a strong incentive to learn from user input. (Most cutting edge features are developed private - very valuable to get that into your tool)

      And getting the data is not hard, they already have it. Risky is indeed a bit making use of that data, as that requires at least some humans (as potential whistleblowers). But you don't even have to tell them, where the data came from.

      Whether they do it? No idea, I assume not, but I see a risk.

      2 replies →

I don't trust any of these companies with my data, and I assume that whatever data they've got is going to be used, one way or another, no matter what they tell you. It would be nice if you could stick a sentinel in your data that if it ever shows up in the models you know they've broken the rules for sure.

  • Yeah. Unfortunately they know you can't catch them on this so they feel absolutely free to do whatever they please

    If such a data sentinel did exist then we might see them change their behavior

Today I registered a free account with Grok, because I simply was curious. Man, training is "off" by default even with the free tier. I was positively surprised. As a European I'm disappointed too.

  • Grok has possibly the worst ToS of any of the AI providers. They are probably different in the EU, but:

    > In choosing to submit, create, generate, record, post, or display Inputs on or through the Service, you grant an irrevocable, perpetual, transferable, sublicensable, royalty-free, and worldwide right to SpaceXAI to use, copy, store, modify, process, adapt, transmit, distribute, reproduce, publish, upload, download, display in public forums, list information regarding, make derivative works of, and distribute such Content, including anything referenced therein, in any and all media or distribution methods now known or later developed, for any purpose, and to aggregate your User Content and derivative works thereof for any purpose, including but not limited to: (i) maintain and provide the Service; (ii) improve our products and the Service and for our other business purposes, such as data analysis, customer and market research, developing new products or features, or identifying or displaying usage or User Content trends; and (iii) perform such other actions to enforce these Terms, comply with our Privacy Policy, comply with applicable law or governmental, court, and law enforcement requests or requirements or keep our Service safe.

    > To the extent the User Content includes a person’s image, likeness, voice, or other similar attributes, you grant SpaceXAI the same rights to use those attributes as part of the User Content as described above. You represent and warrant that you have obtained all rights, licenses, notices, permissions, and consents necessary for SpaceXAI to use that User Content.

    https://x.ai/legal/terms-of-service

  • Lol, I'm fully expecting in 6 months time:

    "A bug in our portal had the setting for training inverted. This means that when you expected us not to be training on your data, we actually were. We know this adversely impacts the trust our users invested in us, so as of today we are crediting all affected accounts with $200 to use on our latest models".

“Opt in by default” would mean that it is not enabled by default. Do you mean opt out?

  • You are "opting in to sharing your prompts for training", by default in this case. My slider says: "Allow the use of your interactions with Vibe to train Mistral's AI models.", it is on by default for everyone on the Team plan, the admin can't centrally turn it off anymore, and any user can toggle it when they want to. This all changed last week.

    I know I'm naive but I expect that when I pay, this stuff is simply off, so I was already surprised by the Pro plan. But I did look out for it there, because Anthropic made this switch some time ago.

    • > You are "opting in to sharing your prompts for training", by default in this case.

      I understand what you're saying here, but maybe "turned on by default" is less confusing for everyone.

    • > You are "opting in to sharing your prompts for training", by default in this case.

      The English term for that is "opt out" not "opt in." To opt is to choose. If something is on by default, you have not opted in. You were forced in, and turning it off means you must opt out. (I.e., choose to be out.)

      normally I wouldn't care about a mistake like this, except that opt in/out are very important concepts in software development and hacker culture. And it reversed the meaning of the original comment in a highly confusing, relevant way.

      7 replies →

  > opt-in by default

Sorry to nitpick but the scheme you’re referring to is called “opt-out.”

> being disappointed that the Pro tier was opt-in to training on prompts by default

"Opt-in" means that the default is non-participation, for example, not training on your prompts. Is it possible that you intended to say "opt-out", which means that the default is participation? That's what the context seems to suggest.

See, for example, https://termly.io/resources/articles/opt-in-vs-opt-out/:

> Data privacy laws like the GDPR and CCPA give individuals the right to opt in or out of different data processing activities.

> · Opt in consent means the user takes an action to show they agree to something,

> · Opt out consent is when they take an action to say no.

Or https://bigid.com/blog/opt-in-vs-opt-out-consent/:

> • Opt-in consent requires users to actively agree before data collection or processing.

> • Opt-out consent allows data collection by default unless the user declines.

This is an important distinction, because confusing the two (as you seem to be doing) can lead you both into unethical fraud and legal liability.