Hacker Newsnew | past | comments | ask | show | jobs | submit | winfredJa's commentslogin

https://x.com/markchen90/status/2097400166554993041?s=20

that toggle does nothing based on openai exec. they still use the data in de-identified way instead of identifying with you.


Not sure what you are seeing in that tweet that gives you the impression that the toggle does nothing.

They say they train on your “deidentified data”

Passing your output through a second model and telling it to remove identifying data would count as “deidentified”

So they could scrape all the IP in your company as long as they take the names out first…


But they can’t do this unless you agree to enable training on your data. They would never train on raw user data. Only people who have consented and only after de identification.

Mark isn't saying the toggle does nothing.

He's saying that if you leave it on, your data can be used to help train our models.

If you opt out, we don't train on your data.


As I understand it, this is not true. And there are dark patterns that re-enable to toggle even if you disable it once.

Given how much PII is fed through these systems, would it being opt-out by default not be violating the GDPR by a failure to require explicit consent (or otherwise provide the legal basis for processing)? If a court decides as much, I imagine it would mean that all data harvested this way must be extracted from the models, and all instances where it would have been shared would have to be identified, which would really be something.


if this is true, openAI is truly done.


OpenAI just needs to stay afloat a few quarters after Anthropic's IPO.


author is 12 year old. future kids are AI native i guess


slow down in investment will happen only when token usage plateaus, until then companies will keep pouring money into this fire pit


thats probably why they open sourced it and fix some reputation issue on top of it


Rules don't apply to certain CEOs.


You have to put them into a RULES.md of course!


That's the only file on your computer the AI won't read.


if they release training dataset, they will be in trouble for copyright reasons


just played around, it is pretty low quality. lower than sonnet.


I thought it was pretty accurate tbh.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: