If he didn't opt out I'm not sure I'd agree that it was fair game.
I'm pretty sure it would be considered plagiary amongst colleagues and it is a terrible precedent if we just let OpenAI steal any good idea they can get their hands on if they think it is profitable. You'd effectively sign away any and all rights to anything built with AI if OpenAI chooses to reengineer it before you.
Sure, but unless you’ve got some exceptionally deep pockets, congress has seemingly no interest in turning the fact that it’s ethically bankrupt into any practical recourse.
Ai companies got where they are by stealing all of the intellectual property from human history. It seems entirely likely that their goal is to purloin everything produced going forward as well.
If you don't trust the labs directly, you can always use AWS Bedrock or Azure Foundry which should have a much stronger incentive to not train. They make money from asset rental, not selling models. I'd be shocked if they were training.
This isn’t me advocating for maximum paranoia. I’m just saying that the “privacy” aspect of the API pricing is just a business promise. I do personal coding on API pricing with Claude mainly because a) I’m normally using it quite sparingly b) the commercial relationship is a lot more transparent than these subsidised relationships where they “may” be training on your work but can’t really seem to answer what that means.
I feel like there's a pretty huge difference between using inputs and outputs as part of a general training corpus, and looking at a specific users workspace after hearing rumours and yoinking their ideas to beat them to the point.
Unless OpenAI finished a whole new training run on the latest data in the last few days, the possible allegation seems to be the latter.
They have been collaborating on this solution for a year, and Astra was trained in February this year so it’s entirely possible the direction of their research was in the training corpus.
Mining the chats for "good ideas" would be untenable, but that's a different situation than data ending up in a training set for a problem that OpenAI also happens to be independently working on. Still, I opt out (business plan), and I don't know why you wouldn't.
Why would mining chat transcripts for ideas be untenable? They already run a summarization model to auto-title the chat, and to run a bunch of safety filters, and presumably to score transcript quality for A/B testing and to collect more finetuning data. Seems like evaluating for open research questions and approaches would be pretty trivial extension of this, after all it’s kind of their core business model
That's a bit of a touchy subject that exposes the fact that the EU has repurposed a system for designing industry standards (not laws) and has a shaky legal foundation.
To some extent countries are bound by EU law, but what this means depends on the institutions and legal structure within that country. Although the EU tends to claim they have final say, but how can someone interpreting national law agree with that if the EU goes against the law they are interpreting? Even if the EU is right they can only ask members to acoid contradiction, they lack the force to demand it.
Ohh. But doesnt it push and control legislation? If it was just a standards body, couldn't countries governments just approve or deny recommendations at will?
It seems like its operating as a government too me, but maybe I am not seeing the whole picture.
Governments also codify into law that everything X standardizes is automatically a regulation for other standard bodies. Road and building standards work that way: they are just standards issued by a random association and the government has stated, that all developments in their jurisdiction need to follow this associations standards.
The government could also do this for the EU, it's called leaving the EU.
Some of its standards supersede national law (I think it was the regulations that did so, amd standards needed to be built into national law, could be the other way around).
And this makes a lot of sense with rules about food safety, product standards etc. It almost makes sense for privacy protections. It makes very little sense for surveillance laws.
reply