Google, DeepMind accused of 'stealing the internet' to create Bard AI chatbot
(2023/07/12)
- Reference: 1689186607
- News link: https://www.theregister.co.uk/2023/07/12/google_alphabet_deepmind_bard_complaint/
- Source link:
Google, DeepMind and parent company, Alphabet, have been accused of "secretly stealing everything ever created and shared on the internet by hundreds of millions of Americans" to build their own AI chatbot, Bard.
Eight pseudonymous individuals – including two minors, aged 13 and 6 – are seeking to lead millions of netizens in a class action effort alleging 10 charges against Alphabet's answer to OpenAI's ChatGPT, among other items in Google's "suite of AI products."
The complaint was filed in a California federal court and alleges the company breached several state and US federal laws including the DMCA, California's Unfair Competition law, and a state invasion of privacy rule. It also accuses Bard's maker of larceny/receipt of stolen property.
[1]
The lawsuit
[4]
[5]
The claim from the unnamed plaintiffs is that the "update" of Google's online privacy policy was an effective doubling-down on its position.
The news comes hot on the heels of similar lawsuits accusing Microsoft-backed OpenAI of privacy breaches and misusing scraped data, as well as of [6]copyright infringement .
[7]
Like the [8]Microsoft and OpenAI class action filed in June, yesterday's lawsuit also namechecks Reg articles – twice – when delving into the technical side of matters. The first time is to explain how Google's immense public [9]C4 dataset ingests material to build its next-gen machine learning systems. The second mention is to further its argument that Google profits from the data, footnoting our [10]April coverage of the internal Google presentation titled "AI-powered ads 2023" outlining Google's plan to roll out generative AI tools to its advertising platform.
DeepMind, still run by co-founder Demis Hassabis after being acquired by Google in [11]2014 , is cited in the suit due to its work on developing the Language Model for Dialogue Applications (LaMDA), considered instrumental in Bard's development as well as in other Google AI products.
[12]Sarah Silverman, novelists sue OpenAI for scraping their books to train ChatGPT
[13]Google says public data is fair game for training its AIs
[14]OpenAI pauses Bing search feature over paywall bypass abilities
[15]Microsoft, OpenAI sued for $3B after allegedly trampling privacy with ChatGPT
The suit claims that Google's moves breach privacy rights and property rights, alleging:
It has very recently come to light that Google has been secretly grabbing everything ever created and shared on the internet by hundreds of millions of Americans. Google has taken all our personal and professional information, our creative and copywritten works, our photographs, and even our emails – virtually the entirety of our digital footprint – and is using it to build commercial Artificial Intelligence (AI) Products like "Bard," the chatbot Google recently released to compete with OpenAI's "ChatGPT." For years, Google harvested this data in secret, without notice or consent from anyone.
The plaintiffs are looking for at least $5 billion, injunctive relief, and implementation of "effective cybersecurity safeguards" to protect the data subjects.
In a statement, Google general counsel Halimah DeLaine Prado said the company had been "clear for years that we use data from public sources – like information published to the open web and public datasets – to train the AI models behind services like Google Translate, responsibly and in line with our AI Principles."
She added: "American law supports using public information to create new beneficial uses, and we look forward to refuting these baseless claims." ®
Get our [16]Tech Resources
[1] https://pubads.g.doubleclick.net/gampad/jump?co=1&iu=/6978/reg_software/aiml&sz=300x50%7C300x100%7C300x250%7C300x251%7C300x252%7C300x600%7C300x601&tile=2&c=2ZK8ih28kPOtZripRgqorXwAAAhE&t=ct%3Dns%26unitnum%3D2%26raptor%3Dcondor%26pos%3Dtop%26test%3D0
[2] https://regmedia.co.uk/2023/07/12/alphabet_class_action.pdf
[3] https://www.theregister.com/2023/07/06/google_ai_models_internet_scraping/
[4] https://pubads.g.doubleclick.net/gampad/jump?co=1&iu=/6978/reg_software/aiml&sz=300x50%7C300x100%7C300x250%7C300x251%7C300x252%7C300x600%7C300x601&tile=4&c=44ZK8ih28kPOtZripRgqorXwAAAhE&t=ct%3Dns%26unitnum%3D4%26raptor%3Dfalcon%26pos%3Dmid%26test%3D0
[5] https://pubads.g.doubleclick.net/gampad/jump?co=1&iu=/6978/reg_software/aiml&sz=300x50%7C300x100%7C300x250%7C300x251%7C300x252%7C300x600%7C300x601&tile=3&c=33ZK8ih28kPOtZripRgqorXwAAAhE&t=ct%3Dns%26unitnum%3D3%26raptor%3Deagle%26pos%3Dmid%26test%3D0
[6] https://www.theregister.com/2023/07/10/in_brief_ai/
[7] https://pubads.g.doubleclick.net/gampad/jump?co=1&iu=/6978/reg_software/aiml&sz=300x50%7C300x100%7C300x250%7C300x251%7C300x252%7C300x600%7C300x601&tile=4&c=44ZK8ih28kPOtZripRgqorXwAAAhE&t=ct%3Dns%26unitnum%3D4%26raptor%3Dfalcon%26pos%3Dmid%26test%3D0
[8] https://www.theregister.com/2023/06/28/microsoft_openai_sued_privacy/
[9] https://www.theregister.com/2023/04/20/google_c4_data_nasty_sources/
[10] https://www.theregister.com/2023/04/21/google_bard_ai/
[11] https://www.theregister.com/2014/01/27/google_deep_mind_buy/
[12] https://www.theregister.com/2023/07/10/in_brief_ai/
[13] https://www.theregister.com/2023/07/06/google_ai_models_internet_scraping/
[14] https://www.theregister.com/2023/07/05/openai_pauses_bing_search/
[15] https://www.theregister.com/2023/06/28/microsoft_openai_sued_privacy/
[16] https://whitepapers.theregister.com/
Eight pseudonymous individuals – including two minors, aged 13 and 6 – are seeking to lead millions of netizens in a class action effort alleging 10 charges against Alphabet's answer to OpenAI's ChatGPT, among other items in Google's "suite of AI products."
The complaint was filed in a California federal court and alleges the company breached several state and US federal laws including the DMCA, California's Unfair Competition law, and a state invasion of privacy rule. It also accuses Bard's maker of larceny/receipt of stolen property.
[1]
The lawsuit
[2]PDF
was keen to note Google's recent [3]update of its privacy policy confirming it scrapes public data from the internet to train its AI models and services – including both Bard and its cloud-hosted products. The suit claims the move was made in response to the FTC's warning that: "Machine learning is no excuse to break the law… The data you use to improve your algorithms must be lawfully collected… companies would do well to heed this lesson."[4]
[5]
The claim from the unnamed plaintiffs is that the "update" of Google's online privacy policy was an effective doubling-down on its position.
The news comes hot on the heels of similar lawsuits accusing Microsoft-backed OpenAI of privacy breaches and misusing scraped data, as well as of [6]copyright infringement .
[7]
Like the [8]Microsoft and OpenAI class action filed in June, yesterday's lawsuit also namechecks Reg articles – twice – when delving into the technical side of matters. The first time is to explain how Google's immense public [9]C4 dataset ingests material to build its next-gen machine learning systems. The second mention is to further its argument that Google profits from the data, footnoting our [10]April coverage of the internal Google presentation titled "AI-powered ads 2023" outlining Google's plan to roll out generative AI tools to its advertising platform.
DeepMind, still run by co-founder Demis Hassabis after being acquired by Google in [11]2014 , is cited in the suit due to its work on developing the Language Model for Dialogue Applications (LaMDA), considered instrumental in Bard's development as well as in other Google AI products.
[12]Sarah Silverman, novelists sue OpenAI for scraping their books to train ChatGPT
[13]Google says public data is fair game for training its AIs
[14]OpenAI pauses Bing search feature over paywall bypass abilities
[15]Microsoft, OpenAI sued for $3B after allegedly trampling privacy with ChatGPT
The suit claims that Google's moves breach privacy rights and property rights, alleging:
It has very recently come to light that Google has been secretly grabbing everything ever created and shared on the internet by hundreds of millions of Americans. Google has taken all our personal and professional information, our creative and copywritten works, our photographs, and even our emails – virtually the entirety of our digital footprint – and is using it to build commercial Artificial Intelligence (AI) Products like "Bard," the chatbot Google recently released to compete with OpenAI's "ChatGPT." For years, Google harvested this data in secret, without notice or consent from anyone.
The plaintiffs are looking for at least $5 billion, injunctive relief, and implementation of "effective cybersecurity safeguards" to protect the data subjects.
In a statement, Google general counsel Halimah DeLaine Prado said the company had been "clear for years that we use data from public sources – like information published to the open web and public datasets – to train the AI models behind services like Google Translate, responsibly and in line with our AI Principles."
She added: "American law supports using public information to create new beneficial uses, and we look forward to refuting these baseless claims." ®
Get our [16]Tech Resources
[1] https://pubads.g.doubleclick.net/gampad/jump?co=1&iu=/6978/reg_software/aiml&sz=300x50%7C300x100%7C300x250%7C300x251%7C300x252%7C300x600%7C300x601&tile=2&c=2ZK8ih28kPOtZripRgqorXwAAAhE&t=ct%3Dns%26unitnum%3D2%26raptor%3Dcondor%26pos%3Dtop%26test%3D0
[2] https://regmedia.co.uk/2023/07/12/alphabet_class_action.pdf
[3] https://www.theregister.com/2023/07/06/google_ai_models_internet_scraping/
[4] https://pubads.g.doubleclick.net/gampad/jump?co=1&iu=/6978/reg_software/aiml&sz=300x50%7C300x100%7C300x250%7C300x251%7C300x252%7C300x600%7C300x601&tile=4&c=44ZK8ih28kPOtZripRgqorXwAAAhE&t=ct%3Dns%26unitnum%3D4%26raptor%3Dfalcon%26pos%3Dmid%26test%3D0
[5] https://pubads.g.doubleclick.net/gampad/jump?co=1&iu=/6978/reg_software/aiml&sz=300x50%7C300x100%7C300x250%7C300x251%7C300x252%7C300x600%7C300x601&tile=3&c=33ZK8ih28kPOtZripRgqorXwAAAhE&t=ct%3Dns%26unitnum%3D3%26raptor%3Deagle%26pos%3Dmid%26test%3D0
[6] https://www.theregister.com/2023/07/10/in_brief_ai/
[7] https://pubads.g.doubleclick.net/gampad/jump?co=1&iu=/6978/reg_software/aiml&sz=300x50%7C300x100%7C300x250%7C300x251%7C300x252%7C300x600%7C300x601&tile=4&c=44ZK8ih28kPOtZripRgqorXwAAAhE&t=ct%3Dns%26unitnum%3D4%26raptor%3Dfalcon%26pos%3Dmid%26test%3D0
[8] https://www.theregister.com/2023/06/28/microsoft_openai_sued_privacy/
[9] https://www.theregister.com/2023/04/20/google_c4_data_nasty_sources/
[10] https://www.theregister.com/2023/04/21/google_bard_ai/
[11] https://www.theregister.com/2014/01/27/google_deep_mind_buy/
[12] https://www.theregister.com/2023/07/10/in_brief_ai/
[13] https://www.theregister.com/2023/07/06/google_ai_models_internet_scraping/
[14] https://www.theregister.com/2023/07/05/openai_pauses_bing_search/
[15] https://www.theregister.com/2023/06/28/microsoft_openai_sued_privacy/
[16] https://whitepapers.theregister.com/
Re: This has been done before...
stiine
You've never read a ToS from beginning to end, have you?
I hope Judge William Alsup gets assigned this case.
Nifty
Sheeks I'm 'stealing' internet content at this moment via my eyeballs and likely to regurgitate something based on it in the future. In a general way where attribution is pointless. If you make it public, it's public.
This has been done before...
Remember the days when corporations sent ships to West Africa and picked up local workers, offering them nice new jobs, "Just jump on the ship, it's a free trip across the ocean, we're looking after your data" and then when they arrived in the Caribbean and America they were given a job (not just their data was sold, but them too) with instructions to grow sugar and harvest it if they wanted "free meals" at the end of the week...
These days "slavery" has been eliminated, and private data no longer exists too so we're not slaves but our lives are not much different, we're just all busy making corporations wealthy again.