News: 1680499868

  ARM Give a man a fire and he's warm for a day, but set fire to him and he's warm for the rest of his life (Terry Pratchett, Jingo)

Google denies Bard trained using OpenAI ChatGPT responses

(2023/04/03)


In brief A Google engineer reportedly quit after warning CEO Sundar Pichai that the company was wrong to train its AI search chatbot Bard on text generated by OpenAI's ChatGPT.

Netizens have posted snippets of their conversations with ChatGPT on a website called [1]ShareGPT . OpenAI prohibits people from using its outputs to train their own models.

AI engineer Jacob Devlin raised concerns that Google would violate OpenAI's terms of service by harvesting data from the website to train its own Bard chatbot, The Information [2]reported last week. Devlin thought the practice was not only wrong, it would make Bard behave too similarly to ChatGPT. After he escalated concerns to Pichai, he reportedly resigned from the company, and joined OpenAI.

[3]

Competition to develop and deploy the most attention-grabbing generative AI products amongst Big Tech companies – especially between Google and Microsoft investee OpenAI – is fierce right now. Directly compiling training data from the outputs of a rival's model would be … awkward. It will be difficult to avoid in the future, as text generated by different models spreads on the internet, but doing it on purpose is undeniably naughty.

[4]

[5]

Google denied it had trained Bard on text produced by ChatGPT. "Bard is not trained on any data from ShareGPT or ChatGPT," a spokesperson from the ad giant [6]told The Verge . The representative, however, declined to comment on whether Google had ever used text generated by ChatGPT to train Bard at all.

Synopsys launches AI tools to design chips

A developer of electronic design automation software, Synopsys, announced a set of AI-powered tools aimed at helping engineers make chips more efficiently.

Microchips are complex systems that can containing billions of transistors, assembled into complex subsystems like CPUs and GPUs. Engineers must ensure those components are carefully arranged.

to do so, engineers use different types of software to refine chip designs . Synopsys has [7]launched a suite of AI tools to tackle system architecture from design to manufacturing.

[8]

"AI design tools are enabling chipmakers to push the boundaries of Moore's Law, save time and money, alleviate the talent shortage, and even drag older chip designs into the modern era," the developer [9]said , quoting analysts from Deloitte.

The software reportedly helps engineers produce chip designs more quickly, predicts where bugs could occur, tests for silicon defects, and speeds up the manufacturing process. Increased productivity levels mean chip shops can produce more and more chips to power today's technology.

[10]Italy bans ChatGPT for 'unlawful collection of personal data'

[11]FTC urged to freeze OpenAI's 'biased, deceptive' GPT-4

[12]This US national lab turned to AI to hunt rogue nukes

[13]UK seeks light-touch AI legislation as industry leaders call for LLM pause

AI hardware startup Cerebras releases seven free open source large language models.

Cerebras, maker of the world's largest chip, trained the GPT-3-based language models using datasets ranging from 111 million to 13 billion parameters in size.

The models were trained under the Chinchilla protocol – a method [14]outlined in a paper published by DeepMind – which figures out how much data a model of a given size should be trained with using limited computational resources.

Cerebras claimed its GPT systems have "faster training times, lower training costs, and consume less energy than any publicly available model to date." They were trained using the company's CS-2 systems – part of its [15]Andromeda AI supercomputer.

[16]

"Artificial intelligence has the potential to transform the world economy, but its access is increasingly gated," the AI hardware startup [17]stated in a blog post. "The latest large language model – OpenAI's GPT4 – was released with no information on its model architecture, training data, training hardware, or hyperparameters. Companies are increasingly building large models using closed datasets and offering model outputs only via API access."

"For [large language models] to be an open and accessible technology, we believe it's important to have access to state-of-the-art models that are open, reproducible, and royalty free for both research and commercial applications."

Cerebras's code describes the models' architecture, weights, and training checkpoints, which were made available on Hugging Face and GitHub under the Apache 2.0 license.

US Federal Trade Commission paying close attention to Big Tech and AI

The US Federal Trade Commission's chairwoman Lina Khan warned she would be keeping a close eye on the AI industry to make sure it isn't controlled by Big Tech.

"As you have machine learning that depends on huge amounts of data and also a huge amount of storage, we need to be very vigilant to make sure that this is not just another site for big companies to become bigger," Khan said this week during an event hosted by the Department of Justice, [18]according to Bloomberg.

Under Khan, the FTC has focused on cracking down on the largest technology monopolies and their potential antitrust and anti-competitive issues. The commission has taken an interest in AI, and has urged companies building the technology to ensure their products are [19]safe and [20]trustworthy if they don't want the regulator breathing down their necks.

"Sometimes we see claims that are not fully vetted or not really reflecting how these technologies work," Khan said. "Developers of these tools can potentially be liable if technologies they are creating are effectively designed to deceive." ®

Get our [21]Tech Resources



[1] https://sharegpt.com/

[2] https://www.theinformation.com/articles/alphabets-google-and-deepmind-pause-grudges-join-forces-to-chase-openai

[3] https://pubads.g.doubleclick.net/gampad/jump?co=1&iu=/6978/reg_software/aiml&sz=300x50%7C300x100%7C300x250%7C300x251%7C300x252%7C300x600%7C300x601&tile=2&c=2ZCqjv0gr2Q3p6INlN08zCgAAAFI&t=ct%3Dns%26unitnum%3D2%26raptor%3Dcondor%26pos%3Dtop%26test%3D0

[4] https://pubads.g.doubleclick.net/gampad/jump?co=1&iu=/6978/reg_software/aiml&sz=300x50%7C300x100%7C300x250%7C300x251%7C300x252%7C300x600%7C300x601&tile=4&c=44ZCqjv0gr2Q3p6INlN08zCgAAAFI&t=ct%3Dns%26unitnum%3D4%26raptor%3Dfalcon%26pos%3Dmid%26test%3D0

[5] https://pubads.g.doubleclick.net/gampad/jump?co=1&iu=/6978/reg_software/aiml&sz=300x50%7C300x100%7C300x250%7C300x251%7C300x252%7C300x600%7C300x601&tile=3&c=33ZCqjv0gr2Q3p6INlN08zCgAAAFI&t=ct%3Dns%26unitnum%3D3%26raptor%3Deagle%26pos%3Dmid%26test%3D0

[6] https://www.theverge.com/2023/3/29/23662621/google-bard-chatgpt-sharegpt-training-denies

[7] https://news.synopsys.com/2023-03-29-Synopsys-ai-Unveiled-as-Industrys-First-Full-Stack,-AI-Driven-EDA-Suite-for-Chipmakers

[8] https://pubads.g.doubleclick.net/gampad/jump?co=1&iu=/6978/reg_software/aiml&sz=300x50%7C300x100%7C300x250%7C300x251%7C300x252%7C300x600%7C300x601&tile=4&c=44ZCqjv0gr2Q3p6INlN08zCgAAAFI&t=ct%3Dns%26unitnum%3D4%26raptor%3Dfalcon%26pos%3Dmid%26test%3D0

[9] https://news.synopsys.com/2023-03-29-Synopsys-ai-Unveiled-as-Industrys-First-Full-Stack,-AI-Driven-EDA-Suite-for-Chipmakers

[10] https://www.theregister.com/2023/03/31/italy_bans_chatgpt_for_unlawful/

[11] https://www.theregister.com/2023/03/30/ftc_openai_gpt4/

[12] https://www.theregister.com/2023/03/30/us_ai_nuclear_hunter/

[13] https://www.theregister.com/2023/03/29/uk_seeks_lighttouch_ai_legislation/

[14] https://www.deepmind.com/publications/an-empirical-analysis-of-compute-optimal-large-language-model-training

[15] https://www.theregister.com/2022/11/15/cerebras_supercomputer_frontier/

[16] https://pubads.g.doubleclick.net/gampad/jump?co=1&iu=/6978/reg_software/aiml&sz=300x50%7C300x100%7C300x250%7C300x251%7C300x252%7C300x600%7C300x601&tile=3&c=33ZCqjv0gr2Q3p6INlN08zCgAAAFI&t=ct%3Dns%26unitnum%3D3%26raptor%3Deagle%26pos%3Dmid%26test%3D0

[17] https://www.cerebras.net/blog/cerebras-gpt-a-family-of-open-compute-efficient-large-language-models/

[18] https://www.bloomberg.com/news/articles/2023-03-27/ftc-reviewing-competition-deception-in-artificial-intelligence

[19] https://www.theregister.com/2023/03/22/ftc_ai_fraud_warning/

[20] https://www.theregister.com/2023/02/28/ftc_warns_ai_hucksters/

[21] https://whitepapers.theregister.com/



The arrogance of tools as support ignorant fools deliver the grandest of follies... and vice versa.

amanfromMars 1

The US Federal Trade Commission's chairwoman Lina Khan warned she would be keeping a close eye on the AI industry to make sure it isn't controlled by Big Tech.

Which then begs the question, if AI can be controlled by anything/anyone other than AI itself, who/what does the US FTC imagine and approve of being in AI command and control?

Y’all might like to consider AI an alien confection way beyond the reach of any divisive human direction or disruptive intrusion ........ and you may like to also realise, for ignorance in such an expanding matter is most definitely not blissful, conspiring and working in concert with others to try to constrain and prevent its primary ascension and succession in support of a former, designedly inequitable counter dependence, will have such system administrations/leading entities suffering punitive negative consequences and/or persistent advanced cyber threats against which there be no possible defence/course for avoidance and non-accountability.

Such is freely offered as sound sensible advice to heed and work creatively and positively with, ....however,.... one must also be aware of these old gold nuggets of human wisdom which just keep on giving and proving themself too right, far too often to be dismissed and thought of as the ravings of a crank or genius ...... Only two things are infinite, the universe and human stupidity, and I'm not sure about the former, with nothing causing more consternation in a group of hypocrites than one honest man having fun with creative intelligence. And good ideas have no borders. :-)

C’est la vie.

"wrong to train its AI search chatbot Bard on text generated by OpenAI's ChatGPT"

Mike 137

Wrong also on very basic technical grounds. The output of ChatGPT is already an artefact of both its inputs and its statistical algorithm. Consequently it's a filtered and distorted version of 'reality' and thus a bad starting point for further statistical manipulation. Cascading chatbots in this way reminds me of the definition of knowledge in The Machine Stops (E M Forster 1909) where it consisted of opinions about opinions about opinions [...] about some distant facts. The result inevitably degenerates into pure verbal fluff (even if the original inputs weren't, which is of course open to question seeing they're trawled from open web content).

Re: "wrong to train its AI search chatbot Bard on text generated by OpenAI's ChatGPT"

Snowy

Like playing a game of AI whispers.

They're wrong, even if not lying

FF22

It's virtually impossible that Bard has not been trained by ChatGPT responses, even if it has not been done deliberately so. That's because answers and content generated by ChatGPT has been all over the web for years now, and since Google was and is unable to differentiate between AI and human generated content (even in cases obvious to humans), the text base Bard was trained on has had to include several examples of ChatGPT-generated content.

That's the real problem with AI, which will be more and more a problem in the months and years to come: they all will feed more and more on each other's output, even if not deliberately done so, inflating and aggregating (literally) each others' flaws and misconceptions, and will gravitate to the same subpar average, as do humans, unfortunately, since content publication has been "democratized". Their answers will become as stupid and unreliable as that of the average facebook commentard's.

Re: They're wrong, even if not lying

Claptrap314

I recall from a recent article in these pages that "AI" fares much better at detecting GAI content than humans--something about spotting statistical artifact. If true, then your supposition would be deeply challenged.

But I DO agree with the conclusion. The net result is not going to be any sort of improvement when it comes to general usage.

If you don't do the things that are not worth doing, who will?