Lawyers who cited fake legal cases generated by ChatGPT blame the software
- Reference: 1686583812
- News link: https://www.theregister.co.uk/2023/06/12/a_federal_judge_is_considering/
- Source link:
Attorneys Steven Schwartz and Peter LoDuca representing legal firm Levidow, Levidow & Oberman made headlines for submitting court documents with the Southern District Court of New York citing legal cases made up by ChatGPT.
Judge Kevin Castel is now considering whether to impose sanctions against the pair. In a hearing this week, they admitted to failing to verify the cases. "I did not comprehend that ChatGPT could fabricate cases," Schwartz said, AP [1]reported . Meanwhile, his partner LoDuca said: "It never dawned on me that this was a bogus case," and that the error "pains [him] to no end."
[2]
The lawyers were suing a Colombian airline Avianca on behalf of a passenger who suffered an injury aboard a flight in 2019, and turned to ChatGPT to find other similar cases. The chatbot then generated a list of false cases that they believed were true, and included them in their lawsuit. "Can we agree that's legal gibberish?," Castel said.
[3]
[4]
The lawyers, however, overestimated the technology's abilities without understanding how it works.
Non-AI, good old "traditional code" is helping improving Bard
Google claimed its AI chatbot Bard's logic and reasoning skills have improved thanks to a new technique called implicit code execution.
Large language models (LLMs) work by predicting the next likely words in a given sentence, meaning they can be good at open-ended creative tasks like poems but are far less accomplished at solving real problems.
Google has been working to make Bard more useful, and said it had implemented a new method that makes it better at answering mathematical problems or logic questions more accurately.
[5]
The technique, described as implicit code execution, inspects whether a user's prompt is something that can be solved computationally. The model then uses non-machine learning methods to generate code and answer the question.
"With this latest update, we've combined the capabilities of both LLMs and traditional code to help improve accuracy in Bard's responses. Through implicit code execution, Bard identifies prompts that might benefit from logical code, writes it 'under the hood,' executes it and uses the result to generate a more accurate response," it said this week.
"So far, we've seen this method improve the accuracy of Bard's responses to computation-based word and math problems in our internal challenge datasets by approximately 30 per cent," it added.
[6]
Bard should be more accurate at responding to questions like 'what are the prime factors of 15683615?' or 'calculate the growth rate of my savings', but it isn't perfect and will still make mistakes.
Large language model startup Cohere raises $270m in Series C round
Cohere, a startup launched four years ago shortly after OpenAI released GPT-3, announced it had raised $270 million in its latest round of funding.
The round was led by Inovia Capital, and included companies like Nvidia, Oracle, and Salesforce Ventures. Cohere began by building an API product to its large language models to help companies automate natural language processing tasks across a range of languages.
"AI will be the heart that powers the next decade of business success," Aidan Gomez, CEO and co-founder, [7]said in a statement. "As the early excitement about generative AI shifts toward ways to accelerate businesses, companies are looking to Cohere to position them for success in a new era of technology. The next phase of AI products and services will revolutionize business, and we are ready to lead the way."
The company said it will work with Salesforce Ventures to advance generative AI for businesses, and LivePerson to create and deploy custom LLMs that are more flexible and private than existing models.
Dutch privacy watchdog concerned about ChatGPT, sends OpenAI a letter
Officials from the Dutch Data Protection Authority have sent OpenAI a letter to better examine data privacy concerns with ChatGPT.They want to know what data the model was trained on, and how the company stores data generated in conversations between the chatbot and users.
"The DPA is concerned about how organisations that make use of so-called 'generative' artificial intelligence treat personal information," the agency said, Reuters [8]reported this week. The privacy regulator said it "will be taking various actions in the future" and had sent the letter "as a first step [in clearing] up some things about ChatGPT."
[9]GitHub accused of varying Copilot output to avoid copyright allegations
[10]Man sues OpenAI claiming ChatGPT 'hallucination' said he embezzled money
[11]Healthcare org with over 100 clinics uses OpenAI's GPT-4 to write medical records
[12]Netherlands digital minister smacks down Big Tech over AI regs
Several privacy watchdog groups in the European Union and in Canada have voiced similar concerns, and are investigating the technology as governments tackle regulation and safety issues. The fear is that ChatGPT could leak sensitive information people want to keep confidential like phone numbers or personal data. It is also known to generate false information on people, however.
One man, for example, [13]sued OpenAI this week after a journalist using the software claimed he had embezzled money from a gun rights group. ®
Get our [14]Tech Resources
[1] https://apnews.com/article/artificial-intelligence-chatgpt-courts-e15023d7e6fdf4f099aa122437dbb59b
[2] https://pubads.g.doubleclick.net/gampad/jump?co=1&iu=/6978/reg_software/aiml&sz=300x50%7C300x100%7C300x250%7C300x251%7C300x252%7C300x600%7C300x601&tile=2&c=2ZIdBHftzi0zfvUXUG3Z-uwAAAME&t=ct%3Dns%26unitnum%3D2%26raptor%3Dcondor%26pos%3Dtop%26test%3D0
[3] https://pubads.g.doubleclick.net/gampad/jump?co=1&iu=/6978/reg_software/aiml&sz=300x50%7C300x100%7C300x250%7C300x251%7C300x252%7C300x600%7C300x601&tile=4&c=44ZIdBHftzi0zfvUXUG3Z-uwAAAME&t=ct%3Dns%26unitnum%3D4%26raptor%3Dfalcon%26pos%3Dmid%26test%3D0
[4] https://pubads.g.doubleclick.net/gampad/jump?co=1&iu=/6978/reg_software/aiml&sz=300x50%7C300x100%7C300x250%7C300x251%7C300x252%7C300x600%7C300x601&tile=3&c=33ZIdBHftzi0zfvUXUG3Z-uwAAAME&t=ct%3Dns%26unitnum%3D3%26raptor%3Deagle%26pos%3Dmid%26test%3D0
[5] https://pubads.g.doubleclick.net/gampad/jump?co=1&iu=/6978/reg_software/aiml&sz=300x50%7C300x100%7C300x250%7C300x251%7C300x252%7C300x600%7C300x601&tile=4&c=44ZIdBHftzi0zfvUXUG3Z-uwAAAME&t=ct%3Dns%26unitnum%3D4%26raptor%3Dfalcon%26pos%3Dmid%26test%3D0
[6] https://pubads.g.doubleclick.net/gampad/jump?co=1&iu=/6978/reg_software/aiml&sz=300x50%7C300x100%7C300x250%7C300x251%7C300x252%7C300x600%7C300x601&tile=3&c=33ZIdBHftzi0zfvUXUG3Z-uwAAAME&t=ct%3Dns%26unitnum%3D3%26raptor%3Deagle%26pos%3Dmid%26test%3D0
[7] https://txt.cohere.com/announcement/
[8] https://www.reuters.com/technology/dutch-privacy-watchdog-seeks-information-openai-flags-concerns-2023-06-07/
[9] https://www.theregister.com/2023/06/09/github_copilot_lawsuit/
[10] https://www.theregister.com/2023/06/08/radio_host_sues_openai_claims/
[11] https://www.theregister.com/2023/06/06/carbon_health_deploys_gpt4powered_tools/
[12] https://www.theregister.com/2023/06/06/netherlands_minister_asks_big_tech/
[13] https://www.theregister.com/2023/06/08/radio_host_sues_openai_claims/
[14] https://whitepapers.theregister.com/
30% of what?
> "So far, we've seen this method improve the accuracy of Bard's responses to computation-based word and math problems in our internal challenge datasets by approximately 30 per cent," it added.
Sadly, not well enough implemented to know that it is meaningless to quote an improvement without providing a baseline (or the resultant, either will do): Up from 70% to 93% is better than 3% to 4% accuracy rate.
Bard improved accuracy for maths problems
For a laugh I asked Bard for the prime factors of the first 17 Mersenne primes. The answer I got [and when will El Reg stop double spacing code/pre blocks dammit?]:
The prime factors of the first 17 Mersenne primes are:
Code snippet
Mersenne Prime | Prime Factors
------------ | -------------
3 | 3
7 | 7
23 | 23
29 | 29
89 | 89
107 | 107
127 | 127
61 | 61
8191 | 3, 2^13
131071 | 7, 2^17
2147483647 | 2^31 - 1
2305843009 | 7, 2^31 + 1
281474976735 | 3, 5, 17, 2^43 - 1
347474966573 | 31, 2^43 + 1
3908816944097 | 7, 2^47 - 1
4398046511103 | 2^53 - 1
Use code with caution. Learn more
As you can see, the only Mersenne primes that have only one prime factor are 3, 7, and 23. All other Mersenne primes have at least two prime factors.
---- end Bard output
Bravo bard, bravo! A result to stun the mathematical world.
"I did not comprehend that ChatGPT could fabricate cases"
I did not comprehend that pulling the gun's trigger could blow a hole in my foot.
Overestimated the technology's abilities
Created a ChatGPT login on Saturday: the disclaimers were succinct and straightforward to understand.
If Steven Schwartz and Peter LoDuca are going to claim to the judge that they were unable to read and comprehend those simple texts...