Researchers say an AI-powered transcription tool used in hospitals invents things no one ever said(apnews.com)

posted 13 days ago

Stopthatgirl7@lemmy.world

technology@lemmy.world

43 commentshide report

Tech behemoth OpenAI has touted its artificial intelligence-powered transcription tool Whisper as having near “human level robustness and accuracy.”

But Whisper has a major flaw: It is prone to making up chunks of text or even entire sentences, according to interviews with more than a dozen software engineers, developers and academic researchers. Those experts said some of the invented text — known in the industry as hallucinations — can include racial commentary, violent rhetoric and even imagined medical treatments.

Experts said that such fabrications are problematic because Whisper is being used in a slew of industries worldwide to translate and transcribe interviews, generate text in popular consumer technologies and create subtitles for videos.

More concerning, they said, is a rush by medical centers to utilize Whisper-based tools to transcribe patients’ consultations with doctors, despite OpenAI’ s warnings that the tool should not be used in “high-risk domains.”

Sort:

Hot Top Controversial New Old

[ - ]

Onno (VK6FLAB)@lemmy.radio

11 points

13 days ago

AI does not mean Artificial Intelligence, it means Assumed Intelligence.

permalink

report

[ - ]

magnetosphere@fedia.io

34 points

13 days ago

“This seems solvable if the company is willing to prioritize it.”

I know how to make the company prioritize it: make Whisper illegal to use (or even promote) until a certain threshold of accuracy is met. This software is absolute garbage at best, and a genuine hazard at worst.

Lame, ineffective “warnings” serve no purpose but to cover OpenAIs ass. Hit them in the wallet, and they’ll pay attention.

permalink

report

[ - ]

QuadratureSurfer@lemmy.world

-1 points

13 days ago

Rather than making it illegal to use, people need to use these tools responsibly. If any of these companies are using almost any kind of AI/machine learning they need to include a human in the loop that can verify that it’s working correctly. That way if it starts hallucinating things that were never said, it can be caught and corrected.

I’ve found that Whisper generally does a better job at translating/transcribing audio than other open source tools out there, so it’s not garbage… But it absolutely is a hazard if you’re trying to rely solely on it for official documents (or legal issues).

As far as promotion goes… It’s open source software, it’s not being sold.

permalink

report

parent

[ - ]

leftzero@lemmynsfw.com

7 points

13 days ago

people need to use these tools responsibly

Have you met people…?

permalink

report

parent

[ - ]

Llewellyn@lemm.ee

4 points

13 days ago

I have an even better idea: make tool creators and / or CEO of the company, using the tool, liable for all tool’s mistakes and hallucinations.

permalink

report

parent

[ - ]

barsoap@lemm.ee

3 points

13 days ago

It is illegal to use in the EU for anything even remotely sensitive. Like, if you subtitle a movie with it and it messes up noone cares, your problem, if you’re doing anything that has any legal implications, from college applications over job interviews to court proceedings, they’ll nail you to the cross. For AI to be used in such domains it has to be certified and AIs certified for even a subset of these things plainly don’t exist.

It’s like with self-driving cars: What OpenAI is producing is pretty much on the level of Tesla’s “full self driving”. It’s not even waymo who have proper autonomy tech certified to operate in a limited area in a benevolent (to venture capital) jurisdiction (some municipality or the other). Wake me when it gets actual approval from actual regulatory bodies actively trying to break it.

permalink

report

parent

[ - ]

TootSweet@lemmy.world

71 points

13 days ago

LLMs in medicine. What could go wrong?

permalink

report

[ - ]

QuadratureSurfer@lemmy.world

10 points

13 days ago

Whisper isn’t a large language model.

It’s a speech to text (STT) model.

permalink

report

parent

[ - ]

Llewellyn@lemm.ee

6 points

13 days ago

Which has the same concept as the LLM under the hood, hasn’t it?

permalink

report

parent

[ - ]

DragonTypeWyvern@midwest.social

7 points

12 days ago

This isn’t a Large Language Model, it’s a Chonky Linguistic Algorithm.

permalink

report

parent

[ - ]

zoostation@lemmy.world

0 points

13 days ago

Oh word?

permalink

report

Technology

!technology@lemmy.world

Create post

This is a most excellent place for technology news and articles.

Our Rules

Follow the lemmy.world rules.
Only tech related content.
Be excellent to each another!
Mod approved content bots can post up to 10 articles per day.
Threads asking for personal tech support may be deleted.
Politics threads may be removed.
No memes allowed as posts, OK to post as comments.
Only approved bots from the list below, to ask if your bot can be added please contact us.
Check for duplicates before posting, duplicates may be removed

Approved Bots

Community stats

17K
Monthly active users
12K
Posts
542K
Comments

Our Rules

Approved Bots

Community stats

Community moderators