AI trained on AI garbage spits out AI garbage.(www.technologyreview.com)

posted 4 months ago

ModerateImprovement@sh.itjust.works

technology@lemmy.world

58 commentshide report

Sort:

Hot Top Controversial New Old

[ - ]

Admiral Patrick@dubvee.org

81 points

4 months ago

As junk web pages written by AI proliferate, the models that rely on that data will suffer.

Good.

permalink

report

[ - ]

Madrigal@lemmy.world

79 points

4 months ago

“On two occasions I have been asked, ‘Pray, Mr. Babbage, if you put into the machine wrong figures, will the right answers come out?’ I am not able rightly to apprehend the kind of confusion of ideas that could provoke such a question.” - Charles Babbage

permalink

report

[ - ]

bionicjoey@lemmy.ca

14 points

4 months ago

The business people adopting AI: “who cares what it’s trained on? It’s intelligent right? It’ll just sort through the garbage and magically come up with the right answers to everything”

permalink

report

parent

[ - ]

RecluseRamble@lemmy.dbzer0.com

1 point

4 months ago

Not so hard to imagine given that these people have always seen technical systems as magic.

permalink

report

parent

[ - ]

CookieOfFortune@lemmy.world

6 points

4 months ago

Of course modern UX design is very much based on getting the right answer with the wrong inputs (autocorrect, etc).

permalink

report

parent

[ - ]

lennivelkant@discuss.tchncs.de

1 point

4 months ago

I believe Robustness was the term I learned years ago: the ability of a system to gracefully handle user error, make it easy to recover from or fix, clearly communicate what was wrong etc.

Of course, nothing is ever perfect and humans are very creative at fucking up, and a lot of companies don’t seem to take UX too seriously. Particularly when the devs get tunnel vision and forget about user error being a thing…

permalink

report

parent

[ - ]

Crazyslinkz@lemmy.world

57 points

4 months ago

Garbage in; Garbage out.

permalink

report

[ - ]

_haha_oh_wow_@sh.itjust.works

19 points

4 months ago

Shit-fueled ouroboros

permalink

report

parent

[ - ]

lemmeout@lemm.ee

4 points

4 months ago

You can’t explain it!

permalink

report

parent

[ - ]

BluesF@lemmy.world

2 points

4 months ago

Recycle the garbage that comes out… Still more garbage out.

permalink

report

parent

[ - ]

Lvxferre@mander.xyz

36 points

4 months ago

Model degeneration is an already well-known phenomenon. The article already explains well what’s going on so I won’t go into details, but note how this happens because the model does not understand what it is outputting - it’s looking for patterns, not for the meaning conveyed by said patterns.

Frankly at this rate might as well go with a neuro-symbolic approach.

permalink

report

[ - ]

CeeBee_Eh@lemmy.world

-6 points

4 months ago

The issue with your assertion is that people don’t actually work a similar way. Have you ever met someone who was clearly taught "garbage’?

permalink

report

parent

[ - ]

Lvxferre@mander.xyz

11 points

4 months ago

The issue with your assertion is that people don’t actually work a similar way.

I’m talking about LLMs, not about people.

permalink

report

parent

[ - ]

CeeBee_Eh@lemmy.world

-11 points

4 months ago

I know you are, but the argument that an LLM doesn’t understand context is incorrect. It’s not human level understanding, but it’s been demonstrated that they do have a level of understanding.

And to be clear, I’m not talking about consciousness or sapience.

report

[ - ]

25 points

4 months ago

Well, you’ve got a timestamped copy of much of the Web that existed up until latent-diffusion models at archive.org. That may not give you access to newer information, but it’s a pretty whopping big chunk of data to work with.

permalink

report

[ - ]

palordrolap@kbin.run

20 points

4 months ago

Hopefully archive.org have measures in place to stop people from yanking all their data too quickly. As least not without a hefty donation or something. As a user it can chug a bit, and I’m hoping that’s the rate-limiting I’m talking about and not that they’re swamped.

permalink

report

parent

[ - ]

Grimy@lemmy.world

7 points

4 months ago

That would go against the principal of the archive imo but regardless, if you take away all means of acquiring data freely, you are just giving companies like OpenAI and Google who already have copies of it an insane advantage.

AI isn’t going away, we need to make sure we have free access to it as to not give our whole economy to a handful of companies.

permalink

report

parent

Technology

!technology@lemmy.world

Create post

This is a most excellent place for technology news and articles.

Our Rules

Follow the lemmy.world rules.
Only tech related content.
Be excellent to each another!
Mod approved content bots can post up to 10 articles per day.
Threads asking for personal tech support may be deleted.
Politics threads may be removed.
No memes allowed as posts, OK to post as comments.
Only approved bots from the list below, to ask if your bot can be added please contact us.
Check for duplicates before posting, duplicates may be removed

Approved Bots

Community stats

17K
Monthly active users
6.1K
Posts
132K
Comments

Our Rules

Approved Bots

Community stats

Community moderators