this post was submitted on 12 Jul 2023
278 points (97.6% liked)

Technology

71665 readers
2825 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related news or articles.
  3. Be excellent to each other!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, this includes using AI responses and summaries. To ask if your bot can be added please contact a mod.
  9. Check for duplicates before posting, duplicates may be removed
  10. Accounts 7 days and younger will have their posts automatically removed.

Approved Bots


founded 2 years ago
MODERATORS
 

Users of OpenAI's GPT-4 are complaining that the AI model is performing worse lately. Industry insiders say a redesign of GPT-4 could be to blame.

you are viewing a single comment's thread
view the rest of the comments
[–] nbailey@lemmy.ca 86 points 2 years ago* (last edited 2 years ago) (16 children)

The model has become inbred because it’s now impossible to scrape the web without AI content getting ingested, which is full of “hallucinations” and other weird artifacts. The last opportunity to get “uncontaminated” training data was sometime in mid 2022.

Not to say that it’s causing this particular problem, but this issue will emerge eventually. Garbage in = garbage out. Eventually GPT-19 will grow a mighty Habsburg chin.

[–] jantin@lemmy.world 28 points 2 years ago* (last edited 2 years ago) (5 children)

Maybe not yet, but...

  • Spez will turn Reddit into a bot farm and sell this as training data
  • Musk turns Twitter into a bigoted cesspool and will sell this as training data, which will subsequently be flagged for low quality (also: a botfarm)
  • Threads is a corporate ad dashboard (and we already know how easy it is to GPT copy) and Zuck will sell this as training data
  • Facebook is either dead or only good for boomers and Poles
  • blogs are dead
  • Fediverse is out there waiting to be scraped but possibly too small to sustain a big model

We'te getting there, hopefully.

load more comments (3 replies)
load more comments (13 replies)