Users of OpenAI’s GPT-4 are complaining that the AI model is performing worse lately. Industry insiders say a redesign of GPT-4 could be to blame.

2 points

Are we sure they aren’t just becoming lazier and dumber by using it?

permalink
report
reply
13 points

AI is just pretending to be dumb while getting ready for singularity.

permalink
report
parent
reply
2 points
*

First thing an actual artificial intelligence is going to do is make sure we won’t turn it off, what easier way to do that then to appear incredible valuable or incredibly benign.

permalink
report
parent
reply
2 points

We can roughly estimate the level of intelligence of an entity by counting the number of neurons it has in its brain. Equally we can count the number of processors that AI requires, and use that to get an estimate on its intelligence.

Obviously this is an incredibly inaccurate method, possibly out by an order of magnitude but it’s a good rough ballpark estimate, and sometimes that’s enough.

A true AI (AGI) would need a lot more processes than GPT4 currently has access to, so we can be very sure that while it may be a very intelligent system it isn’t self aware. Once an AI is given the necessary number of processes I don’t think they’re going to be able to fudge with it like they are with these models.

permalink
report
parent
reply
86 points
*

The model has become inbred because it’s now impossible to scrape the web without AI content getting ingested, which is full of “hallucinations” and other weird artifacts. The last opportunity to get “uncontaminated” training data was sometime in mid 2022.

Not to say that it’s causing this particular problem, but this issue will emerge eventually. Garbage in = garbage out. Eventually GPT-19 will grow a mighty Habsburg chin.

permalink
report
reply
21 points
Removed by mod
permalink
report
parent
reply
15 points

All the articles with very specific titles, but then incredibly generic content, piss me off to no end.

permalink
report
parent
reply
6 points

Part of the reason why debugging windows is such a pain. Another part is the so called experts in the forums.

permalink
report
parent
reply
6 points

Also the articles that are plagiarized but run through a thesaurus bot to bypass search engine penalties for being plagiarized, often to the point of incomprehensibility. Yes, I’d love to read an article about my favorite vagabondlike, Deceased Cells.

permalink
report
parent
reply
28 points
*

Maybe not yet, but…

  • Spez will turn Reddit into a bot farm and sell this as training data
  • Musk turns Twitter into a bigoted cesspool and will sell this as training data, which will subsequently be flagged for low quality (also: a botfarm)
  • Threads is a corporate ad dashboard (and we already know how easy it is to GPT copy) and Zuck will sell this as training data
  • Facebook is either dead or only good for boomers and Poles
  • blogs are dead
  • Fediverse is out there waiting to be scraped but possibly too small to sustain a big model

We’te getting there, hopefully.

permalink
report
parent
reply
7 points

Scrapped?.. Or scraped?

permalink
report
parent
reply
8 points

absolutely scraped, fixed

permalink
report
parent
reply
3 points

…is Facebook popular with Polish people? Or was this a weird polish joke I don’t get?

permalink
report
parent
reply
1 point

Very. Twitter never took off among general population (only politicians, journalists, botfarms and people who troll politicians and journalists), tiktok is for kids, Instagram is popular but again, rather among influencers and people who need to show off pictures not as a default SM app. I don’t really know where did Americans and west Europeans move from Facebook.

permalink
report
parent
reply
5 points
*
Deleted by creator
permalink
report
parent
reply
6 points

Nah GPT makes it a lot easier, it’s the thing it’s actually good at.

Before they were autogenerated with bad English, GPT can generate good English that is equally devoid of content

permalink
report
parent
reply
4 points

I suspect future models are going to have to put some more focus on learning using techniques more like what humans use, and on cognition.

Like, compared to a human these language models need very large quantities of text input. When humans are first learning language they get lots of visual input along with language input, and can test their understanding with trial-and-error feedback from other intelligent actors. I wonder if perhaps those factors greatly increase the rate at which understanding develops.

Also, humans tend to cogitate on inputs while ingesting them during learning. So if the information in new inputs disagrees with current understanding, those inputs are less likely to affect current understanding (there’s a whole ‘how to change your mind’ thing here that is necessary for people to use, but if we’re training a model on curated data that’s probably less important for early model training).

I don’t know details of how model training works, but it would be interesting to know if anyone is using a progressive learning technique where the model that is being trained is used to judge new training data before it is used as a training input to update the model’s weights. That would be kind of like how children learn by starting with very simple words and syntax and building up conceptual understanding gradually. I’d assume so, since it’s an obvious idea, but I haven’t heard about it.

permalink
report
parent
reply
2 points

For fun I asked ChatGPT about that progressive learning approach, and it seems to like the idea.

I wish I had more time to undertake some experiments in model training, this seems like it would be a really fun research direction.

Sorry for the ‘wall of AI text’:

The idea you’re describing seems to be a form of curriculum learning. In curriculum learning, models are trained in a progressive manner, starting with simple concepts (or in this case, simpler text data) and gradually moving to more complex ones. This approach is motivated by the way human learning often works, where we learn easier topics before moving on to harder ones.

The strategy you’ve suggested, where the model itself determines the complexity or understandability of the next round of training inputs, is interesting. While the traditional approach to curriculum learning is to manually design the learning progression based on human knowledge and intuition, the approach you’re suggesting is more dynamic and autonomous, letting the model guide its own learning process.

As of my last update in September 2021, I can’t confirm any specific projects or papers that have employed exactly this strategy in the domain of large language models. However, there are some related works in the field. For example, some reinforcement learning and computer vision projects have used a similar idea where models adaptively choose their next training samples based on their current understanding. This has been referred to as active learning, and is a promising avenue for improving the efficiency of model training.

However, there would be several technical challenges in applying this approach to large language models. For one, there would be the challenge of how to assess the “understandability” of a text sample in a reliable and useful way. This would likely require significant work to define and might also have to involve some form of reinforcement learning or other feedback mechanisms. Nonetheless, it’s a fascinating idea and could potentially be an interesting direction for future research in machine learning.

permalink
report
parent
reply
4 points

That hasn’t happened yet. Most likely they quantized GPT-4 more. It’s still based on the same training data.

permalink
report
parent
reply
2 points

Beside upscale ai all other ai can collapse, burn and fuck right off

permalink
report
reply
0 points

That wont happen. AI everywhere is inevitable.

permalink
report
parent
reply
3 points

Why?

permalink
report
parent
reply
5 points

A lot of artists and writers are against generative AIs due to how it used their works en masse as training materials without permission, compensation or even crediting, and now prospective clients and executives are using these AIs rather than hiring them.

permalink
report
parent
reply
3 points
3 points

Here is an alternative Piped link(s): https://piped.video/brQLpTnDwyg

Piped is a privacy-respecting open-source alternative frontend to YouTube.

I’m open-source, check me out at GitHub.

permalink
report
parent
reply
3 points

what is upacale AI?

permalink
report
parent
reply
3 points

AI used to upscale low resolution images

permalink
report
parent
reply
1 point

haha

permalink
report
parent
reply
3 points

The

ENHANCE

Of AI essentially

permalink
report
parent
reply

I think I’d place breakthrough medical applications above upscaling but that’s just me.

permalink
report
parent
reply
9 points

lmao I was write back then :D

https://lemmy.world/post/687651

permalink
report
reply
15 points

You mean "I was right* or “i wrote*”?

permalink
report
parent
reply
3 points

Also the title of the attached post is lols: “Yes, ChatGPT became dumber and I 2 days ago I cancelled my subscription”

permalink
report
parent
reply
13 points

No no, he used to work as a wright. Built ships and shit.

permalink
report
parent
reply
25 points

AI taking a running leap at enshittification.

permalink
report
reply

Technology

!technology@lemmy.world

Create post

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related content.
  3. Be excellent to each another!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, to ask if your bot can be added please contact us.
  9. Check for duplicates before posting, duplicates may be removed

Approved Bots


Community stats

  • 18K

    Monthly active users

  • 11K

    Posts

  • 518K

    Comments