Researchers create 30 fake student accounts to submit model-generated responses to real exams. Professors grade the 200 or 1500 word responses from the AI undergrads and gave them better grades than real students 84% of the time. 6% of the bot respondents did get caught, though… for being too good. Meanwhile, AI detection tools? Total bunk.

Will AI be the new calculator… or the death of us all (obviously the only alternative).

Note: the software was NOT as good on the advanced exams, even though it handled the easier stuff.

75 points

Not at all surprising. ChatGPT ‘knows’ a course’s content insofar as it’s memorized the textbook and all the exam questions. Once you start asking it questions it’s never seen before (more likely for advanced topics that don’t have a billion study guides and tutorials for) it falls short, even for basic questions that’d just require a bit of additional logic.

Mind you, memorizing everything is impressive and can get you a degree, but when tasked with a new problem never seen before ChatGPT is completely inadequate.

permalink
report
reply
26 points

Right? Can students use the internet on this test? Because the LLMs have the entire internet to search for the answers, and I guarantee you those textbooks and exam questions are online and searchable.

permalink
report
parent
reply
17 points

I wonder how undergrads would do on the same exams given unlimited time and internet access but with LLMs blocked. That’s essentially what the LLMs have.

permalink
report
parent
reply
2 points

The LLMs blocked themselves?

permalink
report
parent
reply
19 points

Memorizing everything is impressive for a human.

It’s less impressive for a computer.

permalink
report
parent
reply
6 points

This is incorrect as was shown last year with the Skill-Mix research:

Furthermore, simple probability calculations indicate that GPT-4’s reasonable performance on k=5 is suggestive of going beyond “stochastic parrot” behavior (Bender et al., 2021), i.e., it combines skills in ways that it had not seen during training.

permalink
report
parent
reply
33 points

I don’t care. Maid robot when

permalink
report
reply
6 points

Like a Roomba?

permalink
report
parent
reply
5 points
*

I want mine with cat ears.

permalink
report
parent
reply
25 points

Now we know how to beat AI. We just have to pass the No LLM Left Behind act.

permalink
report
reply
23 points

I take it that this was social sciences because based on what I have seen so far I don’t think it can even outperform a college kid in maths

permalink
report
reply
17 points

All this moral panic is garbage.

Easily solved by using essays with an unseen question written in exam conditions as assessment instruments.

Literally a pencil and paper solves this problem.

permalink
report
reply
8 points

A lot of students do not perform well under exam conditions due to stress and pressure. Also, unless you’re entirely eliminating coursework, it doesn’t remove the issue.

permalink
report
parent
reply
-2 points

No assessment method is perfectly suited to every student.

Coursework can be similarly adapted.

permalink
report
parent
reply
2 points

Coursework can be similarly adapted.

How?

permalink
report
parent
reply

Technology

!technology@lemmy.world

Create post

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related content.
  3. Be excellent to each another!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, to ask if your bot can be added please contact us.
  9. Check for duplicates before posting, duplicates may be removed

Approved Bots


Community stats

  • 18K

    Monthly active users

  • 10K

    Posts

  • 457K

    Comments