this post was submitted on 19 Aug 2026
479 points (96.7% liked)
Fuck AI
8025 readers
1015 users here now
"We did it, Patrick! We made a technological breakthrough!"
A place for all those who loathe AI to discuss things, post articles, and ridicule the AI hype. Proud supporter of working people. And proud booer of SXSW 2024.
AI, in this case, refers to LLMs, GPT technology, and anything listed as "AI" meant to increase market valuations.
founded 2 years ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments
Given OPs axis definitions.
The max homework score for nonai is 115. But there are AI users scoring 115+? I'm confused why there are no better scores above 115 for nonai. Is that data insignificant? Are students with scores higher than 115 just being flagged as AI? It makes the data look unreliable.
I'd also really like to know if the data range that isn't comparable 115+ is at all meaningful data. Because that's where there is a major dropoff. There is obviously a dropoff of significant before 115. But, just from looking at how the data is being presented with this massive dropoff but with zero data for nonai.
I'd be interested in knowing what fraction of the total population is actually being represented in the 115+ dropoff. The way it's presented it could literally be like 10 students.
This is not a defense of AI use. But more a criticism of how the data is being presented.
Also, Any student could be using AI as a resource similar to how I would use solutions manuals or previous tests to study back in my education. The good students aren't gonna copy it verbatim and get flagged. Their going to get the right answer, and use it to learn the process so they can present their own work and actually learn the material.
AI is trash. But this is nothing different than before AI when students would copy the solutions manuals blindly or their peers work. Those students have always existed. AI just makes it slightly more accessible. But, really only slightly. And, honestly, Chegg was just as easy to copy back in the day.
Edit: This graph isn't in the paper that OP linked. So, probably why it's presented so badly. I'm assuming it's from some click bait article then. There is a similar graph in the paper that is much more clear.
But the paper clarifies that this is an actual survey of students. So it's not just from flagged AI homework. Which is good and bad in its own way.
The paper itself in section 5 supports what I hypothesized above. It's not really about AI making students dumb. It's about making it easier for students that already don't want to learn to finish the homework. Same as the people that would copy solutions manuals.
Quote from section 5.
The paper doesn't argue it from what I read. But I'd also argue there is some bias in the survey from the questions themselves. Or at least how the paper lumps categories. The students scoring well in both exams and homework are less likely to consider their use of AI as meaningful enough to answer "used AI for homework". It's a problem with the survey. The survey is asking the students that actually learned the material to attribute their homework to AI alone in the same way the "copy and paste" AI users would.
The survey would probably benefit from having No AI, Some AI, All AI as it's responses. Even in anonymous surveys people make these own interpretations of the questions in their head. And a student that used AI as a resource to learn is very likely to just choose "No AI" when presented with questions of how they got their final answers to homework. Because in their head they are thinking of the students that copy and pasted AI responses and think "I'm not like that".
It's likely why the double high scoring AI user sample is so low (paper says this). The survey is not allowing the response for this set of users to categorize themselves as using AI without feeling like they are like the copy and paste students. So those people are likely just self categorizing as "no ai" because the survey doesn't allow them to distance themselves from the other population. Surveys are hard to write. Even ones where the users know it's anonymous. Good students that use AI will see themselves as good students and the copy pasters as bad students. Human emotion plays a role and most people will not self categorize themselves further negatively than they feel they should be. They are more likely to select the imperfect category that is more positive.
Bad students don't care though. They'll admit to AI use. They already don't care enough to learn the material. They have no reason to lie. So they'll select "used AI".
If you want a good survey you need very simple and minimal answer categories that the majority of the population will be able to self categorize themselves well. But, you don't want to minimize to such a degree that your survey introduces self categorizing bias. I would argue that's what happened here. The most important category is an extremely low sample size compared to the rest of the data.
And I'd also say it's likely the majority of people in reality. They just self categorized as "no ai" because the survey didn't allow them to categorize themselves more accurately without being associated with what they see as negative.
The entire paper seems to not address this strict categorization in its data that is likely introducing bad response data. But I'd have to look at the actual survey questions to know.
I think it's a good paper that suffers from an overly strict self categorizing bias in its survey data.
The values aren't a score. They normalized the average score to 100 for non-AI in both homework and exam. Then noted the deviation. People using AI to get very high scores (130% of the average) were tank the test (40% of the average) six months later.
Thanks. Before my edit was my original comment before I realized their was a PDF of the paper available. Should have answered my own question. Appreciated.
I disagree that it's no different to before AI. These days you can prompt for the whole paper to be written for you.
I definitely hyper-focused by analysis on homework and exams like math, engineering, science. The study itself separates these into different analysis. The use in writing is actually significantly less, I think because of how obvious the use of AI is in any type of creative writing.
AI is good for bull shitting an email at your work. It's awful for writing of you're trying to not get caught using AI. You can generate a paper with AI. But, if you want to actually write it like a human, check sources, verify claims, and remove hallucinations. It's literally just awful and likely more time consuming than just writing a bull shit paper in your own words to start.
Have you ever read an article or paper of any reasonable length written by AI? It's very obvious and horrible writing that no human would ever write. And the survey was done in China. Natural sounding writing in Chinese is significantly more difficult for AI to generate than in English. Partly due to the dataset bias, but also because the language is just more complex. The "floor" for writing in the Chinese language is much higher from what I know about it an AI writing.
Those are the ones whose exam scores are also on the rise. After a certain point the AI does the thinking for you and you get perfect HW grades, but can't do well on exams.
I disagree. Bad students could fear being punished for using AI as it's clearly cheating.
True. It could factor in the response if the students don't believe the survey is anonymous. It would be dependent on how the survey was given out. Electronic surveys are less likely to have students believe that. They have no way of knowing if the survey is truly anonymous. It's full trust based.
Paper surveys with no names and just simple bubble filling that they see immediately go into a stack of papers usually removes this worry in collecting data. Though this is much more expensive to collect survey data this way. Thankfully, for students this can be done for free in class.
But, yes, this fear of "being honest" could hurt the dataset if the students don't feel it is truly anonymous.
Good survey collectors will collect data both ways in order to compare the two methods to see if their is any bias in one data collection than another.
It's a problem. But it's one likely the paper accounted for already by comparing datasets. The problem I mentioned (positive categorization bias) is a potential flaw to the survey that they did not account for.
Though, I just realized you might be talking about a positive categorization bias for students that performed badly. A student not wanting to admit they used AI so they select "no ai" even though they did use it and scored poorly on exams. This is likely negligible though. A student that performs badly knows that they performed badly. This actually makes the positive selection bias favor them to be honest. People will select "used AI" as it lets their bad performance be attributed to something other than themselves. It's basically the same reasoning for the other bias I mentioned.
A bad exam score student that used AI has a positive personal bias to be honest and attribute their failure to AI use. They are likely to blame AI use for their failure. Answering honestly if they feel anonymous.
A good exam score student that used AI has a positive personal bias to attribute their success to their hard work and less about their use of AI in homework. The survey answer offers them no option to categorize themselves in this way without being categorized in the same way as bad students. So they "lie" and select the "no ai use" option as it better categorizes what they attribute their success to. They attribute it to their hard work. And there is no option that allows them to express this with these binary options.