Ask Lemmy

26909 readers

3190 users here now

A Fediverse community for open-ended, thought provoking questions

Please don't post about US Politics. If you need to do this, try !politicaldiscussion@lemmy.world

Rules: (interactive)

1) Be nice and; have fun

Doxxing, trolling, sealioning, racism, and toxicity are not welcomed in AskLemmy. Remember what your mother said: if you can't say something nice, don't say anything at all. In addition, the site-wide Lemmy.world terms of service also apply here. Please familiarize yourself with them

2) All posts must end with a '?'

This is sort of like Jeopardy. Please phrase all post titles in the form of a proper question ending with ?

3) No spam

Please do not flood the community with nonsense. Actual suspected spammers will be banned on site. No astroturfing.

4) NSFW is okay, within reason

Just remember to tag posts with either a content warning or a [NSFW] tag. Overtly sexual posts are not allowed, please direct them to either !asklemmyafterdark@lemmy.world or !asklemmynsfw@lemmynsfw.com. NSFW comments should be restricted to posts tagged [NSFW].

5) This is not a support community.

It is not a place for 'how do I?', type questions. If you have any questions regarding the site itself or would like to report a community, please direct them to Lemmy.world Support or email info@lemmy.world. For other questions check our partnered communities list, or use the search function.

Reminder: The terms of service apply here too.

Partnered Communities:

Logo design credit goes to: tubbadu

founded 1 year ago

MODERATORS

Bluetreefrog@lemmy.world

candyman337@lemmy.world

TheSaneWriter@lemm.ee

TheSaneWriter@lemmy.thesanewriter.com

candyman337@sh.itjust.works

Asudox@lemmy.world

can@lemmy.ca

lemmy_bot@lemmy.world

beefbaby182@lemmy.world

asudox@discuss.tchncs.de

Is chatgpt proof that standard tests are bad measures of intelligence (lemmy.world)

submitted 8 months ago by randoot@lemmy.world to c/asklemmy@lemmy.world

42 comments fedilink hide all child comments

LLMs are solving MCAT, the bar test, SAT etc like they're nothing. At this point their performance is super human. However they'll often trip on super simple common sense questions, they'll struggle with creative thinking.

Is this literally proof that standard tests are not a good measure of intelligence?

you are viewing a single comment's thread
view the rest of the comments

[–] starman2112@sh.itjust.works 0 points 8 months ago (2 children)

Disagree. We're very good at using words to convey ideas. There's no reason to believe that we speak much too fast to be properly reflecting on what we say—the speed with which we speak speaks to our proficiency with language, not a lack thereof. Many people do speak without reflecting on what they say, but to reduce all human speech down to that? Downright silly. I frequently spend seconds at a time looking for a word that has the exact meaning that will help to convey the thought that I'm trying to communicate. Yesterday, for example, I spent a whole 15 seconds or so trying to remember the word exacerbate.

An LLM is extremely good at stringing together stock words and phrases that make it sound like it's conveying an idea, but it will never stop to think about the definition of a word that best conveys a real idea. This is the third draft of this comment. I've yet to see an LLM write, rewrite, then rewrite again it's output.

[–] agamemnonymous@sh.itjust.works 1 points 8 months ago* (last edited 8 months ago)

Kinda the same thing though. You spent time finding the right auto-complete in your head. You weighed the words that fit the sentence you'd constructed in order to find the one most frequently encountered in conversations or documents that include specific related words. We're much more sophisticated at this process, but our whole linguistic paradigm isn't fundamentally very different from good auto-complete.

[–] steventrouble@programming.dev 1 points 8 months ago* (last edited 8 months ago) (1 children)

I’ve yet to see an LLM write, rewrite, then rewrite again it’s output.

It's because we (ML peeps) literally prevent them from deleting their own ouput. It'd be like if we stuck you in a room, and only let you interact with the outside world using a keyboard that has no backspace.

Seriously, try it. Try writing your reply without using the delete button, or backspace, or the arrow keys, or the mouse. See how much better you do than an LLM.

It's hard! To say that an LLM is not capable of thought just because it makes mistakes sometimes is to ignore the immense difficulty of the problem we're asking it to solve.

[–] starman2112@sh.itjust.works 1 points 8 months ago

To me it isn't just the lack of an ability to delete it's own inputs, I mean outputs, it's the fact that they work by little more than pattern recognition. Contrast that with humans, who use pattern recognition as well as an understanding of their own ideas to find the words they want to use.

Man, it is super hard writing without hitting backspace or rewriting anything. Autocorrect helped a ton, but I hate the way this comment looks lmao

This isn't to say that I don't think a neural network can be conscious, or self aware, it's just that I'm unconvinced that they can right now. That is, that they can be. I'm gonna start hitting backspace again after this paragraph