Technology

60302 readers

4222 users here now

This is a most excellent place for technology news and articles.

Our Rules

Follow the lemmy.world rules.
Only tech related content.
Be excellent to each another!
Mod approved content bots can post up to 10 articles per day.
Threads asking for personal tech support may be deleted.
Politics threads may be removed.
No memes allowed as posts, OK to post as comments.
Only approved bots from the list below, to ask if your bot can be added please contact us.
Check for duplicates before posting, duplicates may be removed

Approved Bots

founded 2 years ago

MODERATORS

L3s@lemmy.world

enu@lemmy.world

technopagan@lemmy.world

1534

Linus Torvalds reckons AI is ‘90% marketing and 10% reality’ (www.tomshardware.com)

submitted 2 months ago by vegeta@lemmy.world to c/technology@lemmy.world

247 comments fedilink hide all child comments

you are viewing a single comment's thread
view the rest of the comments

[–] Rogers@lemmy.ml 1 points 2 months ago (1 children)

The latest llms get a perfect score on the south Korean SAT and can pass the bar. More than pure marketing if you ask me. That does not mean 90% of business that claim ai are nothing more than marketing or the business that are pretty much just a front end for GPT APIs. llms like claud even check their work for hallucinations. Even if we limited all ai to llms they would still be groundbreaking.

[–] clutchtwopointzero@lemmy.world 16 points 2 months ago (1 children)

Korean SAT are highly standardized in multiple choice form and there is an immense library of past exams that both test takers and examiners use. I would be more impressed if the LLMs could show also step by step problem work out...

[–] Rogers@lemmy.ml -5 points 2 months ago (1 children)

Claud 3.5 and o1 might be able to do that; if not, they are close to being able to do that. Still better than 99.99% of earthly humans

[–] Tamo240@programming.dev 8 points 2 months ago (1 children)

You seem to be in the camp of believing the hype. See this write up of an apple paper detailing how adding simple statements that should not impact the answer to the question severely disrupts many of the top model's abilities.

In Bloom's taxonomy of the 6 stages of higher level thinking I would say they enter the second stage of 'understanding' only in a small number of contexts, but we give them so much credit because as a society our supposed intelligence tests for people have always been more like memory tests.

[–] clutchtwopointzero@lemmy.world 1 points 2 months ago* (last edited 2 months ago)

Exactly... People are conflating the ability to parrot an answer based on machine-levels of recall (which is frankly impressive) vs the machine actually understanding something and being able to articulate how the machine itself arrived at a conclusion (which, in programming circles, would be similar to a form of "introspection"). LLM is not there yet