this post was submitted on 23 Jan 2025
241 points (96.9% liked)

Technology

61081 readers
2471 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related content.
  3. Be excellent to each other!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, to ask if your bot can be added please contact us.
  9. Check for duplicates before posting, duplicates may be removed
  10. Accounts 7 days and younger will have their posts automatically removed.

Approved Bots


founded 2 years ago
MODERATORS
 

LLMs performed best on questions related to legal systems and social complexity, but they struggled significantly with topics such as discrimination and social mobility.

“The main takeaway from this study is that LLMs, while impressive, still lack the depth of understanding required for advanced history,” said del Rio-Chanona. “They’re great for basic facts, but when it comes to more nuanced, PhD-level historical inquiry, they’re not yet up to the task.”

Among the tested models, GPT-4 Turbo ranked highest with 46% accuracy, while Llama-3.1-8B scored the lowest at 33.6%.

you are viewing a single comment's thread
view the rest of the comments
[–] A_A@lemmy.world 1 points 3 days ago (1 children)

Suppose A and B are at war and based on every insults they throw at each other, you train an LLM to explain what's going on. Well, it will be quite bad. Maybe this is some part of the explanation.

[–] FlyingSquid@lemmy.world 4 points 3 days ago (1 children)

But that's exactly the problem. Humans with degrees in history can figure out what is an insult and what is a statement of fact a hell of a lot better than an LLM.

[–] A_A@lemmy.world -2 points 3 days ago

it took maybe thousands or even millions of years for nature to create animals that understand who they are and what's going on around them. Give those machines a few more years, they are not all LLMs and they are advancing quite rapidly.
Finally, i completely agree with you that, for the time being, they are very bad at playing historian.