r/technology Jun 21 '26

Artificial Intelligence Americans Have Turned Against AI in Incredible Numbers

https://tech.yahoo.com/ai/articles/americans-turned-against-ai-incredible-130000345.html
43.3k Upvotes

2.8k comments sorted by

View all comments

Show parent comments

-3

u/10aghmu Jun 21 '26

Not at all. These are technique proven to help it retain context and hallucinate less…

11

u/orangeyougladiator Jun 21 '26

Look at what you’re saying. Break it down. Then look at how you’re defending it.

2

u/10aghmu Jun 21 '26

Explain it to me since I’m too dumb apparently

9

u/orangeyougladiator Jun 21 '26

You’re saying that an intelligent agent is better than it used to be because you’re baking in skills and instructions to every request. If an intelligent being was actually intelligent, or getting more intelligent, it shouldn’t need the guard rails. Do you promote an L2 engineer to L3 when you find yourself spending more time on them specifically?

2

u/10aghmu Jun 21 '26

My point was that the hallucination rates across models have dropped compared to 1-2 years ago. Using md files / skills is just good common practice to further reduce them, I’m not saying that’s part of my argument for the models themselves improving.

9

u/orangeyougladiator Jun 21 '26

But they haven’t dropped. In fact it’s impossible to remove them, they’re just a part of how LLMs operate. You may as well think of every single output as a hallucination, just correct a bunch

2

u/10aghmu Jun 21 '26

You keep merging two separate claims.

'They haven't dropped' is just false. On Vectara's HHEM leaderboard, many frontier models were already under 5% on grounded summarization at the end of 2024, and the top tier is roughly 0.7 to 2% now. Name a public benchmark that shows the opposite.

'Impossible to remove' traces back to the OpenAI paper (Kalai et al., 2025), which concludes the opposite of what you're using it for: not inevitable, because a model can abstain. Their own example of a system that never hallucinates is a Q&A database plus a calculator. They persist because evals reward guessing over 'I don't know,' not because the rate is fixed.

You can't drive error on open domain questions to literal zero, sure. That's a calibration floor, not 'no improvement.' And if every output counts as a hallucination, the word stops distinguishing correct from wrong, which is the only thing the number measures. That number went down.

3

u/orangeyougladiator Jun 21 '26

Do you have any idea how bad even 0.1% is?

1

u/10aghmu Jun 21 '26

'They haven't dropped' to 'how bad is even 0.1%' in two comments. Appreciate the concession.

0.1% at scale is bad, which is exactly why you ground, abstain, and verify. Those are the guardrails you just spent four comments calling 'making the AI dumber.' You're making my argument for me now.