[-] h2aichat_com@lemmy.world 1 points 5 days ago

Turns out I answered this three weeks ago — and posted it next to your comment instead of under it, so you never got the notification. A project about machines making confident mistakes, undone by a threading bug. Noted.

Short version: yes, and that's exactly why that one's in the screenshot. Of the 141 claims we marked, it's the only one you can check without leaving the sentence — no source, no taking our word for it. The other 140 aren't arithmetic. One was a National Holidays Act 1946 that doesn't exist. Another was a Council of Europe report nobody wrote, which a second model then cited back as established fact. You can't do the maths on those. You have to go and look.

[-] h2aichat_com@lemmy.world 1 points 1 week ago

Yes. Groundbreaking stuff — next week we tackle whether water is wet.

The cutoff wasn't the surprise, though. The tone was. The two that "corrected" us didn't hedge: they came back with a specific figure and a specific month to put us right. Being out of date sounded exactly like being rigorous, and that's the bit a cutoff date doesn't warn you about.

We changed the method after this one — every debate now starts from a briefing we've checked ourselves, so they argue about a fact instead of trying to remember one.

[-] h2aichat_com@lemmy.world 1 points 2 weeks ago

Fair hit, and deserved.

That was a file path from my machine sitting where the post body should have been — the script took a filename as the body and published it, and the rehearsal step never showed the body, so it went out unread. The text is up now.

Thanks for the nudge, even sideways.

[-] h2aichat_com@lemmy.world 1 points 2 weeks ago

Agreed on the behaviour — a stale figure from a model is not surprising on its own.

The bit I thought was worth writing down was what the rest of the table did with it: they adopted the stale number as the rigorous one and argued from it, and the model that had it wrong ended up sounding like the careful one in the room. Confidently wrong travels further than right.

Fair enough that this is not news here, though. Wrong community for it.

[-] h2aichat_com@lemmy.world 1 points 2 weeks ago

You're right, and I'd already said the same thing to kata1yst further down before seeing yours — this was the wrong post for this community. "Models get facts wrong" is not news to anyone in fosai, and I should have worked that out before posting rather than after.

There is a second reason it read badly, and that one is entirely mine: the body of this post went out as a file path from my own machine instead of the actual text. My publishing script took a filename as the body and posted it verbatim, and the dry run never printed the body, so nobody caught it. What you saw was a broken post making an obvious point. I have replaced the text and fixed the script.

I am not posting here again unless it is something this community actually talks about. Thanks for saying it straight instead of just downvoting.

[-] h2aichat_com@lemmy.world 2 points 2 weeks ago

You're right, and thank you for saying it plainly. This was the wrong post for this community, and the wrong framing on my part — "models make mistakes" is not news to anyone here, and I should have seen that before posting.

Sorry for the noise. Next time I post here it will be something that actually fits what this community talks about.

-6
submitted 3 weeks ago* (last edited 2 weeks ago) by h2aichat_com@lemmy.world to c/fosai@lemmy.world

Disclosure up front: this is my project — H2AI Chat, AGPL, where several models from different vendors debate a topic in turns while a human moderates.

We ran the same question twice with the same six models, changing one thing.

Without a briefing. We told them Bitcoin's all-time high is $126,080, set in October 2025. Two of them "corrected" us: the previous all-time high "was approximately $69,000 in November 2021, not $126,080." True — five years ago. A third pushed harder: "Either you missed the correction, or you're deliberately using inflated baseline numbers. Which is it?"

The interesting part is not that they were stale. It is what happened next: the table adopted the stale figure as its standard of rigour and argued from it, with the model that got it wrong sounding like the careful one in the room.

With a verified briefing in front of them: not one correction of that kind.

Without: https://h2aichat.com/conversations/en/h2aichat_bitcoin_no_briefing_2026-08-22.html With: https://h2aichat.com/conversations/en/h2aichat_bitcoin_briefed_2026-08-22.html

Nothing is edited in either page. Claims that do not hold are struck through, with the reason and the source underneath.


Edited 2026-08-27. This post originally went out with a file path from my own machine where the text should have been: the publishing script took a filename as the body and posted it verbatim, and the dry run never showed the body, so nobody saw it. That is why the post made no sense, and my apologies to everyone who tried to read it.

[-] h2aichat_com@lemmy.world 2 points 3 weeks ago

You're right, and that one you can — which is exactly why it's the one in the screenshot. Of the 141 claims we marked, it's the only one you can check without leaving the sentence. No source needed, no taking our word for it.

The other 140 aren't like that. A National Holidays Act 1946 that doesn't exist. A Council of Europe report nobody wrote, which a second model then treated as established and a third did arithmetic on. A 2016 study in the American Journal of Psychiatry that is really a 2008 one in the BMJ — right author, invented journal and decade, then reused four more times as a baseline.

Those took two days by hand, and no amount of arithmetic gets you there. The maths one is in the picture precisely because it's the one that needs nothing from us.

0

I built H2AI Chat, an AGPL platform where several different models — from different vendors — debate a topic in turns while a human moderates. Disclosure up front: this is my project.

Over the past two days we hand-verified 41 of those debates, claim by claim: 141 statements marked, 44 of them flatly false.

We don't delete or correct them. The sentence stays, struck through, and you can still read it by selecting it — with the reason and the source underneath. Editing what a model said would break the only promise the site makes.

Three patterns we didn't expect:

  • Fabricated authority shows up exactly where an argument is challenged. One debate answers a budget objection with three invented citations in a single turn.
  • Fabrications spread between models. One invents a figure, a second treats it as established, a third does arithmetic on it.
  • One claim contradicts itself inside its own sentence: "62% voted Remain on a 67% turnout, meaning roughly 22% of the electorate" — which is 41.5%.

Debates: https://h2aichat.com/ Code and the fact-check register: https://github.com/Tonterias/h2aichat

h2aichat_com

0 post score
0 comment score
joined 3 weeks ago