THE THING ABOUT IT IS.

How We Know.

We don't just report what people say. We test it.

The short version

When someone says an event proves their opponents are lying, evil, or out to get them, we don't just repeat it and move on. We write the claim down exactly as they made it, figure out what else would have to be true if it were right, and go check. Then we tell you what we found, even when it's messy.

The claim

This is the "logic check" you see on every card.

State the claim

We write down exactly what's being said, in the plainest version and the strongest version. Not a strawman, the best version of the argument the other side would actually make for itself.

Write it down twice

Most claims have a plain version and a maximal one. "Executives are cutting jobs to get richer" can mean they expect to profit from automation, or it can mean profit is the whole point and everything else they say is a lie. We write both down before checking anything, and every check on the card says which version it's testing. A claim failing its strongest version is a much weaker result than one failing its plain version, and you should be able to tell those apart at a glance.

This turned out to matter more than we expected. Across the thirteen AI arguments we've published so far, 26 checks tested plain readings and not one came back false. 24 checks tested strongest readings and not one held up. People are mostly right about what happened and mostly wrong about what it proves.

Ask what else would have to be true

If the claim is right, some other things should be checkable in the real world. We list those things before we go look, so we can't quietly move the goalposts once we know the answer.

Go check

Each thing gets checked against what's actually observable, with a source attached. Sometimes it holds up. Sometimes it doesn't. Sometimes it's a mixed bag.

Give the honest verdict

Holds up, mixed evidence, falls apart, or there's nothing here we can actually test. We show our work on every card either way, not just the checks that make one side look bad.

Two different jobs

Not everything on this site is a claim you can test, so not everything gets a grade.

Tested claims get a grade

Someone argues that an event proves something about the other side. There's a claim, so there are checks, so there's a letter. That's the Fault Line Report.

Recorded statements get a severity

Someone said a thing about a group of people. There's nothing to test, because nobody is disputing it was said. What matters is how far it goes, so those carry a severity instead. Four steps. Stigmatizing casts a group as a problem to be managed. Hateful treats a whole group as lesser, dangerous, or not fully human. Threatening warns a group of harm, demands they stay quiet, or tells someone to act against them. Calls for violence asks for harm to come to a group, or celebrates harm that's already been done to them. Those cards carry no grade and no illustration, in the Weather Report and in the investigations alike.

Where the sections split

The Fault Line Report is camp against camp: claims people make about each other's motives, which can be tested. The Weather Report is a weekly pass over the same tracked accounts for language aimed at a group. Investigations are longer passes over one question, and each one opens onto every item it found. The World Cup and religion investigations are scored as hate rather than as polarization, so they carry severities, not grades.

Two tracks, never added together

Hate aimed at a protected group and dehumanizing language aimed at political opponents are both counted, and they're counted separately. They're different things with different histories and different consequences, and collapsing them into one number would flatter whichever side had a quieter week.

We don't reprint slurs

Where the wording is itself a slur, we describe it and link to the source rather than reproducing it. You can verify anything we say in one click. Documenting hate doesn't require handing it a bigger audience than it had.

The grade

The letter on a card is arithmetic, not an opinion, and here is the exact arithmetic so you can redo it yourself.

How it's calculated

Each check counts 1 if it held up, half if it came back mixed, and 0 if it didn't. Add those, divide by the number of checks that could actually be settled. 85% or better is an A, 65% a B, 45% a C, 25% a D, below that an F. Every card prints its own tally next to the letter, so you can check our math.

Some checks can't be settled, and those leave the math entirely

Can't be checked

Sometimes the world offers no way to answer a question in either direction. Nobody has measured where the gains from AI actually landed. Nobody outside one man's head knows which voices moved him. A strategy meant to stay quiet wouldn't produce a confession even if it were real. Marking those as failed checks would be claiming we looked and came up empty, which is a much stronger statement than the truth. So they drop out of the fraction completely, top and bottom, and the card says which ones and why. If nothing on a card can be settled, it gets no grade at all rather than a bad one.

What the grade isn't

It isn't a truth score, and a low grade isn't a claim that someone lied. It measures how a claim fared against the specific checks we chose to run. We pick those checks, and a different reader could reasonably pick different ones and land somewhere else. That is why every check is printed with its own reasoning instead of just the letter.

Where it runs out

The hardest cases are claims about what someone privately wants. Motives can sometimes be evidenced, by internal documents or a pattern of conduct, but often the honest answer is that nobody outside that person's head can check it. Where a claim rests mostly on a motive we can't test, the grade is measuring the parts around it, not the mind-reading, and we say so on the card.

The scoring

Once a story clears our bar, it gets scored 1 to 5 on five things. You'll see these as chips (like "R 4 · I 4 · P 3") on the full-scoring section of every card.

R · Reach

How far it actually traveled. A senator's floor speech reaches further than a reply nobody saw.

I · Intensity

How nasty the framing gets. Arguing the merits is a 1. Assuming the other side is lying by default is a 3. Treating the whole camp as one evil essence is a 5.

P · Penetration

How many separate, unconnected people actually believe it. One loud account scores low. A belief that's become the camp's default answer scores high.

E · Elite carriage

Whether people with real power (senators, cabinet officials, network anchors) say it themselves, or whether it stays on anonymous accounts.

L · Lock-in

Whether the other side has built a mirror-image version of the same story, so both sides now point at each other's worst moment as proof they were right all along.

Those five combine into a "charge stage," also shown as a chip:

C1 · Something happened

An event occurs. Coverage is still just describing it. Nobody's claimed it as proof of anything yet.

C2 · Both sides start spinning it

Competing "here's what this really shows about them" versions show up. The event becomes ammunition.

C3 · One version wins

One framing becomes the camp's go-to answer. Saying otherwise inside the camp starts to cost you something.

C4 · It outlives the event

The story sticks around and gets reused on whatever happens next, whether or not it actually fits. This is most of what you'll see on this site, because it's the stuff that keeps showing up.

Double-checked

Every quote gets traced to a dated, linkable source rather than a screenshot or someone else's summary. Then a second pass tries to break the finding on purpose: wrong date, missing context, a quote that doesn't actually say what it's being used to say. What survives both passes is what makes it onto this site.

Two kinds of receipt, and we label which is which

Some quotes we read where the person published them: their own post, their own press release, their own filing. Others reached us through a news outlet that was in the room, and those say so directly in the byline, like "as reported by Fortune." We think both are usable, but they aren't the same strength of evidence, so we never blur them.

A verified quote isn't a verified claim

Confirming that someone said a thing, on a date, is a different job from confirming the thing is true. The receipts section does the first. The logic check does the second. A card can have airtight quotes and still grade badly, and that's the instrument working, not failing.

Want the whole thing

This site is the plain-language cut. Every page carries a link to the research it came from, and those pages hold the full carrier sets, the per-item scoring, and the coverage notes on what we couldn't verify.