∴ Syllogism Lab

Can you out-reason an LLM?

Eight short arguments. For each one, decide whether the conclusion follows from the two premises. Then see how you did against the five models from my AAAI 2026 paper.

Judge the logic, not the facts. A false conclusion can follow, and a true one can fail to.
Argument 1 of 8 keys: Y / N

Syllogisms in 30 seconds

Every sentence has one of four shapes. S, M and P are the three terms: the conclusion links S to P, and the middle term M links them in the premises.

TypeShapeName
AAll S are Puniversal affirmative
ENo S are Puniversal negative
ISome S are Pparticular affirmative
OSome S are not Pparticular negative

A mood such as AAA gives the types of the two premises and the conclusion. A figure says where M sits. In the paper, models did best when M was the grammatical subject:

FigurePremisesM as subjectMean accuracy, 5 models
3M–P, M–S2×83.0%
4P–M, M–S1×80.6%
1M–P, S–M1×72.7%
2P–M, S–M0×70.2%

The eight arguments here are my own examples in classical forms; I checked each answer by brute force over every way three non-empty terms can overlap. Model numbers come from the paper: 11,000 WordNet-based items, zero-shot, no chain-of-thought.