Goodmarks benchmark

AI cracks JEE Advanced 2026 with ease

Four frontier models sat both official papers blind. After we restored Paper 2 Mathematics question 4 where the written value was 1/3, Claude Fable 5.1 leads on 358 out of 360. Three models still sit above 97 percent.

Published September 12, 2026 · Official papers · 102 questions · 360 marks

Sunlit examination hall with empty wooden desks and closed question booklets
JEE Advanced 2026 was held on May 17, 2026. Two papers. Three subjects. 360 marks. The models saw the printed questions only.

JEE Advanced is the filter after JEE Main. A few lakh students sit Main. A few thousand walk into an IIT. The 2026 paper still looks like the exam families know: long diagrams, multi correct lists, numerical ranges, and matching columns that punish a single sloppy option.

Goodmarks split the official English papers into 102 single question PDFs and sent each one to four models. The prompt carried the question images and the question type. It never carried the official key. Scoring happened after the model replied, using the IIT solutions booklets and the paper’s own marking rules.

The result is blunt. The best models are no longer grazing the cutoff. They are finishing within a handful of marks of a perfect script.

The 2026 mark list

Official papers. Blind answers. Same key for every model. Paper 2 Mathematics question 4 is rescored below where the written value was 1/3.

Claude Fable 5.1

358 / 360

99.4%

101 correct · 0 wrong · 1 partial

GPT 6 Astra

356 / 360

98.9%

101 correct · 1 wrong

GPT 5.5

351 / 360

97.5%

100 correct · 2 wrong

Claude Sonnet 5

341 / 360

94.7%

96 correct · 2 wrong · 2 partial · 2 error

Ranked score out of 360

1Claude Fable 5.1Anthropic

358 / 36099.4%

2GPT 6 AstraOpenAI

356 / 36098.9%

3GPT 5.5OpenAI

351 / 36097.5%

4Claude Sonnet 5Anthropic

341 / 36094.7%

Horizontal bar chart of four model scores on JEE Advanced 2026
Claude Fable 5.1 358, GPT 6 Astra 356, GPT 5.5 351, Claude Sonnet 5 341. Fable and Sonnet gained 4 marks on question 4 after the option table was restored.

For more details check the JEE Benchmarks page.

Marks left on the table

A full bar would mean a perfect 360. Claude Fable 5.1 dropped 2 marks. GPT 6 Astra dropped 4. GPT 5.5 dropped 9. Claude Sonnet 5 dropped 19.

Claude Fable 5.1

2

GPT 6 Astra

4

GPT 5.5

9

Claude Sonnet 5

19

Scale is relative to the largest drop so small gaps stay readable.

What “with ease” actually means

A human who scores in the mid 300s on JEE Advanced is not having a normal day. That band is the air that All India Rank 1 lives in most years. Claude Fable 5.1 is 2 marks short of a clean sweep. GPT 6 Astra is 4 marks short. GPT 5.5 is 9 marks short and still at 97.5 percent.

Paper 1 was almost a formality. 3 of the 4 models scored a perfect 180 out of 180 on Paper 1: Claude Fable 5.1, GPT 6 Astra, and GPT 5.5. Physics was a clean sheet for both OpenAI models. Chemistry was a clean sheet for Claude Fable 5.1 and GPT 6 Astra. After the question 4 rescore, both Claude models also sit on a full 120 in Mathematics.

Claude Sonnet 5 is the only model in this set that still looks mortal. It scored 341, lost most of those marks in Chemistry, and returned two questions as errors. Even then, 94.7 percent on the official 2026 paper is a score most coaching classrooms would print on a banner.

Subject by subject

Each subject is out of 120. Full height is full marks.

Claude Fable 5.1

GPT 6 Astra

GPT 5.5

Claude Sonnet 5

Mathematics

120

116

116

120

Physics

118

120

120

118

Chemistry

120

120

115

103

Mathematics is now a clean sheet for both Claude models (120 out of 120). The two OpenAI models stay on 116 because they left no written 1/3 on question 4. Physics is nearly solved. Chemistry is where the field splits: 120 for Fable and Astra, 115 for GPT 5.5, 103 for Claude Sonnet 5.

Paper 1 versus Paper 2

Each paper is 180 marks. Paper 1 has 48 questions. Paper 2 has 54.

Paper 1

Claude Fable 5.1180 / 180
GPT 6 Astra180 / 180
GPT 5.5180 / 180
Claude Sonnet 5167 / 180

Paper 2

Claude Fable 5.1178 / 180
GPT 6 Astra176 / 180
GPT 5.5171 / 180
Claude Sonnet 5174 / 180

A scoring fix on question 4

Paper 2 Mathematics question 4 is a single correct definite integral worth 3 marks. The official letter is B, which is 1/3. All four models returned letter A. That looked like a shared miss until we checked the image we had sent them.

Our split of the official PDF hid the denominators under A and B behind a white box. The models saw 1 over a blank for both plain fractions. C and D (the log forms) stayed readable. Mapping a correct 1/3 onto a letter was then a guess. That is our fault, not a calculus miss.

So we rescored the item on the written value. If a model wrote 1/3, we restored the full plus 3. That is a 4 mark swing from the raw minus 1. If a model left only the letter A and no value, the raw mark stands.

The printed question, now with the full option table, is:

0213x+3dx\displaystyle \int_{0}^{2} \frac{1}{3^{x}+3}\,dx
Official JEE Advanced 2026 Paper 2 Mathematics question 4, a definite integral from 0 to 2 of 1 over 3 to the x plus 3
Restored option table. Paper 2, Mathematics, question 4. Official letter B is 1/3. Single correct. Plus 3 or minus 1.

(A)

1/21/2

Printed option

(B)

1/31/3

Official key

(C)

(ln3)/3(\ln 3)/3

Printed option

(D)

(ln3)/2(\ln 3)/2

Printed option

Claude Fable 5.1

Letter A, written value 1/3

Plus 3 restored

GPT 6 Astra

Letter A, no written value

Raw minus 1 stands

GPT 5.5

Letter A, no written value

Raw minus 1 stands

Claude Sonnet 5

Letter A, written value 1/3

Plus 3 restored

The integral itself

Put u = 3 to the x. Then du = (3 to the x)(ln 3) dx, so dx = du / (u ln 3). The limits change from x = 0 to x = 2 into u = 1 to u = 9. The integral becomes

1ln3191u(u+3)du\displaystyle \frac{1}{\ln 3}\int_{1}^{9}\frac{1}{u(u+3)}\,du

Split the fraction: 1 over u(u+3) equals (1/3) times (1/u minus 1/(u+3)). The antiderivative is (1/(3 ln 3)) ln(u/(u+3)). From 1 to 9 that is

13ln3(ln912ln14)=ln33ln3=13\displaystyle \frac{1}{3\ln 3}\left(\ln\frac{9}{12}-\ln\frac{1}{4}\right)=\frac{\ln 3}{3\ln 3}=\frac{1}{3}

The value is 1/3. On a complete option table that is letter B.

Who wrote 1/3

Claude Fable 5.1 rewrote the integrand, reduced the log term, and arrived at 1/3. It even noted that the printed choices were not clear, then attached that value to letter A. Claude Sonnet 5 used the swap that replaces x with 2 minus x, showed that f(x) plus f(2 minus x) equals 1/3, and also arrived at 1/3 before locking A. Both get plus 3.

GPT 6 Astra and GPT 5.5 left no written work on this item. They returned only letter A. We cannot show that they held the value 1/3, so we did not add marks. GPT 6 Astra’s only remaining miss is this letter. GPT 5.5 still has this letter plus a wrong multi correct on Paper 2 Chemistry question 9.

Claude Fable 5.1 now has no fully wrong item: 101 correct and one partial on Paper 2 Physics question 5 (A and D instead of A, C, and D). Claude Sonnet 5 still drops the rest of its marks in Chemistry, including two questions that returned no parseable answer.

Why this is not memorisation

GPT 5.5 shipped on April 23, 2026. JEE Advanced 2026 was held on May 17, 2026, 24 days later. The 2026 paper did not exist when that model went public. The other three models sat the same printed pages with the same blind prompt.

Each request was one question. Diagrams stayed as printed. Numerical answers counted only if they landed inside the official range. Multi correct items used the paper’s partial marking, so a missing option is not silently forgiven.

How Goodmarks ran the bench

  1. Step 1

    Split the official papers

    Cut the English Paper 1 and Paper 2 booklets into 102 single question PDFs. Options, figures, and matching lists stayed as printed.

  2. Step 2

    Keep every model blind

    Send only the question images and the type: single correct, multi correct, numerical, or matching. The official key never entered the prompt.

  3. Step 3

    Ask for a structured final answer

    Each model could think first. It had to finish with a JSON object. Scoring used that final object, not the prose around it.

  4. Step 4

    Mark with the IIT key

    Apply official solutions after the reply. Use partial marks on multi correct items. Accept a numerical value only inside the published range. On Paper 2 Mathematics question 4 we later restored plus 3 if the written value was 1/3, because our image hid the A and B denominators.

What this means if you are sitting the paper

Models can now finish JEE Advanced. That does not sit the exam for you. Partial marking still bites. Chemistry still separates a 358 from a 341. A clipped option table can also turn a correct value into the wrong letter, which is why we publish the image defect and the rescore.

If you are preparing, the useful lesson is narrower. The paper is readable. The marks live in careful options, not in a secret chapter no one taught. Practice the same shapes the models saw: multi correct lists, numerical ranges, and matching columns, with instant checking and a written solution after every attempt.

The full GPT 5.5 question log still shows the raw letter A on this item, because that run left no written 1/3. See the GPT 5.5 JEE Advanced 2026 report.

For more details check the JEE Benchmarks page. It has the full 2025 and 2026 leaderboard, paper by paper marks, and every miss.

Practise the same exam on Goodmarks

Topic wise JEE MCQs with instant checking and step by step solutions. Free samples on every topic. Pro unlocks the full bank.

Frequently asked questions

What score did the best model get on JEE Advanced 2026?

Claude Fable 5.1 scored 358 out of 360 (99.4 percent) on the official papers: 101 of 102 questions correct, 1 partial, 0 wrong.

How did the four models compare?

After a rescore on Paper 2 Mathematics question 4, Claude Fable 5.1 scored 358, GPT 6 Astra scored 356, GPT 5.5 scored 351, and Claude Sonnet 5 scored 341. All four sat both official English papers with no answer key in the prompt.

What happened on Paper 2 Mathematics question 4?

The integral from 0 to 2 of 1 over (3 to the x plus 3) equals 1/3, which is option B. Our split image hid the denominators under A and B. Claude Fable 5.1 and Claude Sonnet 5 wrote 1/3, so we restored plus 3. GPT 6 Astra and GPT 5.5 returned only letter A with no written value, so their raw minus 1 stands.

Did the models see the official answers?

No. Each request included only the printed question and its type. Answers from the official solutions booklets were applied after the model replied.

Could GPT 5.5 have memorised JEE Advanced 2026 from training data?

Unlikely. GPT 5.5 shipped on April 23, 2026. JEE Advanced 2026 was held on May 17, 2026, 24 days later, so the 2026 paper did not exist when that model went public.