WhichAI
← All tasks

Coding

Review a pull request

Critique a diff for correctness, style and edge cases.

Ranked from coding, reasoning, writing scores — no head-to-head votes for this task yet.

Coding · 50%Reasoning · 40%Writing · 10%

Start with this prompt

Written for this task. Fill in the bracketed parts.

Review this diff.

Focus on correctness first: logic errors, unhandled cases, race conditions,
anything that breaks an existing caller. Then readability. Ignore formatting.

For each issue, tell me the severity and the concrete scenario where it bites.
If the change is fine, say so plainly instead of inventing nitpicks.

---
[paste the diff]
  1. Perplexity
    Perplexity AI
    Strong data
    5.8
  2. 2
    Claude
    Anthropic
    Strong data
    5.5
  3. 3
    DeepSeek
    DeepSeek
    Strong data
    5.5
  4. 4
    Copilot
    Microsoft
    Strong data
    5.4
  5. 5
    Gemini
    Google
    Strong data
    5.4
  6. 6
    Grok
    xAI
    Strong data
    5.3
  7. 7
    ChatGPT
    OpenAI
    Strong data
    5.3
  8. 8
    Meta AI
    Meta
    Strong data
    5.2
  9. 9
    Mistral
    Mistral AI
    Strong data
    5.2

Disagree with this ranking?

Vote head-to-head and your answer feeds straight back into this board.

Help rank these →

Try another task