Legal AI India Benchmark
Legal AI India Benchmark measures how accurately an AI model applies Indian law to concrete facts. It contains 100 self-contained micro cases across 13 legal domains, 3 difficulty tiers, and deliberate traps for outdated or repealed law.
Choose the model to test
Your key is sent only to OpenRouter. Save key stores it in this browser’s local storage; it is never included in benchmark results or sent to this website.
Generation and scoring settings
Recommended public-benchmark profile: temperature 0, Top P 1, 400 output tokens, model-default reasoning, and no seed. Any changed value is recorded in the exported JSON.
The model is answering
Cases are isolated and sent one at a time. Each answer is scored immediately against its gold rubric.
Auditable scorecard
Model answer, gold answer, and grade remain side by side. Open any grade to override the judge.
Finalising always downloads a JSON copy. A signed-in administrator can also save it to this server’s leaderboard. Checking administrator status…