P

Pakistan's Trial Courts (District Judiciary)

Pakistan's judiciary nets $38.50 per dollar invested by training judges on an AI legal assistant

Curated & reviewed by Peter Korpak, Founder & Chief Analyst, 100SignalsHow we verify
$38.50 saved per $1 investedReturn on Investment
~1,848 more cases per year per district (+6.3%)Additional Cases Resolved
Rated better in 59% of comparisons vs. 42% for untrained judgesJudgment Quality

Vendor-reported figures — source: the-decoder.com

Pakistan's Trial Courts (District Judiciary)
Metric Before After Impact
Cases Resolved Per Year Per District ~1,848 additional cases +1,848 cases per year (+6.3% increase)
Judgment Quality (Favorable Ratings) 42% 59% +17 percentage points
Return on Investment $38.50 per $1 invested 38.5x return per dollar invested

The Challenge

Pakistan has fewer than two judges per 100,000 residents (versus 22 in the EU and 30 in England and Wales), and by the end of 2024 had a very large backlog of pending cases, with the vast majority in trial courts. Judges work with minimal technology and no support staff, and before the study only about 25% had ever used a large language model like ChatGPT.

The Solution

Researchers from ETH Zurich, Imperial College London, and the New Economic School ran a randomized field experiment across 1,559 judges in 118 courts (roughly half of Pakistan's trial court judges), testing JudgeGPT, a GPT-4-based assistant using retrieval augmented generation over 129,235 documents (128,292 rulings and 943 laws) to generate cited answers. One group received the tool plus six 90-minute training sessions taught by ETH Professor Elliott Ash; a second group got the same tool access but only a general seminar; a control group attended the seminar with no AI access.

Results

Trained judges used JudgeGPT about four times as much as the untrained-but-equipped group (nearly 60 logins and 200+ prompts after 40 weeks versus about 20 logins and fewer than 50 prompts). Districts with more trained judges resolved roughly 1,848 more cases per year (a 6.3% increase) at moderate exposure, with even bottom-quartile districts clearing about 616 more cases. Researchers estimate a return of about $38.50 saved per dollar invested (at least $10 under conservative assumptions), while ruling quality, appeal rates, and work hours held steady or improved, with no detected increase in gender or religious bias.

Key Takeaways

  • Tool access alone barely moved usage or outcomes — targeted training on which tasks suit the AI (and which don't) was the real multiplier, driving substantially higher usage and measurable case-resolution gains.
  • Training shifted judges toward lower-risk tasks like editing and summarizing and away from open-ended legal questions where hallucination risk is higher, while keeping final decisions in human hands.
  • The study ran on GPT-4, a pre-reasoning model, suggesting today's more capable models could yield even larger productivity gains.

Share:

Details

Company Size
Enterprise
Quality
Curated
Source published
Jul 21, 2026
Last verified
Jul 28, 2026

Have a similar implementation?

Share your customer's AI results and link it to your vendor profile.

Submit a case study →