NYAYAANUMANA and INLEGALLLAMA: The Largest Indian Legal Judgment Prediction Dataset and Specialized Language Model for Enhanced Decision Analysis (2025.coling-main)
Copied to clipboard
Shubham Kumar Nigam, Deepak Patnaik Balaramamahanthi, Shivam Mishra, Noel Shallum, Kripabandhu Ghosh, Arnab Bhattacharya
| Challenge: | In India, a significant backlog of cases burdens the legal system. |
| Approach: | They present a corpus of 7,02,945 preprocessed Indian legal cases compiled for LJP . they use a domain-specific generative large language model tailored to the intricacies of the legal system . |
| Outcome: | The proposed dataset surpasses existing datasets like PredEx and ILDC, and improves prediction accuracy and comprehensible explanations. |
Similar Papers
ILDC for CJPE: Indian Legal Documents Corpus for Court Judgment Prediction and Explanation (2021.acl-long)
Copied to clipboard
Vijit Malik, Rishabh Sanjay, Shubham Kumar Nigam, Kripabandhu Ghosh, Shouvik Kumar Guha, Arnab Bhattacharya, Ashutosh Modi
| Challenge: | a system that could assist a judge in predicting the outcome of a case should be explainable. |
| Approach: | They propose to use a corpus of 35k Indian Supreme Court cases annotated with original court decisions to promote research in this area. |
| Outcome: | The proposed system has an accuracy of 78% versus 94% for human legal experts. |
Legal Judgment Reimagined: PredEx and the Rise of Intelligent AI Interpretation in Indian Courts (2024.findings-acl)
Copied to clipboard
| Challenge: | Prediction with Explanation is the largest expert-annotated dataset for legal judgment prediction and explanation in the Indian context . |
| Approach: | They propose to use an annotated legal judgment prediction corpus to improve models' accuracy . they employ transformer-based models tailored for both general and Indian legal contexts . |
| Outcome: | The proposed system improves the accuracy and explanatory depth of models for legal judgments. |
Precedent-Enhanced Legal Judgment Prediction with LLM and Domain-Model Collaboration (2023.emnlp-main)
Copied to clipboard
Yiquan Wu, Siying Zhou, Yifei Liu, Weiming Lu, Xiaozhong Liu, Yating Zhang, Changlong Sun, Fei Wu, Kun Kuang
| Challenge: | Recent advances in deep learning have enabled a variety of techniques to be used to solve the LJP task. |
| Approach: | They propose a framework that leverages the strength of both LLMs and domain-specific models in the context of precedents. |
| Outcome: | The proposed framework leverages the strength of both LLM and domain models in the context of precedents. |
ILSIC: Corpora for Identifying Indian Legal Statutes from Queries by Laymen (2026.findings-eacl)
Copied to clipboard
| Challenge: | Existing studies have focused on the use of court judgments as input for legal Statute Identification (LSI) however, there is little research to explore the differences between court and laypeople data for LSI. |
| Approach: | They create a corpus of laypeople queries covering 500+ statutes from Indian law . they use court case judgements to compare between the two datasets . |
| Outcome: | The proposed corpus of laypeople queries covers 500+ statutes from Indian law . the results show that models trained on court judgements are ineffective . |
LLMs – the Good, the Bad or the Indispensable?: A Use Case on Legal Statute Prediction and Legal Judgment Prediction on Indian Court Cases (2023.findings-emnlp)
Copied to clipboard
Shaurya Vats, Atharva Zope, Somsubhra De, Anurag Sharma, Upal Bhattacharya, Shubham Nigam, Shouvik Guha, Koustav Rudra, Kripabandhu Ghosh
| Challenge: | Large Language Models have touched upon many real-life tasks. |
| Approach: | They apply Large Language Models to two popular tasks: Statute Prediction and Judgment Prediction. |
| Outcome: | The proposed model performs well in Statute Prediction and Judgment Prediction on Indian Supreme Court cases. |
Neural Legal Judgment Prediction in English (P19-1)
Copied to clipboard
| Challenge: | Recent work on legal judgment prediction has focused on Chinese, but only feature-based models have been considered in English. |
| Approach: | They propose a hierarchical version of BERT which bypasses BERT’s length limitation. |
| Outcome: | The proposed model outperforms existing models in binary violation classification, multi-label classification and case importance prediction. |
Legal Judgment Prediction via Topological Learning (D18-1)
Copied to clipboard
| Challenge: | Existing studies focus on a specific subtask of judgment prediction and ignore the dependencies among subtasks. |
| Approach: | They propose a topological multi-task learning framework that incorporates multiple subtasks and DAG dependencies into judgment prediction. |
| Outcome: | The proposed model improves on baselines on all judgment prediction tasks. |
Legal Judgment Prediction based on Knowledge-enhanced Multi-Task and Multi-Label Text Classification (2025.naacl-long)
Copied to clipboard
| Challenge: | Legal judgment prediction (LJP) is an essential task for legal AI, aiming at predicting judgments based on the facts of a case. |
| Approach: | They propose a knowledge-enhanced approach that incorporates 'label-level knowledge' to enhance the representation of case facts for each task and 'task-level' knowledge to improve synergy. |
| Outcome: | The proposed method is effective in comparison to state-of-the-art (SOTA) baselines. |
IL-TUR: Benchmark for Indian Legal Text Understanding and Reasoning (2024.acl-long)
Copied to clipboard
| Challenge: | Legal systems worldwide struggle with exponentially growing legal cases in various courts. |
| Approach: | They propose a benchmark for Indian legal text understanding and reasoning task that includes domain-specific tasks that address different aspects of the legal system. |
| Outcome: | The proposed benchmark for Indian legal text understanding and reasoning aims to address the gap between models and the ground truth. |
HLDC: Hindi Legal Documents Corpus (2022.findings-acl)
Copied to clipboard
Arnav Kapoor, Mudit Dhawan, Anmol Goel, Arjun T H, Akshala Bhatnagar, Vibhu Agrawal, Amul Agrawal, Arnab Bhattacharya, Ponnurangam Kumaraguru, Ashutosh Modi
| Challenge: | Existing systems that process legal documents are lacking high-quality corpora in low resource languages such as Hindi. |
| Approach: | They propose a Hindi Legal Documents Corpus (HLDC) that contains 900K legal documents in Hindi. |
| Outcome: | The proposed model is based on a corpus of more than 900K legal documents in Hindi. |