SOTAVerified

Legal Reasoning

Papers

Showing 1–50 of 92 papers

TitleStatusHype
IndianBailJudgments-1200: A Multi-Attribute Dataset for Legal NLP on Indian Bail Orders—0
Large Language Models Acing Chartered Accountancy—0
CHANCERY: Evaluating Corporate Governance Reasoning Capabilities in Language Models—0
When Fairness Isn't Statistical: The Limits of Machine Learning in Evaluating Legal Reasoning—0
Parameter Efficient Fine Tuning Llama 3.1 for Answering Arabic Legal Questions: A Case Study on Jordanian LawsCode0
LLM-based HSE Compliance Assessment: Benchmark, Performance, and AdvancementsCode0
LEXam: Benchmarking Legal Reasoning on 340 Law Exams—0
Incorporating Legal Structure in Retrieval-Augmented Generation: A Case Study on Copyright Fair UseCode0
SynLexLM: Scaling Legal LLMs with Synthetic Data and Curriculum Learning—0
Engineering the Law-Machine Learning Translation Problem: Developing Legally Aligned Models—0
Continual Pre-Training is (not) What You Need in Domain Adaption—0
KFinEval-Pilot: A Comprehensive Benchmark Suite for Korean Financial Language Understanding—0
An Explicit Syllogistic Legal Reasoning Framework for Large Language Models—0
Evaluating Test-Time Scaling LLMs for Legal Reasoning: OpenAI o1, DeepSeek-R1, and Beyond—0
Adaptively profiling models with task elicitation—0
CaseGen: A Benchmark for Multi-Stage Legal Case Documents GenerationCode1
Towards Robust Legal Reasoning: Harnessing Logical LLMs in Law—0
JUREX-4E: Juridical Expert-Annotated Four-Element Knowledge Base for Legal ReasoningCode1
NitiBench: A Comprehensive Studies of LLM Frameworks Capabilities for Thai Legal Question AnsweringCode0
Logical Lease Litigation: Prolog and LLMs for Rental Law Compliance in New York—0
Elevating Legal LLM Responses: Harnessing Trainable Logical Structures and Semantic Knowledge with Legal ReasoningCode0
LawGPT: Knowledge-Guided Data Generation and Its Application to Legal LLMCode1
Investigating the Shortcomings of LLMs in Step-by-Step Legal ReasoningCode0
Artificial Intelligence and Legal Analysis: Implications for Legal Education and the Profession—0
Domaino1s: Guiding LLM Reasoning for Explainable Answers in High-Stakes Domains—0
Legal Evalutions and Challenges of Large Language Models—0
LAR-ECHR: A New Legal Argument Reasoning Task and Dataset for Cases of the European Court of Human Rights—0
Weak-to-Strong Generalization beyond Accuracy: a Pilot Study in Safety, Toxicity, and Legal ReasoningCode0
Developing a Pragmatic Benchmark for Assessing Korean Legal Language Understanding in Large Language ModelsCode0
Using LLMs to Discover Legal Factors—0
KRAG Framework for Enhancing LLMs in the Legal Domain—0
Can Large Language Models Grasp Legal Theories? Enhance Legal Reasoning with Insights from Multi-Agent CollaborationCode0
LegiLM: A Fine-Tuned Legal Language Model for Data ComplianceCode0
LAPIS: Language Model-Augmented Police Investigation System—0
LeKUBE: A Legal Knowledge Update BEnchmarkCode0
Formalising Anti-Discrimination Law in Automated Decision Systems—0
Bridging Law and Data: Augmenting Reasoning via a Semi-Structured Dataset with IRAC methodology—0
Towards Supporting Legal Argumentation with NLP: Is More Data Really All You Need?—0
Explainable machine learning multi-label classification of Spanish legal judgements—0
Software Engineering Methods For AI-Driven Deductive Legal ReasoningCode0
LawInstruct: A Resource for Studying Language Model Adaptation to the Legal DomainCode1
ECtHR-PCR: A Dataset for Precedent Understanding and Prior Case Retrieval in the European Court of Human RightsCode0
PARAMANU-AYN: Pretrain from scratch or Continual Pretraining of LLMs for Legal Domain Adaptation?—0
Chain of Logic: Rule-Based Reasoning with Large Language Models—0
Advancing Legal Reasoning: The Integration of AI to Navigate Complexities and Biases in Global Jurisprudence with Semi-Automated Arbitration Processes (SAAPs)—0
Aalap: AI Assistant for Legal & Paralegal Functions in India—0
Automated legal reasoning with discretion to act using s(LAW)—0
TMID: A Comprehensive Real-world Dataset for Trademark Infringement Detection in E-CommerceCode0
Enhancing Logical Reasoning in Large Language Models to Facilitate Legal Applications—0
Modeling Legal Reasoning: LM Annotation at the Edge of Human AgreementCode0
Show:102550
← PrevPage 1 of 2Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1GPT-4Balanced Accuracy82.9—Unverified
2GPT-3.5Balanced Accuracy60.9—Unverified
3Claude-1Balanced Accuracy58.1—Unverified
#ModelMetricClaimedVerifiedStatus
1GPT-4Balanced Accuracy59.2—Unverified