LegalOSS158.6ktracked

Benchmarks & Datasets

Evaluation suites and datasets for legal-domain systems.

39 projects
MCP Servers · Benchmarks & Datasets
awesome-legaltech
Vaquill-AI/awesome-legaltech

A curated list of awesome LegalTech resources - open source platforms, AI models, MCP servers, companies, datasets, and tools for the global legal ecosystem.

2161004h ago
Retrieval & RAG · Case Law & Legal Data
open-us-law
Vaquill-AI/open-us-law

Open, structured corpus of US primary law: 2M+ sections of state statutory codes, the US Code, and state constitutions, plus the ingestion pipeline that builds it. Data on Hugging Face (CC BY 4.0); scrapers Apache-2.0. Quarterly snapshots.

571516d agoPython
Retrieval & RAG · Legal AI & NLP
open-legal-answer-benchmark
Vaquill-AI/open-legal-answer-benchmark

Open, reproducible benchmark of US legal-answer quality. Verified questions, a standard-library scorer, and results anyone can rerun.

101mo agoPython
Benchmarks & Datasets
harvey-labs
harveyai/harvey-labs

A benchmark built to evaluate and improve agent capabilities for supporting legal work.

1.3k2272d agoPython
Legal AI & NLP · Benchmarks & Datasets
LLM-and-Law
Jeryi-Sun/LLM-and-Law

Continually updated LLM and law papers

331412d ago
Legal AI & NLP · Benchmarks & Datasets
legalise
b1rdmania/legalise

Open-source governance infrastructure for AI-assisted legal work

2454d agoPython
Legal AI & NLP · Benchmarks & Datasets
LawBench
open-compass/LawBench

Benchmarking Legal Knowledge of Large Language Models

451742y agoPython
Legal Research & Search · Case Law & Legal Data
italia-corpus
ahmeabd/italia-corpus

Machine-readable Italian legal corpus

432251mo ago
Contracts & Analysis · Legal AI & NLP
claude-legal-skill
evolsb/claude-legal-skill

AI-powered contract review skill with CUAD risk detection, market benchmarks, and lawyer-ready redlines. Works with Claude Code, Codex, Cursor, and 26+ tools.

432531mo ago
Legal AI & NLP · Benchmarks & Datasets
pile-of-law
pile-of-law/pile-of-law

We curate a large corpus of legal and administrative data. The utility of this data is twofold: (1) to aggregate legal and administrative data sources that demonstrate different norms and legal standards for data filtering; (2) to collect a dataset that can be used in the future for pretraining legal-domain language models, a key direction in access-to-justice initiatives.

2847.7k1mo ago
Legal Research & Search · Benchmarks & Datasets
prinzbench
prinz-ai/prinzbench

Private benchmark ranking LLMs on legal

13141mo ago
Case Law & Legal Data · Benchmarks & Datasets
open-australian-legal-corpus-creator
isaacus-dev/open-australian-legal-corpus-creator

The code used to create and update the Open Australian Legal Corpus, the first and only multijurisdictional open corpus of Australian legislative and judicial documents.

124201y agoPython
Legal AI & NLP · Benchmarks & Datasets
LexEval
CSHaitao/LexEval

LexEval: A Comprehensive Benchmark for Evaluating Large Language Models in Legal Domain

103131y agoPython
Case Law & Legal Data · Compliance & Privacy
open-australian-legal-corpus
isaacus/open-australian-legal-corpus

Open Australian Legal Corpus ‍⚖️ The Open Australian Legal Corpus by Isaacus, a foundational legal AI research company, is the first and only multijurisdictional open corpus of Australian legislative and judicial documents. Comprised of 229,122 texts totalling over 60 million lines and 1.4 billion tokens, the Corpus includes every in force statute and regulation in the Commonwealth, New South Wales, Queensland, Western Australia, South Australia, Tasmania and Norfolk Island, in… See the full description on the dataset page: https://huggingface.co/datasets/isaacus/open-australian-legal-corpus.

9518.3k6mo ago
Benchmarks & Datasets
hupd
suzgunmirac/hupd

The Harvard USPTO Patent Dataset

89132y agoJupyter Notebook
Legal AI & NLP · Benchmarks & Datasets
lex_glue
coastalcph/lex_glue

LexGLUE: benchmark for legal language understanding

8319.3k2y ago
Benchmarks & Datasets
Multi_Legal_Pile
joelniklaus/Multi_Legal_Pile

Multi Legal Pile is a dataset of legal documents in the 24 EU languages.

675.6k2y ago
Benchmarks & Datasets
eur-lex-sum
dennlinger/eur-lex-sum

The EUR-Lex-Sum dataset is a multilingual resource intended for text summarization in the legal domain. It is based on human-written summaries of legal acts issued by the European Union. It distinguishes itself by introducing a smaller set of high-quality human-written samples, each of which have much longer references (and summaries!) than comparable datasets. Additionally, the underlying legal acts provide a challenging domain-specific application to legal texts, which are so far underrepresented in non-English languages. For each legal act, the sample can be available in up to 24 languages (the officially recognized languages in the European Union); the validation and test samples consist entirely of samples available in all languages, and are aligned across all languages at the paragraph level.

502.2k2y ago
Case Law & Legal Data · Benchmarks & Datasets
multi_eurlex
coastalcph/multi_eurlex

Multilingual EU law classification dataset

461.4k2y ago
Legal AI & NLP · Benchmarks & Datasets
hupd
HUPD/hupd

The Harvard USPTO Patent Dataset (HUPD) is a large-scale, well-structured, and multi-purpose corpus of English-language patent applications filed to the United States Patent and Trademark Office (USPTO) between 2004 and 2018. With more than 4.5 million patent documents, HUPD is two to three times larger than comparable corpora. Unlike other NLP patent datasets, HUPD contains the inventor-submitted versions of patent applications, not the final versions of granted patents, allowing us to study patentability at the time of filing using NLP methods for the first time.

452.3k3y ago
For maintainers

Build one of these? Verify ownership through GitHub and take over your project's page: tagline, categories, maintainer's note.

How claiming works