Open Source Legal Software.
Every project is a real GitHub repository, stats refreshed from the source. One listing per repo, reviewed by the community, claimed by its maintainer.
Interactive EU legislation reader in 24 languages, with article and recital links, CJEU citations, a CLI, REST API and MCP endpoint.
Dataset Card for "IL-TUR" Dataset Description Summary "IL-TUR": Benchmark for Indian Legal Text Understanding and Reasoning is a collaborative effort to establish a modern benchmark for training and evaluating AI/NLP models on Indian Law. IL-TUR consists of 8 foundational tasks, requiring different types of understanding and skills. Apart from English, some tasks involve Indic languages. This dataset repository has been created to unify the data… See the full description on the dataset page: https://huggingface.co/datasets/Exploration-Lab/IL-TUR.
An open-source patent analyzing web application that focuses on usability and user experience
Autonomous contract risk analysis - scores clauses, suggests negotiation language, gives SIGN/NEGOTIATE/REJECT verdict, compares versions, watches folders. English + Turkish.
Rust search engine for French and EU law: BM25 plus vector retrieval on Postgres with ParadeDB and VectorChord, citation graph resolution, a REST API, and a hosted MCP endpoint.
Open standard for legal context
Clio practice management connector
Process markdown with YAML front matter, conditional clauses, cross-references, imports, and generate professional PDFs ready to be shared.
Dataset Card for "LexFiles" Dataset Summary The LeXFiles is a new diverse English multinational legal corpus that we created including 11 distinct sub-corpora that cover legislation and case law from 6 primarily English-speaking legal systems (EU, CoE, Canada, US, UK, India). The corpus contains approx. 19 billion tokens. In comparison, the "Pile of Law" corpus released by Hendersons et al. (2022) comprises 32 billion in total, where the majority (26/30) of… See the full description on the dataset page: https://huggingface.co/datasets/lexlms/lex_files.
Dataset Summary This dataset was introduced in CaseSumm: A Large-Scale Dataset for Long-Context Summarization from U.S. Supreme Court Opinions. The CaseSumm dataset consists of U.S. Supreme Court cases and their official summaries, called syllabuses, from the period 1815-2019. Syllabuses are written by an attorney employed by the Court and approved by the Justices. The syllabus is therefore the gold standard for summarizing majority opinions, and ideal for evaluating other… See the full description on the dataset page: https://huggingface.co/datasets/ChicagoHAI/CaseSumm.
Nomos — a programming language for legal reasoning. Typed rules with jurisdiction and validity dates, LLM-powered fact extraction, defeasible logic, proof trees that cite statutes and cases. Apache-2.0. Experimental.
A proof-of-concept Unicode obfuscation tool for DOCX and PDF documents.
34 skills covering full case lifecycle
Build one of these? Verify ownership through GitHub and take over your project's page: tagline, categories, maintainer's note.