catalog / Legal & Compliance / LexGLUE — Legal Language Understanding Benchmark
EvalLegal & ComplianceFree

LexGLUE — Legal Language Understanding Benchmark

Multi-task benchmark for legal NLP with 7 datasets covering EURLEX classification, contract clause labeling, court judgement prediction, and more.

Установки31k
⟳ upstream main@419a49d · updated 1y ago
Исходный репозиторий
! Grade B · 88/100 · ReviewSecurity assessment
No compromise signals2capabilities surfaced1known CVE10of 20 OWASP controls clear
External endpoints declaredExternal endpoints declaredVulnerable dependencies
scanned 18d agoosv · gitleaks · opengrep · picklescan + heuristicsfull breakdown in the Security tab ↓

LexGLUE — Legal Language Understanding Benchmark

LexGLUE is the legal analogue of GLUE/SuperGLUE — a comprehensive benchmark spanning seven legal NLP datasets and tasks. It standardizes evaluation across EURLEX (EU legislation classification), ECHR (court judgement prediction), LEDGAR (contract provision classification), SCOTUS (US Supreme Court decision area), ECtHR (article violation prediction), ContractNLI (contract NLI), and CaseHOLD (legal holding identification).

Key Features

  • 7 legal NLP tasks in a single evaluation harness
  • Covers EU and US jurisdictions across legislation, contracts, and case law
  • HuggingFace Datasets integration for easy loading
  • Leaderboard tracking state-of-the-art Legal-BERT, RoBERTa-legal, and other models
  • CC-BY-4.0 dataset license with public reproducibility

Quick Start

from datasets import load_dataset

# Load EURLEX classification task
dataset = load_dataset("coastalcph/lex_glue", "eurlex")
print(dataset["train"][0]["text"][:200])
print(dataset["train"][0]["labels"])  # Multi-label list

# Load ECHR court judgement prediction
scotus = load_dataset("coastalcph/lex_glue", "scotus")
print(scotus["test"][0])
npx ai-supply add lex-glue-legal-benchmark

Curated mirror of the open-source LexGLUE (CC-BY-4.0). Get it from the source.

More from @ai-supply

View profile →
Agent
MetaGPT
Multi-agent framework that assigns GPT roles (PM, engineer, QA) to solve complex software tasks end-to-end.
1.0M
Connector
vLLM
High-throughput, memory-efficient LLM inference engine with PagedAttention and continuous batching.
892k
Connector
Meilisearch
Lightning-fast open-source search engine with typo-tolerance, semantic hybrid search, and sub-50ms response times.
811k
Eval
Weights & Biases (wandb)
ML experiment tracking and visualization — log metrics, hyperparameters, models, and media in real time.
784k