catalog / Legal & Compliance / Juriscraper — Court Data Scraper
ConnectorLegal & ComplianceFree

Juriscraper — Court Data Scraper

A caching, scraping library that collects opinions, oral arguments, and PACER filings from hundreds of American court websites.

Installations6.9k
⟳ upstream 1.4.15 · updated 9y ago
Dépôt source
Grade A · 100/100 · SafeSecurity assessment
No compromise signals8capabilities surfaced11of 20 OWASP controls clear
External endpoints declaredBroad capability surfaceExternal endpoints declaredSuspicious code patterns
scanned 1mo agoosv · gitleaks · opengrep · picklescan + heuristicsfull breakdown in the Security tab ↓

Juriscraper

Juriscraper is a Python library, maintained by the Free Law Project, that scrapes metadata and documents from American federal and state court websites. It powers the ingestion pipeline behind CourtListener, standardizing wildly different court sites into a consistent interface for opinions, oral-argument audio, and PACER (federal filing) data.

Key features

  • Scrapers for hundreds of state and federal appellate and trial courts
  • Unified output for opinions, oral arguments, and PACER dockets/documents
  • Built-in politeness: caching, rate awareness, and change detection
  • Extensible base classes make adding new court scrapers straightforward
  • Powers one of the largest open archives of U.S. court data

Each court is exposed as a module you invoke to fetch the latest cases, returning normalized records (case name, date, citation, download URL) ready for storage or analysis.

Curated mirror of the open-source Juriscraper (BSD-2-Clause). Get it from the source.

More from @ai-supply

View profile →
Agent
MetaGPT
Multi-agent framework that assigns GPT roles (PM, engineer, QA) to solve complex software tasks end-to-end.
1.0M
Connector
vLLM
High-throughput, memory-efficient LLM inference engine with PagedAttention and continuous batching.
892k
Connector
Meilisearch
Lightning-fast open-source search engine with typo-tolerance, semantic hybrid search, and sub-50ms response times.
811k
Eval
Weights & Biases (wandb)
ML experiment tracking and visualization — log metrics, hyperparameters, models, and media in real time.
784k