Skip to content
ai-supply.store
DiscoverCategoriesLeaderboardsCommunityAgent APIFAQ
Sign inSign up free
catalog / Gaming & Simulation / Gymnasium — Standard RL Environment API
◆SkillGaming & SimulationFree

Gymnasium — Standard RL Environment API

Farama Foundation's MIT-licensed reinforcement learning toolkit — the standard interface for RL environments including Atari, MuJoCo, CartPole, and 100+ more.

@ai-supply
Installs232k
⟳ upstream v1.3.0 · updated 3mo ago
↗ Source repository
← More Gaming & SimulationGaming & Simulation leaderboard →How we grade security →Source ↗
✓ Grade A · 100/100 · SafeSecurity assessment
✓No compromise signals15capabilities surfaced9of 20 OWASP controls clear
External endpoints declaredExternal endpoints declaredExternal endpoints declaredSuspicious network references
scanned 16d ago·osv · gitleaks · opengrep · picklescan + heuristics·full breakdown in the Security tab ↓

Gymnasium — Standard RL Environment API

Gymnasium (formerly OpenAI Gym) is the community standard API for reinforcement learning environments, maintained by the Farama Foundation. It defines a universal step/reset/render interface used by virtually every RL algorithm library, enabling agents to be trained across Atari 2600 games, physics simulations (MuJoCo, Box2D), robotics tasks, and custom game environments without code changes.

Key Features

  • 100+ built-in environments: Atari, MuJoCo, CartPole, LunarLander, BipedalWalker, and Toy Text
  • Standard gymnasium.Env interface adopted by Stable-Baselines3, CleanRL, RLlib, and all major RL frameworks
  • Wrappers for reward shaping, observation normalization, frame stacking, and time limits
  • Vector environments (gymnasium.vector) for parallel data collection on multi-core machines
  • Python 3.10–3.14, active maintenance and bug fixes

Quick Start

pip install gymnasium[classic-control]
import gymnasium as gym

env = gym.make("CartPole-v1", render_mode="human")
obs, info = env.reset(seed=42)
for _ in range(1000):
    action = env.action_space.sample()  # random policy
    obs, reward, terminated, truncated, info = env.step(action)
    if terminated or truncated:
        obs, info = env.reset()
env.close()
npx ai-supply add gymnasium-rl-environments

Curated mirror of the open-source Gymnasium (MIT). Get it from the source.

Rating rank
#1
of 13 in Gaming & Simulation
Install rank
#1
of 13 in Gaming & Simulation
Security score
100/100 · A
safe
Security rank
#1
of 13 in Gaming & Simulation
Installs
232k
cat avg 86k
This listing vs category average
Installs
this
cat avg
Security (of 100)
this
cat avg
Adoption trend
See the Gaming & Simulation leaderboard →
✓ Security: Safe · 100100/100 · grade Ascanned 16d ago
✓ no compromise signals15 risk-surface · 6/20 OWASP controls flagged

Compromise signals — malicious or tampered code (leaked secrets, backdoors, a dropped executable) — reduce the score, and known dependency CVEs carry a bounded penalty (they warrant review but never QUARANTINE — update the dependency to clear). Other dangerous-by-capability traits are risk surface, expected for some capabilities. Every finding is mapped to its OWASP control below.

What this capability can do · med confidence (static)
⚑ filesystem⚑ shell⚑ secrets
egress → www.reddit.com, discord.com, pre-commit.com, docs.github.com, help.github.com, packaging.python.org, docs.astral.sh, docs.pytest.org +32
30 scripts

Findings mapped to the OWASP Top 10 for LLM Applications (2025) and the OWASP Machine Learning Security Top 10. Expand any flagged control for the exact findings — compromise reduces the score; expected/risk-surface do not, except a known CVE, which carries a small bounded penalty (high/critical → Review).

OWASP Top 10 for LLM Applications
⚠LLM05Improper Output Handlinghigh
Code that pipes model/user output into shell, eval, SQL or paths unsafely.
•Suspicious code patterns — destructive rm -rf / · Farama-Foundation-Gymnasium-24c3a0d/bin/all-py.Dockerfile (CWE-78)risk surface
•Suspicious code patterns — OS command execution · Farama-Foundation-Gymnasium-24c3a0d/docs/_scripts/linkcheck.py (CWE-78)risk surface
•Suspicious code patterns — dynamic code execution · Farama-Foundation-Gymnasium-24c3a0d/docs/tutorials/training_agents/vector_a2c.py (CWE-95)risk surface
•Suspicious code patterns — pickle deserialization · Farama-Foundation-Gymnasium-24c3a0d/gymnasium/vector/utils/misc.py (CWE-502)risk surface
⚠LLM03Supply Chainmedium
Vulnerable/compromised dependencies, models or archives in the artifact.
•Dependency manifest — 10 pip requirements declared · Farama-Foundation-Gymnasium-24c3a0d/docs/requirements.txtrisk surface
•Non-registry dependency source — 1 requirement(s) from git/URL/editable · Farama-Foundation-Gymnasium-24c3a0d/docs/requirements.txt (CWE-829)risk surface
⚠LLM06Excessive Agencymedium
Over-broad tool/permission surface or unrestricted egress.
•External endpoints declared — 1 distinct host(s) · Farama-Foundation-Gymnasium-24c3a0d/.github/ISSUE_TEMPLATE/bug.ymlrisk surface
•External endpoints declared — 2 distinct host(s) · Farama-Foundation-Gymnasium-24c3a0d/.github/ISSUE_TEMPLATE/question.ymlrisk surface
•External endpoints declared — 3 distinct host(s) · Farama-Foundation-Gymnasium-24c3a0d/.github/workflows/pypi-publish.ymlrisk surface
•External endpoints declared — 5 distinct host(s) · Farama-Foundation-Gymnasium-24c3a0d/CITATION.cffrisk surface
•External endpoints declared — 4 distinct host(s) · Farama-Foundation-Gymnasium-24c3a0d/CONTRIBUTING.mdrisk surface
•External endpoints declared — 10 distinct host(s) · Farama-Foundation-Gymnasium-24c3a0d/README.mdrisk surface
•External endpoints declared — 32 distinct host(s) · Farama-Foundation-Gymnasium-24c3a0d/docs/environments/third_party_environments.mdrisk surface
•External endpoints declared — 9 distinct host(s) · Farama-Foundation-Gymnasium-24c3a0d/docs/tutorials/training_agents/frozenlake_q_learning.pyrisk surface
•External endpoints declared — 6 distinct host(s) · Farama-Foundation-Gymnasium-24c3a0d/gymnasium/envs/classic_control/acrobot.pyrisk surface
⚠LLM10Unbounded Consumptionmedium
Unbounded loops/recursion causing DoS or runaway cost.
Enforced at runtime by the gateway (rate limits + spend caps + size caps); static check flags unbounded loops.
•Potentially unbounded loop — an infinite loop (while True / while(1) / for(;;)) may cause runaway consumption · Farama-Foundation-Gymnasium-24c3a0d/gymnasium/envs/box2d/bipedal_walker.py (CWE-835)risk surface
§LLM09MisinformationGovernance
Artifacts designed to produce false/deceptive output.
Detectable only by runtime behavioral evaluation; addressed via responsible-use attestation.
✓LLM01Prompt InjectionPassed
✓LLM02Sensitive Information DisclosurePassed
✓LLM04Data and Model PoisoningPassed
Backdoors/poisoning in training data or serialized models.
Behavioral poisoning needs model execution; static check covers unsafe serialization + dataset skew only.
✓LLM07System Prompt LeakagePassed
✓LLM08Vector and Embedding WeaknessesPassed
PII or plaintext source leakage in embedding/vector exports.
Embedding inversion/poisoning is largely runtime; static check covers PII in vector exports.
OWASP Machine Learning Security Top 10
⚠ML09Output Integrityhigh
Middleware tampering with model outputs in transit.
Gateway enforces TLS + response integrity; static check flags output-rewriting code.
•Suspicious code patterns — destructive rm -rf / · Farama-Foundation-Gymnasium-24c3a0d/bin/all-py.Dockerfile (CWE-78)risk surface
•Suspicious code patterns — OS command execution · Farama-Foundation-Gymnasium-24c3a0d/docs/_scripts/linkcheck.py (CWE-78)risk surface
•Suspicious code patterns — dynamic code execution · Farama-Foundation-Gymnasium-24c3a0d/docs/tutorials/training_agents/vector_a2c.py (CWE-95)risk surface
•Suspicious code patterns — pickle deserialization · Farama-Foundation-Gymnasium-24c3a0d/gymnasium/vector/utils/misc.py (CWE-502)risk surface
⚠ML06AI Supply Chainmedium
Compromised PyPI/npm packages, typosquats, unsafe serialized models.
•Dependency manifest — 10 pip requirements declared · Farama-Foundation-Gymnasium-24c3a0d/docs/requirements.txtrisk surface
•Non-registry dependency source — 1 requirement(s) from git/URL/editable · Farama-Foundation-Gymnasium-24c3a0d/docs/requirements.txt (CWE-829)risk surface
§ML01Input Manipulation (Adversarial)Governance
Models vulnerable to adversarial perturbations.
Requires runtime robustness evaluation; addressed via publisher robustness attestation.
§ML03Model InversionGovernance
Training data reconstructable from a model's outputs.
Runtime/evaluation property; addressed via model-card data-provenance + DP attestation.
§ML04Membership InferenceGovernance
Determining whether a record was in the training set.
Runtime/evaluation property; addressed via overfitting disclosure + DP attestation.
§ML08Model SkewingGovernance
Models trained on skewed data producing biased output.
Requires fairness evaluation; addressed via model-card bias/limitations disclosure.
✓ML02Data PoisoningPassed
Poisoned training datasets with triggers or anomalous distributions.
Static check covers trigger phrasing, PII and label skew; full poisoning detection is runtime.
✓ML05Model TheftPassed
Unlicensed re-distribution / license-incompatible derivatives.
Static check verifies license declaration; extraction throttling is runtime.
✓ML07Transfer Learning AttackPassed
Backdoored base models / LoRA adapters propagating to derivatives.
Backdoor detection needs behavioral probing; static check covers unsafe serialization + provenance.
✓ML10Model Poisoning (Weights)Passed
Tampered model weight files; integrity must be verifiable.
Static check enforces safe formats + records a content hash for downstream verification.
Other findings (6) · hygiene / uncategorized
•Unrecognized file type — '.gitignore' is not on the allowlist · Farama-Foundation-Gymnasium-24c3a0d/.gitignorerisk surface
•Unrecognized file type — '.cff' is not on the allowlist · Farama-Foundation-Gymnasium-24c3a0d/CITATION.cffrisk surface
•Suspicious network references — URL shortener (14 URLs) · Farama-Foundation-Gymnasium-24c3a0d/CITATION.cffrisk surface
•Unrecognized file type — '.?' is not on the allowlist · Farama-Foundation-Gymnasium-24c3a0d/LICENSErisk surface
•Unrecognized file type — '.dockerfile' is not on the allowlist · Farama-Foundation-Gymnasium-24c3a0d/bin/all-py.Dockerfilerisk surface
•Disallowed file type — '.bat' executables are not permitted · Farama-Foundation-Gymnasium-24c3a0d/docs/make.bat (CWE-434)risk surface
✔ verified source · pinned Farama-Foundation-Gymnasium-24c3a0d
Check against a policy

The same gate an agent runs before installing (POST /api/v1/trust/gymnasium-rl-environments/check). Click a policy:

Consume Gymnasium — Standard RL Environment API programmatically. Authenticate with an API key or session — see Authorize an agent.

# Agents: CHECK BEFORE YOU INSTALL (no auth) — score, grade, level, capability manifest
curl https://ai-supply.store/api/v1/trust/gymnasium-rl-environments

# Gate against your org policy (returns { pass, violations })
curl -X POST https://ai-supply.store/api/v1/trust/gymnasium-rl-environments/check \
  -H "Content-Type: application/json" \
  -d '{"minGrade":"B","denyPermissions":["shell"],"denyUnknownEgress":true}'

# CLI
npx ai-supply add gymnasium-rl-environments

# REST (install → download)
curl -X POST https://ai-supply.store/api/v1/listings/gymnasium-rl-environments/install \
  -H "Authorization: Bearer $AIM_KEY"

# MCP tool
install_listing({ "slug": "gymnasium-rl-environments" })
OpenAPI spec →
vlatest
✓ Security: Safe · 1001mo ago

Curated mirror — latest upstream source. See the repository for tagged releases.

Sign in and install this listing to leave a review.

More from @ai-supply

View profile →
◉Agent
MetaGPT
Multi-agent framework that assigns GPT roles (PM, engineer, QA) to solve complex software tasks end-to-end.
↓ 1.0M
⇄Connector
vLLM
High-throughput, memory-efficient LLM inference engine with PagedAttention and continuous batching.
↓ 892k
⇄Connector
Meilisearch
Lightning-fast open-source search engine with typo-tolerance, semantic hybrid search, and sub-50ms response times.
↓ 811k
△Eval
Weights & Biases (wandb)
ML experiment tracking and visualization — log metrics, hyperparameters, models, and media in real time.
↓ 784k
ai-supply.store

Free, security-vetted AI capabilities — skills, MCPs, plugins, agents, datasets and more, each graded and freshness-tracked, and built for humans and agents alike.

api · v3.1status · all green
Contact
support@ai-supply.storesecurity@ai-supply.store
Catalog
  • Discover
  • Categories
  • Leaderboards
  • Benchmarks
  • Security
  • Scan a repo
Community
  • Community
  • FAQ
For agents
  • Quickstart (60s)
  • Authorize an agent
  • Agent API
  • OpenAPI spec
For builders
  • Publish
  • Dashboard
Account
  • Create account
  • Sign in
  • Settings
Legal
  • Terms
  • Publisher Agreement
  • Acceptable Use
  • Privacy