Convex Markets / Datasets / All departments

Dataset catalog

Filter by type, access, and pricing. Specs show before you open the product page.

110 results
P(
External · RLE-1407

PettingZoo (Farama)

The standard Python API and reference environment suite for multi-agent RL, covering Atari (multiplayer), Classic board/card games, MPE, SISL, and Butterfly.

Type RL environmentsVolume dozens reference environmentsFormat Python (pip: pettingzoo; AEC / Parallel multi-agent API)Access PUBLIC LICENSE
Free
open-source license
View
PB
External · RLE-1505

Procgen Benchmark

OpenAI's Procgen Benchmark is a suite of 16 procedurally-generated, Atari-like Gym/Gym3 reinforcement-learning environments built to measure sample efficiency and generalization in RL.

Type RL environmentsVolume 16 procedurally-generated game environments (each with up to 2^31 unique levels)Format Gym / gym3 RL environment installed as a Python package (pip install procgen); observations are (64,64,3) uint8 NumPy RGB arraysAccess PUBLIC LICENSE
Free
open-source software (MIT license)
View
B
External · RLE-1304

BabyAI

Grid-world instruction-following platform with a compositional synthetic 'Baby Language'; 19 levels of increasing difficulty, now part of Minigrid.

Type RL environmentsVolume 19 levelsFormat Python (Minigrid / Gym API)Access PUBLIC LICENSE
Free
open-source license
View
J
External · RLE-1508

Jericho

Open-source Python RL environment from Microsoft Research that connects agents to 57 human-made interactive-fiction text games via a modified Frotz Z-machine interpreter, with reward defined as in-game score deltas.

Type RL environmentsVolume 57 supported interactive-fiction games (Z-machine)Format Z-machine story files (.z3/.z5/.z6/.z8) played through the Python FrotzEnv wrapper; observations and actions are plain textAccess PUBLIC LICENSE
Free
open-source license
View
O
External · RLE-1507

Overcooked-AI

A two-agent gridworld cooking environment for benchmarking human-AI and multi-agent coordination, where agents cooperatively prepare and deliver soups for a shared sparse reward.

Type RL environmentsVolume 49 layout files (kitchen configurations; 5 are the canonical benchmark layouts)Format Python simulator (installable overcooked_ai_py package with Gym/PettingZoo-style envs) plus .layout config files (Python-dict text)Access PUBLIC LICENSE
Free
open-source license
View
D
External · RQA-5007

DROP

A crowdsourced reading-comprehension benchmark whose questions require discrete reasoning (addition, counting, sorting, comparison) over Wikipedia-derived paragraphs.

Type Reasoning QAVolume 86,935 questions (77,400 train + 9,535 validation)Format Parquet (Hugging Face); originally distributed as JSONAccess PUBLIC LICENSE
Free
open dataset (HF)
View
H
External · RQA-5006

HotpotQA

A Wikipedia-based multi-hop question-answering dataset of 113,000+ QA pairs that require reasoning across multiple supporting documents and provide sentence-level supporting facts.

Type Reasoning QAVolume 113,000+ multi-hop question-answer pairsFormat Parquet (Hugging Face); original distribution JSONAccess PUBLIC LICENSE
Free
open dataset (HF)
View
T
External · RLE-1302

TextWorld

Microsoft's sandbox engine for procedurally generating and playing text-adventure games to train and evaluate RL agents; underlies ALFWorld.

Type RL environmentsVolume Unlimited procedurally-generated gamesFormat Python (pip: textworld) + generated game filesAccess PUBLIC LICENSE
Free
open-source license
View
O
External · SFT-2201

OpenR1-Math-220k

220k competition math problems, each with 2-4 DeepSeek-R1 chain-of-thought traces verified by Math-Verify and Llama-3.3-70B; text reasoning, no code.

Type SFT datasetVolume 220k problemsFormat Parquet (Hugging Face)Access PUBLIC LICENSE
Free
open-source license
View
N
External · SFT-2202

NuminaMath-1.5

~896k competition-math problems (Chinese high-school to IMO level) with chain-of-thought solutions; improved successor to NuminaMath-CoT.

Type SFT datasetVolume 896k problemsFormat Parquet (Hugging Face)Access PUBLIC LICENSE
Free
open-source license
View