AI & ML interests

RL Environments at Scale

Recent Activity

AdithyaSKย  updated a bucket 7 minutes ago
FineEnvs/fleurs-bucket
AdithyaSKย  published a bucket 8 minutes ago
FineEnvs/fleurs-bucket
AdithyaSKย  updated a Space about 1 hour ago
FineEnvs/nayana-ocr-env
View all activity

FineEnvs 's collections 6

Repo2RLEnv โ€” Verifiable RL Environments
Repo2RLEnv coding and terminal RL environments in Harbor format. Datasets include per-task quality labels, provenance and generation economics.
Data Agent
Deterministic data-analysis agent tasks from the jupyter-agent dataset โ€” verified answers, no LLM judge. Harbor env suites, plain dataset & SFT.
LaTeX OCR
LaTeX OCR environment, model, dataset, and five-run training comparison across four models, including unstable and stabilized Gemma.
Paint with Code
An RL environment where the agent paints by writing p5.brush sketches, rewarded by an aesthetic preference model looking at the render.
FineEnvs Academy
A curated collection of articles, guides, tutorials, slides, and resources for learning how to build, train, and evaluate RL environments for Agents
Repo2RLEnv โ€” Verifiable RL Environments
Repo2RLEnv coding and terminal RL environments in Harbor format. Datasets include per-task quality labels, provenance and generation economics.
LaTeX OCR
LaTeX OCR environment, model, dataset, and five-run training comparison across four models, including unstable and stabilized Gemma.
Paint with Code
An RL environment where the agent paints by writing p5.brush sketches, rewarded by an aesthetic preference model looking at the render.
Data Agent
Deterministic data-analysis agent tasks from the jupyter-agent dataset โ€” verified answers, no LLM judge. Harbor env suites, plain dataset & SFT.
FineEnvs Academy
A curated collection of articles, guides, tutorials, slides, and resources for learning how to build, train, and evaluate RL environments for Agents