RL.VENDORS
ZE

Zero Proof Labs


About

Zero Proof Labs develops While and the Zero Proof simulation stack for evaluating and post-training tool-using agents. The platform builds simulated worlds from an agent's tools and policy, generates graded trajectories for evals, SFT, and RL, measures behavior with programmatic markers, and can launch hosted SFT, GRPO, DPO, or reward-model training. The same workflow connects production traces, training datasets, holdouts, model versions, and hosted inference so improvements can be measured before and after training.

What They Offer

Commercial or operational services this company provides.

Agent evaluations
Rubrics / verifiers
Open-source tooling
Training datasets
Agent trajectories
Managed infrastructure

Products & Public Artifacts

Zero Proof Simulations
Framework · Open source

Open simulation framework that derives tool worlds from agent specifications and generates graded trajectories for evals, SFT, and RL datasets.

Visit artifact →
While Hosted Training
Tool · Commercial

Hosted post-training and serving for SFT, GRPO, DPO, and reward models with versioned model endpoints.

Visit artifact →
whileai SDK
Framework · Open source

Open-source SDK for agent simulations, grading, evaluation, training-data generation, and post-training workflows.

Visit artifact →

Technical Capabilities

Tool / API / MCP UseProgrammatic VerifiersStateful / Persistent EnvironmentsReal-App / Workplace Simulation

Only publicly documented capabilities are shown. Absence does not imply the capability is unavailable.

Focus Areas

RL EnvironmentsTraining DataRLHF / Post-trainingEvaluations

Domains

Enterprise

Leadership

Jacob Weiss
CEO & Co-Founder
Viet Le
Head of Cryptography
Azriel Chelst
VP of Commercial
Sahana Dhar
AI Research Engineer