Metaphi AI
About
Builds RL environments, agent evaluations, datasets, and long-horizon enterprise benchmarks.
What They Offer
Commercial or operational services this company provides.
Agent evaluations
Rubrics / verifiers
Custom RL environments
Training datasets
Products & Public Artifacts
EnterpriseSWE / COBOLBench / LH-Bench
Benchmark · Commercial
Enterprise benchmark suite spanning long-horizon software engineering, mainframe and COBOL maintenance, design, financial-document reasoning, and other real enterprise workflows.
Visit artifact →Technical Capabilities
Long-Horizon TasksProgrammatic VerifiersCode / Terminal ExecutionReal-App / Workplace SimulationTool / API / MCP Use
Only publicly documented capabilities are shown. Absence does not imply the capability is unavailable.
Focus Areas
CodingRL EnvironmentsTraining DataEvaluations
Domains
Software EngineeringEnterprise
Leadership
Abhishek Chandwani
Co-Founder