BIFROST

Accelerate Robotics Research

An evaluation platform that delivers insights in 30 mins

Trusted by teams building mission critical physical AI systems

NASA
Saronic
Shield AI
Havoc
Honda
NTT Data
Privateer
ST Engineering

Massively Parallelized GPU Evaluation

Run thousands of GPU-accelerated simulations in parallel and get your results in minutes, not overnight.

TEST EVERYTHING
Every model, checkpoint and architecture, evaluated over lunch.
NO SETUP
Skip the GPU orchestration, simulator config, and policy wiring.
SHARDED & VECTORIZED
Simulation optimised at the engine, machine, and cluster level.

Atomic Subgoals & Custom Rubrics

Go past pass-fail metrics and see exactly where your models break down, and where they hold up. Define your own rubrics to grade rollout quality.

ATOMIC SUBGOALS
Every task decomposes into granular subgoals, scored one by one.
CUSTOM RUBRICS
Encode what quality means for your task and grade every rollout against it.
ROLLOUT QA
Catch degraded behavior that aggregate success rates hide.

Agentic Failure Analysis

Agents watch the rollouts, cluster the failures and rank them by impact, so you're not scrubbing through hundreds of videos to find the pattern yourself.

AUTOMATIC CLUSTERING
Failures grouped by scenario, objects, sensor, lighting, and trajectory.
RANKED BY IMPACT
See which modes cost you performance, not which happened most often.
NATURAL LANGUAGE QUERY
Query your failures, instead of manually searching through videos.
TASK COVERAGE

Benchmarks & Environments

Works with any simulator, benchmark or world model. We work with industrials to codify real world work into custom simulation tasks along with nuanced success rubrics.

ENVIRONMENT LIBRARY Custom benchmark
ANY BENCHMARK · ANY SIMULATOR · ANY WORLD MODEL
ACADEMIC
LIBERO
LIBERO-Plus
RoboLab
RoboCasa
RoboMimic
CALVIN
SIMPLER
RoboTwin 2.0
RoboMemory
RLBench
INDUSTRIAL
Tend CNC Machine
Navigate Desert
Clean Kitchen
Load Dump Truck
Search Mine
Land Aircraft
Wire Ethernet Cables
Inspect Jet Engine
Assemble Parts
Mow Lawn
SIMULATORS
NVIDIA Isaac Sim
Unreal Engine
MuJoCo
ManiSkill
Genesis

Make Your Research Loop 100x Faster

ABOUT THE TEAM

We're AI and robotics researchers, graphics engineers and infrastructure engineers. We've spent the last five years building our own physical AI systems and helping some of the world's largest companies deploy theirs. Now we're making that far easier for everyone else.

BACKED BY
Sequoia
Lux Capital
Airbus Ventures
Wavemaker
Carbide
Techstars
FAQ

What is Manifold?

Manifold is a robot policy evaluation platform. One harness runs your policy on every major simulator benchmark, including LIBERO, RoboCasa and your own scenarios, sharded across GPUs so results land in minutes instead of days.

Which models and benchmarks does Manifold support?

Manifold evaluates VLA and embodied AI policies, including OpenVLA, GR00T, pi-0 and Octo class models, on LIBERO, RoboCasa, RoboMimic and CALVIN, with more benchmarks and embodiments on the way.