Pre-release safety evaluation.
Dangerous-capability evaluations before release, followed by reproduction, diagnosis and retesting. Access and scope are agreed campaign by campaign, subject to capacity.
Launch evaluations are free to the lab.
Your model. Our evaluation compute.
At the lab
Target modelWeights stay here
Pacific Compute
Evaluation compute- Attacker models
- Judge models
- Parallel experiments
- Method development
For endpoint-based evaluations, your lab hosts the model and controls access. Unreleased weights stay on your side by default, with only trusted partnerships granting deeper access. Our compute runs the surrounding evaluation work.
Lab-controlled tenancy
Lab-controlled tenancy is designed to keep your engineers in control of the partition, model keys and approved users. We agree configuration, custody and security requirements directly with your lab before work begins.