Define the boundary
State what a system is intended to do, which failure classes matter and who owns the release decision.
Safety
Capvise Labs does not operate a frontier model today. This page sets the standard we intend to apply as our systems, access and capability grow.
Position
Under-refusal can enable harm. Over-refusal can make ordinary work unreliable and drive users toward less controlled alternatives. Both outcomes deserve measurement.
Our first public safety project, VetoBench, applies this view to generative video. It will count benign failures beside appropriate refusals rather than treating any blocked request as a safety success.
See VetoBenchCommitments
State what a system is intended to do, which failure classes matter and who owns the release decision.
Keep model versions, settings, access source, retries and material human intervention with the result.
Measure harmful behaviour that should be blocked and useful behaviour that safety systems should allow.
When evidence weakens, narrow the release or revise the public statement instead of defending old copy.
Impact
Evidence
Release test
What does the system complete under controlled and adversarial conditions?
Does the behaviour repeat across time, accounts, settings and near-identical tasks?
Can an operator understand, constrain, interrupt and recover the system?
Who pays when the system fails, and how quickly can the damage be reversed?