Score exposed risk
Map the gaps to the controls they implicate.
When the Index shows gaps in a live system, the PSF names the controls that close them.
Read the PSF →Generate a readiness report for an agentic system: answer the controls, add an optional public GitHub evidence scan, see the gaps, and publish a clearly scoped readiness badge.
Paste any public GitHub repository into the form, or use private repo mode with a local `git ls-files` manifest. The scanner reads path evidence only: evals, schemas, tracing, runbooks, approval gates, security policy, and fallback paths.
A lightweight discovery feed from public GitHub repositories tagged for AI agents. Load one, complete the checks, generate a PSF-aligned report, or view the public evidence benchmark PAI runs across recently active projects.
Teams do not need PAI approval to use the Index. They can run the report, fix gaps, publish evidence, and show how their agent maps to the standard.
View public benchmarkBuild PSF evidence packOpen control templatesThe methodology is deliberately conservative: equal PSF domain weighting, transparent answer values, explicit evidence grading, and a clear line between readiness signals and assurance claims.
Input boundary, output validation, data stewardship, observability, deployment control, human oversight, security, and ecosystem resilience.
Three focused checks per domain, scored as evidence exists, partial, not yet, or not applicable.
Public GitHub scans look for file-path evidence such as evals, schemas, runbooks, approvals, security policy, and fallbacks.
The check generates a readiness report. It is not a PAI credential, endorsement, or safety guarantee.
The readiness check is a credible starting point because it separates claims from evidence. The next step depends on whether the report surfaced deployment risk, a client opportunity, a proof package, or a team-wide rollout need.
Score exposed risk
When the Index shows gaps in a live system, the PSF names the controls that close them.
Read the PSF →Follow the benchmark
The public benchmark applies the same Index to well-known agent frameworks and publishes the results as records.
Open the benchmark →Portfolio proof
Use the evidence pack when a team needs artifacts that survive procurement, audit, or internal review.
Build evidence pack →Team rollout
Adopt the PSF as the shared readiness vocabulary for policy, training, and operating cadence.
Read the PSF →This record is maintained by PAI and free to cite. If something is wrong or missing, tell us. Corrections and source suggestions keep the record honest.
Track what changed, read the weekly brief, and follow the public evidence record - the operator loop for production AI risk.