Production AI Institute · Public record
PAI Lab public benchmark

Public agent repositories, measured against visible PSF evidence.

PAI scans public GitHub metadata and file paths for signs of production AI discipline: evals, output schemas, observability, deployment gates, human oversight, security policy, and provider resilience. This is evidence coverage, not certification.

Repositories0
Eval evidence0
Human oversight0
Observability0
Evidence coverage table

Recently active public AI agent repositories

Projects are discovered through GitHub repository search, then scanned for visible PSF-aligned evidence in their public file tree. Higher coverage means more evidence was visible to the scanner, not that PAI has certified or endorsed the project.

GitHub public repository search
Repository
Coverage
Grade
Visible evidence
The live benchmark could not retrieve public repositories right now.

Where the repository list comes from

The benchmark uses GitHub's public repository search endpoint and rotates focused queries for AI agent, agentic AI, LLM agent, and MCP server repositories. The run de-duplicates repositories, excludes archived projects and forks when GitHub returns those flags, and sorts the published table by visible PSF evidence coverage.

  • topic:ai-agent archived:false fork:false stars:>=5
  • topic:agentic-ai archived:false fork:false stars:>=5
  • topic:llm-agent archived:false fork:false stars:>=5
  • topic:mcp-server archived:false fork:false stars:>=5
  • "ai agent" in:name,description,readme archived:false fork:false stars:>=5

How teams use this

The benchmark gives maintainers and production AI teams a concrete way to improve visible evidence. A project can publish the missing artifacts, run its own Agent Readiness report, and link to a stable monthly edition when citing broader ecosystem findings.

  • Use the live table to inspect current public evidence patterns.
  • Use immutable editions for citations, journalism, and longitudinal comparison.
  • Use the issue generator to turn a gap into a constructive maintainer task.
  • Use the evidence pack and control templates to publish the missing artifacts.
Evidence pack builderIssue generatorControl templates
Start here: production AI

Foundational reference pages for teams adopting, deploying, and operating production AI with inspectable controls.

What is production AI?AI agent production ready checklistAI adoption operating guideEnterprise AI deployment guideFive-year automation roadmapWorkflowOS open-source PSF studioPSF standard →