Full-stack AI Agent developer building executable evals, tool-security boundaries, context systems, and human-gated workflows.
全栈 AI Agent 开发者,专注可执行评估、工具安全、上下文工程与人在回路工作流。
Not another resource list: an executable production-readiness gate for AI agents, maintained in English and Simplified Chinese.
不是另一份资源清单: 一套可执行、可审查、完整中英文同步的 AI Agent 生产就绪门禁。
| Current evidence / 当前证据 | Verified result / 已验证结果 |
|---|---|
| Release | v0.17.0 |
| Deterministic tests | 147/147 |
| Prompt-injection fixtures | 8/8 across reference, LangGraph.js, and OpenAI Agents SDK runtimes |
| Provenance | Separate producer + SHA-pinned verifier, GitHub OIDC/Sigstore attestation |
| Starter supply chain | Generated workflows pin actions/checkout to the reviewed v7.0.1 commit SHA |
| Machine contracts | Agent Card, Eval Result, fixture, and readiness-profile JSON Schemas |
| Try / 使用 | Integrate / 接入 | Inspect evidence / 查看证据 | 中文 |
|---|---|---|---|
| Web scorecard | Five-minute fail-closed setup | Attested eval provenance | 中文 README |
| OpenAI Agents SDK eval | Public JSON Schemas | Source-linked incidents | 中文 Schema 指南 |
- AI Content Workflow Skills
is a same-maintainer consumer. Its score gate passes at 12/20, while its
draft-onlyprofile and three launch blockers remain visibly failing. - PaperSage #149 is an independent Agent Card adoption PR under maintainer review. It is not counted as adoption unless merged.
- Agent evaluation and reliability / Agent 评估与可靠性
- MCP and tool security / MCP 与工具安全
- Human-in-the-loop workflow design / 人在回路的工作流设计
- Context engineering and multi-agent systems / 上下文工程与多智能体系统
