← Blog index
2026-07-31·Custom Agents

Custom AI Agent Delivery Checklist: Artifacts from Brief to Go-Live

Custom AI Agent projects rarely fail because the model is “not smart enough.” They fail because deliverables were vague—no intent table, no refuse rules, no acceptance metrics until go-live week. Buyers ask about models and price; if the SOW only says “build a support Agent,” both sides will disagree on what “done” means. GeonAI delivers acceptable artifacts; /agents are capability references only (not a public trial of 363+ presets). Email [email protected]. Pilot metrics: /blog/poc-to-production-agent-checklist. Delivery rhythm: /blog/enterprise-ai-agent-delivery-4-steps.

Why artifacts beat “which model”

Models are swappable; boundaries and acceptance are not handshake deals. Freeze the artifact list before routing and model shopping—and cut change orders in half.

Phase 1 — Discovery (kickoff week)

ArtifactPurposeIf missing
One-pager scenarioUsers, entry, successScope explodes
Intent table v0Top intents and priorityCannot design KB/tools
Boundary tables draft (/blog/custom-agent-scope-boundary-in-contract)Do / don’t / human-confirmPost-launch disputes
Systems & data listKB sources, CRM/tickets, ACLIntegration slips

Phase 2 — Design

Phase 3 — Build and joint test

ArtifactNotes
Environment notesTest/staging URLs, accounts, config
Ingestion reportDoc counts, parse failures, version id
API joint-test logCases, error codes, idempotency
Security checklistNo secrets in browser; no universal API tool

Phase 4 — Pilot acceptance

Copy missing-citation, ACL-leak, missed-handoff, and wrong-promise rates into a signed acceptance sheet. Preset pages and demos are not acceptance criteria.

Phase 5 — Go-live handover

  1. Ops runbook (incident, degrade, rollback)
  2. Knowledge release process (who edits, who approves, how to roll back)
  3. Weekly metrics template
  4. Training notes (business owner + frontline)
  5. Residual risks and iteration backlog

Minimum SOW annex pack

Intent table, boundary tables, tool allowlist, eval set, acceptance metrics, runbook, knowledge release process—if three or more are missing, do not claim “production delivered.”

How to brief GeonAI

Share the priority scenario, systems in play, and acceptance metrics you want in the contract. Email [email protected], /pricing, or Live chat. We deliver acceptable artifacts—not a vague “build an Agent.”

Frequently asked questions

Can we launch without an eval set?

Gray release only—do not sign production acceptance. Without gold items you cannot prove better or worse.

Are the 363+ presets deliverables?

No. They illustrate capability types, not your intents, ACL, or tools.

Do we get model weights?

Most projects deliver configuration, knowledge, gateway, and process; weights follow the contract. Private deploy is separate.

Who keeps these documents?

Prefer the customer KB plus contract annexes; GeonAI can host working copies.

How does this relate to the 4-step delivery playbook?

The playbook is the rhythm; this checklist is the file set per phase. Put both in the SOW.

Does a demo environment mean delivery is done?

No. Done means a signed acceptance sheet and handover pack—not one show-and-tell.

custom AI Agentdelivery checklistSOWacceptanceGeonAI