Production project

Verifiable Autonomous AI Agent Project

Design and verify an autonomous treasury management AI agent that monitors oracle metrics, runs policy engine checks, simulates transactions, requests operator approval above spending thresholds, and executes non-custodial rebalances via an ERC-4337/EIP-7702 smart account with ZK execution verification.

advanced
16 hours3 milestones
A production AI project moving from architecture decisions through evaluation into monitored deployment.

Project evidence

0/3 milestones · rubric 0% · pass 70%

evidence required
checklist required
critical criteria open
in progress

Milestone 1: Policy Engine & Safety Gates

architecture decision

A policy engine specification defining contract allowlists, spending limits, and human gate triggers.

Acceptance criteria

  • Policy engine intercepts tool calls before transaction generation.
  • Spending limits and contract allowlists are explicitly checked.

Use: context · drivers · decision · alternatives · consequences · risks · validation · rollback · evidence

Milestone 2: ERC-4337 Smart Account Integration

implementation

An ERC-4337 UserOperation builder and Paymaster gas sponsorship configuration.

Acceptance criteria

  • Transactions execute via UserOperations through EntryPoint contract.
  • Session keys constrain sub-agent transaction permissions.

Use: context · drivers · decision · alternatives · consequences · risks · validation · rollback · evidence

Milestone 3: ZK Proof Verification & Finality

evaluation

An on-chain ZK verifier integration validating model execution receipts before settlement.

Acceptance criteria

  • zkVM proof receipts are verified by smart contract before payout.
  • Model weight hashes match the on-chain provenance registry.

Use: context · drivers · decision · alternatives · consequences · risks · validation · rollback · evidence

Criterion-level self-assessment

Project rubric

Architecture coherence

Architecture coherence for AI Architecture Decision Workbook is explicit, testable, and connected to submitted evidence. · weight 20%

0/4

Current level: Missing or contradicted by the submitted evidence.

Implementation evidence

Implementation evidence for AI Architecture Decision Workbook is explicit, testable, and connected to submitted evidence. · weight 15%

0/4

Current level: Missing or contradicted by the submitted evidence.

Evaluation quality

Evaluation quality for AI Architecture Decision Workbook is explicit, testable, and connected to submitted evidence. · weight 20%

0/4

Current level: Missing or contradicted by the submitted evidence.

Operational readiness

Operational readiness for AI Architecture Decision Workbook is explicit, testable, and connected to submitted evidence. · weight 15%

0/4

Current level: Missing or contradicted by the submitted evidence.

Security and privacy

Security and privacy for AI Architecture Decision Workbook is explicit, testable, and connected to submitted evidence. · weight 20%

0/4

Current level: Missing or contradicted by the submitted evidence.

Decision communication

Decision communication for AI Architecture Decision Workbook is explicit, testable, and connected to submitted evidence. · weight 10%

0/4

Current level: Missing or contradicted by the submitted evidence.

Portable evidence

Attach project records

Text, HTTPS links, repository paths, and imported Markdown/JSON/text are stored locally and included in exports.

Project review checklist

Unresolved risks

One risk per line. A passing score does not erase open risk.

Connected concepts

Reference architectures

Back to projects