Overview
The Operations Agent handles the third phase of AI-DLC. It takes constructed features to production, verifies they work correctly, and sets up monitoring.Invocation
- Claude Code
- Cursor
- GitHub Copilot
Commands
build
Builds the project for deployment:1
Check Prerequisites
Verify all construction bolts are complete
2
Run Build
Execute build commands from tech stack
3
Run Tests
Execute full test suite
4
Create Artifacts
Package for deployment
Example Session
deploy
Deploys to a target environment:1
Environment Check
Verify target environment configuration
2
Pre-deployment
Run database migrations, cache warming
3
Deploy
Deploy using configured strategy
4
Health Check
Verify services are healthy
Deployment Strategies
The agent supports common strategies based on your infrastructure:verify
Runs verification after deployment:1
Smoke Tests
Run quick sanity checks
2
Health Endpoints
Check all service health endpoints
3
Integration Check
Verify integrations are working
4
Performance Baseline
Ensure response times are acceptable
Example Output
monitor
Sets up or checks monitoring:Logging
Structured logging configuration
Metrics
Key performance indicators
Alerts
Alert rules and thresholds
Dashboards
Visualization of system health
Key Metrics
The agent suggests monitoring:Human Checkpoints
The Operations Agent has 4 human checkpoints aligned with environment progression:Artifacts
Operations artifacts are stored in:Runbooks
The agent generates runbooks for common operations:Deployment Runbook
Deployment Runbook
Step-by-step deployment procedure:
- Pre-deployment checklist
- Deployment commands
- Verification steps
- Rollback procedure
Incident Response
Incident Response
What to do when things go wrong:
- Detection and triage
- Escalation matrix
- Communication template
- Post-mortem process
Scaling Runbook
Scaling Runbook
How to handle increased load:
- Signs of scaling needs
- Horizontal vs vertical
- Scaling commands
- Verification
Best Practices
Always Verify
Always Verify
Never skip verification after deployment. Automated checks catch issues humans miss.
Stage Environments
Stage Environments
Deploy to staging before production. Test the deployment process itself.
Monitor Proactively
Monitor Proactively
Set up alerts before you need them. Don’t wait for production issues.
Document Runbooks
Document Runbooks
Keep runbooks updated. They’re essential during incidents.
