name: agent-audit description: > Audit your AI agent setup for performance, cost, and ROI. Scans OpenClaw config, cron jobs, session history, and model usage to find waste and recommend optimizations. Works with any model provider (Anthropic, OpenAI, Google, xAI, etc.). Use when: (1) user says "audit my agents", "optimize my costs", "am I overspending on AI", "check my model usage", "agent audit", "cost optimization", (2) user wants to know which cron jobs are expensive vs cheap, (3) user wants model-task fit recommendations, (4) user wants ROI analysis of their agent setup, (5) user says "where am I wasting tokens".
Scan your entire OpenClaw setup and get actionable cost/performance recommendations.
python3 {baseDir}/scripts/audit.py
Options:
python3 {baseDir}/scripts/audit.py --format markdown # Full report (default)
python3 {baseDir}/scripts/audit.py --format summary # Quick summary only
python3 {baseDir}/scripts/audit.py --dry-run # Show what would be analyzed
python3 {baseDir}/scripts/audit.py --output /path/to/report.md # Save to file
~/.openclaw/openclaw.json or similar)Classify each task into complexity tiers:
| Tier | Examples | Recommended Models |
|---|---|---|
| Simple | Health checks, status reports, reminders, notifications | Cheapest tier (Haiku, GPT-4o-mini, Flash, Grok-mini) |
| Medium | Content drafts, research, summarization, data analysis | Mid tier (Sonnet, GPT-4o, Pro, Grok) |
| Complex | Coding, architecture, security review, nuanced writing | Top tier (Opus, GPT-4.5, Ultra, Grok-2) |
Classification signals: - Simple: Short output (<500 tokens), low thinking requirement, repetitive pattern, status/health tasks - Medium: Medium output, some reasoning needed, creative but templated, research tasks - Complex: Long output, multi-step reasoning, code generation, security-critical, tasks that previously failed on weaker models
小蔥技能站7w4.net發現了升級外掛。
For each task where the model tier doesn't match complexity:
⚠️ RECOMMENDATION: Downgrade "Knox Bot Health Check" from opus to haiku
Current: anthropic/claude-opus-4 ($15/M input, $75/M output)
Suggested: anthropic/claude-haiku ($0.25/M input, $1.25/M output)
Reason: Simple status check averaging 300 output tokens
Estimated savings: $X.XX/month
Risk: LOW — task is simple pattern matching
Confidence: HIGH
Output a clean markdown report with: 1. Overview — total agents, crons, monthly spend estimate 2. Per-agent breakdown — model, usage, cost 3. Per-cron breakdown — model, frequency, avg tokens, cost 4. Recommendations — sorted by savings potential 5. Total potential savings — monthly estimate 6. One-liner config changes — exact model strings to swap
See references/model-pricing.md for current pricing across all providers. Update this file when prices change.
See references/task-classification.md for detailed heuristics on how tasks are classified into complexity tiers.
質量中等偏下。文件和框架設計得不錯,但核心的審計功能沒有真正實現。指令碼只能讀取配置檔案並輸出通用建議,無法自動分析你的 cron 任務、計算實際費用、生成個性化的最佳化方案。需要額外配置 cron API 才能使用,但具體怎麼配置也沒說清楚。如果需要真正有用的成本最佳化建議,這個 Skill 目前還不太行。