EstimateSAFE
Turn Claude Code into a full game dev studio — 49 AI agents, 72 workflow skills, and a complete coordination system mirroring real studio hierarchy.
Overview
Turn Claude Code into a full game dev studio — 49 AI agents, 72 workflow skills, and a complete coordination system mirroring real studio hierarchy.
42a36917b8beOBSERVED · 2026-10-05What it tells the agent
The instruction file, verbatim from the audited commit — this is the text the model reads, and the surface the audit's instruction layer examines. Quoted here so you can judge it without cloning anything.
--- name: estimate description: "Estimate task effort from complexity, dependencies, velocity, risk. Structured estimate with confidence levels." argument-hint: "[task-description]" user-invocable: true allowed-tools: Read, Glob, Grep model: sonnet --- ## Phase 1: Understand the Task Read the task description from the argument. If the description is too vague to estimate meaningfully, ask for clarification before proceeding. Read CLAUDE.md for project context: tech stack, coding standards, architectural patterns, and any estimation guidelines. Read relevant design documents from `design/gdd/` if the task relates to a documented feature or system. --- ## Phase 2: Scan Affected Code Identify files and modules that would need to change: - Assess complexity (size, dependency count, cyclomatic complexity) - Identify integration points with other systems - Check for existing test coverage in the affected areas - Read past sprint data from `production/sprints/` for similar completed tasks and historical velocity. **If there is none**, say so under Notes and Assumptions — "No sprint history found; this estimate is not calibrated to the team's velocity" — and treat that as a risk in the confidence level, not as a silent padding of the figures. --- ## Phase 3: Analyze Complexity Factors **Code Complexity:** - Lines of code in affected files - Number of dependencies and coupling level - Whether this touches core/engine code vs leaf/feature code - Whether existing patterns can be followed or new patterns are needed **Scope:** - Number of systems touched - New code vs modification of existing code - Amount of new test coverage required - Data migration or configuration changes needed **Risk:** - New technology or unfamiliar libraries - Unclear or ambiguous requirements - Dependencies on unfinished work - Cross-system integration complexity - Performance sensitivity --- ## Phase 4: Generate the Estimate ```markdown ## Task Estimate: [Task Name] Generated: [Date] ### Task Description [Restate the task clearly in 1-2 sentences] ### Complexity Assessment | Factor | Assessment | Notes | |--------|-----------|-------| | Systems affected | [List] | [Core, gameplay, UI, etc.] | | Files likely modified | [Count] | [Key files listed below] | | New code vs modification | [Ratio] | | | Integration points | [Count] | [Which systems interact] | | Test coverage needed | [Low / Medium / High] | | | Existing patterns available | [Yes / Partial / No] | | **Key files likely affected:** - `[path/to/file1]` -- [what changes here] ### Effort Estimate | Scenario | Days | Assumption | |----------|------|------------| | Optimistic | [X] | Everything goes right, no surprises | | Expected | [Y] | Normal pace, minor issues, one round of review | | Pessimistic | [Z] | Significant unknowns surface, blocked for a day | **Recommended budget: [Y days]** ### Confidence: [High / Medium / Low] [Explain which factors drive the confidence level for this specific task.] ### Risk Factors | Risk | Likelihood | Impact | Mitigation | |------|-----------|--------|------------| ### Dependencies | Dependency | Status | Impact if Delayed | |-----------|--------|-------------------| ### Suggested Breakdown | # | Sub-task | Estimate | Notes | |---|----------|----------|-------| | 1 | [Research / spike] | [X days] | | | 2 | [Core implementation] | [X days] | | | 3 | [Testing and validation] | [X days] | | | | **Total** | **[Y days]** | | ### Notes and Assumptions - [Key assumption that affects the estimate] - [Any caveats about scope boundaries] ``` Output the estimate with a brief summary: recommended budget, confidence level, and the single biggest risk factor. This skill is read-only — no files are written. Verdict: **COMPLETE** — estimate generated. --- ## Phase 5: Next Steps - If confidence is Low: recommend a time-boxed spike (`/prototype`) before committing. - If the task is > 10 days: recommend breaking it into smaller stories via `/create-stories`. - To schedule the task: run `/sprint-plan update` to add it to the next sprint. ### Guidelines - Always give a range (optimistic / expected / pessimistic), never a single number - The recommended budget should be the expected estimate, not the optimistic one - Round to half-day increments — estimating in hours implies false precision for tasks longer than a day - Do not pad estimates silently — call out risk explicitly so the team can decide - Confidence is Low whenever the approach is undecided or the core requirements are still TBD — a range built on an unmade decision is a guess
Trust audit
SAFEgrade B · trust 89/100 Nothing in the source contradicts what it says it does. Grade A is reserved for packages that have also passed the behavioural sandbox.
| Layer | What it checks | Result |
|---|---|---|
| L0 | Provenance & inventory | PASS |
| L1 | Static analysis of the code | NA |
| L2 | Instruction surface (what it tells the agent) | PASS |
| L3 | Class-specific surface | PASS |
| L4 | Behavioural (sandbox) | SKIPPED |
What the source does
- Filesystem
- none-observed
- Network
- none-observed
- Shell
- none-observed
- Dependencies
- pinned
- Secrets in source
- none-found
Findings (0)
No findings outside the package's declared scope.
Gates applied: no_behavioural_pass.
42a36917b8befull audit observations/trust-audit/skill/donchitos__estimate.json · Report an issue / request a re-scanAudit history
Every audit this skill has had.
| Date | Source | Verdict | Grade | Score | Change |
|---|---|---|---|---|---|
| 2026-10-05 | 42a36917b8be | SAFE | B | 89 | first audit |
Questions
What does the Estimate skill do?
Turn Claude Code into a full game dev studio — 49 AI agents, 72 workflow skills, and a complete coordination system mirroring real studio hierarchy.
Is Estimate safe to install?
The audit found nothing in the source that contradicts what it says it does, and graded it B (89/100). Grade A is held back for packages that have also passed a sandboxed behavioural run, which is why a clean skill reads B.
What can Estimate access on my machine?
The audit observed no filesystem, network or shell use at all in its source.
How current is this page?
The grade is for one exact copy of the source (42a36917b8be), read on 2026-10-05. The repository is watched, and a new audit runs when it changes — this is the first audit.