@uzysjung/agent-harness 26.150.0 → 26.151.0

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (59) hide show
  1. package/README.ko.md +1 -1
  2. package/README.md +1 -1
  3. package/dist/{chunk-NKBUDHPC.js → chunk-3QBHZUVB.js} +99 -50
  4. package/dist/chunk-3QBHZUVB.js.map +1 -0
  5. package/dist/index.js +389 -274
  6. package/dist/index.js.map +1 -1
  7. package/dist/trust-tier-drift.js +1 -1
  8. package/package.json +1 -1
  9. package/templates/CLAUDE.md +145 -164
  10. package/templates/agents/build-error-resolver.md +1 -1
  11. package/templates/agents/plan-checker.md +1 -1
  12. package/templates/agents/reviewer.md +4 -5
  13. package/templates/antigravity/AGENTS.md.template +3 -23
  14. package/templates/codex/AGENTS.md.template +5 -56
  15. package/templates/hooks/protect-files.sh +4 -0
  16. package/templates/opencode/AGENTS.md.template +4 -52
  17. package/templates/opencode/opencode.json.template +0 -8
  18. package/templates/rules/change-management.md +0 -1
  19. package/templates/rules/cli-development.md +1 -1
  20. package/templates/rules/doc-governance.md +2 -0
  21. package/templates/rules/git-policy.md +1 -1
  22. package/templates/rules/ship-checklist.md +3 -3
  23. package/templates/rules/test-policy.md +3 -8
  24. package/templates/settings.json +1 -16
  25. package/templates/skills/agent-introspection-debugging/SKILL.md +1 -1
  26. package/templates/skills/audit-harness-fit/README.md +113 -0
  27. package/templates/skills/audit-harness-fit/SKILL.md +64 -433
  28. package/templates/skills/audit-harness-fit/evals/scenarios.yaml +222 -0
  29. package/templates/skills/audit-harness-fit/references/apply.md +66 -0
  30. package/templates/skills/audit-harness-fit/references/audit.md +160 -0
  31. package/templates/skills/audit-harness-fit/references/populate.md +74 -0
  32. package/templates/skills/audit-harness-fit/references/verification.md +123 -0
  33. package/templates/skills/compaction-handoff/SKILL.md +2 -2
  34. package/templates/skills/model-orchestration/SKILL.md +7 -0
  35. package/templates/skills/natural-korean/SKILL.md +45 -0
  36. package/templates/skills/north-star/references/roadmap-method.md +2 -2
  37. package/templates/skills/{task-brief → objective-brief}/SKILL.md +17 -16
  38. package/templates/skills/recurrence-prevention/SKILL.md +16 -16
  39. package/dist/chunk-NKBUDHPC.js.map +0 -1
  40. package/templates/agents/code-reviewer.md +0 -237
  41. package/templates/agents/security-reviewer.md +0 -108
  42. package/templates/hooks/task-brief-nudge.sh +0 -57
  43. package/templates/skills/audit-harness-fit/references/official-criteria.md +0 -367
  44. package/templates/skills/continuous-learning-v2/SKILL.md +0 -361
  45. package/templates/skills/continuous-learning-v2/agents/observer-loop.sh +0 -362
  46. package/templates/skills/continuous-learning-v2/agents/observer.md +0 -189
  47. package/templates/skills/continuous-learning-v2/agents/session-guardian.sh +0 -150
  48. package/templates/skills/continuous-learning-v2/agents/start-observer.sh +0 -252
  49. package/templates/skills/continuous-learning-v2/config.json +0 -8
  50. package/templates/skills/continuous-learning-v2/hooks/observe.sh +0 -585
  51. package/templates/skills/continuous-learning-v2/scripts/detect-project.sh +0 -322
  52. package/templates/skills/continuous-learning-v2/scripts/instinct-cli.py +0 -1956
  53. package/templates/skills/continuous-learning-v2/scripts/lib/homunculus-dir.sh +0 -31
  54. package/templates/skills/continuous-learning-v2/scripts/migrate-homunculus.sh +0 -68
  55. package/templates/skills/continuous-learning-v2/scripts/test_parse_instinct.py +0 -1420
  56. package/templates/skills/humanize-korean/SKILL.md +0 -228
  57. package/templates/skills/spec-scaling/SKILL.md +0 -89
  58. package/templates/skills/strategic-compact/SKILL.md +0 -145
  59. package/templates/skills/strategic-compact/suggest-compact.sh +0 -54
@@ -3,23 +3,18 @@
3
3
  **How far** to verify scales with the risk of the change; **when** those results must exist
4
4
  belongs to the Delivery rule.
5
5
 
6
- - An ordinary change gets the repository's baseline CI, regression across the affected scope, and
7
- whatever independent verification it warrants. A high-risk one widens that in proportion to what
8
- it touches and what its failure would cost.
6
+ - An ordinary change gets the repository's baseline CI and regression across the affected scope;
7
+ independent verification only where the Delivery rule requires it. A high-risk one widens that in
8
+ proportion to what it touches and what its failure would cost.
9
9
  - High-risk includes at least authentication, authorization, payments and settlement, personal
10
10
  data, data integrity, concurrency, state transitions, and migrations.
11
11
  - Full regression, full E2E, full mutation, and periodic security scanning belong to the CI/CD
12
12
  schedule, not to a per-change decision.
13
- - Test observable behavior — contracts and invariants — at whichever level best exposes
14
- plausible failures.
15
13
  - For a high-risk change, cover normal, boundary, failure, misuse, and recovery paths,
16
14
  omitting one only when failure on it is not plausible.
17
- - Keep tests deterministic: time, randomness, shared state, execution order, network, and
18
- external services.
19
15
  - Use production-compatible dependencies when a substitute's behavioral differences could affect
20
16
  the result. Otherwise, use explicit test doubles or contract tests.
21
17
  - Never use unauthorized production personal data, credentials, or secrets in tests.
22
18
  - Do not hide failures by weakening assertions, deleting or skipping tests, excluding coverage, or
23
19
  adding indiscriminate retries. Change tests only when the intended behavior has changed.
24
- - Coverage signals untested risk; it is not a target. Follow the repository-defined CI gates.
25
20
  - If the affected scope cannot be established confidently, broaden the validation.
@@ -22,26 +22,11 @@
22
22
  {
23
23
  "type": "command",
24
24
  "command": "bash \"$CLAUDE_PROJECT_DIR/.claude/hooks/protect-files.sh\""
25
- },
26
- {
27
- "type": "command",
28
- "command": "bash \"$CLAUDE_PROJECT_DIR/.claude/skills/strategic-compact/suggest-compact.sh\"",
29
- "async": true,
30
- "timeout": 5
31
25
  }
32
26
  ]
33
27
  }
34
28
  ],
35
29
  "PostToolUse": [],
36
- "UserPromptSubmit": [
37
- {
38
- "hooks": [
39
- {
40
- "type": "command",
41
- "command": "bash \"$CLAUDE_PROJECT_DIR/.claude/hooks/task-brief-nudge.sh\""
42
- }
43
- ]
44
- }
45
- ]
30
+ "UserPromptSubmit": []
46
31
  }
47
32
  }
@@ -138,7 +138,7 @@ Good pattern:
138
138
  ## Integration with ECC
139
139
 
140
140
  - Use `verification-loop` after recovery if code was changed.
141
- - Use `continuous-learning-v2` when the failure pattern is worth turning into an instinct or later skill.
141
+ - Use `recurrence-prevention` when the failure pattern is worth a rule, a gate, or a later skill.
142
142
  - Use `council` when the issue is not technical failure but decision ambiguity.
143
143
  - Use `workspace-surface-audit` if the failure came from conflicting local state or repo drift.
144
144
 
@@ -0,0 +1,113 @@
1
+ # audit-harness-fit
2
+
3
+ 현재 합의와 리포의 근거를 기준으로 에이전트 지침·스킬을 정비하고,
4
+ `AGENTS.md` / `CLAUDE.md`의 프로젝트 맥락을 채우는 스킬입니다.
5
+ 후보 수를 제한하지 않고, 같은 원인을 묶어 영향도순으로 보고합니다.
6
+
7
+ ## 디렉터리
8
+
9
+ ```text
10
+ audit-harness-fit/
11
+ ├── SKILL.md
12
+ ├── README.md
13
+ ├── references/
14
+ │ ├── audit.md
15
+ │ ├── verification.md
16
+ │ ├── apply.md
17
+ │ └── populate.md
18
+ └── evals/
19
+ └── scenarios.yaml
20
+ ```
21
+
22
+ `SKILL.md`는 진입점입니다. 감사, 검증 설계, 승인된 적용, 맥락 채우기별로
23
+ 필요한 참조만 읽도록 구성했습니다. 이 README와 evals는 관리·검토용이며
24
+ 일반 실행의 필수 컨텍스트가 아닙니다. 별도 실행 스크립트·모델 API 키·훅은 없습니다.
25
+
26
+ ## 하는 일
27
+
28
+ | 영역 | 결과 |
29
+ |---|---|
30
+ | 불필요한 질문·반복 검증 | 실제 지연을 만드는 지시와 조건부 교체 문안, 유효한 근거 재사용 조건 |
31
+ | 지침 충돌 | 양쪽 원문, 충돌 상황, 확정된 현재 의도와 실제 구현, 수정안 |
32
+ | 과도한 원칙·불필요한 스킬 | 유지·수정·조건 축소·통합·이관·제거·보류 판단과 의존 관계 |
33
+ | 긴 결정 사유·히스토리 | 본문·메뉴에는 현재 지시만, 이력은 필요할 때 읽는 별도 문서로 참조 |
34
+ | 테스트·상위 모델 위임 | 사용자 사용 장면 → 구현 경로 → 충분한 검증, 필요한 판단만 조건부 위임 |
35
+ | 프로젝트 맥락 | 실제 리포로 기존 스캐폴드를 채우거나 갱신하고 미확정 정보는 그대로 표시 |
36
+
37
+ 전체 감사는 다섯 감사 영역을 다룹니다. 특정 영역만 요청하면 범위를 확대하지
38
+ 않습니다. 문제를 억지로 만들지 않으며, 결과가 많아도 임의의 상위 개수로 자르지
39
+ 않습니다. 중요한 미확인 영역과 생략된 상세 내용은 따로 표시합니다.
40
+
41
+ ## 사용
42
+
43
+ 감사만 수행:
44
+
45
+ ```text
46
+ audit-harness-fit으로 현재 적용되는 지침과 Skill을 전체 점검해줘.
47
+ 후보 수를 제한하지 말고 중복을 묶어 영향도순으로 정리해줘.
48
+ 원문·문제 상황·수정안과 확인/미확인 경로를 보여줘. 파일은 수정하지 마.
49
+ ```
50
+
51
+ 확인한 제안 적용:
52
+
53
+ ```text
54
+ audit-harness-fit의 F-02, F-04 수정안을 해당 로컬 지침 파일에 반영해줘.
55
+ 기존 승인 범위와 필수 보호장치를 유지하고, 추가 권한이 필요한 변경은 분리해줘.
56
+ ```
57
+
58
+ 프로젝트 맥락 채우기:
59
+
60
+ ```text
61
+ audit-harness-fit으로 AGENTS.md와 CLAUDE.md의 미완성 프로젝트 맥락을
62
+ 현재 리포 기준으로 채워줘. 공통 원칙·import·기존 사용자 내용은 보존하고,
63
+ 핵심 사용 장면과 검증 기준을 연결해줘. 확인되지 않은 내용은 미확정으로 남겨줘.
64
+ ```
65
+
66
+ 기존 맥락을 갱신할 때는 "미완성 내용을 채워줘" 대신 "프로젝트 맥락을 최신
67
+ 리포와 확정된 결정에 맞춰 갱신해줘"로 요청합니다. 일반 개발 요청마다 전체
68
+ 감사를 실행하거나 스캐폴드를 자동으로 다시 채우는 방식은 사용하지 않습니다.
69
+
70
+ ## 배치와 기존 저장소 반영
71
+
72
+ ZIP 최상위는 `audit-harness-fit/` 하나입니다. 전체 폴더를 사용하는 클라이언트가
73
+ 인식하는 스킬 디렉터리에 배치합니다. 클라이언트별 로딩 경로는 실제 설정에서
74
+ 확인하고, ZIP을 풀었다는 것만으로 등록·로딩이 완료됐다고 판단하지 않습니다.
75
+
76
+ 사용자가 지정한 하네스 저장소의 배포 위치는 다음 형태로 반영할 수 있습니다.
77
+
78
+ ```text
79
+ templates/skills/audit-harness-fit/SKILL.md
80
+ templates/skills/audit-harness-fit/references/...
81
+ ```
82
+
83
+ 기존 폴더와 먼저 비교하고 로컬 수정사항을 보존합니다. 덮어쓰기만으로 옛 파일이
84
+ 삭제되는 것은 아닙니다. 예전 `references/official-criteria.md` 등의 기존 문서는
85
+ 이 실행 경로에서 요구하지 않지만, 참조·테스트·이력 보존 필요를 확인하기 전에는
86
+ 일괄 삭제하지 않습니다. 배포/로컬 사본이 있다면 저장소의 동기화 방식도 확인합니다.
87
+
88
+ 기본 추천 설치는 기존 자산 ID로 유지하는 통합 작업입니다. 이 ZIP에는 설치기
89
+ 코드 변경이 포함되지 않습니다. 설치 후 안내에는 스킬이 실제 설치된 경우에만
90
+ "이 스킬로 프로젝트 맥락을 채워 달라"는 문장을 연결하고, 기존 FILL 지시와의
91
+ 충돌은 함께 정리합니다. 모든 요청에 붙는 자동 감사 훅은 추가하지 않습니다.
92
+
93
+ ## 검토와 제한
94
+
95
+ [행동 시나리오](evals/scenarios.yaml)는 이 스킬을 개정하거나 특정 동작을 점검할
96
+ 때 선택해 쓰는 합성 사례입니다. 별도의 자동 실행기나 실행 결과가 아닙니다.
97
+ 일반 프로젝트 작업의 추가 테스트 게이트로 읽거나 실행하지 않습니다.
98
+
99
+ 필수 테스트·독립 리뷰·보안·데이터 보호·배포 승인은 유지합니다. 감사는 읽기
100
+ 전용이고, 명시된 범위의 수정만 적용합니다. 상위 모델 사용은 실제 도구·권한·
101
+ 데이터·비용 조건이 허용할 때만 가능합니다. 모델의 판단은 테스트 실행 증거가
102
+ 아니며, 사용 가능한 리뷰어가 없으면 필수 리뷰를 통과했다고 표시하지 않습니다.
103
+
104
+ 폴더와 문서 형식 확인은 실제 클라이언트 로딩, 에이전트 행동, 설치기 연동 테스트를
105
+ 대체하지 않습니다. 이 패키지만으로 GitHub 저장소나 사용자의 설치 환경을 바꾸지
106
+ 않습니다.
107
+
108
+ ## 형식 참고
109
+
110
+ - [Agent Skills specification](https://agentskills.io/specification)
111
+ - [OpenAI Skills guide](https://developers.openai.com/api/docs/guides/tools-skills)
112
+
113
+ 위 문서는 형식 참고용이며 매 감사에서 다시 읽어야 하는 입력이 아닙니다.