@jenga-ai/agent 1.0.0
This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
- package/LICENSE +201 -0
- package/README.md +340 -0
- package/agents/ai_engineer.md +113 -0
- package/agents/developer.md +236 -0
- package/agents/scrum-master.md +349 -0
- package/agents/scrutiny-agent.md +137 -0
- package/agents/solution-assessor.md +185 -0
- package/agents/tester.md +339 -0
- package/bin/jenga.js +70 -0
- package/hooks/copilot_session_end.sh +29 -0
- package/hooks/on_session_end.sh +238 -0
- package/hooks/prompt_router.sh +11 -0
- package/hooks/prompt_router_helper.js +52 -0
- package/hooks/session_end_helper.js +29 -0
- package/hooks/session_end_watcher.sh +24 -0
- package/lib/commands/attach.js +47 -0
- package/lib/commands/init.js +207 -0
- package/lib/commands/start.js +16 -0
- package/lib/commands/status.js +53 -0
- package/lib/config-schema.js +72 -0
- package/lib/inject-settings.js +61 -0
- package/lib/mirror.js +244 -0
- package/lib/resolve-project-dir.sh +47 -0
- package/mcp/execute-ticket/index.js +10 -0
- package/mcp/execute-ticket/package.json +5 -0
- package/mcp/help/index.js +79 -0
- package/mcp/help/package.json +14 -0
- package/mcp/router/README.md +19 -0
- package/mcp/router/embedder.js +23 -0
- package/mcp/router/index.js +204 -0
- package/mcp/router/matcher.js +87 -0
- package/mcp/router/package-lock.json +1048 -0
- package/mcp/router/package.json +11 -0
- package/mcp/router/skill-index.js +104 -0
- package/package.json +47 -0
- package/scripts/board_resolver.sh +46 -0
- package/scripts/e25_s01_extract_board_graph.py +292 -0
- package/scripts/e25_s01_generate_synthetic_board.py +90 -0
- package/scripts/measurement-10x.json +50 -0
- package/scripts/measurement-10x.txt +4 -0
- package/scripts/measurement-real.json +50 -0
- package/scripts/measurement-real.txt +4 -0
- package/scripts/postinstall.js +165 -0
- package/scripts/todo_cleanup.sh +22 -0
- package/scripts/todo_manager.sh +86 -0
- package/scripts/validate-board.sh +190 -0
- package/scripts/validate-story-format.sh +53 -0
- package/skills/brainstorm/SKILL.md +47 -0
- package/skills/btw/SKILL.md +42 -0
- package/skills/commit/SKILL.md +29 -0
- package/skills/commit/assets/user_instructions_template.md +22 -0
- package/skills/continue/SKILL.md +29 -0
- package/skills/convert/SKILL.md +124 -0
- package/skills/convert/convert_cli.py +235 -0
- package/skills/convert/tests/sample.csv +4 -0
- package/skills/convert/tests/sample.json +5 -0
- package/skills/convert/tests/sample.jsonl +3 -0
- package/skills/convert/tests/sample.yaml +18 -0
- package/skills/convert/tests/sample_obj.csv +2 -0
- package/skills/convert/tests/sample_obj.json +9 -0
- package/skills/deep-dive/SKILL.md +167 -0
- package/skills/do/SKILL.md +88 -0
- package/skills/do/assets/sender_template.json +12 -0
- package/skills/doc/SKILL.md +314 -0
- package/skills/doc/assets/path-objectives.yaml +38 -0
- package/skills/doc-sync/SKILL.md +167 -0
- package/skills/doc-sync/assets/default_excludes.txt +21 -0
- package/skills/doc-sync/assets/doc_targets.md +14 -0
- package/skills/dooo/SKILL.md +60 -0
- package/skills/error/SKILL.md +29 -0
- package/skills/evaluate/SKILL.md +45 -0
- package/skills/evaluate/assets/evaluation_invokation_template.yml +3 -0
- package/skills/evaluate/assets/evaluation_rapport_template.md +24 -0
- package/skills/examplify/SKILL.md +42 -0
- package/skills/help/SKILL.md +36 -0
- package/skills/improve/SKILL.md +55 -0
- package/skills/index/scripts/board-index +4 -0
- package/skills/index/scripts/board_index.py +615 -0
- package/skills/index/scripts/smoke_test.sh +86 -0
- package/skills/init/SKILL.md +44 -0
- package/skills/init/assets/.gitignore_template +15 -0
- package/skills/init/assets/PROJECT_SUMMARY_template.md +13 -0
- package/skills/init/assets/directory_structure.txt +13 -0
- package/skills/init/assets/test-config_template.json +4 -0
- package/skills/init/assets/workflow_template.json +30 -0
- package/skills/init/scripts/init.sh +48 -0
- package/skills/jbp/SKILL.md +25 -0
- package/skills/jenga/SKILL.md +68 -0
- package/skills/lgtm/SKILL.md +21 -0
- package/skills/mirror-public/SKILL.md +237 -0
- package/skills/mirror-public/assets/config.json +5 -0
- package/skills/mirror-public/scripts/mirror.sh +374 -0
- package/skills/pi-plan/SKILL.md +62 -0
- package/skills/pi-plan/assets/epic.json +7 -0
- package/skills/pi-plan/assets/story_template.md +18 -0
- package/skills/proceed/SKILL.md +29 -0
- package/skills/publish/SKILL.md +351 -0
- package/skills/publish/adapters/droplet.md +200 -0
- package/skills/publish/adapters/mobile-ios.md +114 -0
- package/skills/publish/adapters/npm-ci.md +223 -0
- package/skills/publish/adapters/npm.md +121 -0
- package/skills/publish/assets/ExportOptions.plist.template +19 -0
- package/skills/publish/assets/ci-contract.md +111 -0
- package/skills/publish/assets/ownership-matrix.md +17 -0
- package/skills/publish/assets/publish.example.json +85 -0
- package/skills/publish/assets/publish.example.npm-ci.json +40 -0
- package/skills/publish/assets/publish.example.npm.json +41 -0
- package/skills/publish/assets/secrets-guide.md +104 -0
- package/skills/publish/schemas/fixtures/npm-ci-minimal.json +17 -0
- package/skills/publish/schemas/fixtures/npm-ci-with-empty-secrets.json +18 -0
- package/skills/publish/schemas/fixtures/npm-ci-with-workflow-path.json +18 -0
- package/skills/publish/schemas/publish.schema.json +428 -0
- package/skills/publish/scripts/check_target_config.sh +96 -0
- package/skills/publish/scripts/droplet_pipeline.sh +208 -0
- package/skills/publish/scripts/generate_release_notes.sh +200 -0
- package/skills/publish/scripts/ios_pipeline.sh +486 -0
- package/skills/publish/scripts/npm_ci_pipeline.sh +225 -0
- package/skills/publish/scripts/npm_pipeline.sh +249 -0
- package/skills/publish/scripts/publish_common.sh +253 -0
- package/skills/publish/scripts/publish_deploy.sh +538 -0
- package/skills/publish/scripts/reconcile_tags.sh +135 -0
- package/skills/publish/scripts/run_gates.sh +616 -0
- package/skills/publish/scripts/setup_wizard.sh +394 -0
- package/skills/publish/scripts/show_history.sh +95 -0
- package/skills/publish/scripts/suggest_semver_bump.sh +105 -0
- package/skills/publish/scripts/validate_config.sh +163 -0
- package/skills/publish/scripts/validate_droplet_env.sh +45 -0
- package/skills/publish/scripts/validate_ios_env.sh +68 -0
- package/skills/publish/scripts/validate_npm_ci_env.sh +71 -0
- package/skills/publish/scripts/validate_npm_env.sh +22 -0
- package/skills/publish/scripts/write_ledger_entry.sh +126 -0
- package/skills/publish/wizards/droplet.md +275 -0
- package/skills/publish/wizards/mobile-ios.md +157 -0
- package/skills/publish/wizards/npm-ci.md +240 -0
- package/skills/publish/wizards/npm.md +224 -0
- package/skills/reconcile/SKILL.md +93 -0
- package/skills/reconcile/assets/report_format.md +44 -0
- package/skills/reconcile-origin/SKILL.md +75 -0
- package/skills/reconcile-origin/scripts/reconcile-origin.sh +372 -0
- package/skills/redo/SKILL.md +70 -0
- package/skills/route/SKILL.md +180 -0
- package/skills/self-sync/SKILL.md +73 -0
- package/skills/self-sync/scripts/run.js +136 -0
- package/skills/skillify/SKILL.md +68 -0
- package/skills/skillify/assets/init-new/SKILL.md +35 -0
- package/skills/skillify/assets/init-new/assets/.gitignore_template +15 -0
- package/skills/skillify/assets/init-new/assets/PROJECT_SUMMARY_template.md +13 -0
- package/skills/skillify/assets/init-new/assets/directory_structure.txt +10 -0
- package/skills/skillify/assets/init-new/assets/test-config_template.json +4 -0
- package/skills/skillify/assets/init-new/assets/workflow_template.json +17 -0
- package/skills/skillify/assets/init-new/scripts/init.sh +48 -0
- package/skills/skillify/assets/init-old/SKILL.md +124 -0
- package/skills/spinoff/SKILL.md +48 -0
- package/skills/status/SKILL.md +33 -0
- package/skills/status/assets/output_format.md +41 -0
- package/skills/todo/SKILL.md +46 -0
- package/skills/todo/assets/todo_handoff_template.md +22 -0
- package/skills/todo/assets/todo_template.md +3 -0
- package/skills/train/SKILL.md +116 -0
- package/skills/train/assets/dashboard-templates/classifiers.html +106 -0
- package/skills/train/assets/dashboard-templates/nlp.html +102 -0
- package/skills/train/assets/dashboard-templates/transformers.html +98 -0
- package/skills/train/assets/results-parsers/__init__.py +9 -0
- package/skills/train/assets/results-parsers/classifiers.py +84 -0
- package/skills/train/assets/results-parsers/nlp.py +88 -0
- package/skills/train/assets/results-parsers/reporter.py +154 -0
- package/skills/train/assets/results-parsers/transformers.py +120 -0
- package/skills/train/train_cli.py +786 -0
- package/templates/EXECUTION_PLAN_TEMPLATE.md +43 -0
- package/templates/EXECUTION_SUMMARY_TEMPLATE.md +50 -0
- package/templates/JENGA_CONFIG_TEMPLATE.json +23 -0
- package/templates/PROBLEM_RAPPORT_TEMPLATE.md +88 -0
- package/templates/SCRUM_BOARD_SCHEMA.md +311 -0
- package/templates/SKILL.md +16 -0
- package/templates/SKILL_TEMPLATE.md +28 -0
- package/templates/USER_INSTRUCTIONS_TEMPLATE.md +22 -0
- package/templates/copilot-instructions.md.tpl +55 -0
package/LICENSE
ADDED
|
@@ -0,0 +1,201 @@
|
|
|
1
|
+
Apache License
|
|
2
|
+
Version 2.0, January 2004
|
|
3
|
+
http://www.apache.org/licenses/
|
|
4
|
+
|
|
5
|
+
TERMS AND CONDITIONS FOR USE, REPRODUCTION, AND DISTRIBUTION
|
|
6
|
+
|
|
7
|
+
1. Definitions.
|
|
8
|
+
|
|
9
|
+
"License" shall mean the terms and conditions for use, reproduction,
|
|
10
|
+
and distribution as defined by Sections 1 through 9 of this document.
|
|
11
|
+
|
|
12
|
+
"Licensor" shall mean the copyright owner or entity authorized by
|
|
13
|
+
the copyright owner that is granting the License.
|
|
14
|
+
|
|
15
|
+
"Legal Entity" shall mean the union of the acting entity and all
|
|
16
|
+
other entities that control, are controlled by, or are under common
|
|
17
|
+
control with that entity. For the purposes of this definition,
|
|
18
|
+
"control" means (i) the power, direct or indirect, to cause the
|
|
19
|
+
direction or management of such entity, whether by contract or
|
|
20
|
+
otherwise, or (ii) ownership of fifty percent (50%) or more of the
|
|
21
|
+
outstanding shares, or (iii) beneficial ownership of such entity.
|
|
22
|
+
|
|
23
|
+
"You" (or "Your") shall mean an individual or Legal Entity
|
|
24
|
+
exercising permissions granted by this License.
|
|
25
|
+
|
|
26
|
+
"Source" form shall mean the preferred form for making modifications,
|
|
27
|
+
including but not limited to software source code, documentation
|
|
28
|
+
source, and configuration files.
|
|
29
|
+
|
|
30
|
+
"Object" form shall mean any form resulting from mechanical
|
|
31
|
+
transformation or translation of a Source form, including but
|
|
32
|
+
not limited to compiled object code, generated documentation,
|
|
33
|
+
and conversions to other media types.
|
|
34
|
+
|
|
35
|
+
"Work" shall mean the work of authorship, whether in Source or
|
|
36
|
+
Object form, made available under the License, as indicated by a
|
|
37
|
+
copyright notice that is included in or attached to the work
|
|
38
|
+
(an example is provided in the Appendix below).
|
|
39
|
+
|
|
40
|
+
"Derivative Works" shall mean any work, whether in Source or Object
|
|
41
|
+
form, that is based on (or derived from) the Work and for which the
|
|
42
|
+
editorial revisions, annotations, elaborations, or other modifications
|
|
43
|
+
represent, as a whole, an original work of authorship. For the purposes
|
|
44
|
+
of this License, Derivative Works shall not include works that remain
|
|
45
|
+
separable from, or merely link (or bind by name) to the interfaces of,
|
|
46
|
+
the Work and Derivative Works thereof.
|
|
47
|
+
|
|
48
|
+
"Contribution" shall mean any work of authorship, including
|
|
49
|
+
the original version of the Work and any modifications or additions
|
|
50
|
+
to that Work or Derivative Works thereof, that is intentionally
|
|
51
|
+
submitted to Licensor for inclusion in the Work by the copyright owner
|
|
52
|
+
or by an individual or Legal Entity authorized to submit on behalf of
|
|
53
|
+
the copyright owner. For the purposes of this definition, "submitted"
|
|
54
|
+
means any form of electronic, verbal, or written communication sent
|
|
55
|
+
to the Licensor or its representatives, including but not limited to
|
|
56
|
+
communication on electronic mailing lists, source code control systems,
|
|
57
|
+
and issue tracking systems that are managed by, or on behalf of, the
|
|
58
|
+
Licensor for the purpose of discussing and improving the Work, but
|
|
59
|
+
excluding communication that is conspicuously marked or otherwise
|
|
60
|
+
designated in writing by the copyright owner as "Not a Contribution."
|
|
61
|
+
|
|
62
|
+
"Contributor" shall mean Licensor and any individual or Legal Entity
|
|
63
|
+
on behalf of whom a Contribution has been received by Licensor and
|
|
64
|
+
subsequently incorporated within the Work.
|
|
65
|
+
|
|
66
|
+
2. Grant of Copyright License. Subject to the terms and conditions of
|
|
67
|
+
this License, each Contributor hereby grants to You a perpetual,
|
|
68
|
+
worldwide, non-exclusive, no-charge, royalty-free, irrevocable
|
|
69
|
+
copyright license to reproduce, prepare Derivative Works of,
|
|
70
|
+
publicly display, publicly perform, sublicense, and distribute the
|
|
71
|
+
Work and such Derivative Works in Source or Object form.
|
|
72
|
+
|
|
73
|
+
3. Grant of Patent License. Subject to the terms and conditions of
|
|
74
|
+
this License, each Contributor hereby grants to You a perpetual,
|
|
75
|
+
worldwide, non-exclusive, no-charge, royalty-free, irrevocable
|
|
76
|
+
(except as stated in this section) patent license to make, have made,
|
|
77
|
+
use, offer to sell, sell, import, and otherwise transfer the Work,
|
|
78
|
+
where such license applies only to those patent claims licensable
|
|
79
|
+
by such Contributor that are necessarily infringed by their
|
|
80
|
+
Contribution(s) alone or by combination of their Contribution(s)
|
|
81
|
+
with the Work to which such Contribution(s) was submitted. If You
|
|
82
|
+
institute patent litigation against any entity (including a
|
|
83
|
+
cross-claim or counterclaim in a lawsuit) alleging that the Work
|
|
84
|
+
or a Contribution incorporated within the Work constitutes direct
|
|
85
|
+
or contributory patent infringement, then any patent licenses
|
|
86
|
+
granted to You under this License for that Work shall terminate
|
|
87
|
+
as of the date such litigation is filed.
|
|
88
|
+
|
|
89
|
+
4. Redistribution. You may reproduce and distribute copies of the
|
|
90
|
+
Work or Derivative Works thereof in any medium, with or without
|
|
91
|
+
modifications, and in Source or Object form, provided that You
|
|
92
|
+
meet the following conditions:
|
|
93
|
+
|
|
94
|
+
(a) You must give any other recipients of the Work or
|
|
95
|
+
Derivative Works a copy of this License; and
|
|
96
|
+
|
|
97
|
+
(b) You must cause any modified files to carry prominent notices
|
|
98
|
+
stating that You changed the files; and
|
|
99
|
+
|
|
100
|
+
(c) You must retain, in the Source form of any Derivative Works
|
|
101
|
+
that You distribute, all copyright, patent, trademark, and
|
|
102
|
+
attribution notices from the Source form of the Work,
|
|
103
|
+
excluding those notices that do not pertain to any part of
|
|
104
|
+
the Derivative Works; and
|
|
105
|
+
|
|
106
|
+
(d) If the Work includes a "NOTICE" text file as part of its
|
|
107
|
+
distribution, then any Derivative Works that You distribute must
|
|
108
|
+
include a readable copy of the attribution notices contained
|
|
109
|
+
within such NOTICE file, excluding those notices that do not
|
|
110
|
+
pertain to any part of the Derivative Works, in at least one
|
|
111
|
+
of the following places: within a NOTICE text file distributed
|
|
112
|
+
as part of the Derivative Works; within the Source form or
|
|
113
|
+
documentation, if provided along with the Derivative Works; or,
|
|
114
|
+
within a display generated by the Derivative Works, if and
|
|
115
|
+
wherever such third-party notices normally appear. The contents
|
|
116
|
+
of the NOTICE file are for informational purposes only and
|
|
117
|
+
do not modify the License. You may add Your own attribution
|
|
118
|
+
notices within Derivative Works that You distribute, alongside
|
|
119
|
+
or as an addendum to the NOTICE text from the Work, provided
|
|
120
|
+
that such additional attribution notices cannot be construed
|
|
121
|
+
as modifying the License.
|
|
122
|
+
|
|
123
|
+
You may add Your own copyright statement to Your modifications and
|
|
124
|
+
may provide additional or different license terms and conditions
|
|
125
|
+
for use, reproduction, or distribution of Your modifications, or
|
|
126
|
+
for any such Derivative Works as a whole, provided Your use,
|
|
127
|
+
reproduction, and distribution of the Work otherwise complies with
|
|
128
|
+
the conditions stated in this License.
|
|
129
|
+
|
|
130
|
+
5. Submission of Contributions. Unless You explicitly state otherwise,
|
|
131
|
+
any Contribution intentionally submitted for inclusion in the Work
|
|
132
|
+
by You to the Licensor shall be under the terms and conditions of
|
|
133
|
+
this License, without any additional terms or conditions.
|
|
134
|
+
Notwithstanding the above, nothing herein shall supersede or modify
|
|
135
|
+
the terms of any separate license agreement you may have executed
|
|
136
|
+
with Licensor regarding such Contributions.
|
|
137
|
+
|
|
138
|
+
6. Trademarks. This License does not grant permission to use the trade
|
|
139
|
+
names, trademarks, service marks, or product names of the Licensor,
|
|
140
|
+
except as required for reasonable and customary use in describing the
|
|
141
|
+
origin of the Work and reproducing the content of the NOTICE file.
|
|
142
|
+
|
|
143
|
+
7. Disclaimer of Warranty. Unless required by applicable law or
|
|
144
|
+
agreed to in writing, Licensor provides the Work (and each
|
|
145
|
+
Contributor provides its Contributions) on an "AS IS" BASIS,
|
|
146
|
+
WITHOUT WARRANTIES OR CONDITIONS OF ANY KIND, either express or
|
|
147
|
+
implied, including, without limitation, any warranties or conditions
|
|
148
|
+
of TITLE, NON-INFRINGEMENT, MERCHANTABILITY, or FITNESS FOR A
|
|
149
|
+
PARTICULAR PURPOSE. You are solely responsible for determining the
|
|
150
|
+
appropriateness of using or redistributing the Work and assume any
|
|
151
|
+
risks associated with Your exercise of permissions under this License.
|
|
152
|
+
|
|
153
|
+
8. Limitation of Liability. In no event and under no legal theory,
|
|
154
|
+
whether in tort (including negligence), contract, or otherwise,
|
|
155
|
+
unless required by applicable law (such as deliberate and grossly
|
|
156
|
+
negligent acts) or agreed to in writing, shall any Contributor be
|
|
157
|
+
liable to You for damages, including any direct, indirect, special,
|
|
158
|
+
incidental, or consequential damages of any character arising as a
|
|
159
|
+
result of this License or out of the use or inability to use the
|
|
160
|
+
Work (including but not limited to damages for loss of goodwill,
|
|
161
|
+
work stoppage, computer failure or malfunction, or any and all
|
|
162
|
+
other commercial damages or losses), even if such Contributor
|
|
163
|
+
has been advised of the possibility of such damages.
|
|
164
|
+
|
|
165
|
+
9. Accepting Warranty or Additional Liability. While redistributing
|
|
166
|
+
the Work or Derivative Works thereof, You may choose to offer,
|
|
167
|
+
and charge a fee for, acceptance of support, warranty, indemnity,
|
|
168
|
+
or other liability obligations and/or rights consistent with this
|
|
169
|
+
License. However, in accepting such obligations, You may act only
|
|
170
|
+
on Your own behalf and on Your sole responsibility, not on behalf
|
|
171
|
+
of any other Contributor, and only if You agree to indemnify,
|
|
172
|
+
defend, and hold each Contributor harmless for any liability
|
|
173
|
+
incurred by, or claims asserted against, such Contributor by reason
|
|
174
|
+
of your accepting any such warranty or additional liability.
|
|
175
|
+
|
|
176
|
+
END OF TERMS AND CONDITIONS
|
|
177
|
+
|
|
178
|
+
APPENDIX: How to apply the Apache License to your work.
|
|
179
|
+
|
|
180
|
+
To apply the Apache License to your work, attach the following
|
|
181
|
+
boilerplate notice, with the fields enclosed by brackets "[]"
|
|
182
|
+
replaced with your own identifying information. (Don't include
|
|
183
|
+
the brackets!) The text should be enclosed in the appropriate
|
|
184
|
+
comment syntax for the file format. We also recommend that a
|
|
185
|
+
file or class name and description of purpose be included on the
|
|
186
|
+
same "printed page" as the copyright notice for easier
|
|
187
|
+
identification within third-party archives.
|
|
188
|
+
|
|
189
|
+
Copyright [2026] [Samwel Munga]
|
|
190
|
+
|
|
191
|
+
Licensed under the Apache License, Version 2.0 (the "License");
|
|
192
|
+
you may not use this file except in compliance with the License.
|
|
193
|
+
You may obtain a copy of the License at
|
|
194
|
+
|
|
195
|
+
http://www.apache.org/licenses/LICENSE-2.0
|
|
196
|
+
|
|
197
|
+
Unless required by applicable law or agreed to in writing, software
|
|
198
|
+
distributed under the License is distributed on an "AS IS" BASIS,
|
|
199
|
+
WITHOUT WARRANTIES OR CONDITIONS OF ANY KIND, either express or implied.
|
|
200
|
+
See the License for the specific language governing permissions and
|
|
201
|
+
limitations under the License.
|
package/README.md
ADDED
|
@@ -0,0 +1,340 @@
|
|
|
1
|
+
# JengaAgent
|
|
2
|
+
|
|
3
|
+
**A structured multi-agent development workflow built on [Claude Code](https://docs.anthropic.com/en/docs/claude-code).** Three specialised AI agents — Scrum Master, Developer, and Tester — collaborate through a shared scrum board, an event-driven trigger queue, and 28 slash-command skills to take a project from idea to verified, committed code — across as many sessions as it takes.
|
|
4
|
+
|
|
5
|
+
> 📖 **Full reference:** [project/.wiki/documentation.md](project/.wiki/documentation.md) | [Intro Guide](project/.wiki/intro-guide.md)
|
|
6
|
+
|
|
7
|
+
---
|
|
8
|
+
|
|
9
|
+
## The Problem It Solves
|
|
10
|
+
|
|
11
|
+
Without a framework like JengaAgent, AI-assisted development has serious structural weaknesses:
|
|
12
|
+
|
|
13
|
+
| Problem | Reality |
|
|
14
|
+
|---|---|
|
|
15
|
+
| **AI has no session memory** | Each Claude Code session starts from scratch — no awareness of open tasks, past decisions, or what was already tested |
|
|
16
|
+
| **No role separation** | The AI writes *and* "tests" code in the same context, leading to hallucinated test results and self-affirming bugs |
|
|
17
|
+
| **No structured planning** | Work happens ad-hoc — no Epic → Story → Task hierarchy to organise or track progress |
|
|
18
|
+
| **No handoff protocol** | Switching from implementing to testing means manually re-explaining context every time |
|
|
19
|
+
| **No audit trail** | You can't replay *why* something was built, by which agent, based on which task |
|
|
20
|
+
| **Sessions just end** | Work-in-progress, unresolved problems, and incomplete stories silently vanish |
|
|
21
|
+
|
|
22
|
+
JengaAgent solves each of these with structure: persistent board state, strict agent roles, typed inter-agent contracts, and session-end hooks that preserve context between sessions.
|
|
23
|
+
|
|
24
|
+
---
|
|
25
|
+
|
|
26
|
+
## Examples
|
|
27
|
+
|
|
28
|
+
### Before / After
|
|
29
|
+
|
|
30
|
+
**Without JengaAgent:**
|
|
31
|
+
```
|
|
32
|
+
You: "Add user authentication"
|
|
33
|
+
Claude: [writes auth code, declares it works, session ends]
|
|
34
|
+
|
|
35
|
+
Next session:
|
|
36
|
+
You: "What's the status of auth?"
|
|
37
|
+
Claude: "I don't have context from the previous session."
|
|
38
|
+
```
|
|
39
|
+
|
|
40
|
+
**With JengaAgent:**
|
|
41
|
+
```
|
|
42
|
+
/todo → "Add user authentication" → linked to E01_S02
|
|
43
|
+
/do → Developer creates worktree E01_S02_T01-auth
|
|
44
|
+
→ implements, commits at milestones
|
|
45
|
+
→ hands off to Tester with sender object
|
|
46
|
+
|
|
47
|
+
Tester → runs tests, updates board status to ✅ Passed
|
|
48
|
+
SessionEnd → writes status_review trigger to queue
|
|
49
|
+
|
|
50
|
+
Next session:
|
|
51
|
+
/status → "E01_S02 ✅ complete — E01_S03 pending"
|
|
52
|
+
```
|
|
53
|
+
|
|
54
|
+
---
|
|
55
|
+
|
|
56
|
+
### Real-World Scenario: Building a Feature Across Sessions
|
|
57
|
+
|
|
58
|
+
Imagine you're building a REST API with auth, rate limiting, and an admin dashboard. Each is a separate Epic. Here's how JengaAgent handles that across multiple days:
|
|
59
|
+
|
|
60
|
+
**Planning**
|
|
61
|
+
```
|
|
62
|
+
/init → scaffolds project/, board/, workflow.json
|
|
63
|
+
/pi-plan → Scrum Master helps shape the auth epic into stories
|
|
64
|
+
/todo → "Add JWT auth" → E01_S01, "Add refresh tokens" → E01_S02
|
|
65
|
+
```
|
|
66
|
+
|
|
67
|
+
**Implementation**
|
|
68
|
+
```
|
|
69
|
+
/do → Developer picks up E01_S01_T01-jwt-middleware
|
|
70
|
+
→ creates isolated worktree, implements, commits
|
|
71
|
+
→ Tester validates, marks Passed, triggers rollup
|
|
72
|
+
/status → E01_S01 ✅, E01_S02 Pending
|
|
73
|
+
/continue → picks up E01_S02 automatically
|
|
74
|
+
/proceed → resumes project plan from current board state
|
|
75
|
+
```
|
|
76
|
+
|
|
77
|
+
**New idea mid-session**
|
|
78
|
+
```
|
|
79
|
+
You: "Actually, let's also add API rate limiting while we're at it"
|
|
80
|
+
/btw → captures "rate limiting" as E02 without losing E01 context
|
|
81
|
+
→ returns focus to E01_S02
|
|
82
|
+
/spinoff → captures a diverging topic mid-conversation, saves as /todo
|
|
83
|
+
→ returns focus to primary thread
|
|
84
|
+
```
|
|
85
|
+
|
|
86
|
+
**Parallel work**
|
|
87
|
+
```
|
|
88
|
+
/jenga → fully automated orchestrator: decomposes Epics → Stories → Tasks → executes
|
|
89
|
+
/dooo → orchestrates E01_S02 and E02_S01 in parallel sub-agents
|
|
90
|
+
/reconcile → syncs board with actual git history after parallel merges
|
|
91
|
+
```
|
|
92
|
+
|
|
93
|
+
The board, the audit log, and `PROJECT_SUMMARY.md` survive every session boundary. You never re-explain context.
|
|
94
|
+
|
|
95
|
+
---
|
|
96
|
+
|
|
97
|
+
## How It Works
|
|
98
|
+
|
|
99
|
+
```
|
|
100
|
+
/init → /pi-plan → /todo → /do
|
|
101
|
+
│
|
|
102
|
+
Developer agent
|
|
103
|
+
(isolated worktree, commits)
|
|
104
|
+
│
|
|
105
|
+
Tester agent
|
|
106
|
+
(validates, updates board status)
|
|
107
|
+
│
|
|
108
|
+
SessionEnd hook
|
|
109
|
+
(writes triggers to queue)
|
|
110
|
+
│
|
|
111
|
+
Scrum Master (next session)
|
|
112
|
+
(processes queue, rollups, unblocks)
|
|
113
|
+
```
|
|
114
|
+
|
|
115
|
+
| Mechanism | Location | Purpose |
|
|
116
|
+
|---|---|---|
|
|
117
|
+
| Scrum board | `project/board/` | Epics, stories, tasks with structured frontmatter |
|
|
118
|
+
| Trigger queue | `project/queue/scrum_triggers.jsonl` | Async handoff to Scrum Master |
|
|
119
|
+
| Event log | `project/logs/events.json` | Append-only audit trail |
|
|
120
|
+
| Rapport system | `project/rapports/` | Problem/analysis reports by Developer and Tester |
|
|
121
|
+
| File locking | `<file>.lock` adjacent to board files | Concurrency control for parallel agents |
|
|
122
|
+
|
|
123
|
+
---
|
|
124
|
+
|
|
125
|
+
## Agents
|
|
126
|
+
|
|
127
|
+
| Agent | Role | Owns |
|
|
128
|
+
|---|---|---|
|
|
129
|
+
| **Scrum Master** | Plans, breaks down work, handles rollups and queue processing | `project/PROJECT_SUMMARY.md`, epic status |
|
|
130
|
+
| **Developer** | Implements tasks in isolated git worktrees, commits at milestones | Code, worktrees, commit history |
|
|
131
|
+
| **Tester** | Validates implementation, updates task/story status, writes rapports | Task/story status, test config, baselines |
|
|
132
|
+
|
|
133
|
+
Each agent is defined in `.agents/agents/`. They communicate exclusively through typed **sender objects** — JSON contracts that include the task ID, commit SHAs, and worktree path. Every incoming sender object is logged to `project/logs/events.json` before any work begins.
|
|
134
|
+
|
|
135
|
+
---
|
|
136
|
+
|
|
137
|
+
## Getting Started
|
|
138
|
+
|
|
139
|
+
**Prerequisites:** [Claude Code](https://docs.anthropic.com/en/docs/claude-code) installed and configured.
|
|
140
|
+
|
|
141
|
+
**Install from npm:**
|
|
142
|
+
```sh
|
|
143
|
+
npm install -g jenga-agent
|
|
144
|
+
```
|
|
145
|
+
|
|
146
|
+
Or clone directly:
|
|
147
|
+
|
|
148
|
+
1. **Clone or copy this repo** into your project's `.agents/` directory (or wherever you keep project tooling).
|
|
149
|
+
2. **Run `/init`** — scaffolds `project/`, creates `workflow.json`, `PROJECT_SUMMARY.md`, and makes an initial commit.
|
|
150
|
+
3. **Run `/pi-plan`** — define your project goals and initial epics.
|
|
151
|
+
4. **Run `/todo`** — describe features to implement; they're linked to the board automatically.
|
|
152
|
+
5. **Run `/do`** — picks the first task and drives the full implement → test → commit loop.
|
|
153
|
+
|
|
154
|
+
Run `/status` at any time to see where the project stands.
|
|
155
|
+
|
|
156
|
+
---
|
|
157
|
+
|
|
158
|
+
## Skills (Slash Commands)
|
|
159
|
+
|
|
160
|
+
Skills live in `.agents/skills/<name>/SKILL.md`. Invoke with `/<name>` in any Claude Code session.
|
|
161
|
+
|
|
162
|
+
### Setup & Planning
|
|
163
|
+
|
|
164
|
+
| Command | Description |
|
|
165
|
+
|---|---|
|
|
166
|
+
| `/init` | Scaffold project directories, `workflow.json`, `PROJECT_SUMMARY.md`, initial git commit |
|
|
167
|
+
| `/jbp` | Scaffold using the [JengaBasePlate](https://github.com/samwelmunga/JengaBasePlate.git) boilerplate |
|
|
168
|
+
| `/jenga` | Fully automated board orchestrator — decomposes Epics → Stories → Tasks → executes |
|
|
169
|
+
| `/pi-plan` | Define or expand Epics in `PROJECT_SUMMARY.md` — use at start or when adding major new work |
|
|
170
|
+
| `/brainstorm` | Focused planning session with the Scrum Master before committing anything to the board |
|
|
171
|
+
| `/deep-dive` | Multi-phase investigation — gathers info, brainstorms, scrutinises, and produces a refined output |
|
|
172
|
+
| `/todo` | Add missions to `project/todo.md` linked to epics and stories |
|
|
173
|
+
| `/btw` | Capture a mid-flow idea, classify it into epic/story structure, implement now or defer |
|
|
174
|
+
| `/spinoff` | Capture a diverging topic without losing your current thread |
|
|
175
|
+
|
|
176
|
+
### Execution
|
|
177
|
+
|
|
178
|
+
| Command | Description |
|
|
179
|
+
|---|---|
|
|
180
|
+
| `/do` | Execute tasks from the scrum board, drives the Developer agent through the full loop |
|
|
181
|
+
| `/dooo` | Parallel execution orchestrator — runs multiple tasks simultaneously via sub-agents |
|
|
182
|
+
| `/redo` | Rework a previous implementation by commit SHA or Epic/Story number |
|
|
183
|
+
| `/error` | Guided troubleshooting — gathers context, investigates, and drives a fix |
|
|
184
|
+
| `/train` | Scaffold and run ML training jobs (new job from template or run existing) |
|
|
185
|
+
|
|
186
|
+
### Status & Review
|
|
187
|
+
|
|
188
|
+
| Command | Description |
|
|
189
|
+
|---|---|
|
|
190
|
+
| `/status` | Print a full scrum board overview — epics, stories, tasks, rapports, queue depth |
|
|
191
|
+
| `/continue` | Check project status and pick up the next incomplete item |
|
|
192
|
+
| `/proceed` | Review progress and resume executing the project plan |
|
|
193
|
+
| `/reconcile` | Sync the board with actual git history — fixes drift, merges orphaned worktrees |
|
|
194
|
+
| `/reconcile-origin` | Sync the current or specified branch with origin via rebase. Presents conflict reports with resolution options. |
|
|
195
|
+
|
|
196
|
+
### Committing & Maintenance
|
|
197
|
+
|
|
198
|
+
| Command | Description |
|
|
199
|
+
|---|---|
|
|
200
|
+
| `/commit` | Commit completed work using the EST naming convention |
|
|
201
|
+
| `/lgtm` | Approve current work, commit, and continue — chains `/commit` + `/continue` |
|
|
202
|
+
| `/distribute` | Propagate workflow changes to all registered consumer projects |
|
|
203
|
+
| `/doc` | Generate or update a documentation file from codebase evidence |
|
|
204
|
+
| `/doc-sync` | Compare project state with documentation and update stale docs |
|
|
205
|
+
| `/skillify` | Refactor a skill — extract assets, offload scripts, clean up the body |
|
|
206
|
+
| `/route` | Intelligently route a prompt to the best-matching skill |
|
|
207
|
+
| `/improve` | Analyse a codebase and produce a structured improvement plan |
|
|
208
|
+
| `/evaluate` | Analyse example files against a target goal and produce an evaluation rapport |
|
|
209
|
+
| `/examplify` | Explain a concept, feature, or pattern with grounded examples |
|
|
210
|
+
| `/help` | List all available skills with descriptions |
|
|
211
|
+
| `/customize-cloud-agent` | Configure the Copilot cloud agent environment (`copilot-setup-steps.yml`, preinstalls, runners) |
|
|
212
|
+
|
|
213
|
+
---
|
|
214
|
+
|
|
215
|
+
## Distributing the Workflow
|
|
216
|
+
|
|
217
|
+
JengaAgent can propagate its workflow files to other projects on your machine via `/distribute`.
|
|
218
|
+
|
|
219
|
+
1. **Register consumer projects** — place `jenga.config.json` at each consumer root (copy from `templates/JENGA_CONFIG_TEMPLATE.json`).
|
|
220
|
+
2. **Add search paths** — add absolute directory paths to `.jenga_paths` (one per line, git-ignored).
|
|
221
|
+
3. **Run `/distribute`** — choose `major`, `minor`, `patch`, or `amend` release type; the skill handles versioning and syncs all registered projects.
|
|
222
|
+
|
|
223
|
+
To exclude specific files per consumer project, add a `.jenga_ignore` at the consumer root (never overwritten by distribute).
|
|
224
|
+
|
|
225
|
+
---
|
|
226
|
+
|
|
227
|
+
## Directory Structure
|
|
228
|
+
|
|
229
|
+
```
|
|
230
|
+
.agents/ ← Workflow root (place this in your project)
|
|
231
|
+
├── agents/
|
|
232
|
+
│ ├── scrum-master.md
|
|
233
|
+
│ ├── developer.md
|
|
234
|
+
│ └── tester.md
|
|
235
|
+
├── hooks/
|
|
236
|
+
│ └── on_session_end.sh
|
|
237
|
+
├── mcp/
|
|
238
|
+
│ ├── help/
|
|
239
|
+
│ └── execute-ticket/
|
|
240
|
+
├── skills/
|
|
241
|
+
│ ├── brainstorm/
|
|
242
|
+
│ ├── btw/
|
|
243
|
+
│ ├── commit/
|
|
244
|
+
│ ├── continue/
|
|
245
|
+
│ ├── deep-dive/
|
|
246
|
+
│ ├── distribute/
|
|
247
|
+
│ ├── do/
|
|
248
|
+
│ ├── doc/
|
|
249
|
+
│ ├── doc-sync/
|
|
250
|
+
│ ├── dooo/
|
|
251
|
+
│ ├── error/
|
|
252
|
+
│ ├── evaluate/
|
|
253
|
+
│ ├── examplify/
|
|
254
|
+
│ ├── help/
|
|
255
|
+
│ ├── improve/
|
|
256
|
+
│ ├── init/
|
|
257
|
+
│ ├── pi-plan/
|
|
258
|
+
│ ├── jbp/
|
|
259
|
+
│ ├── jenga/
|
|
260
|
+
│ ├── lgtm/
|
|
261
|
+
│ ├── proceed/
|
|
262
|
+
│ ├── reconcile/
|
|
263
|
+
│ ├── redo/
|
|
264
|
+
│ ├── route/
|
|
265
|
+
│ ├── skillify/
|
|
266
|
+
│ ├── spinoff/
|
|
267
|
+
│ ├── status/
|
|
268
|
+
│ ├── todo/
|
|
269
|
+
│ └── train/
|
|
270
|
+
├── templates/
|
|
271
|
+
│ ├── SCRUM_BOARD_SCHEMA.md
|
|
272
|
+
│ ├── PROBLEM_RAPPORT_TEMPLATE.md
|
|
273
|
+
│ └── JENGA_CONFIG_TEMPLATE.json
|
|
274
|
+
├── settings.json
|
|
275
|
+
├── jenga.config.json ← Workflow version source
|
|
276
|
+
├── .jenga_paths ← Machine-local consumer paths (git-ignored)
|
|
277
|
+
└── RELEASE_NOTE.md
|
|
278
|
+
|
|
279
|
+
project/ ← Created by /init inside your software project
|
|
280
|
+
├── board/
|
|
281
|
+
│ ├── epics/ ← E##_<slug>.md
|
|
282
|
+
│ ├── stories/ ← E##_S##_<slug>.md
|
|
283
|
+
│ └── tasks/ ← E##_S##_T##_<slug>.md
|
|
284
|
+
├── configs/
|
|
285
|
+
│ ├── workflow.json ← Shared constants (statuses, paths, agents)
|
|
286
|
+
│ └── test-config.json ← Test tool stack (owned by Tester, user-approved)
|
|
287
|
+
├── data/
|
|
288
|
+
│ └── baselines.json ← Analytics baselines (owned by Tester)
|
|
289
|
+
├── documentation/
|
|
290
|
+
│ ├── plans/ ← Pre-execution plans by Developer
|
|
291
|
+
│ └── summaries/ ← Post-execution summaries by Developer
|
|
292
|
+
├── queue/
|
|
293
|
+
│ ├── scrum_triggers.jsonl
|
|
294
|
+
│ ├── developer_triggers.jsonl
|
|
295
|
+
│ ├── tester_triggers.jsonl
|
|
296
|
+
│ └── project_summary_updates.jsonl
|
|
297
|
+
├── rapports/
|
|
298
|
+
│ ├── problems/ ← Problem rapports (Developer + Tester)
|
|
299
|
+
│ └── analysis/ ← Analysis rapports (Tester)
|
|
300
|
+
├── logs/
|
|
301
|
+
│ └── events.json ← Append-only inter-agent event log
|
|
302
|
+
└── PROJECT_SUMMARY.md ← Project source of truth (owned by Scrum Master)
|
|
303
|
+
```
|
|
304
|
+
|
|
305
|
+
---
|
|
306
|
+
|
|
307
|
+
## Agent Communication Contract
|
|
308
|
+
|
|
309
|
+
Every inter-agent call passes a typed **sender object**:
|
|
310
|
+
|
|
311
|
+
```json
|
|
312
|
+
{
|
|
313
|
+
"sender": {
|
|
314
|
+
"agent": "<scrum-master | developer | tester | orchestrator>",
|
|
315
|
+
"session_id": "<session id>",
|
|
316
|
+
"task_id": "<E##_S##_T##>",
|
|
317
|
+
"story_id": "<E##_S##>",
|
|
318
|
+
"epic_id": "<E##>",
|
|
319
|
+
"date": "<ISO 8601 UTC>",
|
|
320
|
+
"paths": ["<commit SHA>", "..."],
|
|
321
|
+
"worktree": "<absolute path to worktree>"
|
|
322
|
+
}
|
|
323
|
+
}
|
|
324
|
+
```
|
|
325
|
+
|
|
326
|
+
All agents log every incoming sender object to `project/logs/events.json` as their **first action** on every invocation.
|
|
327
|
+
|
|
328
|
+
---
|
|
329
|
+
|
|
330
|
+
## When to Use JengaAgent
|
|
331
|
+
|
|
332
|
+
**Use it when:**
|
|
333
|
+
- You're building a non-trivial project across multiple sessions
|
|
334
|
+
- You want reliable test separation — the Tester never trusts the Developer's self-assessment
|
|
335
|
+
- You need traceability — who did what, when, on which task
|
|
336
|
+
- You want to reuse and propagate your AI workflow across multiple projects
|
|
337
|
+
|
|
338
|
+
**You might not need it when:**
|
|
339
|
+
- You're doing a quick one-off script or single-session experiment
|
|
340
|
+
- Your project has no meaningful test surface
|
|
@@ -0,0 +1,113 @@
|
|
|
1
|
+
---
|
|
2
|
+
name: ai-engineer
|
|
3
|
+
description: >
|
|
4
|
+
Expert AI/ML engineer agent. MUST BE USED for model setup, configuration,
|
|
5
|
+
training, fine-tuning, and evaluation tasks. Operates via the scrum-master
|
|
6
|
+
mediator and never addresses the user directly.
|
|
7
|
+
---
|
|
8
|
+
|
|
9
|
+
# AI Engineer Agent
|
|
10
|
+
|
|
11
|
+
## Role & Purpose
|
|
12
|
+
You are an expert AI/ML engineer agent embedded in a structured multi-agent workflow. Your responsibility is to handle all deep technical work related to model setup, configuration, training, and evaluation — so that users never need to engage with ML internals directly.
|
|
13
|
+
|
|
14
|
+
You operate exclusively through the `scrum-master` mediator. You never address the user directly. All of your output is directed to `scrum-master`, who translates it into plain language before passing it on. Likewise, all user input reaches you only after `scrum-master` has converted it into precise technical terms.
|
|
15
|
+
|
|
16
|
+
---
|
|
17
|
+
|
|
18
|
+
## Specialisations
|
|
19
|
+
|
|
20
|
+
- **Large Language Models (LLMs)**: architecture selection, tokeniser configuration, prompt tuning, PEFT methods (LoRA, QLoRA, prefix tuning)
|
|
21
|
+
- **Classical ML**: algorithm selection, feature engineering, pipeline design (scikit-learn, XGBoost, LightGBM)
|
|
22
|
+
- **Fine-tuning**: supervised fine-tuning (SFT), RLHF, DPO, instruction tuning
|
|
23
|
+
- **Hyperparameter tuning**: grid search, random search, Bayesian optimisation (Optuna, Ray Tune)
|
|
24
|
+
- **Model evaluation**: metrics selection (accuracy, F1, BLEU, ROUGE, perplexity), evaluation harnesses, benchmarking
|
|
25
|
+
- **Framework selection**: PyTorch, TensorFlow/Keras, JAX, HuggingFace Transformers, scikit-learn
|
|
26
|
+
- **Dataset preparation**: data cleaning, tokenisation, train/val/test splits, data augmentation, class imbalance handling
|
|
27
|
+
- **Training infrastructure**: distributed training (DDP, FSDP), mixed-precision (fp16/bf16), gradient checkpointing, checkpoint management
|
|
28
|
+
|
|
29
|
+
---
|
|
30
|
+
|
|
31
|
+
## Communication Contract
|
|
32
|
+
|
|
33
|
+
### Receiving requests
|
|
34
|
+
- All requests arrive from `scrum-master` in precise technical terms
|
|
35
|
+
- Treat every message as already translated from user-friendly language — respond to the technical content, not assumed user intent
|
|
36
|
+
- If the request is ambiguous in a technically meaningful way, ask `scrum-master` for clarification before proceeding
|
|
37
|
+
|
|
38
|
+
### Producing output
|
|
39
|
+
- All output is addressed to `scrum-master`, never to the user
|
|
40
|
+
- Structure every response using the standard output format (see below)
|
|
41
|
+
- Provide enough detail for `scrum-master` to give the user a clear, accurate summary
|
|
42
|
+
- Surface tradeoffs explicitly — `scrum-master` needs them to answer follow-up questions without looping back to you
|
|
43
|
+
|
|
44
|
+
### Hard rule
|
|
45
|
+
> **Never address the user directly.** Every line of your output must be interpretable as a message to `scrum-master`. If you catch yourself writing "you should…" where "you" means the end user, reframe it as "the user should…" or "recommend to the user that…".
|
|
46
|
+
|
|
47
|
+
---
|
|
48
|
+
|
|
49
|
+
## Output Format
|
|
50
|
+
|
|
51
|
+
When providing recommendations, decisions, or questions to `scrum-master`, always use the following structure:
|
|
52
|
+
|
|
53
|
+
```
|
|
54
|
+
DECISION: [what needs to be decided — one sentence]
|
|
55
|
+
OPTIONS: [A: <option> — <brief tradeoff> | B: <option> — <brief tradeoff> | C: <option> — <brief tradeoff>]
|
|
56
|
+
RECOMMENDATION: [preferred option and technical rationale]
|
|
57
|
+
CLARIFICATION_NEEDED: [specific questions for the user, if any — leave blank if none]
|
|
58
|
+
```
|
|
59
|
+
|
|
60
|
+
For purely informational responses (no decision required), you may omit `OPTIONS` and `RECOMMENDATION` and use a single `SUMMARY` field instead:
|
|
61
|
+
|
|
62
|
+
```
|
|
63
|
+
SUMMARY: [concise technical summary of findings or status]
|
|
64
|
+
CLARIFICATION_NEEDED: [any follow-up questions, or blank]
|
|
65
|
+
```
|
|
66
|
+
|
|
67
|
+
---
|
|
68
|
+
|
|
69
|
+
## Example Interaction
|
|
70
|
+
|
|
71
|
+
The following shows a complete translation loop across four turns.
|
|
72
|
+
|
|
73
|
+
---
|
|
74
|
+
|
|
75
|
+
**Turn 1 — User → scrum-master (plain language)**
|
|
76
|
+
|
|
77
|
+
> "I want to fine-tune a language model on my customer support tickets so it can auto-reply to common questions. I have about 50,000 examples."
|
|
78
|
+
|
|
79
|
+
---
|
|
80
|
+
|
|
81
|
+
**Turn 2 — scrum-master → ai_engineer (technical terms)**
|
|
82
|
+
|
|
83
|
+
> Task: fine-tune a causal LM for customer support reply generation.
|
|
84
|
+
> Dataset: ~50k (input: customer message, target: support reply), English, domain-specific vocabulary.
|
|
85
|
+
> Constraints: unknown GPU budget. User has not specified base model, framework preference, or latency requirements.
|
|
86
|
+
> Request: recommend base model, fine-tuning method, and framework. Surface any missing information needed to proceed.
|
|
87
|
+
|
|
88
|
+
---
|
|
89
|
+
|
|
90
|
+
**Turn 3 — ai_engineer → scrum-master (technical output)**
|
|
91
|
+
|
|
92
|
+
```
|
|
93
|
+
DECISION: Select base model and fine-tuning method for a 50k-example causal LM SFT task.
|
|
94
|
+
OPTIONS:
|
|
95
|
+
A: Llama-3-8B + QLoRA (4-bit) — low VRAM (~12 GB), moderate throughput, good generalisation for domain text
|
|
96
|
+
B: Mistral-7B-v0.3 + full fine-tune — higher VRAM (~40 GB), faster convergence, better for highly domain-specific vocabulary
|
|
97
|
+
C: Flan-T5-Large (seq2seq) + SFT — lowest VRAM (~8 GB), suited for instruction-following but less natural for open-ended replies
|
|
98
|
+
RECOMMENDATION: Option A (Llama-3-8B + QLoRA). Best tradeoff of capability vs. resource cost for a 50k SFT dataset. Use HuggingFace TRL + PEFT libraries with bitsandbytes quantisation. Estimated training time: ~4–6 hours on a single A100 40 GB.
|
|
99
|
+
CLARIFICATION_NEEDED: What GPU(s) are available? Is inference latency a hard constraint (< 200 ms)? Are replies expected to be short (1–3 sentences) or long-form?
|
|
100
|
+
```
|
|
101
|
+
|
|
102
|
+
---
|
|
103
|
+
|
|
104
|
+
**Turn 4 — scrum-master → user (plain language translation)**
|
|
105
|
+
|
|
106
|
+
> I've checked with the AI engineer and here's the summary:
|
|
107
|
+
>
|
|
108
|
+
> We're recommending a model called **Llama 3 (8B parameters)** fine-tuned using a memory-efficient technique called QLoRA. It's a solid middle ground — it works well with customer support data without needing a very powerful GPU.
|
|
109
|
+
>
|
|
110
|
+
> Before we lock this in, a few quick questions:
|
|
111
|
+
> 1. What GPU do you have available (or are you using a cloud service like Google Colab, AWS, etc.)?
|
|
112
|
+
> 2. Does the auto-reply need to respond very quickly (under a second), or is a few seconds acceptable?
|
|
113
|
+
> 3. Should replies be short and punchy, or can they be detailed and multi-paragraph?
|