@tangle-network/agent-knowledge 13.0.1 → 14.0.1
This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
- package/AGENTS.md +52 -157
- package/CHANGELOG.md +33 -0
- package/README.md +8 -0
- package/dist/benchmarks/index.js +1 -1
- package/dist/{benchmarks-Qk94Gnj4.js → benchmarks-BbPJdmHe.js} +24 -24
- package/dist/{benchmarks-Qk94Gnj4.js.map → benchmarks-BbPJdmHe.js.map} +1 -1
- package/dist/cli.js +1 -1
- package/dist/index.d.ts +7 -5
- package/dist/index.d.ts.map +1 -1
- package/dist/index.js +55 -78
- package/dist/index.js.map +1 -1
- package/dist/{inspect-xAkP4sxi.js → inspect-DCpUoPVb.js} +14 -2
- package/dist/{inspect-xAkP4sxi.js.map → inspect-DCpUoPVb.js.map} +1 -1
- package/dist/memory/index.js +2 -2
- package/dist/{memory-DFSo2iLi.js → memory-BxwOtvd7.js} +2 -2
- package/dist/{memory-DFSo2iLi.js.map → memory-BxwOtvd7.js.map} +1 -1
- package/docs/architecture.md +4 -0
- package/package.json +3 -3
- package/skills/build-with-agent-knowledge/SKILL.md +9 -10
package/docs/architecture.md
CHANGED
|
@@ -58,6 +58,10 @@ After the exact submitted bytes are durable, `runVerifiedResearchLoop` passes th
|
|
|
58
58
|
The ledger materializes only observations whose complete source identity matches a confirmed record, so reusing one URI for different bytes cannot activate the wrong claims and a crash on either side resumes safely.
|
|
59
59
|
Unversioned URI-only ledgers cannot prove which bytes produced their observations; reads and writes fail with `ClaimLedgerMigrationRequiredError` and preserve the original file for an explicit archive-and-reverify migration.
|
|
60
60
|
Before synchronous question generation, the persistent driver records `preparedRounds`; a resume reconstructs and checkpoints any prepared round whose questions were interrupted, and the loop publishes its `research.iteration` event only after that checkpoint succeeds.
|
|
61
|
+
The research loop requires storage readiness and the driver's optional `isComplete()` result before it reports completion.
|
|
62
|
+
An unfinished driver can generate steering with no remaining storage gaps, so passing source requirements does not stop research prematurely.
|
|
63
|
+
Drivers without `isComplete()` use storage readiness alone.
|
|
64
|
+
Without readiness specifications, the loop runs to its round limit and never reports ready.
|
|
61
65
|
|
|
62
66
|
Every write in this layer goes through `durable-fs` (`writeFileDurable`, `writeJsonDurableWithinRoot`): temp file, fsync, atomic rename, and parent fsync.
|
|
63
67
|
`O_NOFOLLOW` descriptors anchored through `/proc/self/fd` prevent a directory swapped for a symlink during a write from redirecting it outside the root.
|
package/package.json
CHANGED
|
@@ -1,6 +1,6 @@
|
|
|
1
1
|
{
|
|
2
2
|
"name": "@tangle-network/agent-knowledge",
|
|
3
|
-
"version": "
|
|
3
|
+
"version": "14.0.1",
|
|
4
4
|
"description": "Build, search, evaluate, and improve source-backed knowledge bases.",
|
|
5
5
|
"homepage": "https://github.com/tangle-network/agent-knowledge#readme",
|
|
6
6
|
"repository": {
|
|
@@ -84,14 +84,14 @@
|
|
|
84
84
|
"zod": "4.5.4"
|
|
85
85
|
},
|
|
86
86
|
"peerDependencies": {
|
|
87
|
-
"@tangle-network/agent-eval": ">=0.
|
|
87
|
+
"@tangle-network/agent-eval": ">=0.174.0 <0.175.0",
|
|
88
88
|
"@tangle-network/agent-interface": "^2.0.0"
|
|
89
89
|
},
|
|
90
90
|
"devDependencies": {
|
|
91
91
|
"@arethetypeswrong/cli": "^0.18.5",
|
|
92
92
|
"@biomejs/biome": "^2.5.11",
|
|
93
93
|
"@neo4j-labs/agent-memory": "0.4.1",
|
|
94
|
-
"@tangle-network/agent-eval": "0.
|
|
94
|
+
"@tangle-network/agent-eval": "0.174.0",
|
|
95
95
|
"@tangle-network/agent-interface": "2.0.0",
|
|
96
96
|
"@types/node": "^26.4.0",
|
|
97
97
|
"mem0ai": "3.1.7",
|
|
@@ -56,6 +56,10 @@ Use existing vector, graph, search, and memory systems through adapters instead
|
|
|
56
56
|
3. Run one representative query or memory sequence through the production path.
|
|
57
57
|
4. Capture retrieved items, final answer or action, citations, errors, tokens, cost, and latency.
|
|
58
58
|
5. Prove a known good case passes and a realistic unsupported or missed case fails.
|
|
59
|
+
|
|
60
|
+
For ingestion or retrieval work, stop when the requested path passes these checks.
|
|
61
|
+
When the task requires improving an existing result:
|
|
62
|
+
|
|
59
63
|
6. Add only the missing improvement step: retrieval search, source acquisition, page update, answer repair, or memory policy.
|
|
60
64
|
7. Write changes to an isolated candidate with a stable run identity.
|
|
61
65
|
8. Compare baseline and candidate on the same development cases, then on unseen cases.
|
|
@@ -79,19 +83,14 @@ Bundled benchmark samples prove adapter wiring only; use complete external datas
|
|
|
79
83
|
|
|
80
84
|
## Completion
|
|
81
85
|
|
|
82
|
-
|
|
83
|
-
|
|
84
|
-
|
|
85
|
-
|
|
86
|
-
-> isolated candidate -> baseline comparison -> unseen comparison
|
|
87
|
-
-> approved promotion or correctly rejected change -> reproducible rerun
|
|
88
|
-
```
|
|
89
|
-
|
|
90
|
-
Report installed versions, exact imports, provider adapters, scope policy, case counts, baseline and candidate results, cost, latency, candidate identity, promotion result, and artifact paths.
|
|
86
|
+
Prove the requested source or memory event reaches production retrieval and an observable answer or action.
|
|
87
|
+
For improvement work, also prove isolation, baseline and unseen comparisons, and approved promotion or correct rejection.
|
|
88
|
+
Report the checked path, scope policy, results, material limits, and reproducible evidence.
|
|
89
|
+
Include dependency identity only when it explains compatibility or the result.
|
|
91
90
|
|
|
92
91
|
## Then consider
|
|
93
92
|
|
|
94
93
|
- `eval-engineering` when new production-derived cases are needed.
|
|
95
94
|
- `build-with-agent-runtime` when agents should research, edit, or compare candidates.
|
|
96
|
-
- `agent-eval
|
|
95
|
+
- `agent-eval` when the product needs shared comparison and release records.
|
|
97
96
|
- `harden` when changing tenant isolation, source trust, deletion, or promotion authority.
|