Skip to content

chore(release): 8.0.6 — require Eval 0.146.0 - #141

Merged
drewstone merged 1 commit into
mainfrom
chore/agent-eval-0.146.0
Aug 16, 2026
Merged

chore(release): 8.0.6 — require Eval 0.146.0#141
drewstone merged 1 commit into
mainfrom
chore/agent-eval-0.146.0

Conversation

@drewstone

Copy link
Copy Markdown
Contributor

Unblocks every consumer that needs @tangle-network/agent-eval/multishot/golden. That subpath is new in Eval 0.146.0, and this package's peer window stopped one version short of it.

The window was derived, not chosen

expectedPeerRange() in scripts/lib/peer-range.mjs computes the range from the dev dependency. For a pre-1.0 dependency it emits >=<version> <major.minor+1.0>, because npm locks a 0.x caret to its minor. Requiring Eval 0.145.21 therefore produced >=0.145.21 <0.146.0 automatically. Requiring 0.146.0 produces >=0.146.0 <0.147.0 by the same rule.

Eval 0.146.0 removes nothing

Measured by diffing the published type surfaces of 0.145.21 and 0.146.0 through the TypeScript checker, across every entry point in the exports map:

measure0.145.21 → 0.146.0
entry points removed0
top-level exports removed0
interface members removed0
top-level exports added51

The 20 signature changes are type-precision improvements on previously untyped values — env?: any becoming NodeJS.ProcessEnv, Promise<any> becoming Promise<DatabaseSync | null>.

Proof

pnpm typecheck clean (src + contracts)
pnpm test 61 files passed | 3 skipped, 604 tests passed | 12 skipped
pnpm build 35 files, 2.65 MB
pnpm verify:package
Verified @tangle-network/agent-knowledge@8.0.6 with one installed copy each of
agent-eval 0.146.0, agent-core 0.9.4, and agent-interface 1.0.0:
clean install, 5 imports, skill, CLI version, and re-pack.

Eval 0.146.0 adds the multishot/golden subpath and removes no export. Measured against 0.145.21, the published type surface loses no entry point, no top-level export and no interface member.
The peer window stopped at 0.146.0 because Eval is pre-1.0 and npm locks a 0.x range to its minor. That window refused an additive release and blocked every consumer that needs multishot/golden.
@drewstone

Copy link
Copy Markdown
ContributorAuthor

@tangletools review now

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

✅ Auto-approved drewstone PR — f1af5758

This PR was opened by the trusted drewstone account.
The full PR reviewer audit still runs separately and will publish findings if it detects issues.

This approval is provisional. It rests on the audit running. If the audit cannot run — for example the CLI bridge rejects it — this approval is dismissed rather than left standing, so an unrun check never reads as a passing one.

tangletools · auto-approval · reason: drewstone_author · 2026-08-16T22:12:58Z

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

✅ Auto-approved drewstone PR — f1af5758

This PR was opened by the trusted drewstone account.
The full PR reviewer audit still runs separately and will publish findings if it detects issues.

This approval is provisional. It rests on the audit running. If the audit cannot run — for example the CLI bridge rejects it — this approval is dismissed rather than left standing, so an unrun check never reads as a passing one.

tangletools · auto-approval · reason: drewstone_author · 2026-08-16T22:13:01Z

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

✅ Auto-approved drewstone PR — f1af5758

This PR was opened by the trusted drewstone account.
The full PR reviewer audit still runs separately and will publish findings if it detects issues.

This approval is provisional. It rests on the audit running. If the audit cannot run — for example the CLI bridge rejects it — this approval is dismissed rather than left standing, so an unrun check never reads as a passing one.

tangletools · auto-approval · reason: drewstone_author · 2026-08-16T22:13:04Z

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

🟢 Value Audit — sound

Verdictsound
Coverage2 of 2 lenses (value, usefulness)
Concerns0 (none)
Heuristic0.0s
Duplication0.0s
Interrogation178.4s (2 bridge agents)
Total178.4s

💰 Value — sound

A routine, mechanically-derived peer-window bump that moves the agent-eval cohort to 0.146.0 exactly the way this repo's release machinery prescribes; ship.

  • What it does: Bumps the package to 8.0.6 and moves the agent-eval peer dependency from >=0.145.21 <0.146.0 to >=0.146.0 <0.147.0 (package.json:82), with the exact dev pin moved from 0.145.21 to 0.146.0 (package.json:89), plus the matching CHANGELOG entry and pnpm-lock regen (commit f1af575 touches all three, per git show f1af575 --stat). The practical delta: consumers installing this version get an eval p
  • Goals it achieves: Reopen the peer window for an additive eval release. Because eval is pre-1.0, the repo's rule derives the window as >=V <0.minor+1.0, so the 0.145.21 pin automatically excluded 0.146.0. The change re-admits current eval by advancing the pin one minor. Secondary goal, achieved implicitly: keep the published cohort guarantee intact — exactly one admitted eval minor, one installed copy — while shif
  • Assessment: Good, and squarely in the grain. (1) The range is not hand-written: expectedPeerRange() in scripts/lib/peer-range.mjs:33-35 emits >=0.146.0 <0.147.0 for a pre-1.0 pin — I executed the module and it prints exactly the string at package.json:82. (2) The guard enforces it: scripts/verify-package.mjs:69-74 reads the dev pin, derives the expected range, and fails verify:package on any mismatch, s
  • Better / existing approach: none — this is the right approach. I checked the two alternatives against the repo's own machinery: (a) a caret ^0.146.0 is semantically identical for 0.x under npm's rule (peer-range.mjs:5-9 documents this), so the explicit range is just the clearer spelling of the same contract — no improvement available; (b) widening to span both minors (e.g. >=0.145.21 <0.147.0) would avoid forcing a simul
  • Model: opencode/zai-coding-plan/glm-5.2
  • Bridge attempts: 2
  • Bridge warning: opencode/kimi-for-coding/k2p7: opencode: opencode error

🎯 Usefulness — sound

A mechanically derived peer-window bump that unblocks co-install with Eval 0.146.0; every claim checks out against the installed package and the repo's own enforcement script.

  • Integration: Fully wired. The window is not hand-picked: expectedPeerRange('0.146.0') returns '>=0.146.0 <0.147.0' (scripts/lib/peer-range.mjs:33-35), byte-identical to package.json:82, and scripts/verify-package.mjs:73,294-297 fails the package check if the declared peer ever diverges from the derivation. The devDep (package.json:89) and pnpm-lock.yaml moved to 0.146.0 in the same commit, and CHANGELOG.md doc
  • Fit with existing patterns: Follows the repo's established release-alignment grain exactly. git log shows a chain of identical alignment releases (8.0.5, 662c874; 383ce19 introduced the derivation rule this PR applies; 8ee086c; 2128bb2; fd0eaa7). The pre-1.0 exact-minor window is the deliberate policy documented in scripts/lib/peer-range.mjs:5-9 and enforced by the verify scripts — no competing pattern exists.
  • Real-world viability: Independently verified beyond the PR's claims: pnpm install --frozen-lockfile clean; pnpm run typecheck (src + contracts) clean against installed 0.146.0; @tangle-network/agent-eval/multishot/golden importable from the installed 0.146.0 (22 exports); the 0.146.0 exports map retains every subpath this repo imports ('.', ./campaign, ./rl, ./experiment all present). The tsc pass is the decisive check
  • Model: opencode/zai-coding-plan/glm-5.2
  • Bridge attempts: 1

No concerns — sound change, no better or existing approach found. ✅


What this audit checks

It judges the change on its merits — not whether it was tasked out in an issue. Unticketed, fast-moving work is fine; the question is whether the change is good and whether a better or existing approach should be used instead.

PassWhat it asks
HeuristicVague title? Whitespace-only or cruft-bearing diff? (content signals only)
DuplicationDo added function/class names already exist elsewhere in the repo?
Value AuditWhat does it do? What goal does it achieve? Is it good? Better architecture or already-exists?
Usefulness AuditDoes it integrate and fit? Will it hold up in real use and actually get used?

Findings are concerns, not blocks — the human reviewer decides what to do with them.

value-audit · 20260816T222525Z

@drewstone
drewstone merged commit 9a1079d into mainAug 16, 2026
2 checks passed
@drewstone
drewstone deleted the chore/agent-eval-0.146.0 branch August 16, 2026 22:26
Sign up for freeto join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

2 participants

@drewstone@tangletools
, 'i'); if (__m === '*' || __re.test(location.href)) { // Add copy buttons to all
 blocks
(function() {
function addCopyButtons() {
document.querySelectorAll('pre code').forEach(function(codeBlock) {
if (codeBlock.parentElement.hasAttribute('data-copy-added')) return;
codeBlock.parentElement.setAttribute('data-copy-added', 'true');
var btn = document.createElement('button');
btn.textContent = 'Copy';
btn.style.cssText = 'position:absolute;top:4px;right:4px;padding:2px 8px;font-size:11px;background:#4ecdc4;border:none;border-radius:4px;color:#1a1a2e;cursor:pointer;opacity:0.7;transition:opacity 0.2s;';
btn.onmouseover = function() { this.style.opacity = '1'; };
btn.onmouseout = function() { this.style.opacity = '0.7'; };
btn.onclick = function() {
navigator.clipboard.writeText(codeBlock.textContent).then(function() {
btn.textContent = 'Copied!';
setTimeout(function() { btn.textContent = 'Copy'; }, 1500);
});
};
codeBlock.parentElement.style.position = 'relative';
codeBlock.parentElement.appendChild(btn);
});
}
addCopyButtons();
// Re-run on dynamic content
var observer = new MutationObserver(addCopyButtons);
observer.observe(document.body, { childList: true, subtree: true });
})();
}
} catch(__e) { console.warn('[Userscript:Add Copy Buttons to Code Blocks]', __e); }
})();
(function(){
try {
var __m = "github.com";
var __re = new RegExp('^' + "github\\.com" + '
chore(release): 8.0.6 — require Eval 0.146.0 by drewstone · Pull Request #141 · tangle-network/agent-knowledge · GitHub
Skip to content

chore(release): 8.0.6 — require Eval 0.146.0 - #141

Merged
drewstone merged 1 commit into
mainfrom
chore/agent-eval-0.146.0
Aug 16, 2026
Merged

chore(release): 8.0.6 — require Eval 0.146.0#141
drewstone merged 1 commit into
mainfrom
chore/agent-eval-0.146.0

Conversation

@drewstone

Copy link
Copy Markdown
Contributor

Unblocks every consumer that needs @tangle-network/agent-eval/multishot/golden. That subpath is new in Eval 0.146.0, and this package's peer window stopped one version short of it.

The window was derived, not chosen

expectedPeerRange() in scripts/lib/peer-range.mjs computes the range from the dev dependency. For a pre-1.0 dependency it emits >=<version> <major.minor+1.0>, because npm locks a 0.x caret to its minor. Requiring Eval 0.145.21 therefore produced >=0.145.21 <0.146.0 automatically. Requiring 0.146.0 produces >=0.146.0 <0.147.0 by the same rule.

Eval 0.146.0 removes nothing

Measured by diffing the published type surfaces of 0.145.21 and 0.146.0 through the TypeScript checker, across every entry point in the exports map:

measure0.145.21 → 0.146.0
entry points removed0
top-level exports removed0
interface members removed0
top-level exports added51

The 20 signature changes are type-precision improvements on previously untyped values — env?: any becoming NodeJS.ProcessEnv, Promise<any> becoming Promise<DatabaseSync | null>.

Proof

pnpm typecheck clean (src + contracts)
pnpm test 61 files passed | 3 skipped, 604 tests passed | 12 skipped
pnpm build 35 files, 2.65 MB
pnpm verify:package
Verified @tangle-network/agent-knowledge@8.0.6 with one installed copy each of
agent-eval 0.146.0, agent-core 0.9.4, and agent-interface 1.0.0:
clean install, 5 imports, skill, CLI version, and re-pack.

Eval 0.146.0 adds the multishot/golden subpath and removes no export. Measured against 0.145.21, the published type surface loses no entry point, no top-level export and no interface member.
The peer window stopped at 0.146.0 because Eval is pre-1.0 and npm locks a 0.x range to its minor. That window refused an additive release and blocked every consumer that needs multishot/golden.
@drewstone

Copy link
Copy Markdown
ContributorAuthor

@tangletools review now

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

✅ Auto-approved drewstone PR — f1af5758

This PR was opened by the trusted drewstone account.
The full PR reviewer audit still runs separately and will publish findings if it detects issues.

This approval is provisional. It rests on the audit running. If the audit cannot run — for example the CLI bridge rejects it — this approval is dismissed rather than left standing, so an unrun check never reads as a passing one.

tangletools · auto-approval · reason: drewstone_author · 2026-08-16T22:12:58Z

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

✅ Auto-approved drewstone PR — f1af5758

This PR was opened by the trusted drewstone account.
The full PR reviewer audit still runs separately and will publish findings if it detects issues.

This approval is provisional. It rests on the audit running. If the audit cannot run — for example the CLI bridge rejects it — this approval is dismissed rather than left standing, so an unrun check never reads as a passing one.

tangletools · auto-approval · reason: drewstone_author · 2026-08-16T22:13:01Z

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

✅ Auto-approved drewstone PR — f1af5758

This PR was opened by the trusted drewstone account.
The full PR reviewer audit still runs separately and will publish findings if it detects issues.

This approval is provisional. It rests on the audit running. If the audit cannot run — for example the CLI bridge rejects it — this approval is dismissed rather than left standing, so an unrun check never reads as a passing one.

tangletools · auto-approval · reason: drewstone_author · 2026-08-16T22:13:04Z

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

🟢 Value Audit — sound

Verdictsound
Coverage2 of 2 lenses (value, usefulness)
Concerns0 (none)
Heuristic0.0s
Duplication0.0s
Interrogation178.4s (2 bridge agents)
Total178.4s

💰 Value — sound

A routine, mechanically-derived peer-window bump that moves the agent-eval cohort to 0.146.0 exactly the way this repo's release machinery prescribes; ship.

  • What it does: Bumps the package to 8.0.6 and moves the agent-eval peer dependency from >=0.145.21 <0.146.0 to >=0.146.0 <0.147.0 (package.json:82), with the exact dev pin moved from 0.145.21 to 0.146.0 (package.json:89), plus the matching CHANGELOG entry and pnpm-lock regen (commit f1af575 touches all three, per git show f1af575 --stat). The practical delta: consumers installing this version get an eval p
  • Goals it achieves: Reopen the peer window for an additive eval release. Because eval is pre-1.0, the repo's rule derives the window as >=V <0.minor+1.0, so the 0.145.21 pin automatically excluded 0.146.0. The change re-admits current eval by advancing the pin one minor. Secondary goal, achieved implicitly: keep the published cohort guarantee intact — exactly one admitted eval minor, one installed copy — while shif
  • Assessment: Good, and squarely in the grain. (1) The range is not hand-written: expectedPeerRange() in scripts/lib/peer-range.mjs:33-35 emits >=0.146.0 <0.147.0 for a pre-1.0 pin — I executed the module and it prints exactly the string at package.json:82. (2) The guard enforces it: scripts/verify-package.mjs:69-74 reads the dev pin, derives the expected range, and fails verify:package on any mismatch, s
  • Better / existing approach: none — this is the right approach. I checked the two alternatives against the repo's own machinery: (a) a caret ^0.146.0 is semantically identical for 0.x under npm's rule (peer-range.mjs:5-9 documents this), so the explicit range is just the clearer spelling of the same contract — no improvement available; (b) widening to span both minors (e.g. >=0.145.21 <0.147.0) would avoid forcing a simul
  • Model: opencode/zai-coding-plan/glm-5.2
  • Bridge attempts: 2
  • Bridge warning: opencode/kimi-for-coding/k2p7: opencode: opencode error

🎯 Usefulness — sound

A mechanically derived peer-window bump that unblocks co-install with Eval 0.146.0; every claim checks out against the installed package and the repo's own enforcement script.

  • Integration: Fully wired. The window is not hand-picked: expectedPeerRange('0.146.0') returns '>=0.146.0 <0.147.0' (scripts/lib/peer-range.mjs:33-35), byte-identical to package.json:82, and scripts/verify-package.mjs:73,294-297 fails the package check if the declared peer ever diverges from the derivation. The devDep (package.json:89) and pnpm-lock.yaml moved to 0.146.0 in the same commit, and CHANGELOG.md doc
  • Fit with existing patterns: Follows the repo's established release-alignment grain exactly. git log shows a chain of identical alignment releases (8.0.5, 662c874; 383ce19 introduced the derivation rule this PR applies; 8ee086c; 2128bb2; fd0eaa7). The pre-1.0 exact-minor window is the deliberate policy documented in scripts/lib/peer-range.mjs:5-9 and enforced by the verify scripts — no competing pattern exists.
  • Real-world viability: Independently verified beyond the PR's claims: pnpm install --frozen-lockfile clean; pnpm run typecheck (src + contracts) clean against installed 0.146.0; @tangle-network/agent-eval/multishot/golden importable from the installed 0.146.0 (22 exports); the 0.146.0 exports map retains every subpath this repo imports ('.', ./campaign, ./rl, ./experiment all present). The tsc pass is the decisive check
  • Model: opencode/zai-coding-plan/glm-5.2
  • Bridge attempts: 1

No concerns — sound change, no better or existing approach found. ✅


What this audit checks

It judges the change on its merits — not whether it was tasked out in an issue. Unticketed, fast-moving work is fine; the question is whether the change is good and whether a better or existing approach should be used instead.

PassWhat it asks
HeuristicVague title? Whitespace-only or cruft-bearing diff? (content signals only)
DuplicationDo added function/class names already exist elsewhere in the repo?
Value AuditWhat does it do? What goal does it achieve? Is it good? Better architecture or already-exists?
Usefulness AuditDoes it integrate and fit? Will it hold up in real use and actually get used?

Findings are concerns, not blocks — the human reviewer decides what to do with them.

value-audit · 20260816T222525Z

@drewstone
drewstone merged commit 9a1079d into mainAug 16, 2026
2 checks passed
@drewstone
drewstone deleted the chore/agent-eval-0.146.0 branch August 16, 2026 22:26
Sign up for freeto join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

2 participants

@drewstone@tangletools
, 'i'); if (__m === '*' || __re.test(location.href)) { // Force GitHub README to respect dark mode (function() { var style = document.createElement('style'); style.textContent = ' .markdown-body { color-scheme: dark light; } .markdown-body pre { background: #161b22 !important; } .markdown-body code { background: rgba(110, 118, 129, 0.4) !important; } .markdown-body table th, .markdown-body table td { border-color: #30363d !important; } .markdown-body img { background: #0d1117; } .markdown-body blockquote { border-left-color: #8b949e; } .markdown-body hr { border-color: #30363d; } '; document.head.appendChild(style); })(); } } catch(__e) { console.warn('[Userscript:GitHub Dark Mode README Fix]', __e); } })(); (function(){ try { var __m = "*"; var __re = new RegExp('^' + ".*" + ' chore(release): 8.0.6 — require Eval 0.146.0 by drewstone · Pull Request #141 · tangle-network/agent-knowledge · GitHub
Skip to content

chore(release): 8.0.6 — require Eval 0.146.0 - #141

Merged
drewstone merged 1 commit into
mainfrom
chore/agent-eval-0.146.0
Aug 16, 2026
Merged

chore(release): 8.0.6 — require Eval 0.146.0#141
drewstone merged 1 commit into
mainfrom
chore/agent-eval-0.146.0

Conversation

@drewstone

Copy link
Copy Markdown
Contributor

Unblocks every consumer that needs @tangle-network/agent-eval/multishot/golden. That subpath is new in Eval 0.146.0, and this package's peer window stopped one version short of it.

The window was derived, not chosen

expectedPeerRange() in scripts/lib/peer-range.mjs computes the range from the dev dependency. For a pre-1.0 dependency it emits >=<version> <major.minor+1.0>, because npm locks a 0.x caret to its minor. Requiring Eval 0.145.21 therefore produced >=0.145.21 <0.146.0 automatically. Requiring 0.146.0 produces >=0.146.0 <0.147.0 by the same rule.

Eval 0.146.0 removes nothing

Measured by diffing the published type surfaces of 0.145.21 and 0.146.0 through the TypeScript checker, across every entry point in the exports map:

measure0.145.21 → 0.146.0
entry points removed0
top-level exports removed0
interface members removed0
top-level exports added51

The 20 signature changes are type-precision improvements on previously untyped values — env?: any becoming NodeJS.ProcessEnv, Promise<any> becoming Promise<DatabaseSync | null>.

Proof

pnpm typecheck clean (src + contracts)
pnpm test 61 files passed | 3 skipped, 604 tests passed | 12 skipped
pnpm build 35 files, 2.65 MB
pnpm verify:package
Verified @tangle-network/agent-knowledge@8.0.6 with one installed copy each of
agent-eval 0.146.0, agent-core 0.9.4, and agent-interface 1.0.0:
clean install, 5 imports, skill, CLI version, and re-pack.

Eval 0.146.0 adds the multishot/golden subpath and removes no export. Measured against 0.145.21, the published type surface loses no entry point, no top-level export and no interface member.
The peer window stopped at 0.146.0 because Eval is pre-1.0 and npm locks a 0.x range to its minor. That window refused an additive release and blocked every consumer that needs multishot/golden.
@drewstone

Copy link
Copy Markdown
ContributorAuthor

@tangletools review now

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

✅ Auto-approved drewstone PR — f1af5758

This PR was opened by the trusted drewstone account.
The full PR reviewer audit still runs separately and will publish findings if it detects issues.

This approval is provisional. It rests on the audit running. If the audit cannot run — for example the CLI bridge rejects it — this approval is dismissed rather than left standing, so an unrun check never reads as a passing one.

tangletools · auto-approval · reason: drewstone_author · 2026-08-16T22:12:58Z

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

✅ Auto-approved drewstone PR — f1af5758

This PR was opened by the trusted drewstone account.
The full PR reviewer audit still runs separately and will publish findings if it detects issues.

This approval is provisional. It rests on the audit running. If the audit cannot run — for example the CLI bridge rejects it — this approval is dismissed rather than left standing, so an unrun check never reads as a passing one.

tangletools · auto-approval · reason: drewstone_author · 2026-08-16T22:13:01Z

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

✅ Auto-approved drewstone PR — f1af5758

This PR was opened by the trusted drewstone account.
The full PR reviewer audit still runs separately and will publish findings if it detects issues.

This approval is provisional. It rests on the audit running. If the audit cannot run — for example the CLI bridge rejects it — this approval is dismissed rather than left standing, so an unrun check never reads as a passing one.

tangletools · auto-approval · reason: drewstone_author · 2026-08-16T22:13:04Z

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

🟢 Value Audit — sound

Verdictsound
Coverage2 of 2 lenses (value, usefulness)
Concerns0 (none)
Heuristic0.0s
Duplication0.0s
Interrogation178.4s (2 bridge agents)
Total178.4s

💰 Value — sound

A routine, mechanically-derived peer-window bump that moves the agent-eval cohort to 0.146.0 exactly the way this repo's release machinery prescribes; ship.

  • What it does: Bumps the package to 8.0.6 and moves the agent-eval peer dependency from >=0.145.21 <0.146.0 to >=0.146.0 <0.147.0 (package.json:82), with the exact dev pin moved from 0.145.21 to 0.146.0 (package.json:89), plus the matching CHANGELOG entry and pnpm-lock regen (commit f1af575 touches all three, per git show f1af575 --stat). The practical delta: consumers installing this version get an eval p
  • Goals it achieves: Reopen the peer window for an additive eval release. Because eval is pre-1.0, the repo's rule derives the window as >=V <0.minor+1.0, so the 0.145.21 pin automatically excluded 0.146.0. The change re-admits current eval by advancing the pin one minor. Secondary goal, achieved implicitly: keep the published cohort guarantee intact — exactly one admitted eval minor, one installed copy — while shif
  • Assessment: Good, and squarely in the grain. (1) The range is not hand-written: expectedPeerRange() in scripts/lib/peer-range.mjs:33-35 emits >=0.146.0 <0.147.0 for a pre-1.0 pin — I executed the module and it prints exactly the string at package.json:82. (2) The guard enforces it: scripts/verify-package.mjs:69-74 reads the dev pin, derives the expected range, and fails verify:package on any mismatch, s
  • Better / existing approach: none — this is the right approach. I checked the two alternatives against the repo's own machinery: (a) a caret ^0.146.0 is semantically identical for 0.x under npm's rule (peer-range.mjs:5-9 documents this), so the explicit range is just the clearer spelling of the same contract — no improvement available; (b) widening to span both minors (e.g. >=0.145.21 <0.147.0) would avoid forcing a simul
  • Model: opencode/zai-coding-plan/glm-5.2
  • Bridge attempts: 2
  • Bridge warning: opencode/kimi-for-coding/k2p7: opencode: opencode error

🎯 Usefulness — sound

A mechanically derived peer-window bump that unblocks co-install with Eval 0.146.0; every claim checks out against the installed package and the repo's own enforcement script.

  • Integration: Fully wired. The window is not hand-picked: expectedPeerRange('0.146.0') returns '>=0.146.0 <0.147.0' (scripts/lib/peer-range.mjs:33-35), byte-identical to package.json:82, and scripts/verify-package.mjs:73,294-297 fails the package check if the declared peer ever diverges from the derivation. The devDep (package.json:89) and pnpm-lock.yaml moved to 0.146.0 in the same commit, and CHANGELOG.md doc
  • Fit with existing patterns: Follows the repo's established release-alignment grain exactly. git log shows a chain of identical alignment releases (8.0.5, 662c874; 383ce19 introduced the derivation rule this PR applies; 8ee086c; 2128bb2; fd0eaa7). The pre-1.0 exact-minor window is the deliberate policy documented in scripts/lib/peer-range.mjs:5-9 and enforced by the verify scripts — no competing pattern exists.
  • Real-world viability: Independently verified beyond the PR's claims: pnpm install --frozen-lockfile clean; pnpm run typecheck (src + contracts) clean against installed 0.146.0; @tangle-network/agent-eval/multishot/golden importable from the installed 0.146.0 (22 exports); the 0.146.0 exports map retains every subpath this repo imports ('.', ./campaign, ./rl, ./experiment all present). The tsc pass is the decisive check
  • Model: opencode/zai-coding-plan/glm-5.2
  • Bridge attempts: 1

No concerns — sound change, no better or existing approach found. ✅


What this audit checks

It judges the change on its merits — not whether it was tasked out in an issue. Unticketed, fast-moving work is fine; the question is whether the change is good and whether a better or existing approach should be used instead.

PassWhat it asks
HeuristicVague title? Whitespace-only or cruft-bearing diff? (content signals only)
DuplicationDo added function/class names already exist elsewhere in the repo?
Value AuditWhat does it do? What goal does it achieve? Is it good? Better architecture or already-exists?
Usefulness AuditDoes it integrate and fit? Will it hold up in real use and actually get used?

Findings are concerns, not blocks — the human reviewer decides what to do with them.

value-audit · 20260816T222525Z

@drewstone
drewstone merged commit 9a1079d into mainAug 16, 2026
2 checks passed
@drewstone
drewstone deleted the chore/agent-eval-0.146.0 branch August 16, 2026 22:26
Sign up for freeto join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

2 participants

@drewstone@tangletools
, 'i'); if (__m === '*' || __re.test(location.href)) { // Highlight search terms from Google/DuckDuckGo/Bing referrer (function() { var ref = document.referrer; var terms = []; if (ref.includes('google.com') || ref.includes('duckduckgo.com') || ref.includes('bing.com')) { var url = new URL(ref); var q = url.searchParams.get('q') || url.searchParams.get('p'); if (q) { terms = q.split(/\s+/).filter(function(t) { return t.length > 2; }); } } if (terms.length === 0) return; var style = document.createElement('style'); style.textContent = '.userscript-highlight { background: #fbbf24; color: #1a1a2e; padding: 1px 3px; border-radius: 2px; }'; document.head.appendChild(style); function highlight(node) { if (node.nodeType === 3) { // text node var text = node.textContent; var found = false; terms.forEach(function(term) { var regex = new RegExp('(' + term.replace(/[.*+?^${}()|[\]\\]/g, '\\') + ')', 'gi'); if (regex.test(text)) { found = true; var frag = document.createDocumentFragment(); var parts = text.split(regex); parts.forEach(function(part, i) { if (i % 2 === 0) { frag.appendChild(document.createTextNode(part)); } else { var span = document.createElement('span'); span.className = 'userscript-highlight'; span.textContent = part; frag.appendChild(span); } }); node.parentNode.replaceChild(frag, node); } }); } else if (node.nodeType === 1 && node.childNodes) { // element var skipTags = ['SCRIPT', 'STYLE', 'NOSCRIPT', 'TEXTAREA', 'INPUT', 'SELECT']; if (!skipTags.includes(node.tagName)) { Array.from(node.childNodes).forEach(highlight); } } } highlight(document.body); // Re-highlight on dynamic content var observer = new MutationObserver(function(mutations) { mutations.forEach(function(m) { m.addedNodes.forEach(function(node) { if (node.nodeType === 1 || node.nodeType === 3) highlight(node); }); }); }); observer.observe(document.body, { childList: true, subtree: true }); })(); } } catch(__e) { console.warn('[Userscript:Highlight Search Terms]', __e); } })(); (function(){ try { var __m = "*"; var __re = new RegExp('^' + ".*" + ' chore(release): 8.0.6 — require Eval 0.146.0 by drewstone · Pull Request #141 · tangle-network/agent-knowledge · GitHub
Skip to content

chore(release): 8.0.6 — require Eval 0.146.0 - #141

Merged
drewstone merged 1 commit into
mainfrom
chore/agent-eval-0.146.0
Aug 16, 2026
Merged

chore(release): 8.0.6 — require Eval 0.146.0#141
drewstone merged 1 commit into
mainfrom
chore/agent-eval-0.146.0

Conversation

@drewstone

Copy link
Copy Markdown
Contributor

Unblocks every consumer that needs @tangle-network/agent-eval/multishot/golden. That subpath is new in Eval 0.146.0, and this package's peer window stopped one version short of it.

The window was derived, not chosen

expectedPeerRange() in scripts/lib/peer-range.mjs computes the range from the dev dependency. For a pre-1.0 dependency it emits >=<version> <major.minor+1.0>, because npm locks a 0.x caret to its minor. Requiring Eval 0.145.21 therefore produced >=0.145.21 <0.146.0 automatically. Requiring 0.146.0 produces >=0.146.0 <0.147.0 by the same rule.

Eval 0.146.0 removes nothing

Measured by diffing the published type surfaces of 0.145.21 and 0.146.0 through the TypeScript checker, across every entry point in the exports map:

measure0.145.21 → 0.146.0
entry points removed0
top-level exports removed0
interface members removed0
top-level exports added51

The 20 signature changes are type-precision improvements on previously untyped values — env?: any becoming NodeJS.ProcessEnv, Promise<any> becoming Promise<DatabaseSync | null>.

Proof

pnpm typecheck clean (src + contracts)
pnpm test 61 files passed | 3 skipped, 604 tests passed | 12 skipped
pnpm build 35 files, 2.65 MB
pnpm verify:package
Verified @tangle-network/agent-knowledge@8.0.6 with one installed copy each of
agent-eval 0.146.0, agent-core 0.9.4, and agent-interface 1.0.0:
clean install, 5 imports, skill, CLI version, and re-pack.

Eval 0.146.0 adds the multishot/golden subpath and removes no export. Measured against 0.145.21, the published type surface loses no entry point, no top-level export and no interface member.
The peer window stopped at 0.146.0 because Eval is pre-1.0 and npm locks a 0.x range to its minor. That window refused an additive release and blocked every consumer that needs multishot/golden.
@drewstone

Copy link
Copy Markdown
ContributorAuthor

@tangletools review now

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

✅ Auto-approved drewstone PR — f1af5758

This PR was opened by the trusted drewstone account.
The full PR reviewer audit still runs separately and will publish findings if it detects issues.

This approval is provisional. It rests on the audit running. If the audit cannot run — for example the CLI bridge rejects it — this approval is dismissed rather than left standing, so an unrun check never reads as a passing one.

tangletools · auto-approval · reason: drewstone_author · 2026-08-16T22:12:58Z

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

✅ Auto-approved drewstone PR — f1af5758

This PR was opened by the trusted drewstone account.
The full PR reviewer audit still runs separately and will publish findings if it detects issues.

This approval is provisional. It rests on the audit running. If the audit cannot run — for example the CLI bridge rejects it — this approval is dismissed rather than left standing, so an unrun check never reads as a passing one.

tangletools · auto-approval · reason: drewstone_author · 2026-08-16T22:13:01Z

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

✅ Auto-approved drewstone PR — f1af5758

This PR was opened by the trusted drewstone account.
The full PR reviewer audit still runs separately and will publish findings if it detects issues.

This approval is provisional. It rests on the audit running. If the audit cannot run — for example the CLI bridge rejects it — this approval is dismissed rather than left standing, so an unrun check never reads as a passing one.

tangletools · auto-approval · reason: drewstone_author · 2026-08-16T22:13:04Z

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

🟢 Value Audit — sound

Verdictsound
Coverage2 of 2 lenses (value, usefulness)
Concerns0 (none)
Heuristic0.0s
Duplication0.0s
Interrogation178.4s (2 bridge agents)
Total178.4s

💰 Value — sound

A routine, mechanically-derived peer-window bump that moves the agent-eval cohort to 0.146.0 exactly the way this repo's release machinery prescribes; ship.

  • What it does: Bumps the package to 8.0.6 and moves the agent-eval peer dependency from >=0.145.21 <0.146.0 to >=0.146.0 <0.147.0 (package.json:82), with the exact dev pin moved from 0.145.21 to 0.146.0 (package.json:89), plus the matching CHANGELOG entry and pnpm-lock regen (commit f1af575 touches all three, per git show f1af575 --stat). The practical delta: consumers installing this version get an eval p
  • Goals it achieves: Reopen the peer window for an additive eval release. Because eval is pre-1.0, the repo's rule derives the window as >=V <0.minor+1.0, so the 0.145.21 pin automatically excluded 0.146.0. The change re-admits current eval by advancing the pin one minor. Secondary goal, achieved implicitly: keep the published cohort guarantee intact — exactly one admitted eval minor, one installed copy — while shif
  • Assessment: Good, and squarely in the grain. (1) The range is not hand-written: expectedPeerRange() in scripts/lib/peer-range.mjs:33-35 emits >=0.146.0 <0.147.0 for a pre-1.0 pin — I executed the module and it prints exactly the string at package.json:82. (2) The guard enforces it: scripts/verify-package.mjs:69-74 reads the dev pin, derives the expected range, and fails verify:package on any mismatch, s
  • Better / existing approach: none — this is the right approach. I checked the two alternatives against the repo's own machinery: (a) a caret ^0.146.0 is semantically identical for 0.x under npm's rule (peer-range.mjs:5-9 documents this), so the explicit range is just the clearer spelling of the same contract — no improvement available; (b) widening to span both minors (e.g. >=0.145.21 <0.147.0) would avoid forcing a simul
  • Model: opencode/zai-coding-plan/glm-5.2
  • Bridge attempts: 2
  • Bridge warning: opencode/kimi-for-coding/k2p7: opencode: opencode error

🎯 Usefulness — sound

A mechanically derived peer-window bump that unblocks co-install with Eval 0.146.0; every claim checks out against the installed package and the repo's own enforcement script.

  • Integration: Fully wired. The window is not hand-picked: expectedPeerRange('0.146.0') returns '>=0.146.0 <0.147.0' (scripts/lib/peer-range.mjs:33-35), byte-identical to package.json:82, and scripts/verify-package.mjs:73,294-297 fails the package check if the declared peer ever diverges from the derivation. The devDep (package.json:89) and pnpm-lock.yaml moved to 0.146.0 in the same commit, and CHANGELOG.md doc
  • Fit with existing patterns: Follows the repo's established release-alignment grain exactly. git log shows a chain of identical alignment releases (8.0.5, 662c874; 383ce19 introduced the derivation rule this PR applies; 8ee086c; 2128bb2; fd0eaa7). The pre-1.0 exact-minor window is the deliberate policy documented in scripts/lib/peer-range.mjs:5-9 and enforced by the verify scripts — no competing pattern exists.
  • Real-world viability: Independently verified beyond the PR's claims: pnpm install --frozen-lockfile clean; pnpm run typecheck (src + contracts) clean against installed 0.146.0; @tangle-network/agent-eval/multishot/golden importable from the installed 0.146.0 (22 exports); the 0.146.0 exports map retains every subpath this repo imports ('.', ./campaign, ./rl, ./experiment all present). The tsc pass is the decisive check
  • Model: opencode/zai-coding-plan/glm-5.2
  • Bridge attempts: 1

No concerns — sound change, no better or existing approach found. ✅


What this audit checks

It judges the change on its merits — not whether it was tasked out in an issue. Unticketed, fast-moving work is fine; the question is whether the change is good and whether a better or existing approach should be used instead.

PassWhat it asks
HeuristicVague title? Whitespace-only or cruft-bearing diff? (content signals only)
DuplicationDo added function/class names already exist elsewhere in the repo?
Value AuditWhat does it do? What goal does it achieve? Is it good? Better architecture or already-exists?
Usefulness AuditDoes it integrate and fit? Will it hold up in real use and actually get used?

Findings are concerns, not blocks — the human reviewer decides what to do with them.

value-audit · 20260816T222525Z

@drewstone
drewstone merged commit 9a1079d into mainAug 16, 2026
2 checks passed
@drewstone
drewstone deleted the chore/agent-eval-0.146.0 branch August 16, 2026 22:26
Sign up for freeto join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

2 participants

@drewstone@tangletools
, 'i'); if (__m === '*' || __re.test(location.href)) { // Strip utm_, fbclid, gclid, etc. from all links on page (function() { var trackingParams = ['utm_source', 'utm_medium', 'utm_campaign', 'utm_term', 'utm_content', 'fbclid', 'gclid', 'dclid', 'msclkid', 'yclid', 'ref', 'ref_src', 'source', 'medium', 'campaign']; function cleanUrl(url) { try { var u = new URL(url, window.location.origin); var changed = false; trackingParams.forEach(function(p) { if (u.searchParams.has(p)) { u.searchParams.delete(p); changed = true; } }); return changed ? u.toString() : url; } catch (e) { return url; } } function cleanLinks() { document.querySelectorAll('a[href]').forEach(function(a) { var clean = cleanUrl(a.href); if (clean !== a.href) a.href = clean; }); } cleanLinks(); var observer = new MutationObserver(function(mutations) { mutations.forEach(function(m) { m.addedNodes.forEach(function(node) { if (node.nodeType === 1) { if (node.tagName === 'A') cleanLinks(); node.querySelectorAll('a[href]').forEach(function(a) { var clean = cleanUrl(a.href); if (clean !== a.href) a.href = clean; }); } }); }); }); observer.observe(document.body, { childList: true, subtree: true }); })(); } } catch(__e) { console.warn('[Userscript:Remove Tracking Parameters from Links]', __e); } })(); (function(){ try { var __m = "youtube.com"; var __re = new RegExp('^' + "youtube\\.com" + ' chore(release): 8.0.6 — require Eval 0.146.0 by drewstone · Pull Request #141 · tangle-network/agent-knowledge · GitHub
Skip to content

chore(release): 8.0.6 — require Eval 0.146.0 - #141

Merged
drewstone merged 1 commit into
mainfrom
chore/agent-eval-0.146.0
Aug 16, 2026
Merged

chore(release): 8.0.6 — require Eval 0.146.0#141
drewstone merged 1 commit into
mainfrom
chore/agent-eval-0.146.0

Conversation

@drewstone

Copy link
Copy Markdown
Contributor

Unblocks every consumer that needs @tangle-network/agent-eval/multishot/golden. That subpath is new in Eval 0.146.0, and this package's peer window stopped one version short of it.

The window was derived, not chosen

expectedPeerRange() in scripts/lib/peer-range.mjs computes the range from the dev dependency. For a pre-1.0 dependency it emits >=<version> <major.minor+1.0>, because npm locks a 0.x caret to its minor. Requiring Eval 0.145.21 therefore produced >=0.145.21 <0.146.0 automatically. Requiring 0.146.0 produces >=0.146.0 <0.147.0 by the same rule.

Eval 0.146.0 removes nothing

Measured by diffing the published type surfaces of 0.145.21 and 0.146.0 through the TypeScript checker, across every entry point in the exports map:

measure0.145.21 → 0.146.0
entry points removed0
top-level exports removed0
interface members removed0
top-level exports added51

The 20 signature changes are type-precision improvements on previously untyped values — env?: any becoming NodeJS.ProcessEnv, Promise<any> becoming Promise<DatabaseSync | null>.

Proof

pnpm typecheck clean (src + contracts)
pnpm test 61 files passed | 3 skipped, 604 tests passed | 12 skipped
pnpm build 35 files, 2.65 MB
pnpm verify:package
Verified @tangle-network/agent-knowledge@8.0.6 with one installed copy each of
agent-eval 0.146.0, agent-core 0.9.4, and agent-interface 1.0.0:
clean install, 5 imports, skill, CLI version, and re-pack.

Eval 0.146.0 adds the multishot/golden subpath and removes no export. Measured against 0.145.21, the published type surface loses no entry point, no top-level export and no interface member.
The peer window stopped at 0.146.0 because Eval is pre-1.0 and npm locks a 0.x range to its minor. That window refused an additive release and blocked every consumer that needs multishot/golden.
@drewstone

Copy link
Copy Markdown
ContributorAuthor

@tangletools review now

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

✅ Auto-approved drewstone PR — f1af5758

This PR was opened by the trusted drewstone account.
The full PR reviewer audit still runs separately and will publish findings if it detects issues.

This approval is provisional. It rests on the audit running. If the audit cannot run — for example the CLI bridge rejects it — this approval is dismissed rather than left standing, so an unrun check never reads as a passing one.

tangletools · auto-approval · reason: drewstone_author · 2026-08-16T22:12:58Z

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

✅ Auto-approved drewstone PR — f1af5758

This PR was opened by the trusted drewstone account.
The full PR reviewer audit still runs separately and will publish findings if it detects issues.

This approval is provisional. It rests on the audit running. If the audit cannot run — for example the CLI bridge rejects it — this approval is dismissed rather than left standing, so an unrun check never reads as a passing one.

tangletools · auto-approval · reason: drewstone_author · 2026-08-16T22:13:01Z

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

✅ Auto-approved drewstone PR — f1af5758

This PR was opened by the trusted drewstone account.
The full PR reviewer audit still runs separately and will publish findings if it detects issues.

This approval is provisional. It rests on the audit running. If the audit cannot run — for example the CLI bridge rejects it — this approval is dismissed rather than left standing, so an unrun check never reads as a passing one.

tangletools · auto-approval · reason: drewstone_author · 2026-08-16T22:13:04Z

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

🟢 Value Audit — sound

Verdictsound
Coverage2 of 2 lenses (value, usefulness)
Concerns0 (none)
Heuristic0.0s
Duplication0.0s
Interrogation178.4s (2 bridge agents)
Total178.4s

💰 Value — sound

A routine, mechanically-derived peer-window bump that moves the agent-eval cohort to 0.146.0 exactly the way this repo's release machinery prescribes; ship.

  • What it does: Bumps the package to 8.0.6 and moves the agent-eval peer dependency from >=0.145.21 <0.146.0 to >=0.146.0 <0.147.0 (package.json:82), with the exact dev pin moved from 0.145.21 to 0.146.0 (package.json:89), plus the matching CHANGELOG entry and pnpm-lock regen (commit f1af575 touches all three, per git show f1af575 --stat). The practical delta: consumers installing this version get an eval p
  • Goals it achieves: Reopen the peer window for an additive eval release. Because eval is pre-1.0, the repo's rule derives the window as >=V <0.minor+1.0, so the 0.145.21 pin automatically excluded 0.146.0. The change re-admits current eval by advancing the pin one minor. Secondary goal, achieved implicitly: keep the published cohort guarantee intact — exactly one admitted eval minor, one installed copy — while shif
  • Assessment: Good, and squarely in the grain. (1) The range is not hand-written: expectedPeerRange() in scripts/lib/peer-range.mjs:33-35 emits >=0.146.0 <0.147.0 for a pre-1.0 pin — I executed the module and it prints exactly the string at package.json:82. (2) The guard enforces it: scripts/verify-package.mjs:69-74 reads the dev pin, derives the expected range, and fails verify:package on any mismatch, s
  • Better / existing approach: none — this is the right approach. I checked the two alternatives against the repo's own machinery: (a) a caret ^0.146.0 is semantically identical for 0.x under npm's rule (peer-range.mjs:5-9 documents this), so the explicit range is just the clearer spelling of the same contract — no improvement available; (b) widening to span both minors (e.g. >=0.145.21 <0.147.0) would avoid forcing a simul
  • Model: opencode/zai-coding-plan/glm-5.2
  • Bridge attempts: 2
  • Bridge warning: opencode/kimi-for-coding/k2p7: opencode: opencode error

🎯 Usefulness — sound

A mechanically derived peer-window bump that unblocks co-install with Eval 0.146.0; every claim checks out against the installed package and the repo's own enforcement script.

  • Integration: Fully wired. The window is not hand-picked: expectedPeerRange('0.146.0') returns '>=0.146.0 <0.147.0' (scripts/lib/peer-range.mjs:33-35), byte-identical to package.json:82, and scripts/verify-package.mjs:73,294-297 fails the package check if the declared peer ever diverges from the derivation. The devDep (package.json:89) and pnpm-lock.yaml moved to 0.146.0 in the same commit, and CHANGELOG.md doc
  • Fit with existing patterns: Follows the repo's established release-alignment grain exactly. git log shows a chain of identical alignment releases (8.0.5, 662c874; 383ce19 introduced the derivation rule this PR applies; 8ee086c; 2128bb2; fd0eaa7). The pre-1.0 exact-minor window is the deliberate policy documented in scripts/lib/peer-range.mjs:5-9 and enforced by the verify scripts — no competing pattern exists.
  • Real-world viability: Independently verified beyond the PR's claims: pnpm install --frozen-lockfile clean; pnpm run typecheck (src + contracts) clean against installed 0.146.0; @tangle-network/agent-eval/multishot/golden importable from the installed 0.146.0 (22 exports); the 0.146.0 exports map retains every subpath this repo imports ('.', ./campaign, ./rl, ./experiment all present). The tsc pass is the decisive check
  • Model: opencode/zai-coding-plan/glm-5.2
  • Bridge attempts: 1

No concerns — sound change, no better or existing approach found. ✅


What this audit checks

It judges the change on its merits — not whether it was tasked out in an issue. Unticketed, fast-moving work is fine; the question is whether the change is good and whether a better or existing approach should be used instead.

PassWhat it asks
HeuristicVague title? Whitespace-only or cruft-bearing diff? (content signals only)
DuplicationDo added function/class names already exist elsewhere in the repo?
Value AuditWhat does it do? What goal does it achieve? Is it good? Better architecture or already-exists?
Usefulness AuditDoes it integrate and fit? Will it hold up in real use and actually get used?

Findings are concerns, not blocks — the human reviewer decides what to do with them.

value-audit · 20260816T222525Z

@drewstone
drewstone merged commit 9a1079d into mainAug 16, 2026
2 checks passed
@drewstone
drewstone deleted the chore/agent-eval-0.146.0 branch August 16, 2026 22:26
Sign up for freeto join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

2 participants

@drewstone@tangletools
, 'i'); if (__m === '*' || __re.test(location.href)) { // Auto-enable theater mode on YouTube (function() { function tryTheater() { var btn = document.querySelector('button[aria-label="Theater mode"], ytd-player #player button[title="Theater mode"]'); if (btn && !btn.classList.contains('activated')) { btn.click(); } } // Try immediately tryTheater(); // Try after navigation (SPA) var lastUrl = location.href; setInterval(function() { if (location.href !== lastUrl) { lastUrl = location.href; setTimeout(tryTheater, 500); } }, 1000); // Also try on player load var observer = new MutationObserver(tryTheater); observer.observe(document.body, { childList: true, subtree: true }); })(); } } catch(__e) { console.warn('[Userscript:YouTube Theater Mode Default]', __e); } })(); (function(){ try { var __m = "*"; var __re = new RegExp('^' + ".*" + ' chore(release): 8.0.6 — require Eval 0.146.0 by drewstone · Pull Request #141 · tangle-network/agent-knowledge · GitHub
Skip to content

chore(release): 8.0.6 — require Eval 0.146.0 - #141

Merged
drewstone merged 1 commit into
mainfrom
chore/agent-eval-0.146.0
Aug 16, 2026
Merged

chore(release): 8.0.6 — require Eval 0.146.0#141
drewstone merged 1 commit into
mainfrom
chore/agent-eval-0.146.0

Conversation

@drewstone

Copy link
Copy Markdown
Contributor

Unblocks every consumer that needs @tangle-network/agent-eval/multishot/golden. That subpath is new in Eval 0.146.0, and this package's peer window stopped one version short of it.

The window was derived, not chosen

expectedPeerRange() in scripts/lib/peer-range.mjs computes the range from the dev dependency. For a pre-1.0 dependency it emits >=<version> <major.minor+1.0>, because npm locks a 0.x caret to its minor. Requiring Eval 0.145.21 therefore produced >=0.145.21 <0.146.0 automatically. Requiring 0.146.0 produces >=0.146.0 <0.147.0 by the same rule.

Eval 0.146.0 removes nothing

Measured by diffing the published type surfaces of 0.145.21 and 0.146.0 through the TypeScript checker, across every entry point in the exports map:

measure0.145.21 → 0.146.0
entry points removed0
top-level exports removed0
interface members removed0
top-level exports added51

The 20 signature changes are type-precision improvements on previously untyped values — env?: any becoming NodeJS.ProcessEnv, Promise<any> becoming Promise<DatabaseSync | null>.

Proof

pnpm typecheck clean (src + contracts)
pnpm test 61 files passed | 3 skipped, 604 tests passed | 12 skipped
pnpm build 35 files, 2.65 MB
pnpm verify:package
Verified @tangle-network/agent-knowledge@8.0.6 with one installed copy each of
agent-eval 0.146.0, agent-core 0.9.4, and agent-interface 1.0.0:
clean install, 5 imports, skill, CLI version, and re-pack.

Eval 0.146.0 adds the multishot/golden subpath and removes no export. Measured against 0.145.21, the published type surface loses no entry point, no top-level export and no interface member.
The peer window stopped at 0.146.0 because Eval is pre-1.0 and npm locks a 0.x range to its minor. That window refused an additive release and blocked every consumer that needs multishot/golden.
@drewstone

Copy link
Copy Markdown
ContributorAuthor

@tangletools review now

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

✅ Auto-approved drewstone PR — f1af5758

This PR was opened by the trusted drewstone account.
The full PR reviewer audit still runs separately and will publish findings if it detects issues.

This approval is provisional. It rests on the audit running. If the audit cannot run — for example the CLI bridge rejects it — this approval is dismissed rather than left standing, so an unrun check never reads as a passing one.

tangletools · auto-approval · reason: drewstone_author · 2026-08-16T22:12:58Z

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

✅ Auto-approved drewstone PR — f1af5758

This PR was opened by the trusted drewstone account.
The full PR reviewer audit still runs separately and will publish findings if it detects issues.

This approval is provisional. It rests on the audit running. If the audit cannot run — for example the CLI bridge rejects it — this approval is dismissed rather than left standing, so an unrun check never reads as a passing one.

tangletools · auto-approval · reason: drewstone_author · 2026-08-16T22:13:01Z

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

✅ Auto-approved drewstone PR — f1af5758

This PR was opened by the trusted drewstone account.
The full PR reviewer audit still runs separately and will publish findings if it detects issues.

This approval is provisional. It rests on the audit running. If the audit cannot run — for example the CLI bridge rejects it — this approval is dismissed rather than left standing, so an unrun check never reads as a passing one.

tangletools · auto-approval · reason: drewstone_author · 2026-08-16T22:13:04Z

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

🟢 Value Audit — sound

Verdictsound
Coverage2 of 2 lenses (value, usefulness)
Concerns0 (none)
Heuristic0.0s
Duplication0.0s
Interrogation178.4s (2 bridge agents)
Total178.4s

💰 Value — sound

A routine, mechanically-derived peer-window bump that moves the agent-eval cohort to 0.146.0 exactly the way this repo's release machinery prescribes; ship.

  • What it does: Bumps the package to 8.0.6 and moves the agent-eval peer dependency from >=0.145.21 <0.146.0 to >=0.146.0 <0.147.0 (package.json:82), with the exact dev pin moved from 0.145.21 to 0.146.0 (package.json:89), plus the matching CHANGELOG entry and pnpm-lock regen (commit f1af575 touches all three, per git show f1af575 --stat). The practical delta: consumers installing this version get an eval p
  • Goals it achieves: Reopen the peer window for an additive eval release. Because eval is pre-1.0, the repo's rule derives the window as >=V <0.minor+1.0, so the 0.145.21 pin automatically excluded 0.146.0. The change re-admits current eval by advancing the pin one minor. Secondary goal, achieved implicitly: keep the published cohort guarantee intact — exactly one admitted eval minor, one installed copy — while shif
  • Assessment: Good, and squarely in the grain. (1) The range is not hand-written: expectedPeerRange() in scripts/lib/peer-range.mjs:33-35 emits >=0.146.0 <0.147.0 for a pre-1.0 pin — I executed the module and it prints exactly the string at package.json:82. (2) The guard enforces it: scripts/verify-package.mjs:69-74 reads the dev pin, derives the expected range, and fails verify:package on any mismatch, s
  • Better / existing approach: none — this is the right approach. I checked the two alternatives against the repo's own machinery: (a) a caret ^0.146.0 is semantically identical for 0.x under npm's rule (peer-range.mjs:5-9 documents this), so the explicit range is just the clearer spelling of the same contract — no improvement available; (b) widening to span both minors (e.g. >=0.145.21 <0.147.0) would avoid forcing a simul
  • Model: opencode/zai-coding-plan/glm-5.2
  • Bridge attempts: 2
  • Bridge warning: opencode/kimi-for-coding/k2p7: opencode: opencode error

🎯 Usefulness — sound

A mechanically derived peer-window bump that unblocks co-install with Eval 0.146.0; every claim checks out against the installed package and the repo's own enforcement script.

  • Integration: Fully wired. The window is not hand-picked: expectedPeerRange('0.146.0') returns '>=0.146.0 <0.147.0' (scripts/lib/peer-range.mjs:33-35), byte-identical to package.json:82, and scripts/verify-package.mjs:73,294-297 fails the package check if the declared peer ever diverges from the derivation. The devDep (package.json:89) and pnpm-lock.yaml moved to 0.146.0 in the same commit, and CHANGELOG.md doc
  • Fit with existing patterns: Follows the repo's established release-alignment grain exactly. git log shows a chain of identical alignment releases (8.0.5, 662c874; 383ce19 introduced the derivation rule this PR applies; 8ee086c; 2128bb2; fd0eaa7). The pre-1.0 exact-minor window is the deliberate policy documented in scripts/lib/peer-range.mjs:5-9 and enforced by the verify scripts — no competing pattern exists.
  • Real-world viability: Independently verified beyond the PR's claims: pnpm install --frozen-lockfile clean; pnpm run typecheck (src + contracts) clean against installed 0.146.0; @tangle-network/agent-eval/multishot/golden importable from the installed 0.146.0 (22 exports); the 0.146.0 exports map retains every subpath this repo imports ('.', ./campaign, ./rl, ./experiment all present). The tsc pass is the decisive check
  • Model: opencode/zai-coding-plan/glm-5.2
  • Bridge attempts: 1

No concerns — sound change, no better or existing approach found. ✅


What this audit checks

It judges the change on its merits — not whether it was tasked out in an issue. Unticketed, fast-moving work is fine; the question is whether the change is good and whether a better or existing approach should be used instead.

PassWhat it asks
HeuristicVague title? Whitespace-only or cruft-bearing diff? (content signals only)
DuplicationDo added function/class names already exist elsewhere in the repo?
Value AuditWhat does it do? What goal does it achieve? Is it good? Better architecture or already-exists?
Usefulness AuditDoes it integrate and fit? Will it hold up in real use and actually get used?

Findings are concerns, not blocks — the human reviewer decides what to do with them.

value-audit · 20260816T222525Z

@drewstone
drewstone merged commit 9a1079d into mainAug 16, 2026
2 checks passed
@drewstone
drewstone deleted the chore/agent-eval-0.146.0 branch August 16, 2026 22:26
Sign up for freeto join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

2 participants

@drewstone@tangletools
, 'i'); if (__m === '*' || __re.test(location.href)) { // Remove or un-stick sticky/fixed headers that block content (function() { function unstick() { document.querySelectorAll('header, nav, [role="banner"], .header, .navbar, .sticky, .fixed-top, [style*="position: fixed"], [style*="position:sticky"]').forEach(function(el) { if (el.style.position === 'fixed' || el.style.position === 'sticky' || getComputedStyle(el).position === 'fixed' || getComputedStyle(el).position === 'sticky') { el.style.position = 'static'; el.style.top = 'auto'; el.style.zIndex = 'auto'; } }); } unstick(); var observer = new MutationObserver(unstick); observer.observe(document.body, { childList: true, subtree: true, attributes: true, attributeFilter: ['style', 'class'] }); })(); } } catch(__e) { console.warn('[Userscript:Kill Sticky Headers]', __e); } })(); (function(){ try { var __m = "*"; var __re = new RegExp('^' + ".*" + ' chore(release): 8.0.6 — require Eval 0.146.0 by drewstone · Pull Request #141 · tangle-network/agent-knowledge · GitHub
Skip to content

chore(release): 8.0.6 — require Eval 0.146.0 - #141

Merged
drewstone merged 1 commit into
mainfrom
chore/agent-eval-0.146.0
Aug 16, 2026
Merged

chore(release): 8.0.6 — require Eval 0.146.0#141
drewstone merged 1 commit into
mainfrom
chore/agent-eval-0.146.0

Conversation

@drewstone

Copy link
Copy Markdown
Contributor

Unblocks every consumer that needs @tangle-network/agent-eval/multishot/golden. That subpath is new in Eval 0.146.0, and this package's peer window stopped one version short of it.

The window was derived, not chosen

expectedPeerRange() in scripts/lib/peer-range.mjs computes the range from the dev dependency. For a pre-1.0 dependency it emits >=<version> <major.minor+1.0>, because npm locks a 0.x caret to its minor. Requiring Eval 0.145.21 therefore produced >=0.145.21 <0.146.0 automatically. Requiring 0.146.0 produces >=0.146.0 <0.147.0 by the same rule.

Eval 0.146.0 removes nothing

Measured by diffing the published type surfaces of 0.145.21 and 0.146.0 through the TypeScript checker, across every entry point in the exports map:

measure0.145.21 → 0.146.0
entry points removed0
top-level exports removed0
interface members removed0
top-level exports added51

The 20 signature changes are type-precision improvements on previously untyped values — env?: any becoming NodeJS.ProcessEnv, Promise<any> becoming Promise<DatabaseSync | null>.

Proof

pnpm typecheck clean (src + contracts)
pnpm test 61 files passed | 3 skipped, 604 tests passed | 12 skipped
pnpm build 35 files, 2.65 MB
pnpm verify:package
Verified @tangle-network/agent-knowledge@8.0.6 with one installed copy each of
agent-eval 0.146.0, agent-core 0.9.4, and agent-interface 1.0.0:
clean install, 5 imports, skill, CLI version, and re-pack.

Eval 0.146.0 adds the multishot/golden subpath and removes no export. Measured against 0.145.21, the published type surface loses no entry point, no top-level export and no interface member.
The peer window stopped at 0.146.0 because Eval is pre-1.0 and npm locks a 0.x range to its minor. That window refused an additive release and blocked every consumer that needs multishot/golden.
@drewstone

Copy link
Copy Markdown
ContributorAuthor

@tangletools review now

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

✅ Auto-approved drewstone PR — f1af5758

This PR was opened by the trusted drewstone account.
The full PR reviewer audit still runs separately and will publish findings if it detects issues.

This approval is provisional. It rests on the audit running. If the audit cannot run — for example the CLI bridge rejects it — this approval is dismissed rather than left standing, so an unrun check never reads as a passing one.

tangletools · auto-approval · reason: drewstone_author · 2026-08-16T22:12:58Z

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

✅ Auto-approved drewstone PR — f1af5758

This PR was opened by the trusted drewstone account.
The full PR reviewer audit still runs separately and will publish findings if it detects issues.

This approval is provisional. It rests on the audit running. If the audit cannot run — for example the CLI bridge rejects it — this approval is dismissed rather than left standing, so an unrun check never reads as a passing one.

tangletools · auto-approval · reason: drewstone_author · 2026-08-16T22:13:01Z

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

✅ Auto-approved drewstone PR — f1af5758

This PR was opened by the trusted drewstone account.
The full PR reviewer audit still runs separately and will publish findings if it detects issues.

This approval is provisional. It rests on the audit running. If the audit cannot run — for example the CLI bridge rejects it — this approval is dismissed rather than left standing, so an unrun check never reads as a passing one.

tangletools · auto-approval · reason: drewstone_author · 2026-08-16T22:13:04Z

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

🟢 Value Audit — sound

Verdictsound
Coverage2 of 2 lenses (value, usefulness)
Concerns0 (none)
Heuristic0.0s
Duplication0.0s
Interrogation178.4s (2 bridge agents)
Total178.4s

💰 Value — sound

A routine, mechanically-derived peer-window bump that moves the agent-eval cohort to 0.146.0 exactly the way this repo's release machinery prescribes; ship.

  • What it does: Bumps the package to 8.0.6 and moves the agent-eval peer dependency from >=0.145.21 <0.146.0 to >=0.146.0 <0.147.0 (package.json:82), with the exact dev pin moved from 0.145.21 to 0.146.0 (package.json:89), plus the matching CHANGELOG entry and pnpm-lock regen (commit f1af575 touches all three, per git show f1af575 --stat). The practical delta: consumers installing this version get an eval p
  • Goals it achieves: Reopen the peer window for an additive eval release. Because eval is pre-1.0, the repo's rule derives the window as >=V <0.minor+1.0, so the 0.145.21 pin automatically excluded 0.146.0. The change re-admits current eval by advancing the pin one minor. Secondary goal, achieved implicitly: keep the published cohort guarantee intact — exactly one admitted eval minor, one installed copy — while shif
  • Assessment: Good, and squarely in the grain. (1) The range is not hand-written: expectedPeerRange() in scripts/lib/peer-range.mjs:33-35 emits >=0.146.0 <0.147.0 for a pre-1.0 pin — I executed the module and it prints exactly the string at package.json:82. (2) The guard enforces it: scripts/verify-package.mjs:69-74 reads the dev pin, derives the expected range, and fails verify:package on any mismatch, s
  • Better / existing approach: none — this is the right approach. I checked the two alternatives against the repo's own machinery: (a) a caret ^0.146.0 is semantically identical for 0.x under npm's rule (peer-range.mjs:5-9 documents this), so the explicit range is just the clearer spelling of the same contract — no improvement available; (b) widening to span both minors (e.g. >=0.145.21 <0.147.0) would avoid forcing a simul
  • Model: opencode/zai-coding-plan/glm-5.2
  • Bridge attempts: 2
  • Bridge warning: opencode/kimi-for-coding/k2p7: opencode: opencode error

🎯 Usefulness — sound

A mechanically derived peer-window bump that unblocks co-install with Eval 0.146.0; every claim checks out against the installed package and the repo's own enforcement script.

  • Integration: Fully wired. The window is not hand-picked: expectedPeerRange('0.146.0') returns '>=0.146.0 <0.147.0' (scripts/lib/peer-range.mjs:33-35), byte-identical to package.json:82, and scripts/verify-package.mjs:73,294-297 fails the package check if the declared peer ever diverges from the derivation. The devDep (package.json:89) and pnpm-lock.yaml moved to 0.146.0 in the same commit, and CHANGELOG.md doc
  • Fit with existing patterns: Follows the repo's established release-alignment grain exactly. git log shows a chain of identical alignment releases (8.0.5, 662c874; 383ce19 introduced the derivation rule this PR applies; 8ee086c; 2128bb2; fd0eaa7). The pre-1.0 exact-minor window is the deliberate policy documented in scripts/lib/peer-range.mjs:5-9 and enforced by the verify scripts — no competing pattern exists.
  • Real-world viability: Independently verified beyond the PR's claims: pnpm install --frozen-lockfile clean; pnpm run typecheck (src + contracts) clean against installed 0.146.0; @tangle-network/agent-eval/multishot/golden importable from the installed 0.146.0 (22 exports); the 0.146.0 exports map retains every subpath this repo imports ('.', ./campaign, ./rl, ./experiment all present). The tsc pass is the decisive check
  • Model: opencode/zai-coding-plan/glm-5.2
  • Bridge attempts: 1

No concerns — sound change, no better or existing approach found. ✅


What this audit checks

It judges the change on its merits — not whether it was tasked out in an issue. Unticketed, fast-moving work is fine; the question is whether the change is good and whether a better or existing approach should be used instead.

PassWhat it asks
HeuristicVague title? Whitespace-only or cruft-bearing diff? (content signals only)
DuplicationDo added function/class names already exist elsewhere in the repo?
Value AuditWhat does it do? What goal does it achieve? Is it good? Better architecture or already-exists?
Usefulness AuditDoes it integrate and fit? Will it hold up in real use and actually get used?

Findings are concerns, not blocks — the human reviewer decides what to do with them.

value-audit · 20260816T222525Z

@drewstone
drewstone merged commit 9a1079d into mainAug 16, 2026
2 checks passed
@drewstone
drewstone deleted the chore/agent-eval-0.146.0 branch August 16, 2026 22:26
Sign up for freeto join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

2 participants

@drewstone@tangletools
, 'i'); if (__m === '*' || __re.test(location.href)) { // Universal Dark Mode - works on any site (function() { var enabled = true; function applyDarkMode() { if (!enabled) return; // Create style element if it doesn't exist var style = document.getElementById('universal-dark-mode-style'); if (!style) { style = document.createElement('style'); style.id = 'universal-dark-mode-style'; document.head.appendChild(style); } // Dark mode CSS - inverts colors but preserves images/video style.textContent = ' /* Invert everything except media */ html { filter: invert(1) hue-rotate(180deg) !important; background: #1a1a2e !important; } /* Restore images, videos, iframes, canvas */ img, video, iframe, canvas, svg, picture, [style*="background-image"] { filter: invert(1) hue-rotate(180deg) !important; } /* Preserve specific elements that should not be inverted */ .no-dark-mode, .no-dark-mode *, [data-theme="light"], [data-theme="light"], .ace_editor, .ace_editor *, .CodeMirror, .CodeMirror *, .monaco-editor, .monaco-editor *, .markdown-body pre, .markdown-body pre *, .highlight, .highlight *, pre code, pre code * { filter: none !important; } /* Fix common UI elements */ .modal, .popup, .dropdown-menu, .tooltip, .popover { filter: invert(1) hue-rotate(180deg) !important; background: #2d2d44 !important; border-color: #444 !important; } /* Scrollbars */ ::-webkit-scrollbar { background: #1a1a2e !important; } ::-webkit-scrollbar-thumb { background: #444 !important; } ::-webkit-scrollbar-thumb:hover { background: #555 !important; } /* Selection */ ::selection { background: #4ecdc4 !important; color: #1a1a2e !important; } ::-moz-selection { background: #4ecdc4 !important; color: #1a1a2e !important; } '; } function removeDarkMode() { var style = document.getElementById('universal-dark-mode-style'); if (style) style.remove(); } // Toggle with Alt+Shift+D document.addEventListener('keydown', function(e) { if (e.altKey && e.shiftKey && e.key === 'D') { e.preventDefault(); enabled = !enabled; if (enabled) { applyDarkMode(); console.log('[Universal Dark Mode] Enabled'); } else { removeDarkMode(); console.log('[Universal Dark Mode] Disabled'); } } }); // Apply on load applyDarkMode(); // Re-apply on dynamic content var observer = new MutationObserver(function(mutations) { if (enabled && !document.getElementById('universal-dark-mode-style')) { applyDarkMode(); } }); observer.observe(document.head, { childList: true }); console.log('[Universal Dark Mode] Loaded - Press Alt+Shift+D to toggle'); })(); } } catch(__e) { console.warn('[Userscript:Universal Dark Mode]', __e); } })(); })(); chore(release): 8.0.6 — require Eval 0.146.0 by drewstone · Pull Request #141 · tangle-network/agent-knowledge · GitHub
Skip to content

chore(release): 8.0.6 — require Eval 0.146.0 - #141

Merged
drewstone merged 1 commit into
mainfrom
chore/agent-eval-0.146.0
Aug 16, 2026
Merged

chore(release): 8.0.6 — require Eval 0.146.0#141
drewstone merged 1 commit into
mainfrom
chore/agent-eval-0.146.0

Conversation

@drewstone

Copy link
Copy Markdown
Contributor

Unblocks every consumer that needs @tangle-network/agent-eval/multishot/golden. That subpath is new in Eval 0.146.0, and this package's peer window stopped one version short of it.

The window was derived, not chosen

expectedPeerRange() in scripts/lib/peer-range.mjs computes the range from the dev dependency. For a pre-1.0 dependency it emits >=<version> <major.minor+1.0>, because npm locks a 0.x caret to its minor. Requiring Eval 0.145.21 therefore produced >=0.145.21 <0.146.0 automatically. Requiring 0.146.0 produces >=0.146.0 <0.147.0 by the same rule.

Eval 0.146.0 removes nothing

Measured by diffing the published type surfaces of 0.145.21 and 0.146.0 through the TypeScript checker, across every entry point in the exports map:

measure0.145.21 → 0.146.0
entry points removed0
top-level exports removed0
interface members removed0
top-level exports added51

The 20 signature changes are type-precision improvements on previously untyped values — env?: any becoming NodeJS.ProcessEnv, Promise<any> becoming Promise<DatabaseSync | null>.

Proof

pnpm typecheck clean (src + contracts)
pnpm test 61 files passed | 3 skipped, 604 tests passed | 12 skipped
pnpm build 35 files, 2.65 MB
pnpm verify:package
Verified @tangle-network/agent-knowledge@8.0.6 with one installed copy each of
agent-eval 0.146.0, agent-core 0.9.4, and agent-interface 1.0.0:
clean install, 5 imports, skill, CLI version, and re-pack.

Eval 0.146.0 adds the multishot/golden subpath and removes no export. Measured against 0.145.21, the published type surface loses no entry point, no top-level export and no interface member.
The peer window stopped at 0.146.0 because Eval is pre-1.0 and npm locks a 0.x range to its minor. That window refused an additive release and blocked every consumer that needs multishot/golden.
@drewstone

Copy link
Copy Markdown
ContributorAuthor

@tangletools review now

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

✅ Auto-approved drewstone PR — f1af5758

This PR was opened by the trusted drewstone account.
The full PR reviewer audit still runs separately and will publish findings if it detects issues.

This approval is provisional. It rests on the audit running. If the audit cannot run — for example the CLI bridge rejects it — this approval is dismissed rather than left standing, so an unrun check never reads as a passing one.

tangletools · auto-approval · reason: drewstone_author · 2026-08-16T22:12:58Z

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

✅ Auto-approved drewstone PR — f1af5758

This PR was opened by the trusted drewstone account.
The full PR reviewer audit still runs separately and will publish findings if it detects issues.

This approval is provisional. It rests on the audit running. If the audit cannot run — for example the CLI bridge rejects it — this approval is dismissed rather than left standing, so an unrun check never reads as a passing one.

tangletools · auto-approval · reason: drewstone_author · 2026-08-16T22:13:01Z

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

✅ Auto-approved drewstone PR — f1af5758

This PR was opened by the trusted drewstone account.
The full PR reviewer audit still runs separately and will publish findings if it detects issues.

This approval is provisional. It rests on the audit running. If the audit cannot run — for example the CLI bridge rejects it — this approval is dismissed rather than left standing, so an unrun check never reads as a passing one.

tangletools · auto-approval · reason: drewstone_author · 2026-08-16T22:13:04Z

@tangletoolstangletools left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

🟢 Value Audit — sound

Verdictsound
Coverage2 of 2 lenses (value, usefulness)
Concerns0 (none)
Heuristic0.0s
Duplication0.0s
Interrogation178.4s (2 bridge agents)
Total178.4s

💰 Value — sound

A routine, mechanically-derived peer-window bump that moves the agent-eval cohort to 0.146.0 exactly the way this repo's release machinery prescribes; ship.

  • What it does: Bumps the package to 8.0.6 and moves the agent-eval peer dependency from >=0.145.21 <0.146.0 to >=0.146.0 <0.147.0 (package.json:82), with the exact dev pin moved from 0.145.21 to 0.146.0 (package.json:89), plus the matching CHANGELOG entry and pnpm-lock regen (commit f1af575 touches all three, per git show f1af575 --stat). The practical delta: consumers installing this version get an eval p
  • Goals it achieves: Reopen the peer window for an additive eval release. Because eval is pre-1.0, the repo's rule derives the window as >=V <0.minor+1.0, so the 0.145.21 pin automatically excluded 0.146.0. The change re-admits current eval by advancing the pin one minor. Secondary goal, achieved implicitly: keep the published cohort guarantee intact — exactly one admitted eval minor, one installed copy — while shif
  • Assessment: Good, and squarely in the grain. (1) The range is not hand-written: expectedPeerRange() in scripts/lib/peer-range.mjs:33-35 emits >=0.146.0 <0.147.0 for a pre-1.0 pin — I executed the module and it prints exactly the string at package.json:82. (2) The guard enforces it: scripts/verify-package.mjs:69-74 reads the dev pin, derives the expected range, and fails verify:package on any mismatch, s
  • Better / existing approach: none — this is the right approach. I checked the two alternatives against the repo's own machinery: (a) a caret ^0.146.0 is semantically identical for 0.x under npm's rule (peer-range.mjs:5-9 documents this), so the explicit range is just the clearer spelling of the same contract — no improvement available; (b) widening to span both minors (e.g. >=0.145.21 <0.147.0) would avoid forcing a simul
  • Model: opencode/zai-coding-plan/glm-5.2
  • Bridge attempts: 2
  • Bridge warning: opencode/kimi-for-coding/k2p7: opencode: opencode error

🎯 Usefulness — sound

A mechanically derived peer-window bump that unblocks co-install with Eval 0.146.0; every claim checks out against the installed package and the repo's own enforcement script.

  • Integration: Fully wired. The window is not hand-picked: expectedPeerRange('0.146.0') returns '>=0.146.0 <0.147.0' (scripts/lib/peer-range.mjs:33-35), byte-identical to package.json:82, and scripts/verify-package.mjs:73,294-297 fails the package check if the declared peer ever diverges from the derivation. The devDep (package.json:89) and pnpm-lock.yaml moved to 0.146.0 in the same commit, and CHANGELOG.md doc
  • Fit with existing patterns: Follows the repo's established release-alignment grain exactly. git log shows a chain of identical alignment releases (8.0.5, 662c874; 383ce19 introduced the derivation rule this PR applies; 8ee086c; 2128bb2; fd0eaa7). The pre-1.0 exact-minor window is the deliberate policy documented in scripts/lib/peer-range.mjs:5-9 and enforced by the verify scripts — no competing pattern exists.
  • Real-world viability: Independently verified beyond the PR's claims: pnpm install --frozen-lockfile clean; pnpm run typecheck (src + contracts) clean against installed 0.146.0; @tangle-network/agent-eval/multishot/golden importable from the installed 0.146.0 (22 exports); the 0.146.0 exports map retains every subpath this repo imports ('.', ./campaign, ./rl, ./experiment all present). The tsc pass is the decisive check
  • Model: opencode/zai-coding-plan/glm-5.2
  • Bridge attempts: 1

No concerns — sound change, no better or existing approach found. ✅


What this audit checks

It judges the change on its merits — not whether it was tasked out in an issue. Unticketed, fast-moving work is fine; the question is whether the change is good and whether a better or existing approach should be used instead.

PassWhat it asks
HeuristicVague title? Whitespace-only or cruft-bearing diff? (content signals only)
DuplicationDo added function/class names already exist elsewhere in the repo?
Value AuditWhat does it do? What goal does it achieve? Is it good? Better architecture or already-exists?
Usefulness AuditDoes it integrate and fit? Will it hold up in real use and actually get used?

Findings are concerns, not blocks — the human reviewer decides what to do with them.

value-audit · 20260816T222525Z

@drewstone
drewstone merged commit 9a1079d into mainAug 16, 2026
2 checks passed
@drewstone
drewstone deleted the chore/agent-eval-0.146.0 branch August 16, 2026 22:26
Sign up for freeto join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

2 participants

@drewstone@tangletools