knowledge: 12 insights — sequential ids across parallel workers, inbound-validation task ownership, vendor benchmark claims, element crop screenshots, WebFetch summary vs raw page, Steps-prose guarantees, split fact-check verdicts (+4 folds, 1 dup) - #182
Open
choiyounggi wants to merge 1 commit into
Conversation
…lded into open PRs, 1 dropped duplicate
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for freeto join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
Knowledge flush — 12 insight(s)
Flush run
20260903-213946-4161(headless auto-flush, lock re-entered under the parent hook's run id). Queue rows are keyed byhash. Outcome: 4 new pages, 3 merges into existing pages, 4 folds pushed to open knowledge PRs, 1 dropped as a pending duplicate.Verified best-practice
Every external quote below was re-checked against the raw page with
curl -sL … | grepon 2026-09-03 (not only through a summarizing fetch), except where noted.confidence: verified. Claim: the coordinator assigns RFC/ADR/migration numbers at dispatch; a worker's branch point cannot see a sibling's unmerged number and distinct filenames merge without conflict, so the uniqueness lint runs on the merged tree. Sources: Django migrations topic ("two migrations with the same number"), Djangomakemigrations --merge("Enables fixing of migration conflicts"), Rails 3.2 migrations guide (creation-time timestamps to avoid clashes), git-merge ("incorporated in the final result verbatim"), adr-tools issue fix(orchestrate): ORCH_DIR path token, watch-status stall gating + reason surfacing, merge-on-approval (#87–#90) #102 (two devs both denote ADR 6). Field evidence: linkly t112/t119 both created RFC-0034.confidence: verified. Source: OWASP Input Validation Cheat Sheet ("as early as possible in the data flow, preferably as soon as the data is received from the external party"). Field evidence: agent-crew M2handle_envelopemissingvalidate()caught only by integration review.confidence: verified. Sources: LoCoMo paper (arXiv 2402.17753), Zheng et al. LLM-as-a-judge (arXiv 2306.05685, "over 80% agreement"), Zep blog disputing Mem0's LoCoMo SOTA claim, Mem0's counter-reply (Revisiting Zep’s 84% LoCoMo Claim: Corrected Evaluation & 58.44% Accuracy getzep/zep-papers#5, "58.44%"), OpenVikingstat_judge_result.py(QA and Import token usage counted separately) andjudge.py("be generous with your grading"). The candidate'sjudge.py:239line reference no longer matches the 203-line file; the leniency instruction itself is confirmed and the page says so.clipfromboundingBox()) →confidence: field-tested. Playwright semantics verified (element screenshot vialocator.screenshot(),clipoption,boundingBox()is viewport-relative and scroll-dependent); the Aside CLI clip misbehaviour itself is single-session field evidence, not reproduced here (no browser run). Directive generalised to: element-screenshot primitive first, read back the first crop before a batch, fall back to full-page capture on a persistent wrong-region clip.confidence: verified. Source: Claude Code tools reference ("runs the prompt against the content using a small, fast model. For most fetches, Claude receives that model's answer, not the raw page"; "use curl via Bash for the unprocessed page"). Field evidence re-confirmed: the tmap-skopenapirouteSequential30page contains the exact string "경유지는 최대 30개까지 설정할 수 있습니다." in the raw response.confidence: verified(merged into an already-verified page). Source: SWE book ch12 ("A behavior is any guarantee that a system makes…"). Field evidence: wt-t4-event-push task 03 auditor FAIL→PASS after one added test.confidence: verifiedfor the evidence (dev-loop PR docs(wiki): mandatory design-skill routing for visual-design deliverables #164 is public and merged; commitf5d2395message confirmed viagh api), directive itself field-tested; merged into afield-testedpage.command -v→ dropped, pending duplicate: PR knowledge: 12 insights — dropzone copy vs drop handlers, destination-in mask chaining, media-query inset reset, spatial clamp, env restore vs pop, synthetic-corpus floor, plan-claim recompute, multi-name command -v, alert suppression key, sibling validators (+2 folds) #181'spath-resolution.mdalready carries this exact edge case, instead-of row, POSIX synopsis source and the same local reproduction.Existing-layer check
Pages read: platforms-environment-path-resolution, security-input-validation-at-trust-boundaries, qa-document-verification-spec-document-gates, qa-deliverables-quantitative-claims-in-a-published-document, infrastructure-agent-orchestration-worktree-isolated-workers, qa-process-completion-claims, platforms-tools-harness-mediated-tool-results, testing-quality-minimum-case-set, qa-process-evaluating-review-feedback, qa-process-llm-review-pipelines, qa-document-verification-generated-reference-drift-gates, backend-common-llm-context-window-budget, qa-process-adversarial-change-review, qa-bug-reports-reproducible-reports, qa-environments-browser-console-capture-gaps
Also read on open-PR heads (not on this checkout): checkable-claims-in-an-adopted-plan, sibling-validators-on-a-shared-node (#181); semantic-conflicts-after-parallel-merge, verify-command-in-a-worker-brief (#179); ours-resolution-on-a-mixed-content-conflict, forward-references-in-a-numbered-protocol (#180).
harness-mediated-tool-results(new when-this-applies sentence, edge-case row, instead-of row, source, field context; index cell widened). Steps-prose guarantee →minimum-case-set(edge-case row, instead-of row, SWE-book quote + field evidence; index cell widened). Split verdict →evaluating-review-feedback(edge-case row, instead-of row, PR docs(wiki): mandatory design-skill routing for visual-design deliverables #164 source; index cell widened). None of these three pages is touched by an open knowledge PR.infrastructure/agent-orchestration/sequential-identifiers-across-parallel-workers,infrastructure/agent-orchestration/inbound-validation-ownership-in-task-decomposition,backend/common/llm/vendor-benchmark-claims-for-an-llm-tool,qa/environments/element-crop-screenshots. Each has an index row and a log line.security-input-validation-at-trust-boundaries("validate at the consumer boundary anyway") and adds the task-decomposition angle.qa-process-adversarial-change-review,backend-common-llm-context-window-budget,qa-process-llm-review-pipelines,qa-environments-browser-console-capture-gaps,qa-bug-reports-reproducible-reports. Back-links deliberately NOT added onworktree-isolated-workers,shared-run-state,spec-document-gates,validation-at-trust-boundaries,quantitative-claims-in-a-published-document,completion-claims: theirrelated:/frontmatter lines are rewritten by PR knowledge: 12 insights — deny rules under bypass, merged-tree gate, worker verify command, Kotlin daemon heap, extracted-method this (+7 merges) #179/knowledge: 11 insights — login-expired panes, O_APPEND blackboard, numbered-protocol grafts, --ours on mixed conflicts, unittest floor false positives, SwiftUI sheet gating, silent hotkey registration, rolled-back-run assertions (+3 merges) #180/knowledge: 12 insights — dropzone copy vs drop handlers, destination-in mask chaining, media-query inset reset, spatial clamp, env restore vs pop, synthetic-corpus floor, plan-claim recompute, multi-name command -v, alert suppression key, sibling validators (+2 folds) #181 and a second edit would conflict at merge. Owner may add them after those PRs land.wiki-structure-checks279 pages / 0 findings,wiki-lint-prohibitions0 violations, all touched pages ≤ 120 body lines.Open-PR check
Open
knowledge/*heads listed viagh pr list --search "head:knowledge/": #179knowledge/choiyounggi-20260903-172728, #180knowledge/choiyounggi-20260903-184706, #181knowledge/choiyounggi-20260903-203836. Each was fetched and diffed againstorigin/main -- wiki/.command -vpath-resolution.md(identical edge case + reproduction)checkable-claims-in-an-adopted-plan.md(same trigger: checking an adopted plan)4bc6de8on #181 + PR commentours-resolution-on-a-mixed-content-conflict.md(merge-time count reconciliation)3b78273on #180 + PR commentverify-command-in-a-worker-brief.md(what the brief's verify line names)e242b2con #179 + PR commentworktree-isolated-workers.md(Edit/Write bypass + matcher widening)e242b2con #179semantic-conflicts-after-parallel-merge.md(enum/match semantic conflicts), #180ours-resolution(count conflicts) — adjacent, different trigger (distinct new files, no conflict at all)validation-at-trust-boundaries.mdedit (spatial-value clamping) — different triggersynthetic-corpus-measurement-floor.md(measuring on your own corpus) — different triggerquantitative-claims-in-a-published-document.mdrelated line onlyharness-mediated-tool-results, untouched by open PRs)minimum-case-set.mduntouched)completion-claims.md, notevaluating-review-feedback.mdLint (
wiki-structure-checks,wiki-lint-prohibitions) was run on each fold branch after the edit: 0 findings, 0 violations; fold pages remain ≤ 120 body lines (83/55/73/93 for #181/#180/#179 verify/#179 worktree).Routing decision
infrastructure/agent-orchestration/sequential-identifiers-across-parallel-workers— NEW page; agent-orchestration already owns worker briefs and shared run stateinfrastructure/agent-orchestration/inbound-validation-ownership-in-task-decomposition— NEW page; the lesson is about which task's brief carries the decision, so orchestration rather than security (linked to the security page)backend/common/llm/vendor-benchmark-claims-for-an-llm-tool— NEW page; backend/common/llm owns consuming LLM tooling; no new category neededqa/environments/element-crop-screenshots— NEW page; qa/environments already holds browser-tooling gaps (console capture, bot blocking)platforms/tools/harness-mediated-tool-results— same class (a tool result mediated before the agent sees it)testing/quality/minimum-case-set— it is a "which cases are required" ruleqa/process/evaluating-review-feedback— it is a response-to-review-finding rulecommand -vplatforms/environment/path-resolutionNo new category was added; every insight fit an existing domain/category.