Skip to content
View felmonon's full-sized avatar

Block or report felmonon

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
FELMONON/README.md
Felmon Fekadu builds developer tools and AI reliability systems that leave evidence; an engineering trace passes through observation and verification while a policy gate blocks a regression.

Felmon Fekadu

I build software that leaves evidence.

Developer tools and AI reliability systems that expose hidden behavior, enforce deterministic gates, and make regressions impossible to ignore.

Portfolio · Technical writing · Email · LinkedIn

01 / CLAIM

Most software says “it worked.”

I care about the harder questions: What happened? What changed? What can we prove?

I build the layer that answers those questions.

02 / INSTRUMENTS

Failure modeInstrumentEvidence emitted
Mock driftmsw-inspectorCoverage report + CI gate
Agent trajectory regressionagent-reliability-harnessPolicy findings + baseline diff
Retrieval without evidencedocagent-studioCitations + offline evaluation
Opaque coding-agent qualityagent-fight-clubScored, replayable bouts
Unproven product assumptionsTypeJung · sourceA live, paid user journey

03 / ONE RECORDED TRACE

A six-step trace imports ambiguous agent behavior, verifies schema, budget, safety, and grounding, compares candidate to baseline, blocks a regression, and emits a reproducible report.

A demo shows that something worked once. A trace helps explain why it will keep working.

04 / OPERATING CONSTRAINTS

01 Evidence over adjectives.
02 Deterministic checks in CI. Model judgment offline.
03 Narrow interfaces. Reproducible failures. Useful artifacts.

05 / PROOF LEDGER

RecordPublic evidence
PACKAGE / NPMmsw-inspector-cli v0.3.2 + GitHub Marketplace Action
PACKAGE / PYPIagent-reliability-harness v0.2.2
UPSTREAM14 merged pull requests across 7 repositories / 5 organizations
COMMUNITY5 outside human contributors in msw-inspector
PRODUCTTypeJung, a live full-stack paid product
WRITINGDeterministic checks beat model-judged evals in CI

Dated public proof snapshot, verified 2026-08-01. Every count also links to its live public record.

06 / HANDOFF

Have a failure mode you cannot see yet? Send me the trace.

felmon.tech · felmonon@gmail.com · all repositories

CALGARY / CANADA · TYPESCRIPT / PYTHON · PROFILE DATA · NOT A RÉSUMÉ. A TRACE.

Pinned Loading

  1. msw-inspectormsw-inspectorPublic

    Static analysis for MSW: finds drift between real API calls and mocked handlers, locally and in CI.

    TypeScript 5

  2. agent-reliability-harnessagent-reliability-harnessPublic

    Deterministic policy-as-code and regression gates for AI agent tool-use traces.

    Python

  3. agent-fight-clubagent-fight-clubPublic

    Replay-first arena for benchmarking coding agents on correctness, cost, resilience, and diff quality.

    TypeScript

  4. trace2testtrace2testPublic

    Turns failed AI agent traces into reproducible regression tests and verified fixes.

    TypeScript

  5. docagent-studiodocagent-studioPublic

    Local-first document QA with hybrid retrieval, citation-grounded answers, and offline evaluation.

    Python

  6. jungian-typology-assessmentjungian-typology-assessmentPublic

    A shipped full-stack assessment product with auth, Stripe billing, saved results, and AI-assisted reports.

    TypeScript 1