Skip to content
View rubenmarcus's full-sized avatar
💭
┏( ゜)ਊ゜)┛
💭
┏( ゜)ਊ゜)┛

Block or report rubenmarcus

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
rubenmarcus/README.md

I build agent-ready products — Ruben Marcus, Senior AI Fullstack Engineer

rubenmarcus.devMCP — connect your agentTelegramXLinkedInCVemail


Senior AI Fullstack Engineer in Lisbon, remote worldwide. 14 years shipping — 4+ deep in web3, 2+ building AI developer tools. I build AI-native products, agent tooling, and the harnesses that keep them honest: bounded specs, isolated worktrees, evals that fail closed.

My portfolio speaks MCP. Point your agent at https://www.rubenmarcus.dev/api/mcp and let it read this résumé instead of you — setup below.


Latest projects

ProjectWhat it is
Ralph StarterOpen-source AI coding orchestrator. Swarm mode — race / consensus / pipeline over isolated git worktrees — plus an MCP server and a Figma→code visual validation pipeline.
AutoresearcherBenchmark-driven autonomous research CLI. Divergent agent populations in isolated worktrees, champion merging, a Pareto frontier of candidates, keep/reject validity gating.
AEO.jsAnswer Engine Optimization framework. AI-crawler policy analysis and LLM-ready exports (llms.txt, ai-index.json). Astro + Next plugins.
AEO CheckerThe hosted scanner on top of it — 4,569 scans across 2,259 unique sites.
ECDSA.fail — #1Led AI engineering of an autonomous multi-agent research harness: 9 specialist LLM roles, 7+ providers, role→model routing, fail-closed adapters, spend gates. Research contributor on the resulting publication.
QEC Decoder — #1Top of Optimization Arena's quantum error-correction leaderboard (2,642 EPM, near Bayes-optimal) via a multi-agent campaign with an anti-overfitting evaluation protocol.
CS BrasilBrowser FPS built with an agent gauntlet. WebGL, no install, no netcode. 2,191 players, 154K+ kills, 27 countries in alpha.
Mirofi.shOpen-source multi-agent social-simulation engine (GraphRAG/Zep, OASIS) productized into a hosted SaaS.

Before that: MultiVM Labs / Quantum (post-quantum L1 — main site, block explorer, and an ML-DSA-65 smart wallet on Arbitrum Stylus), Bitte Protocol (top human committer on the production AI runtime, #1 on the agent SDK monorepo, #3 on the wallet), Grover, Zup Innovation / Itaú Open Banking, Santander, and a long agency tail — Under Armour, Centauro, Samsung, Panasonic, Monsanto. Full archive →

Numbers

2.85M+agent messages processed in production (Bitte AI runtime)
24,164unique users on agents I built · 16,703 agents deployed on the runtime
32AI agents and loop roles built · 25 versioned agent skills
4,569AEO scans across 2,259 sites
34K+all-time npm downloads across 8 packages
#1 · #1ECDSA.fail · Optimization Arena QEC decoder
14 yearsshipping · ~2M lines of code, career estimate

GitHub: 2,787 stars, 121 public repos, 1,193 followers

Read me with your agent (MCP)

No account, no auth, no signup. Four read-only tools plus one that emails me a brief.

https://www.rubenmarcus.dev/api/mcp
ToolReturns
get_resumefull resume: experience, skills, proof points, links
get_servicesthe fixed-scope offers
check_availabilitycurrent engagement status
book_introposts a project brief to my inbox
Setup per client

Claude — Settings → Connectors → Add custom connector, name it rubenmarcus, paste the URL. Or in Claude Code:

claude mcp add --transport http rubenmarcus https://www.rubenmarcus.dev/api/mcp

ChatGPT — Settings → Apps & Connectors → developer mode → Create, paste the URL, save.

Cursor (~/.cursor/mcp.json), and any client that takes a streamable-HTTP URL:

{
"mcpServers": {
"rubenmarcus": { "url": "https://www.rubenmarcus.dev/api/mcp" }
}
}

Kimi — add a streamable-HTTP MCP server with the same URL, then run tools/list.

Codex CLI / any stdio-only client (~/.codex/config.toml) — bridge it:

[mcp_servers.rubenmarcus]
command = "npx"args = ["-y", "mcp-remote", "https://www.rubenmarcus.dev/api/mcp"]

Then ask: What has Ruben shipped? · Summarize Ruben's experience with AI agents. · Is Ruben available right now? · Book an intro — I want to build a trading bot.

No MCP client? Plain endpoints.
Resume JSONhttps://www.rubenmarcus.dev/api/resume.json
Resume texthttps://www.rubenmarcus.dev/api/resume.txt
Agent guidehttps://www.rubenmarcus.dev/AGENTS.md
Connect guide, as Markdownhttps://www.rubenmarcus.dev/connect.md
LLM indexhttps://www.rubenmarcus.dev/llms.txt
MCP server cardhttps://www.rubenmarcus.dev/.well-known/mcp/server.json
Agent skillhttps://www.rubenmarcus.dev/.well-known/agent-skills/portfolio-mcp/SKILL.md

Every blog post has a .md twin: append .md to any post URL.

The agent fleet

Three origins: a production fleet that signs transactions, a research command center, and a game harness. Full directory →

Bitte Protocol — the production fleet
AgentChainWhat it does
AI FrameworkEVM · NEAR · Sui · Cardano · MidnightNatural language in, signed transactions out. The runtime everything below runs on.
@bitte-ai/chat + make-agentchain-agnosticAn OpenAPI spec is the agent.
Uniswap agentEthereum / EVMKeyless cross-chain swaps from a chat prompt.
Gnosis PilotGnosisA DeFi yield copilot with live data.
Polymarket agentPolygonA prediction-market analyst that can place the bet.
meme.cooking agentNEAROne prompt, one memecoin.
Solana AssistantSolanaSolana chain data, agent-ready.
Jupiter swap agentSolanaQuote and swap, one endpoint.
Morpho agentEthereumLending and borrowing, spelled out for an LLM.
Aerodrome agentBaseThe full veAERO machine as agent tools (~25 tools).
Walrus agentSuiDecentralized storage by chat.
Sui assistantSuiA Sui explorer that talks back.
ENS agentEthereumNames, records, registrations as tools.
ECDSA.fail — the command center that took #1

Frontier Dissector · Circuit Engineer · Density Analyst · CUDA Engineer · Pod Manager · Research Scout · Orchestrator-Reviewer · Combinator — 9 specialist roles routed across 7+ providers, with contracts, spend gates and fail-closed adapters. How it worked →

CS Brasil — the Gauntlet

Three rules hold the loop together: the ruler is not negotiable (not "looks good" — wins or loses against a CS2 frame, by which measure), the builder never grades its own work, and the loop ends when I stop it, not when an agent declares itself satisfied. Every claim has to carry a number and a file:line.

RoleContract
Measured baselineCaptures every map × 2 aspect ratios × 4 angles before anything changes. No baseline, no A/B — and without A/B the loop is just opinion.
Graphics criticGrades the frame, never the builder's report. Sees pixels and code, nothing else.
Map fidelity criticGrades the map against how the real Brazilian place actually looks.
Weapons criticVisual and feel scored separately.
UI criticMenu and HUD scored separately.
Gameplay criticMovement, bots, combat flow.
Parallel buildersPartitioned by a generated file/range conflict table, so several agents edit one large file at once without colliding.
A/B verifierRe-shoots the same frames and proves the delta, or the round didn't happen.
Regression HunterThe agent that pays for the whole loop: a visual win that breaks the game is a regression.
Bug HunterThe ruler comes before the fix, and a mutation has to prove the ruler bites.
RéguaWrites the invariant, the probe and the gate — 25 consistency criteria that outrank the fidelity bar.
Asset reviewAdversarial critic on every new character, map, model or texture before the front can be called done.
Content & faction pipelinesBuild a team, a real-world map, or a whole faction end to end — roster, crest, cover, 3D character, thumbnail, selection video, original voice.

Critics run in parallel with clean contexts; builders run in parallel across four checkouts of the repo on separate branches, plus dozens of throwaway git worktrees for one-off fixes.

Inside the loop → · The AI harness behind the game →

Agent skills

Versioned method, not loose prompts — 25 skills that encode how the work is actually done. Browse them →

Ralph Starter Orchestration · Autoresearch Benchmark Loop · AEO Delivery System · Benchmark Frontier Archaeology · ECDSA.fail Route Spec · Circuit Engineer Factory · ECDSA Circuit Optimization · Reversible Circuit Validation · Peak Qubit Reduction · Toffoli Reduction · ECDSA.fail Island Hunting · Multi-Agent Research Collaboration · CS Brasil Content Pipeline · Faction Pipeline · Adversarial Asset Review · Régua / Executable Quality Gate · CS Brasil PR Triage · CS Brasil Smoke Check · Gauntlet FPS · Bug Hunt · Blog Voice · Bilingual Publishing · Portfolio Cover System · Frontend Delivery Harness · Portfolio MCP

Writing

Every post ships in EN and PT, and every post has a Markdown twin for agents — append .md to the URL. All posts →

DatePost
2026-08-11My AI harness for frontend: from prompt to pull request
2026-08-11From prompt to product: five ways to build with AI
2026-08-07Inside the Gauntlet loop
2026-08-06This portfolio is agents-welcome. Probably the first.
2026-08-04The AI harness behind the CS Brasil game
2026-08-02A command center for agent swarms, in markdown
2026-07-30I built my portfolio with a fleet of AI agents
2026-07-27How AEO can help your business grow
2026-07-24Building a browser FPS with AI agents
2026-07-21Streaming 2.85M messages: the plumbing of a production agent chat
2026-07-18I rebuilt my agent loop in Mastra. Here's what my runtime gets right.
2026-07-15Routing 9 agent roles across 7 providers: the ECDSA.fail harness
Earlier posts
DatePost
2026-07-11Cross-pollinating LLMs: peer review for machines
2026-07-08Keeping an autonomous research agent honest
2026-07-02Evals are the product
2026-06-25Git worktrees are my agent orchestrator
2026-06-18The swarm that took #1 on ECDSA.fail
2026-06-05Context engineering inside a runtime with 344K chats
2026-05-28How RAG works inside Mirofi.sh
2026-05-12The Mini Shai-Hulud Case and the Real Risk of Dependencies
2026-04-16How I Hit #1 on a Quantum Error Correction Challenge using AI
2026-02-19Automating entire workflows with ralph-starter
2021-05-16Getting started with Next.js + Strapi: Security first
2021-05-07Why use Next.js + Strapi?

On npm

aeo.js · ralph-starter · autoresearcher · make-agent · scanrepo · elendil · new-agent · qday

Stack

  • Languages — TypeScript · JavaScript · Rust · PHP · C#
  • Frameworks — React · Next.js · Svelte 5 · Astro · Vue · Angular · Node · Bun
  • AI & agents — Claude SDK · OpenAI SDK · AI SDK · MCP · Mastra · Ralph loops · custom harnesses
  • Web3 — EVM · NEAR · Sui · Solana · Cardano · Wagmi · Viem · ERC-4337
  • Infra — AWS · GCP · Vercel · Docker · Terraform · GitHub Actions · Turborepo
  • Craft — Figma · Storybook · Tailwind · Three.js · GLSL

Hiring me

Selectively available for full-time roles and freelance contracts. Fixed-scope engagements preferred. I reply within a day or two.

ruben@rubenmarcus.dev · rubenmarcus.dev/contact · or just let your agent call book_intro.

Pinned Loading

  1. csbrasilcsbrasilPublic

    FPS satírico de navegador: 5 facções brasileiras caricatas, 44 personagens, 5 mapas e 26 armas — rounds e Capture the Flag contra bots, direto na aba, sem instalar nada. Three.js + vanilla JS, zero…

    JavaScript 189 25

  2. aeo.jsaeo.jsPublic

    Answer Engine Optimization for the modern web. Make your site discoverable by ChatGPT, Claude, Perplexity & AI search engines. Generates llms.txt, robots.txt, sitemap, JSON-LD & more.

    TypeScript 122 16

  3. portfolioportfolioPublic

    Agent-ready portfolio of Ruben Marcus, AI Fullstack Engineer. A cinematic WebGL site for humans and a protocol surface for agents: a remote MCP server (io.github.rubenmarcus/portfolio, on the offic…

    Astro 22 4

  4. ralph-starterralph-starterPublic

    Bootstrap projects from Figma, Linear, Notion & GitHub specs. The only AI ralph with source integrations.

    TypeScript 104 8

  5. malicious-repositoriesmalicious-repositoriesPublic

    Forked from xndbogdan/malicious-repositories

    collected from LinkedIn scammers

    JavaScript 210 14

  6. 120-perguntas-frontend120-perguntas-frontendPublic

    120 Perguntas Front-end separadas por níveis

    2.1k 154