Skip to content
@Intelligent-Internet

Intelligent-Internet

First Principles Sovereign AI
Intelligent Internet logo

Intelligent Internet

First Principles Sovereign AI

Most AI companies rent intelligence and compete at the UI.We ship the control points that turn intelligence into owned, verifiable work, in the open.


WebsiteBlogHugging FaceSymbioism


🏭 The Agentic Production Line

Agents don't fail for lack of intelligence. They fail when knowledge is stale, retrieval is expensive, output is unverified, and capability never reaches a surface users can own. So we build every stage of the line, not one layer of it:

Control pointWhy it mattersOpen source
01Capability Foundry: create capability, not wrappersIf you only rent frontier APIs, your ceiling is someone else's roadmapII-Medical · II-Search · II-Thought
02Governed Context: turn raw knowledge into machine-usable supplyAgents fail when knowledge is scattered, stale, or outside source boundariesII-Commons · II-Commons-Skills
03Retrieval Fabric: make search cheap, local, inspectableContext is useless if agents can't search before every decision, tool call, or handoffpsql_bm25s
04Work Harness: completion under gates, not just generationAgent output isn't work until it survives validators, evidence review, and replanningII-Agent · II-Researcher · Zenith
05Owned Surfaces: land capability where work happensCapability only compounds when people can run, fork, and extend itCommonGround · CG-Cardbox · opencode-a2a

Models can be rented. UIs can be copied. Control points compound.


🏆 Zenith: #1 on Frontier SWE

The clearest proof that the harness layer matters: on the independent Frontier SWE benchmark, GPT-5.5 running inside Zenith ranks #1 overall, ahead of every frontier model paired with its own native harness. The identical model on its native harness ranks #5. Same model, better control loop.

#ModelHarnessAvg rank ↓Dominance ↑
1GPT-5.5🥇 Zenith2.0692%
2Claude FableClaude Code2.7188%
3Claude Opus 4.8Claude Code5.0671%
4GLM-5.2Claude Code5.3169%
5GPT-5.5Codex (native)5.5368%

Metrics as reported by the Frontier SWE leaderboard. The full 15-entry table is in the Zenith results.

Zenith is our continuous-improvement harness for missions that run for days or weeks, where the dominant failure mode is premature completion. One orchestrator session reads task state each turn and decides whether to spawn workers and testers, register reusable skills, replan, or stop, all over MCP/ACP on top of Claude Code, Codex, or Hermes. In our published ablation across eight long-horizon tasks, Zenith achieves the best mean rank at less than half of RALPH's per-task cost ($176 vs $408).

📄 Technical report: From RALPH to Zenith: Designing Harnesses for Long-Running Agents


🚀 Flagship Projects

ProjectWhat it does
II-AgentStarsOpen general agent framework: browser, code, files, sandboxed execution, documents, slides, multi-model routing
ZenithStars#1 on Frontier SWE. A continuous-improvement harness for long-running agent tasks that turns Claude Code, Codex, or Hermes into a multi-agent mission orchestrator via MCP/ACP
II-ResearcherStarsDeep-research agent: query decomposition, search generation, context compression, self-critique, and cited reports. Scores 84.1 on FRAMES
CommonGroundStarsFrom isolated agents to shared work: records, evidence, handoffs, and decisions that persist beyond one run
psql_bm25sStarsPostgres-native exact BM25: mutable indexes, crash recovery, replication-friendly storage, SQL-native permissions
II-CommonsStarsThe knowledge supply chain: Wikipedia, PD12M, arXiv, and PubMed, parsed, embedded, indexed, and served with provenance

🧠 Open Models & Datasets

Everything on the 🤗 Hugging Face hub, with weights, data, and benchmark traces included.

ReleaseTypeHighlight
II-Medical-8BModelSpecialist medical reasoning with SFT, RL, and safety stages
II-Search-4BModelMulti-hop search and tool-use behavior in a small model
II-Thought-RL-v0Dataset341,795 verified, machine-checkable RL problems across math, code, science, medicine
II-Medical-Reasoning-SFTDatasetPart of 2.2M medical reasoning rows behind the II-Medical series
wikipedia_en · arxiv · pd12mDatasetsPublic knowledge, processed for agents, with citations and source boundaries

📊 At a Glance

🏆 #15,000+🧪 341K🏥 2.2M🤗 9 + 20🏭 5/5
on Frontier SWE (Zenith)GitHub stars across the orgverified RL problems, openmedical reasoning rowsopen models + datasetsproduction-line stages shipped, all open

🧭 Why Open?

We publish the research, the data pipelines, the retrieval infrastructure, the harnesses, and the philosophy, because an intelligence economy only compounds when its production line is inspectable and forkable. Our long-form thesis lives at Symbioism: A Third Path for the Intelligence Age (source, naturally).

Earlier experiments like CoT-Lab, Common Chronicle, and CommonGround-legacy are archived in public. Every stage of the line started as an open experiment; the ones that worked became infrastructure.


Intelligence is our greatest resource. Together, we make it abundant.

ii.inc · Blog · 🤗 Hugging Face · Symbioism

Pinned Loading

  1. ii-agentii-agentPublic

    II-Agent: a new open-source framework to build and deploy intelligent agents

    Python 3.4k 524

  2. CommonGroundCommonGroundPublic

    From isolated agents to shared work

    Python 149 20

  3. zenithzenithPublic

    Zenith: a continuous-improvement harness for long-running agent tasks. Turns Claude Code, Codex, or Hermes into a multi-agent mission orchestrator via MCP/ACP.

    Python 281 40

  4. psql_bm25spsql_bm25sPublic

    PostgreSQL BM25S extension

    PLpgSQL 146 4

  5. ii-researcherii-researcherPublic

    II-Researcher: a new open-source framework designed to aid building search / research agents

    Python 498 75

Repositories

Showing 10 of 23 repositories

People

This organization has no public members. You must be a member to see who’s a part of this organization.

Top languages

Loading…

Most used topics

Loading…