Skip to content
@Inference-Foundry

Inference Foundry

We build high-efficiency local AI tools and conduct research on model execution paradigms.

Inference Foundry

Inference Foundry is an open-source organization building tools and research around LLMs, diffusion models, and efficient local inference. Public pages provide high-level context and project discovery; member-level planning and detailed ownership stay in private org docs.

Organization site:inference-foundry.github.io
Public project index:docs/projects
Detailed member handbook and roster:.github-private(org members only)


Projects

ProjectWhat it isLinks
super-ollamaTerminal-native, in-process local LLM engine (no HTTP in the main UX); llama.cpp via CGo, focus on low overhead and clean teardown.Repo · Roadmap (wiki)
CrucibleOpen research journal and experimental log.Repo
BitForge(planned)Quantization theory, methods, and reproducible experiments across bit-widths and runtimes.Repo TBD — org doc
Lexicon(planned)Open fine-tuned prompt catalog with versioning, licensing, and analysis for reuse.Repo TBD — org doc
Argus(planned)Algorithms to detect AI-generated images using JEPA-based representations.Repo TBD — org doc

For deeper project ownership, plans, and internal notes, use .github-private (members only).


Founders and contact

If you want to collaborate or reach the team, contact us via Discord or founder links above.


Participate


Inference Foundry — open tools and honest measurements.

Pinned Loading

  1. .github.githubPublic

  2. super-ollamasuper-ollamaPublic

    Forked from ollama/ollama

    Go

Repositories

Showing 7 of 7 repositories

Top languages

Loading…

Most used topics

Loading…