Uh oh!
There was an error while loading. Please reload this page.
Uh oh!
There was an error while loading. Please reload this page.
doublewordai
Popular repositories Loading
- control-layer
control-layer PublicThe world’s fastest AI model gateway (450x less overhead than LiteLLM). Unified access to LLMs across endpoints (openAI, self-hosted, etc.) behind a single authentication layer - with API key gener…
- autobatcher
autobatcher PublicDrop-in AsyncOpenAI replacement that transparently batches requests
- deepseek-reddit-agent
deepseek-reddit-agent PublicAn example notebook which shows how you can build a LLM agent that scrapes information from Reddit and summarize key bullets using a self-hosted DeepSeek-R1-Distill-Llama-8B deployed with Titan Tak…
- inference-stack
inference-stack PublicThe Doubleword Inference Stack is the easiest & most performant way to run genAI infrastructure in your private environment.
- inference-lab
inference-lab PublicHigh-performance LLM inference simulator for analyzing serving systems
Repositories
Uh oh!
There was an error while loading. Please reload this page.
doublewordai/documentation's past year of commit activity - control-layer Public
The world’s fastest AI model gateway (450x less overhead than LiteLLM). Unified access to LLMs across endpoints (openAI, self-hosted, etc.) behind a single authentication layer - with API key generation, user management, request logging, and more
Uh oh!
There was an error while loading. Please reload this page.
doublewordai/control-layer's past year of commit activity - batchbench Public
Uh oh!
There was an error while loading. Please reload this page.
doublewordai/batchbench's past year of commit activity Uh oh!
There was an error while loading. Please reload this page.
doublewordai/inference-lab's past year of commit activity - modelexpress Public Forked from ai-dynamo/modelexpress
Model Express is a Rust-based component meant to be placed next to existing model inference systems to speed up their startup times and improve overall performance.
Uh oh!
There was an error while loading. Please reload this page.
doublewordai/modelexpress's past year of commit activity - vllm Public Forked from vllm-project/vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
Uh oh!
There was an error while loading. Please reload this page.
doublewordai/vllm's past year of commit activity - outlet-postgres Public
A plugin for the https://github.com/doublewordai/outlet middleware, for publishing requests & responses through an axum server to postgres
Uh oh!
There was an error while loading. Please reload this page.
doublewordai/outlet-postgres's past year of commit activity - kagent Public Forked from kagent-dev/kagent
Cloud Native Agentic AI | Discord: https://bit.ly/kagentdiscord
Uh oh!
There was an error while loading. Please reload this page.
doublewordai/kagent's past year of commit activity - outlet Public
A high-performance Axum middleware for capturing and correlating HTTP requests and responses with full streaming support.
Uh oh!
There was an error while loading. Please reload this page.
doublewordai/outlet's past year of commit activity - dynamo Public Forked from ai-dynamo/dynamo
A Datacenter Scale Distributed Inference Serving Framework
Uh oh!
There was an error while loading. Please reload this page.
doublewordai/dynamo's past year of commit activity
Top languages
Loading…
Uh oh!
There was an error while loading. Please reload this page.
Most used topics
Loading…
Uh oh!
There was an error while loading. Please reload this page.