Pinned Loading
- iron-batch
iron-batch PublicLLM serving primitives in Rust: paged KV allocation, continuous batching, streaming HTTP.
Rust
- lvm-pass-lens
lvm-pass-lens PublicLoRA fine-tuned a 120B LLM to classify LLVM compiler pass interactions. 53.8% → 82.1% accuracy.
Python
- bit-floor
bit-floor PublicExperimental 1-bit / ternary LLM inference in Rust — weights as XNOR+POPCNT, GPU buffers managed by Drop, no dequantization ever.
Rust
- Axon-Search
Axon-Search PublicHybrid semantic search engine combining BM25 + FAISS retrieval, fused via RRF and cross-encoder reranked, backed by an async politeness-aware crawler.
Python
- Fine-Tuning-Qwen3-8B
Fine-Tuning-Qwen3-8B PublicFine-tuned Qwen3-8B on math — includes the data leak I found in my own benchmark and the GSM8K regression I caused and fixed.
Python
- sift-core
sift-core PublicHybrid search from scratch in Rust: LSM storage, BM25, flat vector index, RRF fusion. Zero deps.
Rust
If the problem persists, check the GitHub status page or contact support.
Uh oh!
There was an error while loading. Please reload this page.

