Uh oh!
There was an error while loading. Please reload this page.
Uh oh!
There was an error while loading. Please reload this page.
Fankserver
Pinned Loading
Repositories
- llm-glossary Public
Plain-language glossary of LLM & inference-serving terminology (MoE, MLA, DSA, KV cache, speculative decoding, ...)
Uh oh!
There was an error while loading. Please reload this page.
fankserver/llm-glossary's past year of commit activity - dgx-spark-laguna-s21-nvfp4-bench Public
Benchmarks for poolside Laguna-S-2.1 (118B-A8B coder) NVFP4 on a single NVIDIA GB10 / DGX Spark (vLLM 0.25.1, TP=1) — DFlash speculative decoding, KV dtype, TP scaling; llama-benchy sweeps with the Qwen-122B depth set for comparison
Uh oh!
There was an error while loading. Please reload this page.
fankserver/dgx-spark-laguna-s21-nvfp4-bench's past year of commit activity - dgx-spark-glm-4.7-awq-bench Public
Benchmarks for GLM-4.7 355B AWQ (QuantTrio W4A16) with vLLM TP=2 across 2x NVIDIA GB10 / DGX Spark over 200G fabric — llama-benchy sweeps, datasets, graphs
Uh oh!
There was an error while loading. Please reload this page.
fankserver/dgx-spark-glm-4.7-awq-bench's past year of commit activity - dgx-spark-nemotron-puzzle-75b-bench Public
DGX Spark (GB10) llama-benchy benchmarks: NVIDIA Nemotron-Puzzle-75B-A9B NVFP4 with MTP speculative decoding on stock vLLM
Uh oh!
There was an error while loading. Please reload this page.
fankserver/dgx-spark-nemotron-puzzle-75b-bench's past year of commit activity - dgx-spark-nemotron-120b-bench Public
Benchmarks for NVIDIA Nemotron-3-Super-120B-A12B (NVFP4, hybrid MoE) with vLLM+MTP on a single NVIDIA GB10 / DGX Spark — llama-benchy sweeps, datasets, graphs
Uh oh!
There was an error while loading. Please reload this page.
fankserver/dgx-spark-nemotron-120b-bench's past year of commit activity - dgx-spark-hy3-295b-bench Public
Benchmarks for Tencent Hy3-295B (Hunyuan 3, NVFP4 MoE) with vLLM+Ray at TP=2 across 2x NVIDIA GB10 / DGX Spark — llama-benchy sweeps, datasets, graphs
Uh oh!
There was an error while loading. Please reload this page.
fankserver/dgx-spark-hy3-295b-bench's past year of commit activity - dgx-spark-qwen3.6-27b-aeon-bench Public
DGX Spark (GB10) llama-benchy benchmarks: Qwen3.6-27B-AEON — 3 vLLM engines, BF16 vs NVFP4, DSpark-DFlash draft head
Uh oh!
There was an error while loading. Please reload this page.
fankserver/dgx-spark-qwen3.6-27b-aeon-bench's past year of commit activity - gfx906-LLM-Inference Public
LLM inference benchmarks, tuning results, and production notes for AMD MI50 (gfx906 / Vega 20) on ROCm — covering Ornith-1.0-35B, Gemma-4-26B, Qwen3.6-35B, MTP, DFlash, EAGLE3, and Docker vs bare-metal comparisons.
Uh oh!
There was an error while loading. Please reload this page.
fankserver/gfx906-LLM-Inference's past year of commit activity - dgx-spark-qwen3.5-122b-bench Public
Benchmarks for Qwen3.5-122B-A10B (hybrid MoE) with vLLM+DFlash on a single NVIDIA GB10 / DGX Spark — llama-benchy sweeps, datasets, graphs
Uh oh!
There was an error while loading. Please reload this page.
fankserver/dgx-spark-qwen3.5-122b-bench's past year of commit activity - vanguard-galaxy-dataexport Public
BepInEx plugin that dumps Vanguard Galaxy's live catalogue (ships, turrets, armor, aspects, factions, conquest ranks) to JSON, plus a Python merger for the wiki's Module:ShipData.
Uh oh!
There was an error while loading. Please reload this page.
fankserver/vanguard-galaxy-dataexport's past year of commit activity
Top languages
Loading…
Uh oh!
There was an error while loading. Please reload this page.
Most used topics
Loading…
Uh oh!
There was an error while loading. Please reload this page.