TMLR 2026 | Mechanistic interpretability: attention-head binding (EB*) as a marker of concept emergence. 7 models, 5 architectures (Pythia 160M–2.8B, OLMo-1B, CRFM GPT-2, SmolLM3-3B, Qwen2.5-1.5B), 41 terms.
-
Updated
Jun 9, 2026 - Python
TMLR 2026 | Mechanistic interpretability: attention-head binding (EB*) as a marker of concept emergence. 7 models, 5 architectures (Pythia 160M–2.8B, OLMo-1B, CRFM GPT-2, SmolLM3-3B, Qwen2.5-1.5B), 41 terms.
PyTorch DDP and multi-GPU LLM training — from a minimal distributed example to NanoGPT speedruns, Muon, H100 profiling, and modded-nanogpt. FBA LAB https://bubblnet.com
OMTR — can LLM memorization be separated from predictability? Preregistered causal probing (activation patching) in Pythia & OLMo on consumer hardware; three honest non-separations, full corrections history
First open-source descriptor-augmented LLM for Neglected Tropical Disease drug discovery | OLMo-7B + QLoRA + DeepChem | Bioactivity & Toxicity prediction for Leishmaniasis, Chagas, Malaria, TB
Research workspace for model diffing between pretrained and post-trained language models.
OpenEuroLLM snapshot of the Berkeley Function Calling Leaderboard evaluation harness and OLMo evaluation orchestration.
OpenEuroLLM tau2-bench evaluation harness with OLMo and Qwen serving, user-simulator, and aggregation scripts.
Paired OLMo continual-training experiment on spaced review and delayed FictionalQA retention
Add a description, image, and links to the olmo topic page so that developers can more easily learn about it.
To associate your repository with the olmo topic, visit your repo's landing page and select "manage topics."