Skip to content
View KhanUzeb's full-sized avatar
💭
God Knows What Am I doing
💭
God Knows What Am I doing

Block or report KhanUzeb

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
KhanUzeb/README.md

Uzeb Khan

AI/LLM Engineer • CS (Data Science) • Delhi

Building intelligent systems with LLMs, retrieval, and agentic workflows.

PortfolioXLinkedInEmail

Visitor Count


About

I build AI systems that combine machine learning with modern LLM workflows — agentic pipelines, retrieval, and multi-agent orchestration.

Focus areas

  • Agentic AI & multi-agent systems
  • Retrieval-Augmented Generation (RAG)
  • Deep learning & computer vision
  • Inference & LLM infrastructure

Open to Agentic AI / LLM Engineer roles & internships.


Tech

LanguagesPythonSQLC++Go
AI/MLPyTorchscikit-learnTransformersOpenCV
BackendFastAPIPostgreSQLRedis
ToolsGitLinuxDocker


GitHub Streak


Pinned Loading

  1. AXIOMAXIOMPublic

    Multi-agent AI system for math, physics, and chemistry problems using LangGraph orchestration with parallel specialist agents

    Python 6

  2. DISPATCHDISPATCHPublic

    Semantic LLM router that classifies prompts into cheap/mid/hard and routes to Groq, OpenRouter, or your upstream. OpenAI/Anthropic-compatible API, MCP server, and chat demo.

    Python 1

  3. VERISVERISPublic

    Veris: agent-agnostic local LLM evaluation and regression testing gate. Runs golden datasets against any callable, scores with DeepEval, compares rolling SQLite baselines, and returns CI-ready exit…

    Python 1

  4. RECALLRECALLPublic

    RAG system that remembers. Answers from episodic memory in ~260ms instead of ~49s. Improved Latency for answers

    Python 3

  5. SIFTSIFTPublic

    SIFT is an autonomous AI research agent using LLM tool calling, web search, and a dynamic reasoning loop. Built with FastAPI, Streamlit, and OpenAI-compatible APIs

    Python 3

  6. KILNKILNPublic

    Fine-tune LLMs on your GPU. CLI tool for QLoRA fine-tuning, DPO alignment, and eval — concurrent jobs, auto-queuing, live metrics.

    Python 2