Skip to content
View heurry's full-sized avatar
🎯
Focusing
🎯
Focusing

Block or report heurry

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
heurry/README.md

Hi, I'm heurry

AI infrastructure and AI agent engineer.

I build and study systems around LLM serving, retrieval-augmented generation, agent workflows, evaluation, and production-oriented AI tooling. My current focus is turning LLM capabilities into reliable software systems: serving backends, retrieval pipelines, tool-using agents, benchmarks, and operational scripts.

Focus

  • LLM serving, inference benchmarking, and deployment workflows
  • Retrieval-augmented generation systems for domain knowledge and code assistance
  • AI agent workflows with tools, memory, planning, and task execution
  • Evaluation, observability, and diagnostics for LLM applications
  • Reproducible engineering: scripts, benchmarks, configs, and deployment notes

Selected Projects

ProjectWhat it showsStack
vllm-ai_infraVehicle diagnostic knowledge retrieval and code-generation platform skeletonPython, LLM, RAG
TwinForgeLearning platform for LLM training, inference, and local AI infrastructurePython, CUDA, LLM infra
AI Agent SystemsTool-using agent workflows for search, code assistance, automation, and task executionPython, LLM, agents
Evaluation & BenchmarksScripts and experiments for measuring LLM serving behavior, latency, and reliabilityPython, Shell, benchmarking

Working Notes

I prefer small, measurable systems over demo-only code:

  • benchmark before tuning
  • keep prompts, tools, configs, and serving scripts versioned together
  • document runtime assumptions and system dependencies
  • make failure modes visible through logs, probes, evaluations, and repeatable tests
  • design agents around clear tool contracts instead of opaque prompt-only behavior

Toolbox

PythonPyTorchvLLMRAGAI AgentsCUDADockerLinuxShellFastAPILangChainLlamaIndex

Contact

  • GitHub: @heurry
  • Interests: AI infrastructure, LLM applications, RAG, AI agents, evaluation, and automation

Popular repositories Loading

  1. vllm-ai_infra vllm-ai_infraPublic

    This repository contains the initial implementation skeleton for a vehicle diagnostic knowledge retrieval and code generation platform.

    Python 2

  2. ros1- ros1-Public

    将在一张中的双目图像分割为左右两张图片发布出去

    CMake 1

  3. pdf_seclect_pic pdf_seclect_picPublic

    选择出pdf中带有图片的页码

    Python 1

  4. flowVQA flowVQAPublic

    将flowVQA的数据提取边界框信息

    Python 1

  5. PDF-Mask2Former PDF-Mask2FormerPublic

    Python 1

  6. TwinForge TwinForgePublic

    面向云原生微服务场景的分布式基础设施管理平台:配置管理、服务治理、可观测监控、CI/CD 自动化、弹性扩缩容与 AIOps 故障诊断。Go 控制面 + React 控制台 + Python AI Service,统一控制台完成服务管理、资源观测、发布追踪与故障诊断。

    TypeScript 1