FazBrowse GitHub Viewer | Trending |
URL:
| Home
Tools: [Original HTTPS Page]

heurry · GitHub

🎯
Focusing
🎯
Focusing

Block or report heurry

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Close all issues, pull requests, and discussions opened by this user Content in all repositories owned by your account will be closed.
Add an optional note
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
heurry/README.md

Hi, I'm heurry

AI infrastructure and AI agent engineer.

I build and study systems around LLM serving, retrieval-augmented generation, agent workflows, evaluation, and production-oriented AI tooling. My current focus is turning LLM capabilities into reliable software systems: serving backends, retrieval pipelines, tool-using agents, benchmarks, and operational scripts.

Focus

  • LLM serving, inference benchmarking, and deployment workflows
  • Retrieval-augmented generation systems for domain knowledge and code assistance
  • AI agent workflows with tools, memory, planning, and task execution
  • Evaluation, observability, and diagnostics for LLM applications
  • Reproducible engineering: scripts, benchmarks, configs, and deployment notes

Selected Projects

Project What it shows Stack
vllm-ai_infra Vehicle diagnostic knowledge retrieval and code-generation platform skeleton Python, LLM, RAG
TwinForge Learning platform for LLM training, inference, and local AI infrastructure Python, CUDA, LLM infra
AI Agent Systems Tool-using agent workflows for search, code assistance, automation, and task execution Python, LLM, agents
Evaluation & Benchmarks Scripts and experiments for measuring LLM serving behavior, latency, and reliability Python, Shell, benchmarking

Working Notes

I prefer small, measurable systems over demo-only code:

  • benchmark before tuning
  • keep prompts, tools, configs, and serving scripts versioned together
  • document runtime assumptions and system dependencies
  • make failure modes visible through logs, probes, evaluations, and repeatable tests
  • design agents around clear tool contracts instead of opaque prompt-only behavior

Toolbox

Python PyTorch vLLM RAG AI Agents CUDA Docker Linux Shell FastAPI LangChain LlamaIndex

Contact

  • GitHub: @heurry
  • Interests: AI infrastructure, LLM applications, RAG, AI agents, evaluation, and automation

Popular repositories Loading

  1. vllm-ai_infra vllm-ai_infra Public

    This repository contains the initial implementation skeleton for a vehicle diagnostic knowledge retrieval and code generation platform.

    Python 2

  2. ros1- ros1- Public

    将在一张中的双目图像分割为左右两张图片发布出去

    CMake 1

  3. pdf_seclect_pic pdf_seclect_pic Public

    选择出pdf中带有图片的页码

    Python 1

  4. flowVQA flowVQA Public

    将flowVQA的数据提取边界框信息

    Python 1

  5. PDF-Mask2Former PDF-Mask2Former Public

    Python 1

  6. TwinForge TwinForge Public

    面向云原生微服务场景的分布式基础设施管理平台:配置管理、服务治理、可观测监控、CI/CD 自动化、弹性扩缩容与 AIOps 故障诊断。Go 控制面 + React 控制台 + Python AI Service,统一控制台完成服务管理、资源观测、发布追踪与故障诊断。

    TypeScript 1


Back | FazBrowse Home | New Git URL