Skip to content
View esonhjz's full-sized avatar
🎯
Focusing
🎯
Focusing
  • UC Berkeley
  • Berkeley, CA
  • 22:49 (UTC -12:00)

Block or report esonhjz

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
esonhjz/README.md
Moe Counter

👤 About Me

  • 🎓 Status: First-year Statistics student at UC Berkeley.
  • 🧑‍💻 Identity: A developer building ML systems, computer-vision pipelines, and interactive character applications.
  • 🚣 Sports: A passionate Canoe Polo athlete.
  • 🎯 Focus: Exploring multimodal AI, efficient LLM inference, model evaluation, and interactive systems.
  • 🎮 Daily Life: Coding, watching anime, gaming, and turning ideas into reality.

Here is a simple index of the projects I am currently working on.

💻 Self-Developed Projects

  • 🛟 H20Saver: A PyTorch/Ultralytics YOLO11x fine-tuning pipeline for drowning-target detection, including dataset analysis, training, validation, and ONNX export.
  • 🔀 airi-llm-router: A FastAPI and asyncio gateway for local LLM inference, with payload-aware queues and GPU-memory backpressure.
  • 🖥️ OwO-Desktop: A Windows Live2D desktop character application with WebSocket controls, a Python LLM/TTS backend, and expression, motion, and lip-sync integration.
  • 🦙 llama-cpp-moe-vram-benchmarks: An automated llama.cpp benchmark/report exploring CPU/GPU MoE offload, 4-bit KV cache, and Flash Attention on consumer hardware.

👥 Team Projects

  • 🌸 My-Spiritual-Solace: A collaborative Live2D dynamic model & resource sharing repository bringing favorite anime characters to life on screen.
  • 🌐 Layer0-kol: Team collaboration project.

🛠️ Skills & Technologies

Python, C, C++, Java, Bash, PowerShell, JavaScript, TypeScript, Node.js, HTML, CSS, Markdown, PyTorch, OpenCV, FastAPI, Docker, Git, GitHub, CMake, Windows, Linux, VS Code, Obsidian, Gmail, and Photoshop

Also Working With

  • ML & Data: Jupyter notebooks, CUDA, ComfyUI, ONNX, Pandas, NumPy, pynvml
  • Native & Interactive: llama.cpp, Direct3D 11, Live2D Cubism Native SDK, IXWebSocket, nlohmann/json, Asyncio

Pinned Loading

  1. H20Saver H20Saver Public

    High-efficiency Drowning Target Detection System 🏊

    Python 20 4

  2. airi-llm-router airi-llm-router Public

    A high-concurrency, localized LLM gateway designed with Python Asyncio and FastAPI. Built to decouple multi-modal traffic from the Airi core system and manage dynamic routing.

    Python 2

  3. ewww-759/Layer0-kol ewww-759/Layer0-kol Public

    Python 1

  4. My-Spiritual-Solace My-Spiritual-Solace Public

    Making beloved characters come alive on the screen - A shared repository for Live2D models.

    2

  5. OwO-Desktop OwO-Desktop Public

    A Windows Live2D desktop character application with Direct3D 11 rendering and WebSocket control

    C++ 2

  6. llama-cpp-moe-vram-benchmarks llama-cpp-moe-vram-benchmarks Public

    Run MoE models on consumer GPUs by offloading Attention to GPU and shunting experts to CPU — 75% less VRAM, 9× faster than brute force.

    PowerShell 1