Algorithm engineer exploring ML, neural networks, and large models.
- Love machine learning
- Love neural networks
- Love large models
- Blog: lustar-blog.com
- Email: luxing999yh@gmail.com
- Email: 2463553352@qq.com
Algorithm engineer exploring ML, neural networks, and large models.
Reproducible solo solution workspace for the CUHK-X Large Model Track
Python 2
Lightweight OpenAI-compatible LLM gateway: multi-upstream routing, managed credentials, and real-time token accounting with a live dashboard.
Go 2
Forked from ray-project/ray
Ray is an AI compute engine. Ray consists of a core distributed runtime and a set of AI Libraries for accelerating ML workloads.
Python
Forked from pegainfer-project/pegainfer
Pure Rust + CUDA LLM inference engine — no PyTorch, OpenAI-compatible, serves Qwen3 to Kimi-K2
Rust
Small-GPU VLM inference experiments on a 4 GB RTX 3050 Ti: vLLM memory boundaries, engine comparison, and prefix caching, with raw data and measurement pitfalls
Python