|
A 54M Chinese LLM built from scratch with native PyTorch, from pretraining to full SFT. |
A lightweight terminal coding agent with tools, MCP, hooks, and Docker sandboxing. |
|
A compact vLLM extension for speculative decoding, chunked prefill, fair scheduling, and prefix caching. |
|




