Новый проект
amitshekhariitbhu/llm-inference-engineering
Learn LLM Inference Engineering step by step - from KV cache, PagedAttention, and continuous batching to vLLM, SGLang, and GPUs.
- Stars
- ★ 221
- Forks
- ⑂ 27
- Язык
- Markdown
- Статус
- В канале 1 янв. 1 г.
AI-разбор
Learn LLM Inference Engineering step by step - from KV cache, PagedAttention, and continuous batching to vLLM, SGLang, and GPUs.