Medium / Medium.com – Telegram

Medium / Medium.com

1.23K subscribers

106K links

Just main page of medium.com fresh from the oven

Download Telegram

About

Blog

Apps

Platform

Medium / Medium.com

1.23K subscribers

Medium / Medium.com

KV Cache Manager: The Key Idea Behind It and How It Works

#llms #pagedattention #kvcachemanager #kvcache #vllm #virtualmemory #kvblocks #gpuworkers

https://hackernoon.com/kv-cache-manager-the-key-idea-behind-it-and-how-it-works

KV Cache Manager: The Key Idea Behind It and How It Works

The key idea behind vLLM’s memory manager is analogous to the virtual memory [25] in operating systems.

14 views17:45

Medium / Medium.com

The Distributed Execution of vLLM

#llms #vllm #megatronlm #memorymanager #spmd #modelparallel #kvcachemanager #kvcache

https://hackernoon.com/the-distributed-execution-of-vllm

The Distributed Execution of vLLM

vLLM is effective in distributed settings by supporting the widely used Megatron-LM style tensor model parallelism strategy on Transformers

19 views00:30