Hacker News
new
|
past
|
comments
|
ask
|
show
|
jobs
|
submit
login
lukasc-ch
3 months ago
|
parent
|
context
|
favorite
| on:
KVarN: Native vLLM backend for KV-cache quantizati...
... and it's on llama.cpp that to this guy!
https://www.reddit.com/r/LocalLLaMA/comments/1txlhxu/i_imple...
lukasc-ch
3 months ago
[–]
This is awesome! Let's give them some stars: -
https://github.com/huawei-csl/KVarN
(original repo, vLLM implementation) -
https://github.com/Anbeeld/beellama.cpp
(llama.cpp implementation + awesome evals)
Guidelines
|
FAQ
|
Lists
|
API
|
Security
|
Legal
|
Apply to YC
|
Contact
Search: