First commit.
This commit is contained in:
@@ -0,0 +1,5 @@
|
||||
llama.cpp is good for cpu/gpu inferencing
|
||||
vllm is good for tensor parallelism
|
||||
ollama sits on top of llama.cpp and shares most of its advantages and shortcomings
|
||||
|
||||
#ai #llama #vllm
|
||||
Reference in New Issue
Block a user