HN Mail
Subscribe
CPP
Llama.cpp Under the Hood
4 points
|
0 comments
Transformers now runs llama.cpp quants
4 points
|
0 comments
Custom Models in Oh My Pi: vLLM, Llama.cpp, SGLang and More
2 points
|
0 comments
Show HN: Llama.cpp fork with 2-4x multiGPU speed for MoE models bigger than VRAM
1 points
|
2 comments
Using Llama-cpp-Python grammars to generate JSON
1 points
|
0 comments
Show HN: IttyBittyAI – Qwen2.5:0.5B running on a 2017 Asus smartphone
2 points
|
0 comments
Show HN: A benchmark comparison of C++ vs. Node.js performance
5 points
|
1 comments