Comparison
llama.cpp vs LMDeploy
A like-for-like, spec-level comparison. Both entries are verified against their docs and repo.
llama.cpp
C/C++ inference for GGUF models on any hardware.
LMDeploy
Efficient inference and serving for open models.
| Spec | llama.cpp | LMDeploy |
|---|---|---|
| Role | Inference engine | Inference engine |
| Tags | inference-engine | inference-engine |
| License | MIT | Apache-2.0 |
| Open source | Yes | Yes |
| Self-hostable | Yes | Yes |
| MCP support | No | No |
| Pricing | free | free |
| Price | Free | Free |
| Usage cost | No model cost | No model cost |
| Models | local, multi | multi |
| Languages | cpp | python, cpp |
| GitHub stars | 119.2k | 7.9k |
| Last activity | 2026-07-03 | 2026-07-03 |