Articles
Liquid AI Ships LFM2.5-230M with llama.cpp, MLX, vLLM, SGLang, and ONNX Support for On-Device Inference⭐7
Liquid AI launches LFM2.5-230M, a 230M-parameter model for on-device inference with llama.cpp, MLX, vLLM, SGLang, and ONNX support. Achieves 213 tokens/s on Sam...
Run a vLLM Server on HF Jobs in One Command⭐10
Deploy a vLLM inference server on Hugging Face Jobs in one command—no manual setup needed. Full guide for scalable LLM serving.
