#vllm

Vllm: 5 AI articles covering vllm news, analysis, and research

Articles

IFM Unveils K2 Horizon: Six Apache 2.0 Models from 0.9B to 375B
Speculative Decoding on CPUs: Nearly 4x Faster Token Generation
Pokee AI Releases Pokee-Isaac 28B: A 10M-Token Context Agentic
Liquid AI Ships LFM2.5-230M with llama.cpp, MLX, vLLM, SGLang,
Run a vLLM Server on HF Jobs in One Command