#cpu inference

Cpu Inference: 4 AI articles covering cpu inference news, analysis, and research

Articles

Accelerating LLM Inference via Vector Index Based Output Embeddings
Speculative Decoding on CPUs: Nearly 4x Faster Token Generation
Liquid AI Releases LFM2.5-Encoder-230M and LFM2.5-Encoder-350M:
LFM2.5-Encoders for Fast Long-Context Inference on CPU