« All posts

LFM2.5-Encoders: Fast Long-Context Inference on CPU

LFM2.5-Encoders released on Hugging Face offer fast performance for long-context inference on CPUs.

Two new models, LFM2.5-Encoder-230M and LFM2.5-Encoder-350M, have been released on Hugging Face. These encoders deliver performance comparable to larger models while maintaining speed with longer inputs. For engineers, this means the ability to perform document-scale tasks on existing hardware, even on CPU.

This synthesis was produced from its source by AI; there is no human editor or manual review step. How we work