Optimum Intel
Model ● Covered in 4 stories + Follow
Optimum Intel is an open-source optimization library developed through a partnership between Intel and Hugging Face that enables efficient inference of machine learning models on Intel hardware. The tool applies quantization, pruning, and other optimization techniques to reduce model latency and size, allowing models like SetFit, Phi-2, and BGE embeddings to run on Intel Xeon CPUs and consumer laptops with minimal accuracy loss. Recent applications include achieving 7.8x faster SetFit inference, running Microsoft's Phi-2 language model on standard laptops, and delivering 4.5x speedups for retrieval-augmented generation pipelines.
Updated 8 August 2026
Specifications
No specifications recorded yet.
Latest developments
Q2 2024
Q1 2024
- A Chatbot on your Laptop: Phi-2 on Intel Meteor Lake
- CPU Optimized Embeddings with 🤗 Optimum Intel and fastRAG
Q2 2022
Relationships
Products & technology
- Hugging Face develops this model · 2 sources
- Integrated with SetFit · 1 source
- Intel develops this model · 1 source
- Phi-2 integrated with this model · 1 source