TLDRocket
Sign in

Hardware & Infrastructure

256 summarised stories in Hardware & Infrastructure, each linking back to the original source. Browse all topics →

Thursday, 29 February 2024

Text-Generation Pipeline on Intel® Gaudi® 2 AI Accelerator

Hugging Face 2 years ago 13

Intel Habana released a text-generation pipeline for running Llama 2 models on Gaudi 2 accelerators, available through the Optimum Habana library version 1.10.4. The pipeline supports three model sizes (7 billion, 13 billion, and 70 billion parameters) and can generate text from prompts with configurable sampling parameters like temperature and top_p values. Users can integrate the pipeline directly into Python scripts or use it with LangChain for building applications, requiring only Hugging Face account access and agreement to Llama 2's community license.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.