TLDRocket
Sign in

Tools & Coding

975 summarised stories in Tools & Coding, each linking back to the original source. Browse all topics →

Monday, 15 January 2024

Accelerating SD Turbo and SDXL Turbo Inference with ONNX Runtime and Olive

Hugging Face 2 years ago 16

ONNX Runtime has optimized inference for SD Turbo and SDXL Turbo image generation models through CUDA and TensorRT execution providers on NVIDIA GPUs. ONNX Runtime achieved throughput gains as high as 229% for SDXL Turbo and 120% for SD Turbo compared to PyTorch across tested batch sizes and step counts. The optimizations enable faster image generation in C#, Java, and Python, with optimized model versions now available on Hugging Face.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.