🤗 Hugging Cast S2E6 - Scale LLMs with Intel Gaudi and Xeon

From the creator

Hugging Cast is a live show about building AI with open source. In this episode, Regis, Ella, Ilyas and Jeff show you how you can accelerate and scale your Gen AI workloads using the latest Intel AI Accelerators, Gaudi 3 and Xeon CPUs, easily with our open source libraries Optimum Intel, Optimum Habana, TGI Gaudi and more. Last we show you how you can run your own benchmarks easily using Optimum Benchmark. Useful resources: https://github.com/huggingface/optimum-intel https://github.com/huggingface/optimum-habana https://github.com/huggingface/tgi-gaudi https://huggingface.co/docs/optimum/main/en/intel https://huggingface.co/docs/optimum/main/en/habana Discussed examples: https://huggingface.co/blog/intel-starcoder-quantization https://huggingface.co/blog/cost-efficient-rag-applications-with-intel https://huggingface.co/blog/universal_assisted_generation https://huggingface.co/blog/setfit-optimum-intel

Choose to Build with AI
Matched to Llama

AI Maker Residence 3

The third AI workshop taught by our legendary teacher, Nick Sarafa. This is a full-day hands-on training workshop for purposeful co-creation with AI using Claude Code. Imagine having access to hundreds of billions of dollars of computing power and knowing exactly how to make it work for you through the power of super intelligence.

◆ Fri 09 Oct 2026 ◆ KOKO Cafe, London ◆ With Nick Sarafa
AI Maker Residence 3
Live event
AI Maker Residence 3
Fri 09 Oct 2026

More like this

Running one yourself?

List your AI event,
wherever it is.

A meetup, a workshop, a hackathon, a conference. Any city, or online. Tell us about it and it lands in front of people already learning this stuff.