Technique

Quantization

Quantization is a method for shrinking large machine learning models, allowing them to run on smaller devices like phones and laptops. It’s especially relevant for anyone working with large language models (LLMs) who wants to optimise performance without sacrificing too much accuracy. Videos under this topic often cover practical techniques, such as compressing models using Python, fine-tuning LLMs like Mistral Small, and leveraging QLoRA for working with limited GPU resources.

Choose to Build with AI
Matched to Quantization

AI Maker Residence 3

The third AI workshop taught by our legendary teacher, Nick Sarafa. This is a full-day hands-on training workshop for purposeful co-creation with AI using Claude Code. Imagine having access to hundreds of billions of dollars of computing power and knowing exactly how to make it work for you through the power of super intelligence.

◆ Fri 09 Oct 2026 ◆ KOKO Cafe, London ◆ With Nick Sarafa
AI Maker Residence 3
Live event
AI Maker Residence 3
Fri 09 Oct 2026
People watching Quantization also follow
Everything, filed properly

Browse by
what it's about

Choose To Studio

Want this working for you?

Let us build it for you. Design and engineering from the people who shipped platforms to billions of users. AI-native, live in weeks, and yours outright at the end.