11:04
Model compression involves techniques that shrink large machine learning models to make them faster and more efficient without sacrificing too much accuracy. By using methods like quantisation, as highlighted in NVIDIA’s video about their 4-bit format, practitioners can make AI run twice as fast on local hardware, including innovative setups like using a Mac as a local AI box. The Python code in the 'Compressing Large Language Models' video offers insights into practical applications of these techniques.
11:04
24:04
16:30
The third AI workshop taught by our legendary teacher, Nick Sarafa. In one of the last events we did an asset manager raised an additional £25M on their fund within a space of 9 months. This is a full-day hands-on training workshop for purposeful co-creation with AI using Claude Code. Imagine having access to hundreds of billions of dollars of computing power and knowing exactly how to make it work for you through the power of super intelligence. One person did.
Apply once › get screened › meet the company
Let us build it for you. Design and engineering from the people who shipped platforms to billions of users. AI-native, live in weeks, and yours outright at the end.
One a week, never sold on, and one click to stop.
We use a few cookies to keep the site running. Necessary ones are always on. Analytics + marketing are off until you say otherwise.
Read our cookie policy · privacy.