Full program
A person interacting with a compact AI model
AlphaEdge logo
Théo Hubert
For those who implement · 11:15 · 40 min

Small language models, fine-tuning & compression

Key facts

Track
Those who implement
Time
11:15 → 11:55
Duration
40 min
Format
Talk
Language
French

“My model doesn’t fit on an embedded chip”

Description

Today, most AI runs in the cloud. Does it have to? This talk shows in practical terms how to run a capable language model locally, on a PC or a board like the Jetson Nano.

In a constrained environment, arbitrarily shrinking a model isn’t enough. We’ll look at how SLMs (Small Language Models) maintain good performance through different approaches: trimming, quantization, pruning and fine-tuning. We’ll then go further by comparing these optimization techniques with another approach: designing architectures built natively for the edge, and therefore suited to compute constraints from the start.

  • SLM
  • Model compression
  • Embedded AI

On stage

COLLECTION COMPLETE

You found them all.

Well done! You just completed the GENAI DAYS sticker easter egg. A special reward is waiting for you on the day of the event. It is reserved for GENAI DAYS attendees and will only be handed out on site.

Claim my reward