ViewTube

ViewTube
Sign inSign upSubscriptions
Filters

Upload date

Type

Duration

Sort by

Features

Reset

571,874 results

Perimeter Institute for Theoretical Physics
Quantization Explained | Perimeter Institute for Theoretical Physics

Some of the most important breakthroughs in physics came about due to the discovery that energy is quantized. This video ...

4:36
Quantization Explained | Perimeter Institute for Theoretical Physics

114,428 views

3 years ago

Julia Turc
How LLMs survive in low precision | Quantization Fundamentals

In this video, we discuss the fundamentals of model quantization, the technique that allows us to run inference on massive LLMs ...

20:34
How LLMs survive in low precision | Quantization Fundamentals

73,871 views

1 year ago

KodeKloud
LLM Quantization Explained

LLM quantization is how a 70B model that needs 140GB of memory gets small enough to run on a normal GPU. Every model you ...

4:18
LLM Quantization Explained

18,804 views

1 month ago

Julien Simon
Deep Dive: Quantizing Large Language Models, part 1

Quantization is an excellent technique to compress Large Language Models (LLM) and accelerate their inference. In this video ...

40:28
Deep Dive: Quantizing Large Language Models, part 1

24,838 views

2 years ago

Airtrain AI
What is LLM quantization?

In this video we define the basics of quantization and look at how its benefits and how it affects large language models.

5:13
What is LLM quantization?

39,052 views

2 years ago

Umar Jamil
Quantization explained with PyTorch - Post-Training Quantization, Quantization-Aware Training

In this video I will introduce and explain quantization: we will first start with a little introduction on numerical representation of ...

50:55
Quantization explained with PyTorch - Post-Training Quantization, Quantization-Aware Training

59,834 views

2 years ago

Matt Williams
Optimize Your AI - Quantization Explained

Run massive AI models on your laptop! Learn the secrets of LLM quantization and how q2, q4, and q8 settings in Ollama can save ...

12:10
Optimize Your AI - Quantization Explained

534,527 views

1 year ago

Julia Turc
Reverse-engineering GGUF | Post-Training Quantization

The first comprehensive explainer for the GGUF quantization ecosystem. GGUF quantization is currently the most popular tool for ...

25:07
Reverse-engineering GGUF | Post-Training Quantization

68,793 views

1 year ago

Akash Murthy
5. Quantization - Digital Audio Fundamentals

In this video, on our quest to create a discrete signal out of a continuous signal, we will begin the discussion on how amplitude ...

9:29
5. Quantization - Digital Audio Fundamentals

107,385 views

6 years ago

BlueSpork
DeepSeek R1: Distilled & Quantized Models Explained

This video explores DeepSeek R1, how distilled versions and quantization make it more accessible, and the trade-offs between ...

3:47
DeepSeek R1: Distilled & Quantized Models Explained

29,193 views

1 year ago

Professor Dave Explains
Quantization of Energy Part 1: Blackbody Radiation and the Ultraviolet Catastrophe

So we know that physics got turned upside down at the turn of the 20th century, but how did that all begin? What was the first thing ...

6:43
Quantization of Energy Part 1: Blackbody Radiation and the Ultraviolet Catastrophe

1,244,545 views

9 years ago

Alex Ziskind
Everything looks fine at 4-bit

I quantized one model 8 ways to find the exact level it starts making things up. Take your personal data back with Incogni!

18:26
Everything looks fine at 4-bit

132,610 views

3 months ago

Zachary Huang
Give me 30 min, I will make Quantization click forever

Text:* https://github.com/The-Pocket/PocketFlow-Tutorial-Video-Generator/blob/main/docs/llm/quantization.md 0:00:00 ...

32:42
Give me 30 min, I will make Quantization click forever

10,898 views

9 months ago

Channels new to you

Devsplainers
LLM Quantization Explained: The Q4 Quality Trap

Your local LLM may be running a quantized file you never chose. Q4, Q8, GGUF formats, KV cache precision, and Ollama or LM ...

9:21
LLM Quantization Explained: The Q4 Quality Trap

56,071 views

2 weeks ago

Cloud Codes
GGUF vs AWQ vs GPTQ: LLM Quantization Methods Explained

Quantization is often sold as "smaller = cheaper = better." But the three things people conflate—VRAM savings, inference speed, ...

9:48
GGUF vs AWQ vs GPTQ: LLM Quantization Methods Explained

855 views

3 months ago

Efficient NLP
Quantization vs Pruning vs Distillation: Optimizing NNs for Inference

Try Voice Writer - speak your thoughts and let AI handle the grammar: https://voicewriter.io Four techniques to optimize the speed ...

19:46
Quantization vs Pruning vs Distillation: Optimizing NNs for Inference

71,088 views

3 years ago

Professor Cunningham
Converting Analog Data to Binary, Sampling, Quantization (AP Computer Science Principles Unit 1)

Strap in, this one's gonna get a bit bumpy. Converting from analog data to digital is a three step process. "Sampling" involves ...

20:23
Converting Analog Data to Binary, Sampling, Quantization (AP Computer Science Principles Unit 1)

19,546 views

3 years ago

MIT OpenCourseWare
Quantization of the energy

MIT 8.04 Quantum Physics I, Spring 2016 View the complete course: http://ocw.mit.edu/8-04S16 Instructor: Barton Zwiebach ...

23:19
Quantization of the energy

30,488 views

9 years ago

Tales Of Tensors
LLM Quantization Explained: GPTQ, AWQ, QLoRA, GGUF and More

00:00 Introduction to LLM Quantization 02:15 What is Quantization? 04:45 Post-Training Quantization (PTQ) vs. QAT 07:30 GPTQ ...

30:14
LLM Quantization Explained: GPTQ, AWQ, QLoRA, GGUF and More

3,846 views

6 months ago

Adam Lucek
Quantizing LLMs - How & Why (8-Bit, 4-Bit, GGUF & More)

Quantizing models for maximum efficiency gains! Resources: Model Quantized: ...

26:26
Quantizing LLMs - How & Why (8-Bit, 4-Bit, GGUF & More)

29,358 views

1 year ago

Show more