Upload date
All time
Last hour
Today
This week
This month
This year
Type
All
Video
Channel
Playlist
Movie
Duration
Short (< 4 minutes)
Medium (4-20 minutes)
Long (> 20 minutes)
Sort by
Relevance
Rating
View count
Features
HD
Subtitles/CC
Creative Commons
3D
Live
4K
360°
VR180
HDR
38,624 results
Two AI model files can both say “4-bit” and still have completely different quality, memory requirements and performance.
87,825 views
4w ago
Applied AI Course: https://arpitbhayani.me/applied-ai System Design for SDE-2 and above: https://arpitbhayani.me/masterclass ...
23,551 views
2w ago
Stop worrying about GPU memory constraints. Learn how AI model quantization reduces model size and speeds up inference ...
378 views
I quantized Qwen3.8-27B — and used the process to look at what “4-bit” actually means in practice. Quantization is usually ...
15,532 views
3w ago
In this video, I test Qwen3.8-27B GSQ-RCO IQ2_XS on an Apple M4 Mac Mini with 24GB RAM and explore what makes this ...
17,358 views
ISTA vs Unsloth: which Qwen3.8 27B quant should you actually download? I tested seven variants on an RTX 5090, and ISTA's ...
1,476 views
9h ago
Your local LLM may be running a quantized file you never chose. Q4, Q8, GGUF formats, KV cache precision, and Ollama or LM ...
56,578 views
In this video I will check out the GSQ RCO quant of Qwen 3.8 27B from ISTA DAS Lab Austria, can this quantization method ...
67,761 views
How do you choose the right local AI model — and what do labels like 14B, INT8, Q6 and GGUF actually mean? In this video, I ...
1,967 views
RebelUI: https://github.com/RealRebelAI/RebelUI/tree/main BUYMEACOFFEE: buymeacoffee.com/realrebelai #quantization ...
3,733 views
Welcome to Module 8. This is the deployment module — the one where everything we have built over the last seven modules ...
14 views
On the surface Skopje looks like a simple quantiser module. Two knobs, two channels, two outputs. But the real power isn't on the ...
3,815 views
7 views
NVIDIA Model Optimizer GitHub by NVIDIA: https://github.com/NVIDIA/Model-Optimizer NVIDIA Model Optimizer helps engineers ...
119 views
1d ago
Two for one! 00:00 Intro 00:27 Logic Pro 07:34 Logic Pro Swing 09:27 Logic Pro Input Quantization 11:10 Ableton Live 15:30 ...
262 views
A fast-paced explanation of quantization's major concepts - covering bit-width, symmetric/asymmetric, granularity, ...
182 views
deeplearning #computervision.
705 views
Four bit quantization can shrink a model's raw weight memory by about 75 percent. The catch is that “smaller” does not always ...
433 views
Summary* The video explains how large language models represent words as numbers in vectors and use quantization by ...
179 views
98 views
Show more