Upload date
All time
Last hour
Today
This week
This month
This year
Type
All
Video
Channel
Playlist
Movie
Duration
Short (< 4 minutes)
Medium (4-20 minutes)
Long (> 20 minutes)
Sort by
Relevance
Rating
View count
Features
HD
Subtitles/CC
Creative Commons
3D
Live
4K
360°
VR180
HDR
11,475 results
In this episode, I talk with Ryan Vogel about Jev, a new type of AI built for classification. Ryan shows how Jev takes an input plus ...
458,874 views
2 days ago
Pithagoras : https://github.com/thecodacus/pithagoras #localai #voiceagent #gpt6 #astra #jarvis #localagent #llamacpp #voiceai.
68,136 views
7 days ago
What local AI models can you run on a regular computer, without a powerful desktop GPU? I ran six LLMs, from 5B to 35B ...
60,988 views
4 days ago
Spark-X2.5-4B is a compact 4-billion parameter model configured with a native 1048576-token context window. While fitting a 4B ...
22,656 views
Strix Halo mini PC vs Mac Studio: AMD's $3.5K unified-memory box claims to run 200B parameter models cheaper, but real ...
27,119 views
6 days ago
AMD CEO Lisa Su just made local AI a lot more interesting. For years, running powerful AI models meant paying for cloud GPUs, ...
30,755 views
A person on Reddit asked those who spent more than $10000 on a local AI setup whether they regret it. Meanwhile a used RTX ...
11,001 views
This video shows how I built a home AI server with only used parts. I walk you through the build process, installing the software ...
40,033 views
5 days ago
We evaluate inference speed (~68–70 tokens/sec) of Ternary Bonsai 2 27B and put its agentic coding capabilities to the test with ...
81,626 views
3 days ago
... for Your Local AI Box https://www.youtube.com/watch?v=tjFNIQBFKMs • The Stack (YouTube): Xiaomi's AI Cube vs DGX Spark ...
14,581 views
I Spent $5000 On Local AI Hardware — Here's What Won | RTX 5090 vs Mac Studio | local AI hardware | self hosted llm | AI ...
6,855 views
If you have been wondering how local AI video generation works on low VRAM setups, these results provide a realistic look at ...
4,775 views
Looking for the best GPU for your Local AI setup, LLM inference, or fine-tuning without overspending? In this video, we break ...
83,962 views
... Patreon: https://www.patreon.com/cw/LukesDevLab #localllm #localai #homelab #llamacpp #homelab #openai #qwen #flash ...
113,229 views
Bonsai 2 packs a 27B language model into about 5.9 GB, with CPU execution and partial GPU offloading opening up even more ...
36 views
1 hour ago
In this video we take apart FreeToken, the Berkeley and Austin engine that runs mixture of experts models up to 753B parameters ...
7,069 views
#LocalAI #AIcoding #GPTOSS #Qwen3 #Qwen35 #OpenCode #Ollama #AMD #RX9060XT #LLM #LocalLLM.
395 views
RTX 5090 local AI inference just got up to 1.9× faster according to NVIDIA, thanks to new llama.cpp optimizations. But does that ...
1,000 views
Discover how to run serious AI at home without breaking the bank, as Lon Seidman reveals his tricks for turning old data center ...
12,816 views
Your four gigabyte GPU can help fix a checkout bug. The wrong model could turn that small repair into your entire evening.
16,529 views
Show more