Upload date
All time
Last hour
Today
This week
This month
This year
Type
All
Video
Channel
Playlist
Movie
Duration
Short (< 4 minutes)
Medium (4-20 minutes)
Long (> 20 minutes)
Sort by
Relevance
Rating
View count
Features
HD
Subtitles/CC
Creative Commons
3D
Live
4K
360°
VR180
HDR
470 results
Bosgame M5 mini PC with 128GB unified memory: can a $2999 local AI desktop actually run 120B models faster than an RTX ...
7,303 views
15h ago
Want to run an LLM on your own hardware? Cedric Clyburn demonstrates self-hosted AI inference with vLLM, from loading an ...
365 views
2h ago
What is the cheapest, and most valuable, way to reach 48GB of memory for local AI in 2026? In this video, I compare nearly every ...
17,904 views
1d ago
... Very Efficient 5:03 - Animation & Video Transcode 6:35 - AI Workload Testing (To 70b & More) 10:28 - AI Processing Efficiency ...
197,036 views
Do local LLMs really run faster on Linux than on Windows? I moved my main machine from Windows 11 to Kubuntu and re-ran ...
616 views
This week on AppStories, John interviews Federico about his review of the M5 Ultra Mac Studio and it local AI capabilities.
756 views
22h ago
Timestamps: 00:00 From Local AI agent to Production 00:42 Why Use Render for AI apps 02:10 Render platform overview 03:28 ...
261 views
Tired of local AI models dragging down your CPU because it isn't leveraging your Snapdragon NPU? In this tutorial, we explore ...
5 views
7h ago
Bonsai 2 packs a 27B language model into about 5.9 GB, with CPU execution and partial GPU offloading opening up even more ...
665 views
21h ago
M5 Ultra 和RTX 5090 跑Local AI,到底该买容量还是速度?结论先说:截至2026-09-16,M5 Ultra 零售机尚未交付,独立同 ...
10 views
1h ago
Subscribe for more self-hosted AI + homelab that runs on hardware you own. #localai #selfhosted #homelab #homeassistant ...
6 views
14h ago
For the first time, Tommy Lam and Peter Muessig team up to show how to bring local AI into your OpenUI5 applications. Tommy ...
180 views
I tested the new M6 Mac mini against an M4 MacBook Air and my fully loaded M4 Max Mac Studio in real video editing, ...
57,747 views
What's unified memory, and why are Apple, AMD, and NVIDIA racing to pack your next laptop with it? Windows Weekly breaks ...
1,683 views
19h ago
Intel Arc long-context AI performance may have received a massive boost through experimental sparse Flash Attention.
38 views
20h ago
DeepSeek v4.1 is hitting 400 tokens per second. We break down the architecture enabling this LLM performance. The speed of ...
14 views
PandatAi connects large language models with the Pandat CALPHAD engine — describe your materials problem in plain ...
55 views
17h ago
This talk presents innovative approaches to building local capacity through human–AI collaboration technologies that enable ...
16 views
Mind Control Helmet Specifications: Core Processing & AI Engine • Single Board Computer: Raspberry Pi 5 • AI Hardware ...
60 views
US data center and information-processing hardware spending has now surpassed housing investment, a line that has never ...
11 views
23h ago
Show more