ViewTube

ViewTube
Sign inSign upSubscriptions
Filters

Upload date

Type

Duration

Sort by

Features

Reset

470 results

The Stack
A 4090 Can't Run This Model. This Box Can

Bosgame M5 mini PC with 128GB unified memory: can a $2999 local AI desktop actually run 120B models faster than an RTX ...

15:38
A 4090 Can't Run This Model. This Box Can

7,303 views

15h ago

IBM Developer and IBM Technology
How to Self-Host an LLM: Local AI Inference with vLLM

Want to run an LLM on your own hardware? Cedric Clyburn demonstrates self-hosted AI inference with vLLM, from loading an ...

11:06
How to Self-Host an LLM: Local AI Inference with vLLM

365 views

2h ago

RepoChad
Best Ways to Get 48GB VRAM for Local AI

What is the cheapest, and most valuable, way to reach 48GB of memory for local AI in 2026? In this video, I compare nearly every ...

17:35
Best Ways to Get 48GB VRAM for Local AI

17,904 views

1d ago

Hardware Canucks
Apple M5 Ultra vs The Fastest PC

... Very Efficient 5:03 - Animation & Video Transcode 6:35 - AI Workload Testing (To 70b & More) 10:28 - AI Processing Efficiency ...

17:12
Apple M5 Ultra vs The Fastest PC

197,036 views

1d ago

Next Tech and AI
Linux vs Windows for Local LLMs — I Tested AMD, Intel and NVIDIA

Do local LLMs really run faster on Linux than on Windows? I moved my main machine from Windows 11 to Kubuntu and re-ran ...

19:30
Linux vs Windows for Local LLMs — I Tested AMD, Intel and NVIDIA

616 views

1d ago

MacStories
Local AI Gets Serious on the M5 Ultra Mac Studio | AppStories

This week on AppStories, John interviews Federico about his review of the M5 Ultra Mac Studio and it local AI capabilities.

39:21
Local AI Gets Serious on the M5 Ultra Mac Studio | AppStories

756 views

22h ago

Sonny Sangha
How to Ship Apps & AI Agents Without Managing Infrastructure (with Durable AI Workflows)

Timestamps: 00:00 From Local AI agent to Production 00:42 Why Use Render for AI apps 02:10 Render platform overview 03:28 ...

34:49
How to Ship Apps & AI Agents Without Managing Infrastructure (with Durable AI Workflows)

261 views

2h ago

Kawaii Nezumi
Run Local AI Models on Snapdragon NPU with GenieX

Tired of local AI models dragging down your CPU because it isn't leveraging your Snapdragon NPU? In this tutorial, we explore ...

10:52
Run Local AI Models on Snapdragon NPU with GenieX

5 views

7h ago

Coding Horizon
Local AI: (Qwen 3.8) Bonsai 2.0 Runs On 0GB VRAM

Bonsai 2 packs a 27B language model into about 5.9 GB, with CPU execution and partial GPU offloading opening up even more ...

8:03
Local AI: (Qwen 3.8) Bonsai 2.0 Runs On 0GB VRAM

665 views

21h ago

Kaka | Agent Infra & AIOps
M5 Ultra vs RTX 5090 跑 Local AI:现在先别买

M5 Ultra 和RTX 5090 跑Local AI,到底该买容量还是速度?结论先说:截至2026-09-16,M5 Ultra 零售机尚未交付,独立同 ...

6:23
M5 Ultra vs RTX 5090 跑 Local AI:现在先别买

10 views

1h ago

BigIron AI
Build a Private Voice Assistant — No Alexa, No Cloud (STT → LLM → TTS)

Subscribe for more self-hosted AI + homelab that runs on hardware you own. #localai #selfhosted #homelab #homeassistant ...

10:48
Build a Private Voice Assistant — No Alexa, No Cloud (STT → LLM → TTS)

6 views

14h ago

UI5
UI5ers live #49: Build AI-Powered UI5 Apps Locally with LM Studio

For the first time, Tommy Lam and Peter Muessig team up to show how to bring local AI into your OpenUI5 applications. Tommy ...

43:13
UI5ers live #49: Build AI-Powered UI5 Apps Locally with LM Studio

180 views

22h ago

Stephen Robles
Mac mini just leapt forward

I tested the new M6 Mac mini against an M4 MacBook Air and my fully loaded M4 Max Mac Studio in real video editing, ...

8:54
Mac mini just leapt forward

57,747 views

1d ago

TWiT Tech Podcast Network
The Future of PCs: Unified Memory Wars!

What's unified memory, and why are Apple, AMD, and NVIDIA racing to pack your next laptop with it? Windows Weekly breaks ...

14:22
The Future of PCs: Unified Memory Wars!

1,683 views

19h ago

NewEraAi
Intel Arc Nearly Tripled Long-Context AI Speed

Intel Arc long-context AI performance may have received a massive boost through experimental sparse Flash Attention.

8:43
Intel Arc Nearly Tripled Long-Context AI Speed

38 views

20h ago

Noise to Nodes
The DeepSeek V4.1 Flash Advantage: Why It's the Best AI Model for You

DeepSeek v4.1 is hitting 400 tokens per second. We break down the architecture enabling this LLM performance. The speed of ...

7:23
The DeepSeek V4.1 Flash Advantage: Why It's the Best AI Model for You

14 views

22h ago

CompuTherm: Pandat Software
PandatAi: AI-powered Materials Design Platform

PandatAi connects large language models with the Pandat CALPHAD engine — describe your materials problem in plain ...

1:55
PandatAi: AI-powered Materials Design Platform

55 views

17h ago

NIEHS
Scaling Local Capacity Through Human AI Collaboration-Hazards Risk Reduction and Response - 8/25/26

This talk presents innovative approaches to building local capacity through human–AI collaboration technologies that enable ...

59:25
Scaling Local Capacity Through Human AI Collaboration-Hazards Risk Reduction and Response - 8/25/26

16 views

20h ago

DeepSea Developments
I built a "Mind-control helmet" with a Raspberry Pi 5

Mind Control Helmet Specifications: Core Processing & AI Engine • Single Board Computer: Raspberry Pi 5 • AI Hardware ...

11:05
I built a "Mind-control helmet" with a Raspberry Pi 5

60 views

21h ago

thehype.
ai morning #74 — data centers just passed housing in us spending

US data center and information-processing hardware spending has now surpassed housing investment, a line that has never ...

9:11
ai morning #74 — data centers just passed housing in us spending

11 views

23h ago

Show more