ViewTube

ViewTube
Sign inSign upSubscriptions
Filters

Upload date

Type

Duration

Sort by

Features

Reset

8,630,115 results

bycloud
7 Popular LLM Benchmarks Explained [OpenLLM Leaderboard & Chatbot Arena]

Check out my website here! https://leaderboard.bycloud.ai/ In this video, I will be going through and explain the benchmarks for ...

5:50
7 Popular LLM Benchmarks Explained [OpenLLM Leaderboard & Chatbot Arena]

30,534 views

2y ago

IBM Technology
LLM & AI Agent Benchmarks vs Reality: Why AI Applications Break

Learn more about LLM Benchmarks here → https://ibm.biz/~e64ktvs52 Your AI model scored high, but does it actually work?

15:01
LLM & AI Agent Benchmarks vs Reality: Why AI Applications Break

33,692 views

1mo ago

Better Stack
AI Benchmarks Are Fake!?

AI models are gaming their own benchmarks, and Cursor's reward hacking research just proved how often it happens.

5:39
AI Benchmarks Are Fake!?

5,907 views

2mo ago

Execute Automation
Mac mini M6 vs M5 Max AI Benchmark — Qwen3.8-27B, TTFT, Prefill & Agentic Coding

In this video, I'm testing how the new Mac mini M6 handles serious local AI workloads using Qwen3.8-27B-Splash in LM Studio, ...

10:08
Mac mini M6 vs M5 Max AI Benchmark — Qwen3.8-27B, TTFT, Prefill & Agentic Coding

27,129 views

5d ago

AI Engineer
The Art & Science of Benchmarking Agents — Vincent Chen, Snorkel AI

ARC AGI 3 launched a few weeks before this talk with every task human solvable and frontier models under 1%. That gap is the ...

23:25
The Art & Science of Benchmarking Agents — Vincent Chen, Snorkel AI

4,457 views

3mo ago

Chase AI
The Sol 6.1 Benchmarks Are STUPID, So I Tested It vs Sonnet 5.5

Higgsfield: https://higgsfield.ai/s/chase-h-ai-PhSvkH ⚡Master Claude Code + Codex: https://www.skool.com/chase-ai FREE ...

15:25
The Sol 6.1 Benchmarks Are STUPID, So I Tested It vs Sonnet 5.5

2,957 views

2h ago

NetworkChuck
i got one....and it's FAST!!!

Dip your toe into local AI at Micro Center, the AI destination: https://micro.center/322535 Austin, Texas: sign up now for a FREE ...

21:53
i got one....and it's FAST!!!

533,335 views

6d ago

Lex Clips
Limits of AI benchmarks | Demis Hassabis and Lex Fridman

Lex Fridman Podcast full episode: https://www.youtube.com/watch?v=-HzgcbRXUK8 Thank you for listening ❤ Check out our ...

3:51
Limits of AI benchmarks | Demis Hassabis and Lex Fridman

5,837 views

1y ago

Prospectus Lab
Understanding AI Benchmark Scores

In this video, we break down the launch of Anthropic's Claude Opus 4.6 and its benchmark scores, particularly the 80.8% on ...

8:45
Understanding AI Benchmark Scores

904 views

2mo ago

Theo - t3․gg
Which AI Models Are Worth Using

We're ranking each model currently out by AI labs, and nearly model you'd reasonably use today goes on the board, so lets break ...

36:36
Which AI Models Are Worth Using

171,558 views

1mo ago

Vectro AI
AI Benchmarks Explained for Beginners. What Are They and How Do They Work?

Ever wonder how we actually measure if one AI is "smarter" than another? It's not just a feeling; there's a whole system of ...

7:00
AI Benchmarks Explained for Beginners. What Are They and How Do They Work?

1,966 views

1y ago

AI Coding Daily
I Tested NEW GPT-6-Astra on Coding Benchmarks

Extra video for Premium members: "GPT-6-Astra in ChatGPT App: Browser Use and Light Level" ...

16:33
I Tested NEW GPT-6-Astra on Coding Benchmarks

34,960 views

3w ago

Caleb Writes Code
GPT-6 Astra.. full analysis..

Zo Computer: https://zo-computer.cello.so/2XNkhpgqHRy OpenAI released GPT-6 Astra, what the AI industry has been waiting for ...

11:08
GPT-6 Astra.. full analysis..

474,149 views

3w ago

Tina Huang
Every AI Model Explained In 20 Minutes (Update)

Sponsored by Viktor, the AI employee that lives in Slack and Microsoft Teams and connects to 3200+ tools. Hire Viktor for your ...

20:24
Every AI Model Explained In 20 Minutes (Update)

148,443 views

1mo ago

IBM Technology
What are Large Language Model (LLM) Benchmarks?

Want to play with the technology yourself? Explore our interactive demo → https://ibm.biz/BdKetJ Learn more about the ...

6:21
What are Large Language Model (LLM) Benchmarks?

26,125 views

2y ago

AI Explained
Gemini 3.1 Pro and the Downfall of Benchmarks: Welcome to the Vibe Era of AI

Do we have a new best AI model, or do we have the downfall of benchmarks in general, as a way of capturing machine ...

18:50
Gemini 3.1 Pro and the Downfall of Benchmarks: Welcome to the Vibe Era of AI

110,006 views

7mo ago

n8n
We Ranked AI Models by Their Performance in n8n

n8n now has an Official AI Benchmark. A free community resource for choosing the best model for your use cases. Link to the ...

3:29
We Ranked AI Models by Their Performance in n8n

4,089 views

7mo ago

Learn Meta-Analysis
R9700 AI Pro: Benchmarks and First Impressions

Is the R9700 a "good deal"? I compare a R9700 32gb to RTX 4060 8gb. I look at a variety of Nvidia, Qwen, and Gemma models ...

14:02
R9700 AI Pro: Benchmarks and First Impressions

6,821 views

8d ago

Fireship
Open-weight AI just hit 2.8 trillion parameters…

Get 20% off Mobbin Pro to help your agent design UIs that don't suck - https://mobbin.com/fireship Moonshot just released Kimi K3 ...

5:09
Open-weight AI just hit 2.8 trillion parameters…

1,032,714 views

2mo ago

Show more