TRANSKRIPTEnglish

Why you NEED to be running local AI models (FULL beginners guide)

21m 28s4,210 ord581 segmentsEnglish

FULLSTÄNDIGT TRANSKRIPT

0:00

I'm about to show you the future of AI,

0:02

AI agents, and Open Claw. Over the past

0:05

2 months, I've spent over $50,000 to

0:08

use, test, and learn about local AI

0:11

models. What I learned, I think, can

0:13

dramatically change your life and save

0:15

you tons of money, even if you're on a

0:16

cheap computer. You don't need to buy

0:18

Mac Studios like me. In this video, I

0:20

will cover everything local AI models.

0:23

I'll cover what computers you need, what

0:25

local AI models even are, which models

0:28

you can run, what use cases you can do

0:30

today, and how this lets you use Open

0:33

Claw completely for free. I'll also show

0:36

you a glimpse of the future that I am

0:37

100% confident is what's going to

0:39

happen. By the end of this video, you'll

0:41

be a local AI master, and you'll be

0:43

running your own local super

0:45

intelligence on your computer. So, let's

0:47

lock in and get into it. So, this might

0:50

be my most important video yet. I'm so

0:52

excited to take you through what I've

0:53

learned over last couple months. Even if

0:55

you have no idea what local AI models

0:57

are, you're going to get so much out of

1:00

this video. So, let's start off with why

1:02

local AI is so important and why you

1:04

need to be using it. This is what you're

1:07

probably doing today. If you're on

1:08

ChatGPT or Claude or using any of the AI

1:12

frontier models that everyone knows

1:13

about, you're using a cloud model. That

1:16

means you're using AI that are running

1:18

on these big servers that might be

1:20

underground or one day in space because

1:22

of Elon or might be on an island. But,

1:24

because you're using these models that

1:26

are running on these servers, that means

1:27

a lot of different things. One, it's ex-

1:30

pensive. You're paying for every token

1:32

you use. Every time you send a prompt to

1:35

these servers, it's doing a bunch of

1:37

calculations, and they're charging you

1:39

for each one of those calculations.

1:41

These are where these massive API bills

1:44

are coming from. These are where the

1:46

$200 month subscription plans are coming

1:48

from and all your API usage. It's very

1:51

expensive to be running AI models on

1:54

these billions of dollars of chips. It

1:56

also has many other downsides, zero

1:58

control. A lot of people complain all

2:00

the time they feel like their AI models

2:01

are getting stupider. In reality, they

2:04

probably are. These AI companies are

2:06

constantly dialing the knobs and

2:08

changing things to try to save money.

2:10

You have zero control over the AI models

2:13

running on these servers. You have zero

2:15

privacy. Every message you send to

2:18

ChatGPT or Claude or Gemini or any AI

2:21

model you're using on the cloud, those

2:24

employees can read those logs. Nothing

2:27

you say is private and secure. Every

2:29

question you ask about your health or

2:31

maybe if you're a sicko and you have

2:32

your own AI girlfriend, they can read

2:36

all of those messages you're sending.

2:38

It's also laggy. You need to be

2:39

connected to the internet. If you don't

2:41

have great internet, it can take a while

2:42

to get your prompt sent there and sent

2:44

back. So, there's a high latency. And on

2:46

top of that, it's just not scalable. A

2:48

lot of people been learning this lately

2:49

with open claw. Maybe you connected to

2:51

Opus 46 API, you send a bunch of

2:53

prompts, you look at your API bill and

2:55

whoops, you spent a thousand dollars

2:57

over the last day cuz you sent a bunch

2:58

of prompts. It's not scalable at all.

3:00

And if you want super intelligence

3:02

working for you 24/7, it's going to cost

3:05

you millions of dollars. But, with all

3:07

that being said, you do get one benefit,

3:09

which is you get frontier AI. You're

3:12

getting the best AI models, they're

3:13

running on these servers, and you're

3:15

getting the best performance. That is

3:17

probably what you're used to today. But,

3:19

where I strongly, strongly believe the

3:22

future is going and what I actually

3:24

believe you will be doing in the next 12

3:27

months is you will be using local AI

3:29

models. What are local AI models? These

3:32

are AI models instead of all these

3:34

complex multiplication equations

3:36

happening on servers across the world,

3:39

they're happening locally on the Mac

3:41

mini on your desk or the Mac Studio or

3:44

the old dusty Lenovo laptop, whatever

3:46

you're using, the models run locally.

3:49

And that has a tremendous amount of

3:50

benefits. First of all, it's completely

3:52

free, right? You're not paying for

3:54

tokens. It is just the cost of the

3:56

electricity going into the computer you

3:58

have plugged into the wall. It's fully

3:59

customizable. If If you want to take a

4:01

local model and make it sound like you,

4:03

or make it rap like Kendrick Lamar, you

4:06

can do that. They are fully

4:07

customizable. It's also fully secure and

4:10

private, so every message you're sending

4:12

to your local AI running on your

4:15

computer on your desk stays on your

4:17

computer. It does not go to the

4:19

internet. Nobody can read your prompts

4:22

or your messages back and forth. If you

4:24

want to get freaky deaky and make your

4:25

own AI boyfriend or girlfriend, you can

4:27

do that and no one will read those

4:29

messages. Not that I would know anything

4:31

about that. But also, zero latency.

4:34

There is no messages going to the

4:36

internet. It's all staying on your

4:37

device, so you literally can unplug this

4:40

from the internet and the AI would still

4:42

work. You can be on an airplane, vibe

4:45

coding to your heart's content, and it

4:47

doesn't matter because there's no

4:48

internet. It's all local on your

4:50

computer. And here's the best part.

4:52

Here's why I'm bought in, and here's why

4:54

I think everyone will be using local AI

4:56

in the next 12 months. Because it's

4:58

local, because it's free, it is

5:00

extremely scalable, which means you can

5:03

have AI doing work for you 24/7 365. I

5:07

have right now, and I'm going to demo

5:09

this later in the video, so make sure to

5:10

stick around for this. I have right now

5:13

four local AI models doing things for me

5:16

continuously, going on the internet and

5:18

scraping websites, writing me content,

5:20

writing me newsletters, writing code for

5:22

me, just doing things at all times of

5:25

the day. It's like I have multiple

5:27

employees working for me. This is an

5:29

advantage I have over all of my

5:32

competition because they are not running

5:34

local AI models. And if you do the

5:36

things I'm about to show you in this

5:38

video, you will have the same crazy

5:41

advantage over everyone else in the

5:42

marketplace as well. That's why it's

5:44

super critical to stick here till the

5:46

end. Now, the one downside, what's the

5:48

one downside of all of this? Local

5:50

models aren't quite as smart as the

5:52

frontier models. I'd say they're about 6

5:55

months behind at all times. So, so 6

5:58

months ago was like Opus 4.5, Sonnet

6:00

4.5, around that realm. The local models

6:03

are about there. Now, if you think back

6:05

to 6 months ago when Opus 4.5 came out,

6:08

it absolutely blew people's minds. So,

6:10

we're still we're at that point when it

6:12

comes to local models. So, it's still

6:14

really, really strong. So, that brings

6:16

us to our next point, which is what

6:18

computers do you need? Do you need to

6:20

run out and buy $50,000 worth of Mac

6:23

Studios and DGX Sparks like me? Well,

6:25

the answer to that is no. You can

6:27

literally run local models on any

6:30

computer you have. So, if you have an

6:32

old crappy laptop in your closet from

6:34

like college or something, you can take

6:36

that out and run local models. If you

6:37

have the new $600 Mac Mini that everyone

6:40

was running out and buying a few months

6:42

ago, you can run local models on that.

6:43

That was a very good purchase. Now, are

6:45

the models you're running on these

6:47

cheaper, smaller machines going to be

6:49

Opus 4.5 level? Well, no, but there's

6:52

still use cases you can run. You can

6:54

still do things like memory management

6:57

for your open claw. Having a very small

7:00

local model deciding which memories get

7:03

loaded into context for your open claw

7:06

or your AI agent or whatever is still a

7:08

really powerful use case that you can

7:10

run on your $600 Mac Mini. And I'll go

7:13

through the exact models you should be

7:14

downloading for each device in a second.

7:16

But, even if you have these old dusty

LÅS UPP MER

Registrera dig gratis för att få tillgång till premiumfunktioner

INTERAKTIV VISARE

Titta på videon med synkroniserad undertext, justerbart överlägg och fullständig uppspelningskontroll.

REGISTRERA DIG GRATIS FÖR ATT LÅSA UPP

AI-SAMMANFATTNING

Få en omedelbar AI-genererad sammanfattning av videoinnehållet, nyckelpunkter och slutsatser.

REGISTRERA DIG GRATIS FÖR ATT LÅSA UPP

ÖVERSÄTT

Översätt transkriptet till över 100 språk med ett klick. Ladda ner i valfritt format.

REGISTRERA DIG GRATIS FÖR ATT LÅSA UPP

BEGREPPSKARTA

Visualisera transkriptet som en interaktiv begreppskarta. Förstå strukturen med ett ögonkast.

REGISTRERA DIG GRATIS FÖR ATT LÅSA UPP

CHATTA MED TRANSKRIPT

Ställ frågor om videoinnehållet. Få svar från AI direkt från transkriptet.

REGISTRERA DIG GRATIS FÖR ATT LÅSA UPP

FÅ UT MER AV DINA TRANSKRIPT

Registrera dig gratis och lås upp interaktiv visning, AI-sammanfattningar, översättningar, begreppskartor och mer. Inget kreditkort krävs.

Why you NEED t… - Fullständigt Transkript | YouTubeTranscript.dev