HomeVideos

Steal My Exact AI OS Setup (5 simple tips)

Now Playing

Steal My Exact AI OS Setup (5 simple tips)

Transcript

822 segments

0:00

The number one most common question I've

0:01

been getting lately is about how I have

0:03

my AI operating system organized.

0:05

Questions about how I should be routing

0:07

files and how I should have wikis

0:09

organized and folders organized and

0:10

where do I put client projects and so

0:12

many questions around the organization

0:14

of it. Because the problem is if you

0:15

don't have it organized, then your agent

0:17

is going to start to hallucinate things

0:18

not only to you, but potentially in your

0:20

skills and when you're having it build

0:22

automations and stuff like that that

0:23

could get pretty bad. So today I want to

0:25

talk to you guys about these five

0:27

different tricks that I've been using

0:28

that have helped me keep my AI operating

0:30

system super accurate, super upto-date,

0:32

and it allows me to add more and more

0:33

data every week without sacrificing

0:35

quality or memory. So let's not waste

0:37

any time and just get straight into

0:38

today's video. All right, so here is an

0:40

older version of my Herku project that I

0:42

pulled in just to show you guys a quick

0:44

skill that I built that I'll be giving

0:45

you all for free. So look at this. I'm

0:47

going to go in here and I'm going to run

0:48

this OS audit skill. Now what's going to

0:51

happen here is it's going to look

0:52

through my entire project. It's going to

0:53

read everything. It's going to look at

0:54

all of the routing rules, and it's going

0:56

to tell me all of the areas where there

0:58

are things that are weak, where I need

0:59

to make some improvements, where I need

1:01

to update some data, that kind of stuff.

1:04

Essentially, the deliverable of this

1:06

audit is an audit. It says, "Hey, here

1:08

are 10 things I noticed. Here are the 10

1:09

fixes. Do you want me to do them? Yes or

1:11

no." It won't actually do anything yet.

1:13

This is just kind of an exploratory

1:15

phase. It will also create this folder

1:17

at the root of your project if you don't

1:18

have one. If you do, it's called audits.

1:20

and then it will just chuck in a

1:21

markdown file that tells you what it

1:22

found and what the fixes are. And by the

1:24

way, if you guys want to get this skill

1:26

for free, just go to my free school

1:27

community. Link for that is down in the

1:28

description. Come into here, click on

1:30

classroom, click on all YouTube

1:31

resources, and you'll find it in there.

1:32

So, while this is running, what I want

1:34

to do real quick is talk about why this

1:35

is so important, and then we'll come

1:36

back and we'll look at the actual audit.

1:38

So, what I want to talk about today

1:40

before we get into these five tips are

1:42

the different methods of context

1:43

failures. So, four different failure

1:45

modes of context and then two different

1:47

context types. So understanding these

1:49

six different things is going to make

1:51

the audit make more sense and it's going

1:52

to just help you in your day-to-day when

1:54

you're using your AIOS way more. So the

1:56

first thing is the four failure modes,

1:58

poisoning, bloat, confusion, and clash.

2:00

Basically meaning when your agent tells

2:02

you something that you know is incorrect

2:03

or maybe you don't even know, maybe you

2:05

find out later that it's incorrect. When

2:06

it makes a mistake because of the

2:07

context, there are these four reasons.

2:09

So let's start with poisoning. Poisoning

2:11

means you have a false fact somewhere in

2:13

the context. So imagine this is the

2:15

context. Imagine this is a false fact

2:17

that gets dropped in amongst these green

2:19

facts which are the right ones. The

2:21

agent looks in there and it will display

2:23

that back to you or put it in the email

2:25

to the customer or whatever it is

2:26

because it was in the context. So the

2:28

agent inherently is not going to

2:30

intentionally lie. It's probably going

2:32

to grab something and then use that as

2:34

data. But the problem is the data set

2:36

was poisoned with an incorrect fact. Now

2:39

luckily poisoning is the easiest one to

2:40

fix because basically it's just a matter

2:42

of having some verification. So making

2:44

sure that it fact checks everything with

2:46

a web search or maybe two web searches

2:48

or fact checks and cross- checks across

2:50

your live database or maybe if it's not

2:52

100% confident it has to just human in

2:54

the loop. You know what I mean? So

2:55

poisoning is the easiest to fix. Now

2:57

let's take a look at bloat. Bloat is

3:00

when there's so much stuff. There's just

3:02

way too much data. And this is what I

3:04

think a lot of you guys start to feel as

3:06

you scale up your AIOS's and you start

3:07

to use them more and more. Now, this one

3:09

is tough because as you can see, we all

3:11

know about context rot as far as the

3:12

window, but we also know about the idea

3:14

of needle in the haystack, right? The

3:17

agent is going to look at the data that

3:18

it's currently loaded in in order to

3:20

make some sort of decision. And if it

3:22

has way too much to look at, then it's

3:24

going to be really hard for it to

3:24

actually pull out what's relevant and

3:26

what's not. And there's just going to be

3:28

some stuff that bleeds in, and you

3:29

probably don't want it to. The tough

3:31

thing about bloat is that it's a little

3:32

bit harder to fix. But when I talk about

3:34

expertise versus situational, that's

3:36

something that's really going to help us

3:37

out here. So just hold on to that

3:38

thought for a sec. So anyways, that's

3:40

bloat. Then we have confusion. This is

3:42

where there is something in the

3:44

database. So it's a little bit like

3:46

poisoning, a little bit like bloat, but

3:47

it's basically an irrelevant fact or

3:49

something is completely missing. This is

3:51

more of the classic hallucination

3:52

because it tries to fill in the wrong or

3:55

missing data with its own, which is I

3:57

know it sounds very similar to

3:58

poisoning, but poisoning is more so it

4:00

grabs a fact that is just inherently

4:02

wrong and confidently delivers it.

4:03

Whereas this one is it gets confused

4:05

because of the facts that it sees and

4:07

because of facts that it potentially

4:08

knows are missing and it just answers

4:10

instead. And then the last one is a

4:12

clash. Basically meaning you've got old

4:14

data or you've got two different data

4:16

sources that have clashing information.

4:18

So it doesn't know which one to choose

4:19

between. Very simple example. In March

4:22

your policy was always refund. In June

4:24

your policy is now to never refund. So

4:26

when a refund question comes up the

4:27

agent doesn't know which source to

4:29

trust. So maybe it will trust the old

4:30

one. Maybe sometimes it trusts the new

4:32

one. maybe it just makes something else

4:33

completely up. Now, this next piece that

4:35

I wanted to talk about is expertise

4:38

versus situational context and this is

4:39

also very important and it relates back

4:41

directly into these four things. So, if

4:45

you guys remember when you think about

4:46

the whole idea of your second brain and

4:48

if you use my framework on building your

4:49

AOS which is the four C's where we

4:51

basically split this up into having

4:53

context, connections, capabilities and

4:55

cadence and you think of the context and

4:57

the connections as the second brain

4:59

piece of your AIOS. I think of context

5:01

as expertise and I think of connections

5:03

as situational. So let me break that

5:06

down so it you know actually makes more

5:07

sense. So expertise context is the

5:10

things that you actually need all the

5:12

time. So you know who are you? What are

5:15

your goals? What does your business do?

5:17

That's expertise context. The way I like

5:19

to think about this is an analogy of a

5:21

principal and a teacher. Pretend they're

5:23

trying to make a seating chart. The

5:24

principal knows how classrooms should

5:26

run. They know where the whiteboard is

5:29

and where the, you know, the door is and

5:30

they understand how you should build a

5:32

good seating chart, but the teacher has

5:34

the situational context of every

5:36

student. Which student has bad vision

5:38

and needs to sit closer? Which two

5:39

students are loud and if they sit next

5:41

to each other, they're going to laugh

5:42

the whole class. The teacher has that

5:43

situational context that you load in

5:45

just in time. So here, the way I think

5:47

about this is expertise context is the

5:49

rulebook, right? Your policies, your

5:52

pre-loaded information that needs to be

5:53

in every single run. Sort of like the

5:55

system prompt. Whereas the situational

5:57

context is things that you need just in

6:00

time. So for example, let's say you have

6:02

this AIOS and you're curious about um

6:04

you know a customer support ticket that

6:05

came in yesterday. You don't need that

6:08

customer support ticket to always live

6:09

in the context. There's just no reason

6:11

for that. What would that do? Well, that

6:13

would add in probably a bunch of bloat

6:15

and confusion, right? Or maybe even some

6:17

clashing, too. So, what you do is just

6:19

in time when you realize you have to

6:21

answer something about this customer,

6:22

for example, like Thursday at 2 p.m.,

6:25

then you would go use that live lookup,

6:27

pull the data in because you need it,

6:29

and then you'd actually be able to

6:31

enhance or augment the agent's response

6:33

because it had real-time situational

6:36

context. So, hopefully those six

6:38

different concepts make sense. All of

6:40

that contributes to this audit that

6:42

we're currently running right now in my

6:44

AIOS. And besides that, it just really

6:46

helps you think about the way you prompt

6:48

your agent, you organize your agent, and

6:50

what you give your agent as far as like

6:52

here's data you should always have in

6:53

AIOS and here's data that you only need

6:55

to grab every once in a while. Real

6:57

quick guys, a quick message from today's

6:58

sponsor, Hyper Agent. So, a couple of

7:01

videos back I showed you the council

7:02

that I built on Hyper Agent and enough

7:04

of you guys asked for it that I actually

7:06

put the whole thing up there for you to

7:07

just go and grab. So, a quick refresher,

7:10

Hyper Aagent is built by the Air Table

7:12

team, and it spins up a real cloud

7:14

machine for every single agent. You get

7:15

a full browser, a shell, and all the

7:17

tools to go do actual research. So, I

7:20

used that to build a council of five

7:21

agents, each with its own persona. So,

7:24

for example, this week I dropped a

7:25

course idea into Slack, and then the

7:27

council went at it. One played the

7:29

skeptical buyer, one pulled what

7:30

competitors already charge, and one

7:32

stress tested the math. Then they came

7:34

back to me with a verdict, where it was

7:36

weak, the cheapest way to test it before

7:38

I build anything, and a few things I

7:39

honestly didn't want to hear. Now, you

7:41

could go build this council yourself,

7:42

just like I did, but now you don't have

7:43

to because I published my exact council

7:45

on the Hyper Agent Marketplace. So, you

7:47

can sign up, you can install it, and you

7:48

can run your own idea through it in just

7:50

one click. And my link gets you $1,000

7:53

in free credits to start, which is down

7:54

in the description. So, go get your

7:56

ideas roasted, and let's get back to the

7:58

video. Okay, so this audit finished up

8:00

in about 2 minutes. So, let me go to the

8:01

top and let's just read through this

8:03

real quick. So, here is the OS audit on

8:07

um July 22nd. It was my third run today.

8:09

I was testing it out, making some

8:10

tweaks, but anyways, the knowledge here

8:12

is current through June 29th, so almost

8:14

a full month ago. What we have here is

8:16

the routing integrity is red. There's an

8:18

OTAA mis route and upped.

8:21

We have index truth also red. Red being

8:23

bad, green being good. Index says 55

8:26

folders, but disk has 79. Index has 52

8:29

rows. You can see freshness is a red,

8:31

bloat duplication is a yellow, hygiene's

8:33

a red, context placement red. A bunch of

8:35

problems that we need to solve here with

8:36

our AIOS. It tells us basically why it

8:39

found all of this information. And what

8:41

would wrong answer you today? So any

8:43

postjune 29th business questions would

8:45

give me a confident June state answer,

8:47

which would be wrong. If I asked about

8:49

how Q3 OTAAS are going, it would misoute

8:51

and it would give me wrong info. If I

8:53

asked what's in the projects, it would

8:55

give me stale data. So, it's showing me

8:56

all of these issues that are currently

8:58

living because it did a full audit of

9:00

our OS. And then it gives us a fix list

9:03

which says await approval. But we could

9:05

say, okay, yeah, let's go do A, B, C,

9:07

and D right now to get everything back

9:09

up to date and everything actually

9:10

working the way that it should be. So,

9:12

look what it recommends. Finish

9:13

inflight. So, commit the cleanup

9:15

decisions. Rename the row scale blah

9:16

blah blah. Those are all things that we

9:18

could say, "Yep, Opus 4.8, go ahead and

9:21

do that." We have routing plus index

9:23

truth. So, just fixing some of the

9:24

stuff. We have data catchup. So

9:26

ingesting more wiki from my Q&As's and

9:29

from my meetings and stuff like that.

9:31

And then it also suggests durability. So

9:33

having a weekly cron for fireflies which

9:35

is you know my meeting transcripts,

9:37

YouTube polls, archive sweeps, quarterly

9:39

reruns. So it's helping us constantly

9:41

build this thing to make it better and

9:43

better. So that was the actual audit

9:44

deliverable. Let me real quick show you

9:46

guys the actual skill. So in mycloud in

9:49

my skills, if I go to my OS audit and I

9:51

pull up the skillmd, I can do control

9:53

shiftv. if you guys are in Visual Studio

9:55

Code, so we can just preview it a little

9:57

bit better. We have the skill, the

9:58

description, and the argument hint. And

10:00

basically, I'm not going to read all

10:01

this, but let's take a look at what I'm

10:02

telling it. So, OS audit, is your AIOS

10:05

still true? This is your operating

10:06

manual. Indexes and wiks are claims

10:08

about what exists and what's current.

10:10

The audit checks every claim against

10:11

reality. Read only, never fix, or rename

10:14

or delete. Just give a report of what

10:17

needs to be changed. And this works on

10:18

any cloud code project. So, here is why

10:21

I wanted to tell you guys about those

10:22

different things. Here it goes through

10:23

the difference between poisoning, bloat,

10:25

confusion, and clash. And then it also

10:27

goes through expertise versus

10:28

situational context. And then we get

10:30

into the actual steps that this audit

10:31

will be running. So step zero is prior

10:34

report and recent evidence. So look for

10:35

earlier reports inside of the audit

10:37

folder. If one doesn't exist, then

10:39

create that audit folder real quick. You

10:40

can see that for a large project of 100

10:42

plus folders to fan out one explore sub

10:45

agent per check below, giving each the

10:47

checks instructions verbatim plus the

10:49

project route, and then you merge their

10:51

reports. So utilizing some sub aents

10:52

here the bigger that your project gets.

10:55

So the very first check is routing

10:56

integrity. Basically does everything it

10:58

points to exists. So it will read

11:00

through the cloudmd the local the agent

11:02

MD. It'll read through these different

11:04

routing projects or sorry routing files

11:06

that are basically a table of contents

11:08

because it's probably more than just one

11:10

that you have in here especially if

11:11

you've added in some things like a

11:12

karpathy wiki or even multiple karpathy

11:15

wikis. It will also look in the reverse

11:16

direction. It will see if there are

11:18

things that are completely misouted,

11:19

which sometimes happens, and it will

11:20

spot check everything. Then it will go

11:22

to index truth. Do the indexes match the

11:25

disk. Then it will go to freshness. Are

11:27

all the data feeds current? So it will

11:29

look and see if they are fresh,

11:31

drifting, frozen, retired, or if they're

11:33

on demand. And it will go through all of

11:35

your different connections and

11:36

understand when do you need to pull

11:38

things in or, you know, do you only do

11:40

that when kind of like a situational

11:42

context presents itself. It will look at

11:44

memory. it will go into looking for

11:46

bloat, duplication, organization. So,

11:48

that's basically how this works. I'm not

11:50

going to read this entire thing out, but

11:51

the skill is downloadable for free

11:53

inside of my free school community. The

11:54

link for that is down in the

11:55

description. Okay, so now that we've

11:57

seen a quick example and some background

11:59

context on why I built it like this and

12:01

why that's important, let's talk about

12:03

these five hacks. So, this first one is

12:07

CloudMD as a router. So, a lot of you

12:10

guys are probably when you're building

12:11

very specific projects, you're using

12:13

CloudMD sort of like a system prompt.

12:14

sort of like, hey, here's your

12:16

background. Here's what you do, blah

12:17

blah blah. And for the most part, that

12:19

is correct. But the way that I have my

12:22

um Herk 2 AIOS set up is I have it all

12:26

under one massive folder. So, if you see

12:28

right here, let's just call this one my

12:30

Herk 2 project. Now, what I have inside

12:32

of Herk 2 is a bunch of folders.

12:34

Literally, just a bunch of folders. And

12:36

then some of those folders have their

12:38

own folders inside. And some of those

12:39

folders have their own folders inside

12:40

there. And that's how I have mine set

12:42

up. And the reason I like to do it this

12:44

way is because I like to just be able to

12:46

push this main folder to GitHub and

12:49

everything backs up. So for me, that was

12:50

the easiest and I like it that way. You

12:52

don't have to do it, but that's honestly

12:53

the way that I recommend. That also

12:55

means I could CD into this directory and

12:57

this has its own cloudmd and so does

12:59

this one and so does this one. So that's

13:00

where I can set more like specific

13:02

project level sort of system prompts if

13:04

I want. But this main cloudmd that I use

13:06

up here, I treat this almost purely as a

13:09

router. So let me show you guys that. If

13:11

I go into my Herk 2 and I go to the

13:12

cloudMD and I preview this here, what do

13:15

we see is basically this is routing. So

13:17

I start off saying, "Yeah, hey, you're

13:18

Nate Herk's AI operating system. Your

13:20

job is to help him spend less time in

13:21

operations, blah blah blah." But here is

13:23

where you actually go to find data. You

13:25

know, here's where things live. If you

13:26

need this, you go here. If you need

13:29

this, you go here. If you need this, you

13:30

go here. If you need this, you go here.

13:32

All of this is basically just a routing

13:34

table. And then I get into like the

13:35

knowledge base. Here's the wiki path.

13:37

Here's the hot cache. Here's the index.

13:38

Here's the GP fallback. Right? Here's

13:40

the memory system. Here's this. Here's

13:41

this. Here's this. Here are the tools.

13:43

Here's this and this. You know, there's

13:44

obviously more than this, but here's

13:46

where my API keys live. Here's where

13:47

skills and agents live. Here's where

13:49

decisions live. Here's templates. Here's

13:51

references. Here's projects. Here's

13:53

other worlds. And other worlds are

13:54

basically massive other repos that all

13:57

that all have their own individual

13:58

GitHub repo because I like to keep

14:00

everything synced, right? And so, as you

14:02

can see, my cloudmd is essentially just

14:04

a master routing file, just a master

14:07

table of context. And so I just wanted

14:09

to show you guys like the way that I

14:12

kind of have mine set up with my main

14:13

folders. And once again, this isn't the

14:16

most optimal way and there's probably

14:17

some ways that I could change this up,

14:19

but this works for me right now. So in

14:21

my Herk 2, right, I've got obviously a

14:22

clamd I've got my cloud with all of my

14:25

pretty much global skills and global sub

14:28

aents and my settings here that apply to

14:30

basically this whole project, unless of

14:32

course I CD into one of these projects.

14:34

But, you know, I've got my brainstorms

14:36

folder, which is anytime I run a grill

14:38

me session, it saves it here. I've got

14:40

the herk brain, which is like my first

14:42

overall wiki. I've got my other worlds,

14:44

which are all of my other worlds. These

14:46

are massive projects that I would

14:48

consider standalone cla projects. So,

14:50

maybe if you have clients and stuff,

14:51

this is where you could put that sort of

14:52

stuff. And then I've got a folder called

14:54

projects, which is the largest one.

14:55

Every single little time that I spin up

14:57

a new chat and I start brainstorming

14:59

about something or I want to create some

15:00

sort of deliverable, it will save it in

15:02

here. So, I'll show you guys that in a

15:03

sec. And then I've also got at the root

15:05

level something called brand assets

15:06

where I have like brand guidelines,

15:08

logos, pictures of me, things like that.

15:10

So just to show you guys that in here,

15:12

I'm on the desktop app this time.

15:14

Obviously got, you know, different

15:15

things. I've got agents, I've got cloud,

15:17

I've gotcodex. So keep in mind all we're

15:19

building here is a bunch of files and

15:20

folders, which means you can plug in

15:22

Hermes, you can plug in codecs, you can

15:23

plug in anything else. So just keep that

15:25

in mind. Right now obviously we're kind

15:26

of focusing on claude. But either way,

15:28

here's the brand assets one, right? I've

15:30

got some fonts, I've got some pictures,

15:31

I've got, you know, PGs here, stuff like

15:33

that. If I go to my brainstorms, this is

15:35

where I have some of my brainstorming

15:37

sessions from the grill me. I can see

15:39

I've got some audits, I've got my

15:40

decision log, I've got other things

15:41

here. Here's my herb brain wiki. Here is

15:44

my other worlds with all of these other

15:45

projects. And pretty much all of these

15:47

other projects have their own GitHub

15:49

reposiated with them. And let's see,

15:52

here's my projects one. So, this is

15:53

where it starts to get big, right? Like

15:55

I said, almost everything that I do

15:57

inside of this project goes in here. So

15:59

that's why there's so many things in

16:00

here. And what you'll notice, guys, is

16:02

there's also one right here called

16:03

YouTube videos, which if I open this up,

16:06

this is also massive because every

16:07

single one of these videos, obviously,

16:09

I've got different things. So, you know,

16:10

in here, I've got like transcripts, I've

16:13

got, you know, scripts, I've got tests,

16:15

I've got, you know, sometimes I run a

16:16

bunch of um like research and I run a

16:19

bunch of like Python skills for this

16:20

kind of stuff. So anyways, the point

16:22

being there's just so many ways I can

16:25

drill down in here. This is the way that

16:26

I currently have mine set up. Obviously,

16:28

I'm trying to optimize it every day and

16:30

playing around with different things,

16:31

but this is the way that I've got it set

16:33

up with one big projects folder with a

16:35

lot of my work in there. And then I save

16:37

other things where it makes sense. But

16:39

like I said, it doesn't really matter if

16:40

you have it super flat, so you've got

16:42

tons of projects at your root, or it

16:44

doesn't matter if you have one main one

16:45

that you drill down everything into. All

16:47

that matters is that you have the right

16:49

routing rules in place so that you and

16:51

your agents can find it. Now, moving on

16:53

to number two, have AI audit itself. So,

16:56

just what I did there, as you guys saw

16:58

with this audit, just do this. I

17:00

obviously built a skill around this, but

17:01

what I used to do is like at the end of

17:03

every week when I made a bunch of

17:04

changes or the end of every month, I'd

17:05

be like, "Hey, look through everything.

17:08

Like, open up every file, look through

17:10

my routing rules, and just make sure

17:12

that everything's still accurate. Make

17:13

sure that this all makes sense. And you

17:14

know what? If you don't like how this is

17:16

set up based on your best practices or

17:18

based on the way that I talk to you

17:19

every day, then suggest some changes and

17:21

let's just, you know, keep iterating and

17:23

keep scaling this thing up because I

17:24

know I add a lot of data to you every

17:27

single week. And right now, you guys

17:29

probably know that I have a few

17:30

different capacity wikis. I've got my

17:32

two main ones, right? I've got my main

17:33

one for my YouTube transcripts and I've

17:35

got my main one for my meeting

17:37

transcripts. And I originally had all of

17:39

that in one master wiki. And then it was

17:41

actually cloud code that suggested, hey,

17:43

you know, I've noticed that you're

17:44

putting in these two specific types of,

17:47

you know, meeting transcripts or

17:48

specific types of data every single

17:50

week, like on a specific cadence. You

17:52

might as well just split those up so

17:53

that it makes it easier for me to search

17:55

through data. So, I'm getting less of

17:57

that, you know, what do we call it? I'm

17:59

getting less of that bloat. I'm getting

18:00

less of that confusion. And I'm able to

18:02

answer you not only quicker and more

18:04

accurately, but also cheaper because I'm

18:05

not spending so many tokens, so many of

18:07

your tokens looking through everything.

18:09

So yes, I've made sort of a more formal

18:11

audit process with that skill that you

18:13

guys can check out, but you can also

18:15

just have a conversation, right? I think

18:17

a really important mindset shift is to

18:18

realize that everyone's building their

18:20

own second brains. Everyone's doing a

18:21

little bit differently and there's not a

18:23

right way. There's the only way that

18:25

you're doing this wrong is if you're

18:27

constantly getting wrong answers and

18:29

you're not doing anything about it. But

18:30

besides that, like don't get so stressed

18:32

out about, oh, you know, I saw Nate do

18:34

it this way or I saw this other YouTuber

18:35

do it this way. Just do it however it

18:37

works for you. I think a really good

18:38

test is could you pull up your file

18:40

explorer, you know, could you pull up

18:42

your file explorer right here? Could you

18:44

find like think of something that you

18:46

did and see if you could find it without

18:47

searching, without asking Claude, see if

18:49

you could find it. And if you can follow

18:51

your different folders and you can find

18:52

the path yourself, then it's probably

18:54

set up pretty intuitively where an agent

18:56

could also do so, especially if it had

18:57

some routing rules in place. So, I'm

18:59

constantly making deliverables in here

19:01

and then I'm just finding them manually

19:02

because I want to make sure that this is

19:03

still intuitive to me and that I

19:05

actually understand my own second brain,

19:07

my own AIOS. Okay, moving on to number

19:10

three. Build automations to update data.

19:12

If you guys remember here, right here in

19:15

the audit, it said durability weekly

19:17

cron for these different things. What

19:19

you'll notice is as you start to use

19:21

this thing more and more, you're

19:22

probably going to be telling it

19:23

frequently, hey, can you go pull that

19:25

data or hey, can you go do this? And you

19:27

might realize that this isn't sort of

19:29

like a just in time sort of situational

19:30

context thing. Like for my wiki, every

19:33

single Monday I have a Q&A, I want that

19:35

in there. Every single Tuesday when I

19:37

meet with my leadership team, I want

19:38

that in there. So why not just set up

19:40

crrons to automatically pull that stuff

19:41

in there so that if I for some reason

19:43

forget to after one of my meetings, I

19:45

don't have to worry about it. And next

19:46

time I go talk to my agent, it will

19:49

already be there. And I know that sounds

19:51

simple, but it's something that I've

19:52

noticed a lot of people aren't doing. So

19:54

definitely set up some sort of crons to

19:56

pull in the data that you want to always

19:58

be living inside of your local project.

20:01

As you guys know, that's super simple.

20:02

All you need is basically the API key

20:04

and you can say, "Hey, you know what,

20:05

Claude, go do this for me. Set up this

20:07

cron, set up this routine or help me

20:09

push this, you know, script onto modal

20:11

or whatever it is." You can set that up

20:13

really easily with natural language.

20:15

Number four is to segment knowledge. So

20:18

as you guys know, I have two different

20:19

wikis. I've got my YouTube one, I've got

20:21

my U meeting transcript one, and I

20:23

segmented them because they were both

20:25

starting to grow, and I planned to keep

20:27

growing those and growing those. So, the

20:29

more you can segment stuff out, the

20:31

better. If you find yourself having

20:33

little like nodes of knowledge that are

20:35

going to keep growing that are very

20:36

distinct and different, then it's so

20:38

much easier to split those up because

20:40

now your agent knows, okay, if I have to

20:43

find something related to, you know, um,

20:45

Apify for some reason, let me check if

20:47

Nate's ever made a video about that. and

20:49

I know exactly to go to his YouTube

20:51

video transcripts wiki rather than

20:53

having to search through potentially

20:54

five or 10x more files in order to find

20:57

what I'm looking for. So, it's just

20:58

about how can you narrow the actual

21:00

context that your agent is going to be

21:02

looking through? And a very common

21:04

question I get around this whole

21:05

segmenting knowledge idea is where do

21:07

client projects go and like should I put

21:09

everything under one folder? And the way

21:12

that I answer this question is basically

21:13

that if I was working with a bunch of

21:15

different clients, I would 100% keep

21:17

information internally inside of my AIOS

21:20

and I would segment that by client. So

21:22

I'd probably have one folder in here.

21:24

I'd have one folder here called like,

21:25

you know, clients. And then I'd have

21:28

client A, client B, client C. And then

21:31

in there I'd have different files that

21:33

refer to what I'm working on with this

21:35

client. But the caveat here is if I was

21:38

working on something like a very

21:39

specific deliverable for this client, I

21:41

would probably build that out over here

21:44

because this would maybe be like my

21:45

client-f facing you know um repo and

21:49

what I do is I would say okay in here

21:52

this is where it lives right this lives

21:53

in my it's typing in a different color.

21:56

This lives in my C users you know Nate H

21:58

whatever inside of a folder called

22:00

client A. But what I'd have in here

22:02

would be the internal knowledge like,

22:03

hey, you know, we started working with

22:05

this client on June 22nd. He signed the

22:06

contract on June 25th. This was our

22:08

project price. Here are the, you know,

22:10

discovery calls. Here are the here's the

22:12

scope of work. But here is where I'd

22:14

have the external stuff, which would be

22:16

like the actual deliverables that I'd be

22:18

giving this client so that I could, you

22:19

know, potentially have him be a

22:21

collaborator on this repo or I could,

22:23

you know, push this to a certain

22:24

environment, whatever it is. I would

22:25

keep it segmented a little bit away from

22:27

my actual internal AIOS, but I would

22:30

100% still be giving my AIOS context of

22:32

this because it still needs to know

22:34

about this engagement. It just maybe

22:36

doesn't need to own all of the stuff

22:38

inside of it when I'm pushing my whole

22:39

repo, things like that. So, it's very

22:41

situational and hopefully this is, you

22:43

know, hopefully you understand like that

22:44

there's no right or wrong way to do it.

22:47

There's just wrong answers that you can

22:49

get and that's wrong. And then what I

22:51

have here for number five is to

22:52

backtrack. I've I've obviously stressed

22:55

the importance of like when you realize

22:57

something is wrong, when you realize it

22:59

searched for five minutes for something

23:00

it should have found instantly or it

23:01

searched and said, "Hey, I don't have

23:03

access to that, but you know it's in

23:04

there." You have to correct it, right?

23:06

You say, "Hey, you know, you actually

23:08

have access to that, so please make sure

23:09

that never happens again. Make sure

23:10

you're checking things first." But more

23:12

importantly, have it backtrack. I found

23:14

a lot of success in that when I say,

23:15

"You know what? You told me you didn't

23:17

have access to this, but I know you do."

23:19

So, go look through what you did, where

23:22

you searched, and help me figure out why

23:24

you didn't find that data right away.

23:27

And then once it does that, it says,

23:28

"You know what? I made a mistake. I'm

23:29

sorry about that. This is what I did.

23:31

This is where I should have looked, and

23:32

I would have gotten that much faster."

23:33

Now that you've had it prove its own

23:35

mistake and tell you what it did wrong

23:37

and how to fix it, then just have it fix

23:39

it. Have it update the routing. Have it

23:41

maybe even move stuff or reorganize

23:42

stuff based on what it found when it was

23:44

backtracking. And that works a lot

23:46

better than just saying, "Make no

23:48

mistakes. Don't let that happen again.

23:49

So anyways, I know that second brains

23:52

and AI operating systems very very hot

23:54

topic right now and as they should be

23:56

and a big problem that we're looking at

23:58

we're talking to businesses about that

24:00

obviously I know every business is

24:01

having is once people start building

24:03

their own second brains and their own AI

24:05

operating systems how do you sync all

24:06

the data together at the team level at

24:09

the department level and that's

24:10

something that I'm looking into very

24:11

actively and playing around with. I

24:12

don't have a great answer for you guys

24:13

right now besides the fact that I don't

24:16

think it's a tech problem. I think it's

24:17

a people problem. I think that you could

24:19

do this with Google Drive, with Notion,

24:21

with GitHub. I think the problem is

24:23

having people habit shift to syncing

24:25

data, knowing what to pull in, making

24:27

their agents read stuff, permissioning.

24:29

I think it's a people and a habit

24:30

problem that we need to figure out

24:32

before we start pushing like, hey, this

24:34

is the right way to do it. So, anyways,

24:36

like I said, looking into it. I'll

24:37

definitely be bringing more content on

24:38

YouTube as we start to, you know, figure

24:40

this out a little bit more. But that's

24:42

something to be aware of. But I do think

24:44

that the best thing you can be doing

24:45

right now is mastering and understanding

24:47

your own systems so that when you do

24:49

start to bring this to a team level,

24:51

it's a bit easier because you've already

24:52

walked the walk. So, that is going to do

24:55

it for today. I hope you guys enjoyed

24:56

this one and if you learned something

24:57

new, please give it a like. It helps me

24:58

out a ton. And as always, I appreciate

24:59

you guys making it to end of the video

25:01

and I'll see you on the next one. Thanks

25:02

guys.

Interactive Summary

This video provides a comprehensive guide on organizing an AI operating system (AIOS) to prevent hallucinations, improve accuracy, and maintain memory efficiency. The author introduces an 'OS Audit' skill to identify weaknesses, explains the four main modes of context failure (poisoning, bloat, confusion, and clash), and differentiates between 'expertise' and 'situational' context. Furthermore, the video outlines five actionable tips for maintaining an effective AIOS: using CloudMD as a master routing file, regularly auditing the system, building automations for data updates, segmenting knowledge into distinct categories, and using backtracking to correct AI mistakes.

Suggested questions

4 ready-made prompts