A truly 'local' AI model must actually compute on your own machine — running software locally that just calls out to a remote API doesn't count as local.
Scott clarifies the common confusion around 'local' AI: many people think running a local app makes it local AI, when in fact the actual inference is happening via API calls to remote servers. ✦ AI generated
Scott Tolinski · Syntax · 2026-07-08 · original ↗
starts at this moment · 11:26
“In episode 1009 you mentioned that nobody knows what a local model is. Can you do an explainer on what local models are?”
Well, local models are just uh models that you're running on your local machine. I think some people really disconnect that and think, 'Ooh, I'm running the software locally, therefore it's local.' But it like when reality they're going off and doing API calls.
verbatim transcript · starts at 11:26
11:26mentioned that nobody knows what a local model is. Can you do an explainer on what local models are? Well, local models are just uh models that you're running on your local machine. I think some people really disconnect that and think, "Ooh, I'm running the software locally, therefore it's local." But it like when reality they're going off and doing API calls. That was the whole thing >> Yeah, all these people are like, "Just
11:52use a local thing." It runs a local use open code and deep seek for local. Like, that's not local. It's going to China. >> Yeah, so it's it's literally running on your machine. And I will point you to the CJ did a really incredible video called your guide to local AI. And so instead of just recreating that here, cuz you know we could I could just recreate CJ's awesome video on the fly
12:15here, right? I'm just going to point you to that cuz it's really super good and he even uses a piece of bread to show you how big his computer is. So that's fun. >> [laughter] >> Yeah, the the local model stuff is really interesting. We'll we'll keep coming back to it because there's like this like idealistic thing of yes, I would love to run local models. It's it's private and whatever
12:38and it's on my machine and whatever. And then there's the actual like you don't realize like I again, I'm saying people don't know what local models are and now I'm going to say like people don't realize how much computer is actually being thrown at this stuff until you try to run something on your own machine that is of of the quality that you're expecting from like an Opus.
13:02>> I will have Yeah, I'll have a lot more opinions on this once I get a local AI rig and get deep into it because uh that's coming for me. It's just I'm waiting for some next round of hardware to be released. It's just Yeah. >> The local AI stuff makes a lot of sense when you're doing purpose-built stuff. Um when you're doing toxicity detection, when you're trying to detect parts of
13:25your face, There are thousands of small models on hugging face that you can run in the browser. Look up Transformers.js. We had Xenova on the the podcast where there is a lot of really cool stuff that you can run in real time in like that is is good very good in the browser, but they are all dedicated to uh specific things, you know, doing speech detection, text-to-speech,