Never Pay For Claude Upgrades Again — Transcript
Full transcript
- 0:00Let's be honest with each other. Claude
- 0:01usage limits absolutely suck. I'm sure
- 0:04you know the feeling. You're deep in a
- 0:05Claude prompting session and all of a
- 0:07sudden, bang, you've hit your usage
- 0:09limit. And it doesn't matter if you're
- 0:11on the $20 plan, the $100 plan, or even
- 0:13the $200 plan, which I'm on, it still
- 0:16seems to be an issue. Up until 3 weeks
- 0:18ago, I was still running into rate
- 0:20limits on the $200 plan and this was
- 0:22happening just after an hour session of
- 0:24Claude code. So, I started questioning,
- 0:26why is this actually happening? And I
- 0:28started digging deep into how Claude
- 0:29actually works and I worked out that I
- 0:31was using Claude wrong. And since then,
- 0:33for the last 3 weeks, I haven't hit a
- 0:35single Claude usage limit. And I
- 0:37literally use Claude 6 hours a day. So,
- 0:39in today's video, I'm going to give you
- 0:40my full framework to never hit your
- 0:43Claude usage limits again. After
- 0:44watching this video, you're going to be
- 0:45able to extract the maximum value out of
- 0:48your Claude subscription. And some of
- 0:49these tricks, I haven't heard a single
- 0:51other creator talk about yet. If you
- 0:53don't know who I am, I'm Miles
- 0:54Deutscher, I'm 25 years old and I've
- 0:56been using Chat GPT since it first came
- 0:58out 3 years ago. Since then, I've been
- 1:00going deep down the AI rabbit hole and
- 1:02I've used AI to help power my agency
- 1:04business to over $20 million in revenue
- 1:06and my digital product business to over
- 1:07$5 million in profit. As an AI
- 1:10enthusiast, I created this channel to
- 1:11help share my learnings with you. So,
- 1:13hopefully, you can also use AI to
- 1:15achieve your goals as well. All right,
- 1:16so the first step to never hitting your
- 1:18Claude limits again is all in your
- 1:20workflow. A really simple thing you can
- 1:22fix is being very intentional with how
- 1:24you prompt Claude. Most people figure
- 1:26out what they want whilst talking to
- 1:28Claude. And although it's a great
- 1:29brainstorming partner, don't use the
- 1:31most powerful model like Opus 4.7 for
- 1:34brainstorming. Make sure you really
- 1:35think about your prompt first before you
- 1:37use Claude. Or if you have to brainstorm
- 1:39with Claude, simply use a cheaper model,
- 1:41which we'll go through later in this
- 1:42video. Model selection is a huge trick
- 1:45that most people are missing and if
- 1:46you're just using the default model,
- 1:48that's probably one of your biggest
- 1:49issues. The next thing you need to
- 1:50understand is how Claude actually works.
- 1:53Certain tasks take up more tokens. For
- 1:55example, if I just chat to Claude, "Hey
- 1:57Claude, how are you?" and ask it
- 1:59questions and go back and forth, this
- 2:00actually isn't using up many tokens.
- 2:03Where you start running into token
- 2:04issues is whenever Claude has to build
- 2:06something. So, if you're creating a
- 2:08drawing, if you're coding, if you're
- 2:10creating an artifact, a dashboard,
- 2:12that's what burns through tokens. So,
- 2:14you're actually better off before
- 2:15building something, doing more planning.
- 2:17So, going back and forth to make sure
- 2:19the build is going to be right. Let's
- 2:20say you're building a finance dashboard,
- 2:22going back and forth to make sure the
- 2:24reason and architecture is fully correct
- 2:26before you physically build it. Because
- 2:28having to rebuild the dashboard is going
- 2:30to chew up much more tokens than simply
- 2:33speaking about building the dashboard.
- 2:35That's a major learning for me now. So,
- 2:36whenever I go and I build something, I
- 2:38always plan beforehand. If you're a
- 2:40coder, you can actually go into Claude
- 2:41Code specifically, and you can click
- 2:43plan mode. This will by default do no
- 2:46code and only focus on planning. So, if
- 2:48you're building stuff on Claude Code,
- 2:50going into plan mode first, doing all of
- 2:52your planning before you actually code,
- 2:55is going to save you a lot of tokens.
- 2:57So, that simple mindset shift, just from
- 2:59having the understanding of how tokens
- 3:00are actually burnt, makes a big
- 3:02difference. All right, now let's get
- 3:03into one of the major needle movers,
- 3:06chat length. Long chats are a silent
- 3:08killer. If you use a single chat and you
- 3:11keep talking to it, you have to
- 3:13understand how Claude works. It keeps
- 3:15having to go and scrape thousands of
- 3:17prior messages to understand the
- 3:19necessary context for giving you a
- 3:22response. So, what I prefer to do is
- 3:24actually set up projects. I have one for
- 3:25this YouTube channel, AI Edge, for
- 3:26example, and every time I start a new
- 3:28task like a video analysis or a new
- 3:31video script, I open up a brand new
- 3:34chat. This way, Claude is using up less
- 3:36tokens trying to understand the context.
- 3:38And the other really powerful thing you
- 3:40can do here is in your project folders,
- 3:42put in your memory and instructions. So,
- 3:45tell Claude exactly what the purpose of
- 3:46these chats are. For example, for AI
- 3:48Edge, it has a memory of exactly what AI
- 3:50Edge is, my goals with the business, my
- 3:52goals with the channel. Then in
- 3:53instructions, I tell Claude exactly what
- 3:56I want from it, how I want it to format
- 3:57its answers. Here's a trick if you're
- 3:59really trying to extract the maximum
- 4:00value from each plan. In your
- 4:02instructions, you can actually write in
- 4:04something like be cognizant of response
- 4:07length, try and speak in a concise
- 4:09manner, try and act in an efficient
- 4:11manner. If you just write that in to
- 4:13your instructions, your responses by
- 4:15default, whenever you're in that
- 4:16project, will actually be shorter, thus
- 4:18saving you tokens. Oh, and by the way,
- 4:20so you remember all of this stuff going
- 4:22forward, there'll be a free link in the
- 4:23description below which contains all of
- 4:25the tips from this video in a single PDF
- 4:28that you can then actually put into
- 4:30Claude to help it configure your chats.
- 4:32So, you simply claim that by going to
- 4:33link in the description, signing up for
- 4:35the AI newsletter, and you'll unlock the
- 4:36Instagram where the assets can be found
- 4:38inside the Google Drive. So, to
- 4:40summarize the chat setup, I'll say it
- 4:41like this. Three highly focused chats
- 4:44beats one long chat every day of the
- 4:46week if you set up your memory
- 4:48correctly. So, I've spoken about the
- 4:49memory system for the project folders
- 4:52themselves if you're just using the
- 4:53chats. Something that you may want to
- 4:55consider that I've done another video
- 4:57on, my second brain video, is actually
- 4:59setting up custom memory cuz you have to
- 5:01understand, you don't have full control
- 5:04if you're just using the chats over
- 5:05which memory Claude saves. It's
- 5:07essentially a black box. Yes, it is
- 5:09technically updating the memory, but you
- 5:11don't know exactly what and how it's
- 5:14updating. So, the best workaround for
- 5:16this is to actually create a local
- 5:17folder on your computer, like your AI
- 5:19brain folder, which stores all of your
- 5:21local memory. So, in a folder, you want
- 5:23to configure it something like this.
- 5:25Have an instructions.md file that
- 5:27basically has the instructions for
- 5:29Claude to follow. You'll notice there's
- 5:30a very important line here. Whenever new
- 5:33information surfaces during a session,
- 5:35the relevant MD file must be updated.
- 5:37Don't let facts slip through. This is so
- 5:40important because the instructions are
- 5:42actually getting the memory to auto
- 5:44update. Because without this, Claude is
- 5:46very random in terms of what it actually
- 5:48adds to the memory file. If you have a
- 5:50line like this in instructions, it means
- 5:52memory is constantly updating. Don't
- 5:54worry because in the free guide below,
- 5:55I'm going to give you the exact prompt
- 5:57to set up your memory and folder system
- 5:59in the correct way. And if you want to
- 6:01watch the dedicated video on that after,
- 6:03I've got a video on my second brain
- 6:04guide, which also connects this to
- 6:06Obsidian. But for now, all you need to
- 6:07know is that you need an instructions
- 6:09file and a memory MD file. Then, if you
- 6:12go back into Claude and you utilize
- 6:14co-work, the cool thing about co-work is
- 6:16you can work inside a specific folder.
- 6:18So, I can choose my AI brain folder, and
- 6:21every single time I'm typing a prompt
- 6:23that's connected to that folder on
- 6:24co-work, it's going to be tapping in to
- 6:26the instructions and all of the specific
- 6:29memory to that folder. And the great
- 6:30thing about this is that you actually
- 6:32own your own memory folder. It's not in
- 6:34a Claude black box, so if you do want to
- 6:36ever switch models in the future, or if
- 6:38you want to plug it into another tool,
- 6:40you actually own your memory. I
- 6:41personally think memory sovereignty, so
- 6:43owning your own memory, is the most
- 6:44important thing in AI because memory is
- 6:47transportable. If you have your memory,
- 6:49you're going to be a very effective AI
- 6:51user. The people that I find that aren't
- 6:52very effective with AI don't own their
- 6:54own memory and don't have the systems in
- 6:56place to update their memory. The good
- 6:58thing is, it's super simple and I've
- 6:59actually developed a prompt for it to
- 7:01save you time. All right, that was layer
- 7:03one, your workflow. Let's talk about
- 7:05layer two, and this is just so important
- 7:07and I don't see enough people doing it.
- 7:09This is model stacking. If you are using
- 7:11Opus 4.7, the most expensive model to
- 7:14ask questions, do basic scraping, and
- 7:16basic research, you're just burning
- 7:18through tokens. There are other models
- 7:20that are 90% as capable on these basic
- 7:23tasks that you can default to, and then
- 7:25only use Opus 4.7 for the work where you
- 7:29really need a smart model to shine. I'll
- 7:30show you what I mean. So, if you go into
- 7:32create a new chat on Claude, you can
- 7:34select your model. So, Opus 4.7 is
- 7:36clearly the smartest. So, if I was
- 7:37making a big business decision, if I was
- 7:39coding a financial dashboard, and if I
- 7:41was analyzing really important data, of
- 7:43course, yes, I will use this. Still with
- 7:46the framework that I showed you in step
- 7:47one. However, the other models are still
- 7:49pretty good. Specifically, Sonnet.
- 7:51Sonnet for most real work is great. If
- 7:52you're analyzing a PDF, if you want to
- 7:55update some text, if you want to do a
- 7:56bit of writing, Sonnet in my opinion is
- 7:58fantastic. And Haiku is still pretty
- 8:00capable for quick tasks like extracting
- 8:02data, scraping. There are probably
- 8:04people that use Open Clo that are
- 8:05watching this video. I do too. For
- 8:07certain workflows, Haiku is a really
- 8:09quick and easy way to scrape data from
- 8:10the internet. And the way I do it is I
- 8:12use the cheaper models for the scraping,
- 8:14for the repetitive tasks, and then I'll
- 8:16use Opus 4.7 as my curator. It's the
- 8:19model that curates all of the research
- 8:21from these other models. If you're not
- 8:22using Open Clo, you can still implement
- 8:24the strategy on the desktop application
- 8:27of Claude itself through simply being
- 8:29cognizant of what model you're using,
- 8:31using Sonnet 4.6 if you think that the
- 8:33task is less important or slightly
- 8:35easier, and then just manually switching
- 8:37to Opus 4.7 when you think the task is
- 8:39slightly more advanced. On Claude Code
- 8:41as well, if you're a coder, this is
- 8:43something to be very cognizant of.
- 8:45There's actually an effort level that
- 8:46you can select. If you're on extra high
- 8:48or max, you're going to chew through
- 8:50tokens. Yes, you're going to get a
- 8:51better result, but let's say you're only
- 8:53using Claude Code to update local
- 8:55folders, you're not going to need max
- 8:56effort. You can get away with medium or
- 8:58even low in some cases. I would only use
- 9:00extra high or max effort if you're doing
- 9:02a task where you really need precision
- 9:04and output. And Claude Code, just like
- 9:06the chats, allows you to select the
- 9:07model. Now, I'm not going to lie, Opus
- 9:094.7 is way better than the other models,
- 9:11but if you need something quickly vibe
- 9:13coded, like let's say, you know, a
- 9:14little dashboard or just a interface or
- 9:17a web portal or something relatively
- 9:18easy, you don't need to use it. I would
- 9:20just prioritize using it for the more
- 9:21advanced use cases. And here's the
- 9:23reality as well. You don't just need to
- 9:25use Claude. I use Quen and Kimi on my
- 9:28Open Clo. I use Grok for real-time news,
- 9:31which is built into my X subscription. I
- 9:33use Gemini as well, which is built into
- 9:34my my subscription. You don't just need
- 9:36to use Claude. So, just be cognizant of
- 9:38when you're using the model. If you're
- 9:39really running into issues and let's say
- 9:41you're on one of the lower plans, just
- 9:43make sure you use Claude for the work
- 9:44where you need it. Understand where
- 9:45Claude shines. It really shines when it
- 9:47comes to interactive dashboards, when it
- 9:48comes to coding. But if you just want a
- 9:50little bit of news, you can use Grok.
- 9:52You can have another cheaper model. You
- 9:53can even use GPT. For voice prompting
- 9:55and brainstorming, I do that all the
- 9:56time and then I use Claude as my curator
- 9:58because I know what it's good at. So
- 10:00just understand the strengths of each
- 10:01model and combine multiple models in
- 10:03your workflow. I actually did a tools
- 10:04video on this where I break down my use
- 10:06case for my top seven AI models and in
- 10:08that video, I break down the strengths
- 10:10of each model. And I feel like that's
- 10:12also one of the reasons I was able to
- 10:13reduce my token expenditure cuz I've
- 10:15started leaning on other models for
- 10:17things where those models genuinely
- 10:18shine. Because although Claude is great,
- 10:20other models actually do some things
- 10:21better. So, keep that in mind next time
- 10:23you're doing a task. All right, now you
- 10:25understand the second step, which is
- 10:27selecting the right model for your use
- 10:29case, let's talk about something
- 10:31underrated, tool splitting. The
- 10:33important thing to understand is that
- 10:34although Claude is one application, it's
- 10:36not just one product. You have the chat,
- 10:39you have co-work, and you have Claude
- 10:40Code. And you need to treat each
- 10:42individually. And you even have extra
- 10:43products now like Claude Design with its
- 10:45own separate usage limit. Now, Claude
- 10:47treats the token limits for Claude Code
- 10:50and the Claude Chat the same with design
- 10:52being separate. But this is where you
- 10:54need to understand who you are as a user
- 10:55because if you are a heavy Claude Code
- 10:57user and you're chewing through your
- 10:59overall plan with Claude Code, then what
- 11:01you can actually do is hook up the API
- 11:03to Claude Code. It will run separately
- 11:06through the API. You need to go onto
- 11:07Anthropic's website to the API section
- 11:10to create one. And then this will
- 11:12actually now be separate from your chat.
- 11:14If you're a normal user who only uses
- 11:15Claude Code sometimes, you can get away
- 11:17with having it under your main plan. The
- 11:19only thing I'd urge is that you actually
- 11:21go into claude.ai/settings/usage
- 11:26and actually look at where your usage
- 11:27is. So just being aware of where your
- 11:29limits are. For example, right now, I've
- 11:31used up 12% of all models and it resets
- 11:33on Friday at 9:00 a.m. so I'm all good.
- 11:35But I've actually run out of Claude
- 11:37design cuz I was going crazy on that
- 11:38yesterday and that's a completely
- 11:40separate topic because Claude design is
- 11:42so usage-intensive. And now if I want
- 11:44more tokens for Claude design, I
- 11:46actually need to go in and I need to
- 11:48purchase pay-as-you-go usage credits,
- 11:50they call it, for Claude design. This is
- 11:52also a strategy where instead of you
- 11:54having to upgrade your entire plan, if
- 11:56you notice that you're about to hit your
- 11:57limits in let's say 24 hours and you
- 12:00just need to do a few more prompts, you
- 12:02can actually just buy more usage credits
- 12:04instead of upgrading and spending the
- 12:05extra $100, jumping from 100 to $200,
- 12:08you might only need, you know, 5 to 10
- 12:10extra dollars worth of credits to get
- 12:12you over that hump. So
- 12:13claude.ai/upgrade,
- 12:14this will allow you to manage your plan.
- 12:16So just being aware of it, I think is a
- 12:17big thing and actually checking in on
- 12:19your limits so, you know, you're not
- 12:20blindsided when something happens. And
- 12:22this will can also change your behavior.
- 12:23If you notice you're at 75% and you need
- 12:25the model for work tomorrow, then maybe,
- 12:28you know, you won't go crazy vibe coding
- 12:30application the night before. I think
- 12:31that awareness is really key and
- 12:32although it's annoying, it's something I
- 12:33think a lot of people are missing. Now
- 12:35unfortunately, Claude is actually
- 12:36getting stingier with their limits. This
- 12:38is a conversation for another day but
- 12:40energy and compute is a real problem for
- 12:42them and that's why we've actually seen
- 12:44degraded performance across Claude over
- 12:46the last month or so. You've probably
- 12:47noticed Claude in some cases has become
- 12:49dumber and although the new models are
- 12:50great, like 4.7, it's also really
- 12:53usage-intensive. And they've actually
- 12:54taken it a step further and now as you
- 12:56can see in the pro plan, you no longer
- 12:58have access to Claude code. You can only
- 13:00get the Claude code plan from the $100 a
- 13:02month plan or the $200 a month plan. So
- 13:04$17 subscription users no longer have
- 13:07access to Claude code. I kind of get why
- 13:09they're doing this. They are obviously
- 13:11making a bet that serious users who want
- 13:13to use Claude code are willing to pay
- 13:14$100 and more and that they need to
- 13:16conserve energy somewhere so they would
- 13:18rather, you know, ice out the pro cohort
- 13:21but it still does suck and this is just
- 13:23something we need to accept about the
- 13:24current state of AI. Prices are not
- 13:27going to get better. They are only going
- 13:28to get worse. So, I do feel like a video
- 13:30like this actually might be one of the
- 13:31most important ones of the year, so you
- 13:33can save money longer term, because this
- 13:35problem is only going to get worse. You
- 13:36need to tweak your usage habits now,
- 13:39unless you've just got money to burn,
- 13:40because these features are only going to
- 13:42become more restrictive. I mean, we're
- 13:43already seeing features like Claude 2
- 13:45design, which are amazing, the average
- 13:46person can't afford. I blew through my
- 13:48entire plan in a 12-hour session, and if
- 13:51I want to keep using it now, I'm
- 13:52probably going to spend $50 to $100 a
- 13:54day if I want to use it seriously. So,
- 13:56it is a problem, and if you're on a
- 13:57budget, you might want to reconsider, as
- 13:59I mentioned in the last step, which
- 14:01tools you're actually using. And if you
- 14:02do use an interface like Hermes Agent or
- 14:04Open Claude, you might want to consider
- 14:06even running local source models if you
- 14:08want to front-load an investment into
- 14:09hardware, like a solid Mac Studio that
- 14:11can actually run local models, then
- 14:13you're literally running an open-source
- 14:14model on your computer, and then you're
- 14:15not actually paying a subscription, but
- 14:17you have to bear the upfront hardware
- 14:18costs. So, your architecture will depend
- 14:20on what you want to get out of AI, but I
- 14:22think now we actually really need to be
- 14:23cognizant of where our subscription
- 14:25money is going. I still think Claude is
- 14:27the best subscription to have overall,
- 14:29but I think no longer is it a viable
- 14:31option for absolutely everything just
- 14:33due to the sheer cost issue that we are
- 14:35experiencing. I think it's good to have
- 14:37other subscriptions or even utilize
- 14:39other free tools to spread the load up.
- 14:41Now, I want to give you one of my
- 14:42biggest tips in the video. If you really
- 14:44want to save usage cost, if you have
- 14:46tasks that you do on a repetitive basis,
- 14:48if there are things that you do in your
- 14:50day-to-day or in your weekly life,
- 14:51simply create Claude skills for them.
- 14:53It's a bit of work up front, but what
- 14:55this will then generate is the ability
- 14:57for you to hit {forward slash} and then
- 14:59select a skill, for example, creative
- 15:01session, which now loads up my morning
- 15:03brief skill that I follow every day. So,
- 15:06I've front-loaded the work, I've trained
- 15:08Claude to do something, be my morning
- 15:09assistant, and now whenever I want to do
- 15:11my morning session, I simply enter the
- 15:13skill. You can do the same for script
- 15:15writing. You can train a Claude skill to
- 15:16write like you. You can do the same for
- 15:18portfolio analysis. You can train Claude
- 15:21to analyze the portfolio in the way that
- 15:23you want with the exact parameters that
- 15:24you want. And the reason why this
- 15:25actually saves tokens over time is
- 15:27because Claude knows exactly what to do.
- 15:29You're not constantly correcting it and
- 15:31re-prompting, which is what ends up
- 15:33burning tokens. So, I've actually done a
- 15:34full guide on Claude skills, how to
- 15:36develop skills from scratch. I'll leave
- 15:37that in the top corner if you want to
- 15:39check it out. And that's another feature
- 15:40of Claude that you can utilize to
- 15:42significantly reduce your token
- 15:43expenditure. Now, let's say you're in a
- 15:44situation where you've done everything
- 15:46I've told you about in today's video,
- 15:48and you are still running into usage
- 15:50limits. Let's just be frank with each
- 15:52other. Maybe it means you need to
- 15:54upgrade your subscription. Now, I don't
- 15:55think you need to jump to the $200 plan
- 15:57if you haven't implemented anything I've
- 15:59spoken about today. But if you have and
- 16:01you're still running into issues, you
- 16:02might just be a power user of AI, and
- 16:05that is just the cost of using AI right
- 16:06now. So, if you really want to use
- 16:08Claude, if you love Claude code, if you
- 16:09love Claude co-work, and you're still
- 16:11running into issues on the $17 plan, you
- 16:13may have to upgrade to 100. And if
- 16:14you're still running into issues like I
- 16:16was on the $100 plan, you may have to
- 16:18upgrade to 200. That is just the name of
- 16:20the game right now, unfortunately. Of
- 16:22course, as I said, you can use other
- 16:23models in order to alleviate some of
- 16:25that burden, but if you're a Claude
- 16:27power user, this is just the cost of
- 16:29using the best frontier model in the
- 16:30world right now. What I will say,
- 16:31though, is don't just jump into the max
- 16:33plan if you haven't optimized. Optimize
- 16:35first and then make subscription
- 16:37upgrades later. That's obviously the
- 16:38best way to go about it. If you want all
- 16:40of my tips condensed into one framework
- 16:42from today's video, make sure to use the
- 16:44link in the description below to claim
- 16:46your free PDF. It's going to contain all
- 16:48the tips and tricks, including the set
- 16:50up prompt for an efficient instructions
- 16:52and memory system. Thanks everyone for
- 16:53your time today. Hopefully, I was able
- 16:55to help you out, save you some money,
- 16:57and help you extract [music] more out of
- 16:58your Claude. I will see you in the next
- 17:00video. Have a lovely rest of your day.
- 17:02Peace out.
About this transcript
This page contains the full transcript of Never Pay For Claude Upgrades Again by AI Edge, generated from the public captions YouTube serves with the video. The transcript has 3,896 words across 550 segments, with the original timestamps preserved so you can click any line to jump to that moment in the embedded player.
What you can do with it
Use the transcript to take notes, quote the speaker, build a study guide, generate a summary with ChatGPT or Claude via the YouTube Summary tool, or export it as a timed subtitle file with YouTube to SRT. You can also re-open it in the transcriber to translate the transcript into 100+ languages.
Free YouTube transcript tool
YouTube2Text is a free YouTube transcript generator — no signup, no daily limit. Paste any YouTube link and get the full transcript instantly, with timestamps, click-to-jump, translation to 100+ languages, AI prompts for ChatGPT, Claude, and Gemini, and exports to TXT, SRT, VTT, or Markdown.