Getting the transcript
Reading the captions from YouTube. A video nobody has opened here before takes 10 to 30 seconds; this page fills in on its own.
Getting the transcript
Reading the captions from YouTube. A video nobody has opened here before takes 10 to 30 seconds; this page fills in on its own.

Sharbel A. · @sharbelxyz
This video has no Most replayed graph yet: YouTube shows one only once a video has enough views. These are the moments viewers replayed most in Sharbel A.'s most watched videos.
Most replayed moment at 13:32
3.9x that video's typical replay level
built-in memory, which is great. Here is how I would think about memory providers. Memo is interesting if you want a dedicated memory layer for personalized AI agents. It focuses on extracting, storing, linking, and retrieving memories efficiently. Their
Said at 13:25
Most replayed moment at 6:37
5.8x that video's typical replay level
trading strategy. Tests it. If it's better, it keeps it. If it's worse, it discards it and tries again. So, that's exactly what I built. Okay, here's the system we have at play. I gave it two years of crypto data,
Said at 6:31
Most replayed moment at 1:53
3.7x that video's typical replay level
let's install it together. Okay, so for step one, we need to actually start by installing Bullpen's CLI so that we can actually do everything that we want to do. And for that, let's first open Claude and put in the dangerously skip
Said at 1:47
The graph counts replays. It does not show where viewers stopped watching.
Words
2,640
Runtime
17:36
Speaking pace
150wpm
Reading time
11min
150 words per minute, below the 160 25th percentile of 349 measured videos. That distribution comes from the 349-video hook study.
Opening (first 30 seconds)
Hermes says it's free, but I had a client paying $5,000 a month for it. That's $60,000 a year, and he had no idea why. See, almost everybody falls into this trap at some point. Even if you're not using Hermes that much. This is because everything you run Hermes on gets pushed through its most expensive methods, even the simple stuff that never needed it. And it's why most people quietly burn thousands of
75 words, the words spoken in the first 30 seconds at 150 words per minute.
Free, no signup. See how the first 30 seconds hold attention, with rewrites.
Sentence shape
| Measure | This transcript |
|---|---|
| Sentences | 174 |
| Average words per sentence | 15.2 |
| Longest sentence | 67 words |
| Questions asked | 7 |
| Sentences containing a number | 17 |
Most used terms
Filler phrases
25 in total: like 9 · actually 6 · uh 6 · basically 1 · kind of 1 · literally 1 · you know 1.
A literal whole-word count of the same phrase list the Prepublish browser extension uses, so a phrase inside another word is not counted and a phrase used in its ordinary sense still is. It is a count and not a judgement.
Free, no account. See where attention is likely to drop, with a rewrite for each weak line. The free check shows the scores and the one issue costing the most. Or run it on the words above first.
Free · No login · See a sample audit first if you prefer.
What this transcript is
Every word below is the caption track YouTube publishes for this video, pulled from the video itself and reproduced unchanged. It is not Prepublish's writing, not a summary, and not a re-transcription: it is the video's own published captions. English captions, generated automatically by YouTube, in the video’s original language. Source: the video on YouTube. A channel that would rather this page did not exist can ask for its removal through the contact page, and it is removed.
Hermes says it's free, but I had a client paying $5,000 a month for it. That's $60,000 a year, and he had no idea why. See, almost everybody falls into this trap at some point. Even if you're not using Hermes that much. This is because everything you run Hermes on gets pushed through its most expensive methods, even the simple stuff that never needed it. And it's why most people quietly burn thousands of dollars with it.
So, today I'm going to give you the exact steps to fix this. We'll go over exactly how to run Hermes for free, how to stop it from turning [music] into a trap, and if you do ever need to pay for it, the setups I'd actually pay for that don't waste your money. Let's get started. Hermes itself is free. However, the models you plug into it is where the cost comes from. And the way you route tasks is what decides whether Hermes feels cheap, expensive, or completely unusable.
If you use a premium model to summarize a tiny note, that is waste. If you use a weak free model to debug a real code base, that is also waste because now the cost is your time. The goal is not to be cheap for the sake of being cheap. The goal is to use the right model for the right job. That is why I think about Hermes in three tiers. Free, affordable, and optimal. Free is for learning and simple workflows. Affordable is the best setup for most users.
Optimal is for people who want Hermes to be part of their daily operating system. Let's start with number one, free. Okay, starting off with the first tier, free. Yes, you can run Hermes agent completely for free. However, free is not cheap. Wait, what? The free setup is simple in concept. Install Hermes locally on a local machine, then connect it to either a local model or a free model route. If you want the private version, you run a local model through something like Ollama.
If you want the easiest free version, you use whatever free hosted model access is available at the time, but you should assume those free hosted models can change, get rate limited, or disappear. And with that, you now see the cost of free is you need uh a local machine first and foremost, and you know, you could run Hermes on a cheap machine, but you can't run local models on a cheap machine. So, in order to be able to use Hermes completely for free, free of charge, like zero-dollar cost, then you would probably need a $10,000 machine.
Maybe a little less, maybe a little more depending on how fancy-schmancy you want to get, but free is great for learning Hermes. If you want to go as most affordable as possible, it is good for understanding how sessions work, in case you have uh a local machine laying around and you want to plug it into uh the the cheapest models possible, it can help you test memory, test skills, uh play around with Hermes, ask it to summarize short notes, and learn how the tool works overall.
If Hermes is turning a short note into a checklist, it doesn't need the strongest model in the world. If it is summarizing a small transcript, pulling out action items, or helping you draft a basic reply, free or local can be enough. If the workflow is private and simple, local can actually be the better choice here. But, this is the line I would draw. Free is perfect for learning Hermes. Free is not the setup I would trust to run a business.
Free models are usually weaker, slower, have smaller context windows, and are overall less reliable with tools. That means they can struggle with long sessions, complex instructions, coding tasks, and multiple step workflows. So, if Hermes is editing a production repo, debugging a confusing error, or making a decision where a bad answer costs you time or money, I would not rely on the free setup. Free gets you started, but it is not where Hermes becomes powerful.
Let's move on to tier two. The setup I would recommend to most people is the affordable setup, because this is where Hermes starts becoming genuinely useful without making you scared to use it. My affordable setup would be if you have a local machine, then use your local machine. If you don't, then you can get a 5 to $8 a month VPS cloud subscription, so that it can run your Hermes instance. But, most people have some kind of local device, local machine running, which most machines, even as cheap and dusty as they can be, if you can keep it on, then it can do the job.
So, my affordable setup would be a $20 a month Codex subscription linked to Hermes, plus Open Router for smaller task. Codex handles the work that needs a strong model. Open Router handles the smaller tasks that do not need premium reasoning. This is the key shift. Stop asking, "How do I make every task free?" and instead start asking which tasks actually deserve the expensive model. That one question is what makes Hermes affordable.
In this setup, Codex is for serious work. Coding, debugging, repo edits, project planning, important reasoning, and multi-step tasks where Hermes has to use tools properly. Open router, on the other hand, is for smaller work. Think summaries, extraction, formatting, classification, lightweight cron jobs, and low-risk assistant tasks. Here's the simple version. If Hermes is summarizing a YouTube transcript, for example, cheap model.
If Hermes is turning meeting transcript into bullets, cheap model. But if Hermes is editing your codebase, debugging a broken deploy, planning a complex feature, or running a tool-heavy workflow where one-bad-step can waste your afternoon, use Codex. This works because most daily tasks are not equally hard. A lot of work is just transformation. The moment Hermes has to reason across files, or make changes that matter, that is when you need and want the stronger model.
That is the difference between being cheap and being stupid. On to tier three, what I like to call the optimal tier. The optimal setup is for people who want Hermes to become part of how they work every day. This is the setup you will want eventually as you scale your work with Hermes. If you're using Hermes as a real assistant, a coding partner, a content operator, or a business workflow layer, this is where I'd use the $100-a-month Codex subscription linked to Hermes, plus Open Router for smaller tasks.
It's very similar to the tier before it, but instead of the $20 subscription, we're using the $100. It gives us five x more usage. In case you're an insane user of Hermes for any reason, you can use the $200 subscription for Codex, which is 20 times more usage, but myself, someone who works full-time in this space, I myself use the $100 subscription a month. It will be more than enough. For 99.99% of people, the $100 Codex subscription is enough.
But even in the optimal setup, I would not use the most expensive model for everything. That is still wasteful. The optimal setup is stronger model access for the tasks that matter, with cheaper routing for everything else. So, what belongs on the premium model? Serious coding, debugging, complex repo edits, multi-step tool use, private business workflows, and anything where you are going to ship the output or act on the answer.
If Hermes is changing files, interpreting errors, uh coordinating a workflow, that is where you want the better model. Basically, where you can't afford to risk a mistake. What still belongs on cheap models, even in the optimal setup? Things like transcript clean-up, simple summaries, cron jobs, a lot of your cron jobs can run on cheaper models, and low-risk tasks. The best part is you can literally ask Hermes, "Hey Hermes, based on my tasks and crons, what should I route to, for example, Kimmy K2, and what should I route to my Codex plan?
Help me save on cost without losing quality." And the great thing is your own Hermes can help you audit your own personal workflow, so that you can get personalized answers. Uh for example, for me personally, I just asked it this and it audited my workflow and it found a bunch of new crons, things that I've set up recently or within the last month that I can route to Kimmy K2. I can give it more models. For example, it can do the analysis itself.
That is the great part about Hermes. The optimal setup is not about flexing the most expensive subscription. It is about reducing friction. When Hermes becomes part of your daily workflow, you do not want to constantly think, "Can I afford this task?" You want important work to get the best model and small work to go somewhere cheaper. That is the difference between a toy setup and an operating setup. Free is for learning, affordable is for everyday usefulness, and optimal is for when Hermes becomes part of your actual work.
Now, here's the actual routing framework I would use because this is the part that saves money in real life. Use free or local models when cost or privacy matters more than quality. Use cheap open router models like GLM 5.2 or Kimmy K2. Those are great value for money models right now, by the way, that are almost as good as GPT or Claude models for a fraction of the cost. I know a lot of people that use these as their main models and they feel just as good sometimes.
You can use those when the task is low risk and reversible. Test them out. Try them out yourself. You can ask them summarize this, format this, classify these messages, or just run these crons for me. And use Codex when the task is high risk, long context, or tool heavy. Things like edit this repo, debug this error, plan this feature, all of this workflow, run a multi-step research task, or use tools and verify the result.
Those are strong model tasks. If you remember one rule, make it this. Cheaper models are for reversible work. Strong models are for work you will act on. That one sentence will save you more money than obsessing over every model leaderboard. Now, let's talk about what actually burns money in Hermes, because this is where the surprise bills come from. Hermes can get expensive because it is powerful. It remembers things, it uses tools, it can run long sessions, even call on sub agents, and run scheduled jobs every single day.
All of that is useful, but all of that can also increase tool calls and background usage. So, if you want Hermes to stay cheap, the first rule is to set spending caps. This is my OpenRouter account, and you can just see over the last month I've spent $40, but I use OpenRouter for other things in our business as well. If I go to API keys, you can see for Hermes alone I've spent so far this month $2.3, and that was mostly a lot of crons I've run through Kimmy K2, and some through NanoBanana for image generation.
If you use OpenRouter, I strongly recommend you create an API key just for Hermes, just like I have, and give it a monthly or weekly spend limit. Do not use one unlimited key ever. Separate keys make it way easier to see what is actually costing you money. You can click new, you can name it Hermes, you can put a credit limit, say $100 a month, and you can set it for whatever frequency you want, monthly expiration, you can decide on that yourself.
I don't personally use expirations. I just renew them manually every now and then, whenever I feel like it, but that is it. You click create, copy the API key that it just generated, and boom, you're good to go. Second, do not run expensive models in the background unless you have a very good reason for. If a cron is checking something, summarizing a small feed, or formatting a report, it probably does not need the premium model.
Third, start a fresh session when the context gets bloated. Long sessions can become expensive because the model keeps seeing more and more context. If you're done with a task, start a new session. Do not make every little request part of the same giant conversation. Fourth, disable tools and skills you do not need. Tools are one of the reasons Hermes is useful, but if every task has access to everything, the agent can spend time thinking about things it does not need.
This is where creating dedicated agent profiles for every task you need in Hermes makes sense. Because with Hermes, you have something called Hermes profiles, and with Hermes profiles, you're able to turn on and off certain skills for every different profile that you set up and you can customize. And fifth, do not automate nonsense. This sounds obvious, but it is the biggest hidden cost. People set up agents to run every hour or every minute on tasks that do not matter.
They create daily reports that they don't end up reading. They run background jobs because it feels cool. Trust me, I've been there. But if a job doesn't save you time or make you money, then it's probably not needed. If a workflow would not be worth doing manually once, it is probably not worth automating every day. That is the real cost control. Not just cheaper models, better judgment. So, here is my final recommendation.
If you're brand new to Hermes, start free. Install Hermes, connect it to a free or local model route, and learn the interface. Do not pay for a big setup before you even know what you want Hermes to do. If you want Hermes to become useful, move to the affordable setup. Use the $20 a month Codex subscription linked to Hermes for serious work, then use Open Router for smaller tasks. That is the setup I would recommend to most people because it gives you enough power to actually use Hermes without making every tiny task feel expensive.
And if Hermes becomes part of your daily work, move to the optimal setup. Use the $100 a month Codex subscription linked to Hermes. Keep Open Router for smaller tasks, set limits, and route intelligently. That is the version I would use if Hermes is helping you with coding, content, research, operations, and real business workflows every single day. The mistake is thinking there is one perfect setup. There is not. There is the setup for learning, the setup for using, and the setup for operating.
Free gets you started. Affordable gets you useful. Optimal gets you leverage. So, yeah, you can run Hermes agent for free, but the better question is, what do you want Hermes to do? If you just want to learn, free is enough. If you want it to help you every day, use the affordable setup. If you want it to become part of your workflow, pay for the strong model where it matters, and stop wasting it where it does not. That is how you run Hermes without wasting money.
If you enjoyed this video, make sure to subscribe because I have a ton more content coming your way just like this. And the algorithm gods predict that you are very likely to enjoy this video next. So, click it and I'll see you there.
The words are the caption track's own and nothing is reworded or re-transcribed. Paragraph breaks are placed between sentences so the text reads as prose.
Free tools for your own script: paste a draft and see where it stands before you record it.
Paste your draft and see where viewers are likely to drop off, with a rewrite for each weak line.
Paste the first 30 seconds of your own draft for a hook score and rewrites.
Check your draft against YouTube's advertiser-friendly guidelines before you record it.
Read this channel's public videos and transcripts, and download a writing brief for it.