How One Fable Chat Manages 100+ Codex Agents — Transcript
Full transcript
- 0:00Hey guys, I've
- 0:02built out quite an interesting system
- 0:04over the past few days that is just made
- 0:07to be as simple as possible for having a
- 0:11for having one orchestration agent
- 0:14proactively push your tasks and goals
- 0:18forward.
- 0:19So this is based off of
- 0:22the the agent that I use is like the
- 0:23central orchestration agent. It is Fable
- 0:27and I'm managing all of my tasks from
- 0:30within notion but you could do it from
- 0:33you know, I previously worked from
- 0:35linear. I've worked from GitHub issues
- 0:37which in fact was was sinking over to
- 0:39linear originally when I I made this
- 0:41system just
- 0:43two days ago. It was with like markdown
- 0:46files and it was kind of just like local
- 0:48tickets, but I've decided to and I I am
- 0:51actually finding that I do keep like
- 0:52defaulting back to notion
- 0:54and I think as like a kind of having a
- 0:57central workspace for agents.
- 1:00I do think it it it really works well
- 1:02for that job. So I'll walk you through
- 1:05like the kind of entire system just to
- 1:07give you like a kind of high-level
- 1:08overview and then I can give like a
- 1:10little demonstration of it as well. I'll
- 1:12try to keep this video shorter than than
- 1:15all of the previous one. So
- 1:17you have one main Claude chat and I'm
- 1:21currently choosing to run on on Fable
- 1:25and
- 1:26the the main Claude chats it
- 1:29chat it does not do any work itself.
- 1:32It simply acts as the agent that you
- 1:35speak to. It's like a chief of staff and
- 1:37it
- 1:39is just going to offload
- 1:41all of that work to sub agents.
- 1:45So you use the main Claude chat in this
- 1:48case Fable for planning for talking kind
- 1:51of brainstorming for discussing ideas
- 1:54for clarifying the scope of work and
- 1:56this is going to be the main agent and
- 1:58it's I mean going in my opinion it's
- 2:01it's the smartest agent
- 2:03within the fleet. But again, you could
- 2:05swap it out for any agent that you
- 2:07prefer. It it it really doesn't matter.
- 2:11And what it's going to do is it's going
- 2:12to
- 2:13offload all of the tasks, to delegate
- 2:16all of the tasks to Codex,
- 2:18to Codex app server. The reason I built
- 2:20this system is I really hate speaking to
- 2:225.6 soul. I in all honesty just find it
- 2:26really difficult to work with. The
- 2:28outcomes are good, but the
- 2:31just the general kind of personality,
- 2:34the
- 2:35yeah, just the the the the general
- 2:37experience of working with GPT 5.6 soul
- 2:39I would say isn't a pleasant one, but it
- 2:42kind of gets the job done. So, I'm using
- 2:44Fable which is much easier to
- 2:46communicate with. Like basically I was
- 2:48just kind of finding myself getting a
- 2:50little bit like annoyed and let's say
- 2:52like more stressed than I would be when
- 2:54dealing with GPT 5.6 soul. So, instead
- 2:57I've put Fable there as like the middle
- 2:59layer which also does all of the the
- 3:02planning and architecture stuff as well
- 3:04and GPT does as well. So, now you have
- 3:06the benefit of two agents
- 3:08that are able to interact with each
- 3:09other which in my opinion it leads to
- 3:12much higher token spend, but it also
- 3:13leads to much better outcomes. Um so,
- 3:16all of the actual work, the
- 3:17implementation work is done by 5.6 soul.
- 3:20Um again, it's using Open AI's official
- 3:23Codex app server which just gives you a
- 3:25adjacent API where you can kind of send
- 3:28messages into a Codex chat. You can
- 3:30archive chats, you can rename them, you
- 3:32can
- 3:33approve permissions, you can do anything
- 3:35basically that's available within the
- 3:37app um programmatically.
- 3:40And then what I have is every 20
- 3:43minutes, so again to go over the whole
- 3:45architecture, every 20 minutes I have a
- 3:47worker which is just like a cron job
- 3:50that triggers a
- 3:52Claude agent. It's not using Fable. I
- 3:55think it was either Sonnet or maybe
- 3:57Haiku, actually. And it lists our recent
- 4:01chats in in Codex, and it just makes
- 4:03sure that nothing's blocked, that
- 4:05nothing needs our attention. So, it
- 4:07makes sure that all work is still moving
- 4:09forward, and then it will send that back
- 4:11to our orchestration agent. So,
- 4:14if there's something that needs our
- 4:15attention, if there's something
- 4:18Yeah, anything. The job of this
- 4:20orchestration agent is to push work
- 4:22going forward. And so, this worker just
- 4:24lets the main agent here know, the chat
- 4:26down here, that like something's
- 4:28blocked. Maybe it needs a decision from
- 4:30me. Maybe the orchestration agent can
- 4:33solve it himself if we've previously
- 4:34solved something similar and it's logged
- 4:36in in decisions.
- 4:38It it it kind Again, this is like how
- 4:40I've built this system, and I'm kind of
- 4:42comfortable with giving more autonomy to
- 4:45Fable. But, you may not want to do that.
- 4:47You may want to say like, you know, all
- 4:49decisions have to have to have to be be
- 4:51made by you. So, it's just a kind of
- 4:53personal preference and kind of how much
- 4:55you uh
- 4:56you trust the agent.
- 4:58Uh so, as I mentioned, like I'm working
- 5:00from Notion tasks. I'll I'll show you
- 5:02how that dashboard look looks in In
- 5:04fact, I'll just show you right now just
- 5:05to um
- 5:08to give you an idea. So, what I'm
- 5:09currently doing That should be priority
- 5:11five. What I am currently doing then is
- 5:16I'm just trying to see what ones I can
- 5:18open. So, okay, I'm This is it's just a
- 5:22simple like Notion Notion database.
- 5:25Here's like a few like different tasks
- 5:27I'm working on. I was just checking
- 5:28there's nothing like confidential there.
- 5:30So, this is the priority
- 5:33tab. And so, in order just to not
- 5:36overwhelm the Claude agent with like
- 5:40tons of different things going on. I
- 5:43mean, technically you could it can
- 5:44probably handle it. It's more for my own
- 5:46sanity, to be completely honest, but
- 5:48I've limited it to only be working on
- 5:50like five active tasks at one time. In
- 5:53my backlog, I've got about 900 tasks,
- 5:56but here I want a super clean view of
- 5:59basically what is going on. Now, you can
- 6:02also like click into it and each task
- 6:04can have like its own
- 6:07acceptance criteria and this like
- 6:10progress circular icon thing like
- 6:13automatically increases as different
- 6:15tasks get
- 6:17get ticked off. Agents then leave
- 6:19comments. I'm not going to open these
- 6:21because this is for it's hedge fund
- 6:23stuff, so I just want to
- 6:25to leave it off. In fact, this one's
- 6:26fine. Like agents leave comments, so
- 6:28like if you want to get like a quick
- 6:30update on what's going on, they come in
- 6:32every 30 minutes. It just gives a
- 6:34summary, so you can kind of get an
- 6:35overview of each task for what's going
- 6:38on.
- 6:38And on the far right-hand side, this is
- 6:41the agent status.
- 6:43So,
- 6:44it it in progress means there is
- 6:47actively activity happening that there
- 6:49is an active chat working on this and
- 6:52pushing it forward. If it's blocked, it
- 6:54means it could need something from me.
- 6:57You know, so needs plan, that will
- 6:59automatically be kicked off. It will
- 7:00start writing a plan and then it will
- 7:02ask me for information. If it's done,
- 7:04it's asking me if I want to approve the
- 7:05spec. In progress, work is actively
- 7:07moving forward. Blocked, something's
- 7:09wrong. So, the Claude's chat will
- 7:11automatically pick that up. I'll show
- 7:12you that in a second. It's kind of a so
- 7:15I don't think I really really need to
- 7:16explain it. And then this on the last
- 7:18side shows the last time that this chat
- 7:22was successfully picked up. So, if it's
- 7:25in progress here, it means that 19
- 7:27minutes ago when the previous run went,
- 7:30the worker that I mentioned that runs
- 7:31every
- 7:32every 20 minutes,
- 7:34it said yeah, this is in progress and
- 7:36it's moving forward. So, that should
- 7:37update in in a minute as well. So, it
- 7:39just lets let know
- 7:41that this status kind of is up up to
- 7:43date. And again, you could make that
- 7:45work around more
- 7:47more frequently or or so it down
- 7:49depending on how many tokens you want to
- 7:51spend.
- 7:52In terms of like this system, I'm not
- 7:55using Claude for any dev work at all.
- 7:59And I've used up I in fact it has built
- 8:02out some of this system, but it's not
- 8:03really dev work. It's It's It's kind of
- 8:05simple stuff. Mostly just like markdown
- 8:07files and kind of some coordination
- 8:10layers. And I've used up about 70% in
- 8:14I think it was 3 days ago I started
- 8:16building it when when whenever I
- 8:17published that last video. Anyway, so
- 8:19you're basically going to be paying
- 8:20about $200 a month to Claude to just
- 8:24have it as a communication layer,
- 8:26which I'm personally fine with of course
- 8:28for
- 8:29I think a lot a lot of most people they
- 8:32may not be happy with doing it, but for
- 8:35me, you know, you It's like paying a
- 8:37manager to kind of sit on top of the
- 8:38team. And like why would you do that in
- 8:40an organization? It's so that everything
- 8:42doesn't come to you. So that there's a
- 8:44bit of like a bridge between you. So
- 8:46that you have the manager make sure that
- 8:49work is happening. Like this is
- 8:50literally like a micro manager. Every 20
- 8:52minutes it's going and checking all of
- 8:55the active work that I want to be
- 8:57completed and making sure that it's it's
- 8:59running.
- 9:00So then it has hooks. So these are just
- 9:03like kind of little prompts that fire on
- 9:06different events. So turn end means when
- 9:09Claude sends a response to me, it's the
- 9:12end of his turn and it's going to launch
- 9:15like two sub-agents essentially that
- 9:18just go and say
- 9:20you know, if if we message here saying
- 9:22let's pause that task or let's start
- 9:24that task or Fable tells me that task is
- 9:27finished, it's going to read those
- 9:29recent messages and it's going to go and
- 9:31update the task in Notion as being
- 9:33finished. Or if there's nothing to
- 9:34update, it's not not going to do
- 9:36anything. Yeah, it's also going to check
- 9:38are there any blocked items and raise it
- 9:41to the orchestrator. So, when we message
- 9:43now, I showed you there was that blocked
- 9:45task in in Notion, it should get picked
- 9:47up automatically.
- 9:48Let me also just say like this, this has
- 9:51been built in 2 days. I'm getting value
- 9:53out of it already, but it's like a work
- 9:55in progress, but I think I will like
- 9:57open source this um system just for for
- 10:00you to go and like build on top of.
- 10:02It may not even need to be open source,
- 10:04but and you can just like
- 10:06I don't know, copy this video like just
- 10:07give a prompt to an agent to copy it and
- 10:09it would be be very easy to do.
- 10:12Um then also on session start, this
- 10:15happens when you start a new Claude
- 10:18session, but I was asking Claude some
- 10:20other specifics and apparently it does
- 10:22also happen if there's like a long
- 10:24delay, it can also trigger a new session
- 10:26start. I can't remember exactly what
- 10:28what it said, but um
- 10:30in that case, it's going to fetch all of
- 10:33our our priorities
- 10:35that are currently listed in Notion and
- 10:38it's going to bring them locally and
- 10:39then Fable knows that is the work that
- 10:41it needs to push forward and it's also
- 10:44going to check for any work that isn't
- 10:45moving forward and it's going to
- 10:48it's going to highlight that to me. So,
- 10:50anything that is blocked that is moving
- 10:52forward in the prompt, I've told it it
- 10:54needs to fail like very loudly. It it's
- 10:57not acceptable for work not to be moving
- 11:00forward. If it's in that list of
- 11:02priorities, it means and and the like
- 11:05due date is today, it means this needs
- 11:08to be pushed forward. This is a priority
- 11:10for the organization, for the business,
- 11:11this this need needs to get done. Then
- 11:14as I've mentioned as well, there's the
- 11:15watcher, sorry, there's actually two.
- 11:17So, one of them is monitoring as I
- 11:19mentioned already Codex app server.
- 11:22This is checking for block chats for
- 11:24maybe something times out, maybe
- 11:26something goes wrong. I ran out of like
- 11:28hard disk space earlier because I've got
- 11:30so many work trees on on this machine
- 11:34Um and so it like flagged that as an
- 11:36issue. The other thing to say the thing
- 11:40that I will say I do prefer about
- 11:41Anthropic is their mobile app for
- 11:46is their mobile app for Claude because
- 11:48your
- 11:50your Claude code chats are in sync with
- 11:53with the mobile app and whereas that is
- 11:55like an issue on
- 11:57on GPT on the Codex app if you message
- 12:00on the mobile app
- 12:02it's not going to appear
- 12:05in the same chat on the desktop. It just
- 12:08exists but it just doesn't render so
- 12:10it's like a front end thing. It's like a
- 12:12minor thing but it is a little bit you
- 12:15can kind of forget what where exactly
- 12:16you left off. So it's just one thing
- 12:18that I have noticed I do prefer and I
- 12:20find like the thinking to be a bit
- 12:22better and a bit kind of more stable on
- 12:24Claude. So this entire system this
- 12:26entire operation is made for kind of
- 12:29being out and about and I know there's
- 12:31loads of other people working on this
- 12:32like remote control
- 12:34problem as well. I was just playing with
- 12:35another tool earlier today Orca for like
- 12:38remote control orchestration. I've used
- 12:40tmux for it as well. I do like their
- 12:42solution.
- 12:44Again kind of like what I said earlier
- 12:47is I in the previous video is like I'm
- 12:50just trying to keep things as complex as
- 12:52necessary and as simple as possible and
- 12:56it's like a battle that I'm kind of
- 12:57fighting every single day. And with this
- 13:00system
- 13:01there is no external complexity. I'm
- 13:04using Claude natively
- 13:07and then I'm using the Claude app
- 13:08natively. It can send me push
- 13:10notifications on my phone. It does a
- 13:12very good job of that. So if any work
- 13:14gets blocked whilst this is operating on
- 13:16on my local machine if anything goes
- 13:19wrong here with any of these pull
- 13:20requests I get a message in the official
- 13:23in the official uh app and I can just
- 13:26respond to it you know, if I'm out
- 13:27walking, if I'm out for dinner, if I'm
- 13:30you know, drinking coffee, like what
- 13:31whatever I'm doing.
- 13:32The work isn't blocked. This is just
- 13:35acting as like a watcher, as a manager
- 13:37of all of my Codex chats, of all of my
- 13:39Codex sessions. Um
- 13:42or again, whatever implementation agent
- 13:43you want. It's monitoring all of your
- 13:45agents 24/7 throughout the day. And you
- 13:48have one of the smartest models do it
- 13:50doing this for you. projects.md Sorry, I
- 13:53will say as well, actually, I took the
- 13:55original inspiration of this idea
- 13:58from a guy that I saw he was an OpenAI
- 14:02employee, in fact, and he had actually
- 14:05open-sourced the system. It was a little
- 14:07bit different to this, but it it was
- 14:09like the general thesis with was the
- 14:11same. And of course, he's working at
- 14:12OpenAI, so he was pushing to to do this
- 14:14with um using uh 5.6 Soul as Soul as
- 14:19your
- 14:20um chief of staff, whereas I'm instead
- 14:24kind of saying I like 5.6 Soul for
- 14:26implementation, but in terms of like the
- 14:28chief of staff, I don't think it's very
- 14:31good. Uh I I think you're much better
- 14:33off using something like Fable, which is
- 14:35just much nicer to communicate with. Uh
- 14:37the other thing is
- 14:40I think Claude is far, far better at
- 14:44Let's say like sub-agent orchestration.
- 14:47And so, by that, I mean natively within
- 14:50Claude at launching loads of different I
- 14:52mean, GPT has got better at it, but I do
- 14:55prefer I I I just from what I've seen, I
- 14:57do think that that Claude is actually
- 14:59better at that. Uh so, Claude can kind
- 15:01of spin up agents to kind of watch over
- 15:03things, you know, if you've got really
- 15:05high priority stuff, it can just put a
- 15:06high core agent to go and watch um the
- 15:09the implementation and make sure that
- 15:11everything's kind of uh moving forward.
- 15:13But anyway, yeah, projects is just my
- 15:15active GitHub pro Sorry, this is the
- 15:17folder structure as well, just so you
- 15:18can kind of see what I'm in here. So,
- 15:20the only thing that isn't represented
- 15:22here is hooks because it's it's done
- 15:23natively within
- 15:25within Claude.
- 15:27And sorry, just before I forget, I'll
- 15:29also write out the
- 15:32build skill skills directory. Okay. So,
- 15:36yeah, so there's a projects.json,
- 15:38which at the moment I'm working on about
- 15:41five to eight things. I've got two
- 15:43active projects.
- 15:46And then I've got a cup just other side
- 15:48projects that are just things that I
- 15:49like to do for fun, things I find
- 15:51interesting, you know, on a Saturday or
- 15:53Sunday I'll often just like dedicate one
- 15:55or two days a week to explore new ideas,
- 15:57to try out new things. Some of them are
- 15:59just like one one-time projects as well
- 16:01that like just little tools that I kind
- 16:02of build for myself. So, projects.json
- 16:05just contains every project. It It tells
- 16:09Claude if it's local only,
- 16:11as in doesn't need to deploy on GitHub,
- 16:13doesn't need to use work trees and all
- 16:15of that, or is it just a local project.
- 16:19And just kind of some things about like
- 16:20my any kind of rules for working
- 16:22specifically in that project that it
- 16:24should be aware of. If the project is
- 16:26like a big priority for me and and
- 16:28that's kind of pretty much it. So, it's
- 16:30just a very lightweight JSON. And then
- 16:32forward slash projects
- 16:34is a symlink, so it's just a it contains
- 16:37all of the different projects that I'm
- 16:39working on at the moment in subfolders.
- 16:43And then decisions.md
- 16:45I actually changed into a folder, so
- 16:48it's decisions forward slash
- 16:50decisions.md and then each project has
- 16:55its own has its own decision log as
- 16:58well. And the reason that I like to do
- 17:01this is because if I have a preference
- 17:03in doing something or I have
- 17:05let's say I make an architecture
- 17:07decision on like something like no,
- 17:09don't don't build that with airtable
- 17:12because we're in the process of
- 17:13migrating to another database. If I give
- 17:15a
- 17:16a preferred approach of doing things
- 17:18anything like that. The the rule is if
- 17:21it's anything that the agent can't
- 17:23understand from the code as in it's
- 17:25being like verbal communication
- 17:28that isn't in the code or in the doc
- 17:30somewhere, it gets logged into or
- 17:32anything that isn't obvious to the
- 17:33agent.
- 17:36It gets logged in a decisions.md folder.
- 17:39I try I'm trying anyway to keep these
- 17:41like as small and lean as possible just
- 17:45to not blow the
- 17:46the agent's context window every time he
- 17:48opens this this file. It's it's the most
- 17:51basic type of memory that that you can
- 17:53have. I'll probably look at improving it
- 17:55later on.
- 17:56But yeah, it's just like the agents can
- 17:58kind of look over it if they're not sure
- 18:00what to do before coming to me with a
- 18:01question. They can do a quick look over
- 18:03the decisions and just see if we've
- 18:05dealt with with something like this
- 18:07before.
- 18:09Then the next thing again, this is just
- 18:11acting as like an orchestration layer.
- 18:13It does no building itself. Everything
- 18:15gets handed off to app server, but it
- 18:17does do planning and kind of
- 18:19clarification of all types of work. So
- 18:22you can have kind of management skills.
- 18:24So I've got I'm not even sure how many
- 18:27skills I have, but one of the ones that
- 18:28I made is the build skill
- 18:31which I kind of mentioned in the last
- 18:33video. This gets automatically invoked
- 18:36whenever we're going to build something.
- 18:38And it's basically a wrapper that sits
- 18:41on top of the over superpowers skill
- 18:45set. So it it runs brainstorming and
- 18:47then it runs through implementation. The
- 18:49only difference in this skill
- 18:51again, it literally wraps and just says
- 18:53call the brainstorming skill set. When
- 18:56it gets to the
- 18:59the specification part of the
- 19:01superpowers workflow,
- 19:03it goes through an adversarial audit
- 19:07loop that I mentioned last time. So it's
- 19:08going to spin up two sub agents. One of
- 19:10them is going to be Opus 5, and one of
- 19:13them is going to be GP T5.6 Soul.
- 19:16With the Opus 5 agent, I then also get
- 19:18Soul to actually order its work because
- 19:21I
- 19:22I do like Opus 5, but I just find that
- 19:25if if you run its findings through GPT,
- 19:29um GPT will kind of
- 19:31find a load of things that Opus has a
- 19:34tendency to over exaggerate. So, I use
- 19:36Opus as a second opinion, but I I filter
- 19:38it down anyway. So, anyway, one
- 19:41GPT 5.6 Soul and then one
- 19:44uh Opus 5 agent both audit the plan
- 19:49and look for gaps. They look for
- 19:50weaknesses. They look for things that
- 19:52should be clarified, for things that
- 19:53agents missed, for
- 19:55um you know, dangerous behavior, for
- 19:57code that could be like reused, or logic
- 19:59that that
- 20:01um that already exists that doesn't need
- 20:03to be be rebuilt, all types of stuff. It
- 20:05its job is to basically attack the plan,
- 20:07and then it will come back and give the
- 20:09orchestration agent here the um its
- 20:12findings. The orchestration agent will
- 20:14ask me, I'll clarify, and then it will
- 20:16go back, and it will just loop over this
- 20:18adversarial
- 20:19review process. Now,
- 20:21I mentioned this in the previous video,
- 20:23but again, depending on what you're
- 20:24building
- 20:26is going to depend how many loops you
- 20:28you you want to do with this audit
- 20:30process because it can get expensive on
- 20:32tokens, especially if you run 5.6 Soul
- 20:35on on high reasoning. I've blown through
- 20:38three uh pro accounts credit this week
- 20:40because I've been building
- 20:43changes into the actual trading system
- 20:45that that I'm running, and of course,
- 20:47it's if you're you're dealing with real
- 20:49money, I much prefer to do like 10
- 20:51adversarial iterations, but for
- 20:53something small, you may only need to do
- 20:56one, uh and for other stuff, you may do
- 20:58three. So, again, it's going to be like
- 20:59personal kind of preference.
- 21:01And so, yeah, I that is like the build
- 21:03skill. The other thing as well is um
- 21:07what I was finding is that often the
- 21:09agents could default into using
- 21:12superpowers.
- 21:13And on some of the work that I do and so
- 21:16it would try to like build everything
- 21:17and superpowers is in itself a bit of a
- 21:21like I love it. It's my favorite set of
- 21:23skills but it can be a little bit of
- 21:24like overkill or too much of an
- 21:26elaborate workflow for like really small
- 21:28changes. Like I was asked to put in like
- 21:30a button and that makes one change in
- 21:32the database and it went through this
- 21:34whole superpowers workflow and it took
- 21:35about
- 21:36I don't know. It was about 3 hours for
- 21:38something that that
- 21:40really was like it just wasn't
- 21:41necessary. It was complete overkill.
- 21:44So that's one thing. So the build skill
- 21:45as well also audits the task and tries
- 21:49to understand is this something worth
- 21:51using superpowers on or is this just
- 21:53like a really quick fix that can just be
- 21:55validated with like a test or two.
- 21:58And then also on some of the work that
- 22:00I'm doing as well, it can be
- 22:02you know, about auditing stuff and I had
- 22:06cases where it was trying to build a
- 22:07solution before we had even actually
- 22:09audited what what what the actual
- 22:11problem was like investigated what what
- 22:13the problem was. Like it would be more
- 22:14like you know, things going wrong with
- 22:17within the the fund that's in like
- 22:19something isn't operating properly. Go
- 22:21and investigate what's going wrong and
- 22:22then let's actually harden that that
- 22:24system. So this kind of would choose
- 22:28what type of agent does it route to?
- 22:31Does it launch an investigation first?
- 22:33Does it go straight into like a
- 22:34superpowers workflow? Does it just do
- 22:36like a
- 22:37small fix? And so that was another thing
- 22:40there. I've got one or two other skills
- 22:42but honestly off the top of my head I
- 22:44completely forgotten what what they are.
- 22:46Anyway, I'll give you an example of like
- 22:47the system in action. So if we come back
- 22:50into
- 22:51if we come back into the tasks here
- 22:54I don't know why my shortcut's not not
- 22:57working there.
- 22:59Anyway,
- 23:01right. So we've got a few things that
- 23:02are blocked.
- 23:04So, what we should be able to do here is
- 23:05if we come we've got fable on I can
- 23:07literally just say hi. So, I shouldn't
- 23:09need to prompt. Yeah, okay, right. So,
- 23:12without even saying this is what I mean
- 23:14about I just wanted this system to
- 23:15basically be
- 23:17as proactive as possible. So, it's gone
- 23:21and seen
- 23:24that uh
- 23:251554
- 23:27on notion task 1554 which is this needs
- 23:31a user decision. So, it works in
- 23:32priority order by the way.
- 23:35Now,
- 23:36what I I I I didn't like it before kind
- 23:39of overwhelmingly with like a ton of
- 23:41information. So, look now it's looking
- 23:43through like all of the local files.
- 23:46It's going through the projects and the
- 23:48ledger of work which is the local
- 23:50history of things that we've built
- 23:52just to kind of check like where we got
- 23:54on to to to kind of check its notes. Um
- 23:57and
- 23:58again, just
- 24:00the whole thing of basically telling
- 24:03this agent like
- 24:05you do not build anything. You apply all
- 24:09of your intelligence simply into pushing
- 24:11the project forward. Also, sorry. It's
- 24:14kind of very similar to the concept that
- 24:16open claw and Hermes or MS
- 24:20agent were kind of built on top of of
- 24:23being like a proactive personal
- 24:25assistant. So, it's kind of built off of
- 24:26that, but this is more for like a chief
- 24:28of staff.
- 24:30I I've I did previously try a system
- 24:32like this with Hermes agent and just
- 24:34wasn't able to um
- 24:36to to get it to work at this same level.
- 24:38So, again, yeah, it's just going to push
- 24:39the work forward. It's just saying,
- 24:41"Look, this is blocked clean mergeable."
- 24:43The one thing that I really hate is when
- 24:45it sends me these like super long
- 24:47messages. So,
- 24:48uh
- 24:49I'm still trying to like perfect the
- 24:52communication again because
- 24:54I don't like how long is it going to
- 24:55take me to read this many I I don't
- 24:57know. Maybe like 30 seconds or 40
- 24:59seconds, but if it's just more concise,
- 25:02if it's more clear and concise, again,
- 25:04you if if you're dealing with hundreds
- 25:06of these messages per day, if you can
- 25:08shave off like 30 seconds off of
- 25:11hundreds of messages per day or even
- 25:13more than that, again, just on the
- 25:15amount of emotional bandwidth that you
- 25:17need to apply to each message, if it can
- 25:19just Yeah, exactly. Merge PR, check
- 25:21screen.
- 25:22Um
- 25:24Okay, so now he's going to go and tell
- 25:26that agent, again, this will be a Codex
- 25:28agent that's done the work. It's going
- 25:29to send that API request over to back
- 25:32back to the Codex, sorry. So, this one,
- 25:34it's needs user decision, so that should
- 25:36now become unblocked.
- 25:38It it needs to send the message and then
- 25:40the hook runs after he's sent his final
- 25:42message, which needs another agent to
- 25:43run. So, there's going to be like a
- 25:46maybe like a 30-second, maybe like a
- 25:481-minute delay after it sends its final
- 25:51turn, because that's when the hook runs.
- 25:53So, it's not like real time. You could
- 25:56like
- 25:57the reason that I'm also using hooks is
- 25:59because again it uses Haiku. So, number
- 26:01one, you save on context, but
- 26:03that is all all of the kind of um
- 26:07system side of things is happening on
- 26:09the back end without this Fable agent
- 26:11needing to even think about it, because
- 26:13otherwise he can forget about stuff,
- 26:15part of his kind of context, which I I'm
- 26:18like obsessed with preserving these
- 26:21agents' context as as much as possible.
- 26:25is going to be used up on kind of
- 26:27maintaining the system, which is exactly
- 26:28what we don't want. This Fable agent's
- 26:30only job is to work with me on pushing
- 26:34all all all all all all of the the work
- 26:36for And previously, when I was doing
- 26:38this, I would either have
- 26:41about 30
- 26:43terminals open running individual
- 26:46sessions
- 26:48or
- 26:50running natively within the Codex local
- 26:53app. Uh because there have been like a
- 26:56lot of improvements in there recently
- 26:58that I do like and I have have been
- 27:00using, but this is just such is just a
- 27:03much more peaceful way of working, I
- 27:05think is the easiest way of describing
- 27:07it where again it's costing like $200 a
- 27:10month to um
- 27:12just to have this the this level of
- 27:14intelligence that sits in between you
- 27:18as as the user. So, it just reduces the
- 27:22amount of like nonsense that that you
- 27:24have to deal with. It refines down all
- 27:26of the chats and I think that it's a I'm
- 27:29biased because I built it, but I I think
- 27:31that this is like the cleanest um
- 27:35the the cleanest way of working that
- 27:37I've come across for manage for for
- 27:39building a ton of stuff. So, like
- 27:41there's a lot of people working on
- 27:42building out these like software
- 27:44factories
- 27:46at at the moment and I think that this
- 27:48is like a really interesting way of
- 27:50doing it. Again, I'm not the the person
- 27:52that has kind of come up with this idea.
- 27:54I've just taken a a few different things
- 27:56that I I saw this open AI guy the other
- 27:58day who had like a kind of sim- similar
- 28:01thing of having a chief of staff and
- 28:04having one subfolder with all of your
- 28:06projects. And and I thought it was so
- 28:08interesting to come from a guy who works
- 28:10at open AI to say, "Don't initiate Codex
- 28:13within each project like you normally
- 28:15would. Just put him one layer above and
- 28:17he's going to know what what what what
- 28:19to go into." Which I was quite surprised
- 28:21to hear hear them say actually. So, I
- 28:24thought like why not? That does make
- 28:25sense especially for for myself working
- 28:28on multiple projects. Let's
- 28:30Let's give it a try. So, anyway,
- 28:31basically now I just sit here all day
- 28:33just working with like one
- 28:37one chat here. You you could have like
- 28:39multiple going on and like as always the
- 28:43the problem that you you still have is
- 28:45like you know, okay, the problem that
- 28:47you have is like your
- 28:49how much you're willing to pay on AI
- 28:51subscriptions.
- 28:53I personally get tremendous value out of
- 28:56the subscription, so I'm happy to take
- 28:58another subscription. So, my limiting
- 29:00factor in terms of scale
- 29:03it is down to
- 29:04how
- 29:06how many decisions I can make per day.
- 29:09Like the one after that, I guess would
- 29:11be like how much time I have per day,
- 29:13but really it comes down to like how how
- 29:15many
- 29:16good decisions I can make per day before
- 29:19I I get very fatigued, and I'm I'm
- 29:21really quite obsessed with this at the
- 29:25moment. But again, with this system, I
- 29:26can go out on on my phone, and it it's
- 29:29just going to ping me whenever it gets
- 29:30stuck. So, it knows what the priorities
- 29:32are. It knows
- 29:34what to push forward. I've already done
- 29:35the planning work. If there's no plan,
- 29:37it's going to proactively ask me and
- 29:40start making it, and then it's going to
- 29:41commit the plan. And once it's got the
- 29:43plan, it can kind of go go ahead with
- 29:45the with with the jobs. The other reason
- 29:48I built this as well, I don't know if
- 29:50anyone else has experienced this, but
- 29:51when working with Superpowers
- 29:53brainstorming, there was like a really
- 29:55big architecture refactor that I'm still
- 29:58doing that I mentioned already on the
- 30:01on the trading system to move it into
- 30:03being a queue-based system in instead of
- 30:05to of of sending orders at real time. It
- 30:08just kind of makes more sense. Anyway,
- 30:10it's a huge refactor, and it it involved
- 30:13it's about eight different pull
- 30:14requests, and it was it's basically just
- 30:17too much work for Superpowers to handle
- 30:21because it it's it's like a huge
- 30:23project, and so it kept stalling. And
- 30:25so, I wrote that plan
- 30:28about a week ago, and it's still not
- 30:30implemented with
- 30:31these agents. So, that's why originally
- 30:34I because they kept stalling because
- 30:36like they would have to compact their
- 30:37context so much. So, that's why I got
- 30:40this watcher agent in place in the first
- 30:42place, and then I thought, "Okay, this
- 30:44makes complete sense. Let me actually
- 30:45just build a um
- 30:47a system around this." So, yeah, I will
- 30:49upload the I'll put this on GitHub. If
- 30:52it's interesting for you or just to play
- 30:54around with it, by all means, you you
- 30:56can use it for free, or you can even
- 30:57just take a screenshot of this and send
- 30:59it to your agent and kind of customize
- 31:01it.
- 31:02I think it's very good. It's very, very
- 31:04flexible. It's very simple. So, there's
- 31:06kind of really a limit. Like, literally,
- 31:08all we're talking about is send API
- 31:10requests to
- 31:12um to OpenAI agent, have a tasks
- 31:16database. So, again, that could be
- 31:17GitHub, it can be Notion, it can be
- 31:19Linear, what whatever you prefer to use.
- 31:21Move all of your projects in into one
- 31:23subfolder, which is where this Chief of
- 31:25Staff agent works from, and then we're
- 31:28literally just talking about a couple of
- 31:29hooks
- 31:31that basically monitor the work that's
- 31:32going on and make sure that it keeps
- 31:34getting pushed
- 31:36pushed forward. So, um yeah, anyway, I
- 31:38hope this was helpful. Thank you so much
- 31:40for watching. Any questions, let me know
- 31:42in the comments.
About this transcript
This page contains the full transcript of How One Fable Chat Manages 100+ Codex Agents by Nath Aston, generated from the public captions YouTube serves with the video. The transcript has 5,746 words across 858 segments, with the original timestamps preserved so you can click any line to jump to that moment in the embedded player.
What you can do with it
Use the transcript to take notes, quote the speaker, build a study guide, generate a summary with ChatGPT or Claude via the YouTube Summary tool, or export it as a timed subtitle file with YouTube to SRT. You can also re-open it in the transcriber to translate the transcript into 100+ languages.
Free YouTube transcript tool
YouTube2Text is a free YouTube transcript generator — no signup, no daily limit. Paste any YouTube link and get the full transcript instantly, with timestamps, click-to-jump, translation to 100+ languages, AI prompts for ChatGPT, Claude, and Gemini, and exports to TXT, SRT, VTT, or Markdown.