Orchestrating Your Life with AI Agents

A systems approach, not a coding one. hsv.ai talk, Sept 16, 2026. Daniel Mayo.

Watch on YouTube (57 min). Chapters:

  1. 0:00 How I use AI day to day
  2. 3:10 Four ways to use AI for ongoing projects
  3. 5:05 Where chat apps break when a project outlives a conversation
  4. 11:24 How I got here: coding agents for everyday life
  5. 18:22 The view from 30,000 feet: one folder of workspaces
  6. 24:09 Who does the work, where the record lives, what can be replaced
  7. 37:18 One real workspace, one month: the trip loop
  8. 50:44 Try it: one workspace for one real project
  9. 52:08 Q&A

The slides I showed (PDF), all eight, including the live example from a real workspace (slide 7). Everything below is written notes from the talk and the questions from the room, not the slides themselves.

The short version: give each long-running project in your life one folder of plain text files that you own. Point your AI assistant at the folder. The assistant reads the files at the start of a session and updates them at the end. Assistants and models get swapped. The folders stay.

Repo with templates and setup for Claude Code, Codex, and Gemini CLI (MIT):
github.com/rdmayo21/agentic-workspaces

The problem

A trip. A health question that lasts a year. A side project. The people you owe replies to. My test for any way of using AI on these, three months later: where does this stand, what did I decide, and why?

The same four failures showed up in every tool I tried once a project outlived a conversation:

You re-explain

Every session starts from the model's memory of you, not the project's record.

You can only append

Things change. The project folder just gets one more conversation.

The record isn't yours

It sits in a vendor's app, in a vendor's format, on a vendor's timeline.

The "why" is gone

You can find that you kept the hotel and changed the bus. Not why.

One more constraint: whatever fixes this can't become a job. Keeping the system running has to stay small.

How I got here

StepWhat it wasWhat broke
1. Projects in the chat appsA folder of related conversations with some memory across them.I couldn't update the project as things evolved. I just kept adding conversations.
2. A skill per area of lifePut the context in a skill file, call it when the topic comes up, let the assistant update it.A skill is a standard operating procedure. It was never built to hold a record over months.
3. Now: the record in a folder, the skill a pointerOne folder of plain files per project. The skill shrinks to a few lines: go here, read these first.Cost: a few minutes a session.

What a workspace is

FileWhat it holds
AGENTS.mdHow the assistant should work here
STATUS.mdWhere things stand
DECISIONS.mdWhat was decided, and why, dated
NEXT-ACTIONS.mdWhat needs doing
RESEARCH.mdFindings, with sources and dates
references/The actual source documents

Three kinds of parts, and it helps to keep them straight:

Owned: keep for years

The workspace folders and a private git backup, synced at the end of every session.

Replaceable: pick per task

The agentic harness (Claude Code, Codex, whatever is next), the model, the tools, and the ten-line pointer skill.

You: the only constant

Ask, decide, approve. A few minutes a day.

Getting in takes one command: /trip at the desk, the same thing from the phone. Once there are several workspaces, one more workspace reads all of them every week and ranks what needs me. Think of it as the assistant at the front desk, with the others as the specialists.

Try it: one workspace for one real project

# 1. Clone it as your workspace system. Keep YOUR remote private.
git clone https://github.com/rdmayo21/agentic-workspaces ~/ai-workspaces
cd ~/ai-workspaces
git remote set-url origin <your-private-remote>

# 2. Wire it into your tools
python3 skills/new-ai-workspace/scripts/workspace.py bootstrap

# 3. Create your first workspace
python3 skills/new-ai-workspace/scripts/workspace.py create lisbon-trip \
  --description "Plan and run a one-week trip to Lisbon in May 2027."

# 4. Back it up
python3 skills/new-ai-workspace/scripts/sync.py now

Or just tell your coding agent: "go to that repo and set it up for me." Then open a new session, type /lisbon-trip (Claude Code, Gemini CLI) or $lisbon-trip (Codex), and say what you want to work on. The README has the full quickstart.

Use it for a month. Then tell me what I'm missing.

Questions from the room

Paraphrased from the recording. A few of the better answers came from the audience, and they are marked.

1. Can I tell the AI at logoff to write a memory file, then reload it next time?

Yes, and that instinct is the whole idea. A workspace makes it a habit with a fixed shape: the assistant updates status, decisions, and next actions at the end of the session, and reads them at the start of the next one. You never manage a "memory file" by hand.

2. What if I keep my own file on a separate machine and paste it back in each time?

That works, and it is the fourth approach on the first slide: you own the context. The cost is the manual upkeep. The workspace version is the same ownership with the assistant doing the bookkeeping, inside files you can open and correct any time.

3. Does this involve MCP?

Only at the edges. MCP servers are tools the harness uses to reach email, a calendar, a drive, the web. They sit on the replaceable side. The workspace itself is plain files and needs no MCP at all.

4. You said a skill could hold years of records, then said it broke within months. Which is it?

Both. It worked at first, then broke as the record grew and changed. A skill is a procedure: a fixed way to do something, like checking email. A long project is a record that changes. Stuffing the record into the procedure file is what failed. Now the skill is about ten lines that point at the folder.

5. On the overview slide, are those boxes directories or commands?

Both. Each box is a folder on my machine and a slash command that points at it. One small setup skill creates the two together, so I never wire them by hand.

6. Which agentic harness do I need?

Any of them. I move between Claude Code and Codex depending on which is better that month. The repo sets up pointer skills for Claude Code, Codex, and Gemini CLI. Each has its own skill format, so the setup script writes one for each.

7. If I invoke a workspace mid-conversation, won't my earlier chat bleed into it?

It can. The habit: start a fresh session (/clear or a new chat) before you call a workspace, and keep throwaway chatting outside any workspace. From the audience: a skill can also run in its own subagent with its own context window and hand back only the result, which keeps the main context clean.

8. Is there a way to park a conversation, go talk about something else, and come back?

Not a true stack. /compact shrinks the conversation to its essentials so it costs less context, and a new session gives you a clean slate. The better answer is the workspace itself: if the state is in the files, you can drop the session and pick the project up later without losing anything.

9. Do you use the biggest model for everything?

No. I use the strongest model when creating a workspace or making a real decision. Reading a workspace and telling me where things stand does not need a top model.

10. Does this save tokens?

Probably not, and I don't mind. I use more tokens now and I finish projects I had put off for months or years. From the audience: one session that starts with the full record can replace twenty chats that each start from zero, so some of it balances out.

11. How do you use this from your phone?

You need to reach the harness running on your own machine, because that is where the skills and the folders are. Codex has a remote function in its mobile app. For Claude Code, I had Claude Code build me a small private web app that talks to the session on my desktop. The vendors' own mobile options were the weakest part of the stack when I gave this talk, and I expect that to change.

12. Can you create and use skills from the web or mobile apps?

From the audience: you can, through profile settings and a per-session toggle, and it takes a lot of clicks. On the desktop and CLI harnesses a skill is just a file in a folder, which is why everything here assumes those.

13. If I write my own MCP server, is it still interchangeable?

Yes. MCP is a shared standard, so the same server works across harnesses. Skills are the same story with small format differences per tool.

14. You said it made a phone call for you. What was that?

The trip workspace flagged that my hotel's front desk closes before my flight lands. I asked the assistant to call them. It had no way to, so it found an AI calling service, placed the call, and came back with the answer: they would leave a key. That was an experiment with a tool, approved by me, not something the workspace does on its own. The useful part was the workspace catching the problem weeks ahead.

Honest limits

It costs a few minutes a session to load and update the record. It has broken on me in several ways, and the fixes went back into the templates. The assistant can still record a discussed option as if it were a decision, which is why DECISIONS.md exists as a file you can read and correct. And nothing updates itself from the world: if your flight moves, you have to tell it, unless you have given it a tool that can see the change.

Feedback

Questions, bug reports, and "this didn't work for me" are all welcome in GitHub Discussions on the repo. For a private note, DM me on X: @caveatemptorst1.

Thanks to hsv.ai and CohesionForce for hosting, and to everyone who asked questions.