EP2 · Starter kit · free

ComfyUI with no GPU: Meta's AI agent rented one for me

Episode 1 used a terminal and Claude Code on my own laptop. For episode 2 I gave the same job to Meta's Muse, which runs commands on its own cloud computer, so there was no terminal on my side. It made the picture below on a rented Modal L4 GPU. This is the whole run, written down, with the real numbers.

  • ≈95 sexecution, Muse-reported
  • ≈3 minsend → image
  • $0.024usage it reported
A vintage film camera on a wooden desk by a window, generated with Krea 2 Turbo

Krea 2 Turbo via ComfyUI · made by Muse · seed 14899. The small ring text around the lens is garbled — the agent flagged that itself.

Watch the episode

The whole run, start to finish: the one paragraph I pasted, the walls Muse hit on its own cloud computer, the image it made, and the bill at the end.

Six minutes forty seconds. There was no terminal on my side at any point.

What I actually did

One prompt, one image, and an agent cleaning up its own mess. The short version first; the catches are below.

I gave Muse one paragraph. It downloaded the free kit, unzipped it, read the setup notes, created its own Python environment, worked around a proxy, started ComfyUI with the Krea 2 Turbo model on a rented GPU, made one image, shut the container down, and then went and read its own billing report. On the cached rerun of 2026-10-03, the execution stage took about 95 seconds by Muse's own report, and about 3 minutes passed from the moment I sent the prompt to the image landing. Muse's settled usage figure was about $0.024.

I am not going to call that "free". Modal required a payment card on file before any GPU would start — I tested that on a brand-new account and it refused. What my card was charged for that workspace that month was $0.00, because credits covered the usage, but that is a credit being applied, not a promise about your account. Extra usage can cost money.

WHAT CHANGED VS EP1

No terminal of your own

In episode 1 the agent ran in a terminal on my Windows laptop. Muse runs on Meta's cloud Linux computer, so nothing of it touches your machine.

WHAT DIDN'T CHANGE

The same walls

A blocked Python install, a network proxy, and no browser to log in with. Different agent, different error messages, same category of problem.

WHAT YOU STILL DO

Approve the login

Both agents needed me for exactly one thing: approving the login in my own browser. Everything else they handled.

What you need before you start

Three things, and one of them is the reason this is not a "free GPU" video.

  • Access to Muse, and a plan that allows it. Muse is Meta's personal agent. For availability and plans, go by Meta's own announcement: at the time of writing it names the United States and Canada, and what your own account actually gets is down to your account. My run used Muse's free usage; Muse also has paid plans. I can't promise yours will match.
  • A Modal account, with a payment card on file. This is the part people dislike, so it goes first. GPU access on Modal needs a card even to use the monthly credit — I re-tested this on 2026-10-03 with a fresh account and no card, and the run stopped with Modal's own message: "Please add a payment method to use L4 GPU functions." No GPU ran, nothing was downloaded, nothing was billed. That is Modal's rule, not the agent's.
  • The free kit. Same kit as episode 1, and it now carries the fixes this run needed: the environment, the proxy, and the permission notes. You can get it further down this page.

Not affiliated with or sponsored by Meta, Modal or Krea. I use my own accounts and pay Modal directly.

What it cost, in the two figures that matter

There are two different numbers here, from two different sources, and mixing them up is the easiest way to be wrong about this.

  • Execution: about 95 seconds. Muse reported that stage itself on the cached rerun of 2026-10-03 (its own stage timings: ComfyUI ready in about 25 s, image about 50 s, container about 75 s). That is the run being described here, not a full session.
  • Send to image: about 3 minutes. From the moment I sent the prompt (10:19 PM) to the image landing (10:22 PM).
  • Usage reported by Muse: about $0.024 (its settled figure, reported 10:23 PM: L4 about $0.0173, CPU about $0.0056, memory about $0.0012 — the three add up to about $0.0241). This is a usage figure on the Modal account, not an amount charged to a card.
  • Charged to my card for that workspace: $0.00 that month. That came from the Modal billing overview page (total spend this month, credits applied) — a different source from the figure above. The two are separate and should never be added together.
  • The very first full session: about 13 minutes (2026-09-30, first recorded frame to finished image). Most of that was me approving things, not compute. That was a cold run; the 95 seconds above is a cached rerun.
  • An L4 GPU listed at about $0.80 an hour when I checked on 2026-09-30 — the GPU only. CPU, memory and storage are billed separately, and prices change.

Here is the discipline that produced those numbers. Muse did not quote a price straight away — it said billing hadn't settled, went away, checked again later, and came back with the total and a breakdown. If your agent quotes you a cost in the first minute, it is reading a running estimate, and that estimate moves. Read the settled figure, and say which source you are quoting.

And read this one twice: a monthly credit is an allowance, not a free tier. It applies because a card is on file, it resets, it can change, and it runs out. On my account the usage limit was one I had set myself ($30). I don't know what a brand-new account starts with, so open your own billing page and read it before you run anything.

The spending controls, before anything else runs

If you take one action from this page, take this one — before you hand an agent anything.

  1. Open the billing page and read two settingsModal's usage and billing page shows a usage limit and a separate, harder spend limit. They are two different controls. On my page the usage limit was set to $30 and the page said it could be raised as high as $42.50 — that $42.50 is the ceiling the page offers, not the value your account starts at.
  2. Decide both numbers yourselfI set the usage limit to $30 to match the monthly credit, and left the spend limit at the account's own default rather than inventing a value. I haven't tested exactly what each control stops, so read Modal's page and choose yours deliberately — do not assume either starts at zero.
  3. Log in once, in your own browserMuse cannot log in for you and should not try. It hands you a link; you open it and approve. My prompt also forbids printing any token and touching payment or account settings. Read the prompt before you send it, not after.

I'm not claiming any setting guarantees a zero bill. Credits covered my test, and extra usage can cost money.

The whole prompt, as sent

This is the entire instruction. There is no follow-up, no numbered step list and no configuration after it.

Download the starter kit from this link: [kit link]
Unzip it, then read SKILL.md and SETUP-CHECKLIST.md.
Use Modal to run ComfyUI with the Krea 2 Turbo model on a GPU and make one image:
"a vintage film camera on a wooden desk by a window, soft morning light, real photograph, 85mm lens".
Rules: wait for my approval before any login or network permission; never print a token; do not touch payment or account settings.
When you're done, tell me what it cost after Modal's billing has settled.

Three lines in there are doing deliberate work, and they are worth stealing for any agent you hand a cloud account to:

RULE 1

Ask before the login

"Wait for my approval before any login or network permission" keeps the one genuinely sensitive step — granting access to your account — in your hands, not the agent's.

RULE 2

Never print a token

A cloud login ends in a token. A transcript with a token in it is a leaked token. The ban is cheap to state and worth stating.

RULE 3

Cost after it settles

"Tell me what it cost after Modal's billing has settled" is the difference between quoting a moving estimate and quoting a bill. Episode 1's agent got this wrong.

Do it yourself, step by step

Five steps. Only the first two and the last one are yours.

  1. Get Muse access and confirm your plan allows this.Go by Meta's announcement for availability and plans, and check your own account rather than assuming. Muse is the whole reason there is no terminal on your side.
  2. Make the Modal account, add a card, set the limits.A card on file is required for GPU access; then open the billing page and set the usage limit yourself. This is the only step with money in it, which is why it comes before the fun part.
  3. Download the kit and give Muse the prompt.The kit is further down this page. Paste the prompt above, with the kit link in the first line. Then step back: Muse works through it on its own screen, and you can watch its task list update.
  4. Answer the login and permission questions when they appear.You approve the login in your own browser. On the first session I also got Muse's network-permission prompts, one after another, and I kept choosing "allow once" — it asked six times or more. There is a note in the kit about when "always allow" is reasonable (only for the addresses you have actually checked).
  5. Check the picture, then the bill.Look at the image with a critical eye — mine had garbled text on the lens, which the agent flagged before I did. Then wait a few minutes and read the settled usage on Modal's page, not the first figure you see.

If your agent runs in a terminal on your own machine instead, the episode 1 route still applies, and the same kit covers it — including the Windows encoding setting, which belongs to that route only, not to Muse.

What Muse ran into

These are specific to an agent working on a borrowed cloud computer rather than your laptop. All of them are in the kit so the next agent to try this can skip them.

ENVIRONMENT

Python refused the install

The system Python on that machine is managed by the OS and will not accept packages — the same "externally-managed-environment" wall you hit on a modern Linux desktop. Muse worked around it by building its own environment instead of forcing the system one.

NETWORK

A proxy in the way

The cloud computer sits behind a proxy, so the cloud tool could not reach its own API. The fix is a specific extra package for the client. You would never guess this from your own laptop, where there is no proxy.

LOGIN

No browser to log in with

The login flow normally opens a browser. A cloud computer has none, so Muse tried a handful of different ways to surface a link before one worked — and I could see every attempt in the chat, which is exactly what you want.

PERMISSIONS

The same prompt, six times

Each new attempt at reaching Modal's servers triggered the same network-permission question. Choosing "allow once" means the next attempt asks again. That is not a bug; it is the agent being sandboxed, and it is why the run takes longer than the compute does.

Which agents this was actually tested on

Only these three, and only as listed. Everything else is untested, and I would rather say so than guess.

  • Claude Code — episode 1, Windows terminal, 2026-09-30. It set up ComfyUI and Krea 2 on a rented GPU from one paragraph, and hit a Windows encoding crash and a GitHub rate limit on the way.
  • Meta Muse — its own cloud computer, 2026-10-03. The run this page is about: environment, proxy and login walls, described above.
  • Codex CLI — 2026-10-04, a cached rerun. Already logged in to Modal with the model weights cached, it made an image in about 2 minutes 33 seconds. Its settled usage for that run was about $0.0251 — a usage figure, not a card charge, and not something to compare agents by.

Not tested: Codex's desktop client and its approval prompts, a first-time Modal login from scratch on any of these, a cold start with no cached weights, the path where the account has no card on file (beyond confirming it stops), and every other agent. The kit says "works with an agent that can read files and run commands", because that is the actual requirement — it does not claim every agent will.

The free kit

Same kit as episode 1, updated with what this run needed. Version 2.

  • SETUP-CHECKLIST.mdThe 5 human steps, including the $30 cap
  • AGENT-PROMPT.mdThe exact paragraph + a shorter one for the kit
  • SKILL.mdSkill file for Claude and other agents
  • krea2_comfy.pyModal + ComfyUI + Krea 2 Turbo script
  • prompts_batch.jsonThe 14 prompts from the batch run
  • LICENSE-NOTES.mdWhat you can do with the images

Plain text and Python only. No installers, nothing to run blind.

Where these numbers come from

So you can check them rather than take my word for it. All from 2026-10-03 unless stated.

  • ≈95 s — Muse's own stage timings on the cached rerun (ComfyUI started in about 25 s, image about 50 s, container about 75 s). It is the execution stage, not the whole session.
  • ≈3 min — the screen timeline of that same run, from the prompt going out at 10:19 PM to the image arriving at 10:22 PM.
  • $0.024 — Muse's settled report at 10:23 PM, with a breakdown by GPU, CPU and memory.
  • $0.00 — the total spend shown on Modal's own billing overview for that workspace that month, with credits applied. Read off the page, not reported by the agent.
  • ≈13 min — the first full session, 2026-09-30, from the first recorded frame to the finished image; mostly approval waits.
  • $0.0251 — Codex CLI's settled usage for its 2026-10-04 cached rerun, read from Modal's billing for that run. Not a card charge.
  • $42.50 — the ceiling the Modal usage page says the limit can be raised to, read on 2026-10-03. What a new account starts with: not checked.

Every run is different. Prompts, model, resolution, hardware and how long the container idles all move the total. Treat these dated figures as what happened on those days, not as what you should expect to pay.

Want the version with a terminal?

Episode 1 is the same job driven by Claude Code on a Windows laptop, with the spending cap written out step by step. Read the first episode page.