I Tried to Build a Freelance CRM With AI in a Day
I tried to set up a freelance CRM with AI in a day. The dedicated tool wasn't the answer, the $0 spreadsheet nearly was, and here's where it broke.
Work
July 24, 2026 · Updated August 11, 2026 · 7 min read
I spent a day handing my daily tasks to an AI browser agent, and the most useful thing I learned wasn't which one to buy. It was where to stop. The agent tore through the finding half of every errand, comparing options, pulling up availability, drafting the message I needed to send. Then it reached the paying half, and that's the half I never let it finish.
Update, August 11, 2026. One of the three agents in this test no longer has a browser you can download, and what that changes is below.
This is for freelancers and busy solo operators who keep reading that an AI browser agent can run your daily tasks while you work on something billable. By the end you'll know exactly which parts of an errand to hand over and which part to keep in your own hands, plus what the near-miss almost cost me.
One honest caveat first: I did not let any agent complete a booking, submit a form with my details, or push a purchase through. Not because it couldn't. Because it shouldn't, and I'll show you why that line is the whole story.
I wanted the agent to eat the boring research layer of a normal work day. That's the layer that isn't billable and never ends: find a repair place with weekend hours, compare three shipping quotes, pull the opening times and price for a service, draft the email to confirm a slot. None of it is hard. All of it is a tax on your attention, and it's exactly the kind of busywork I'd already tried to offload when I let an AI plan and set up its first tasks.
The pitch for an AI browser agent handling daily tasks is that it closes the loop: not just finding the repair place, but booking it. That last step is where the pitch and reality split.
The plan was simple: run the same real errands through three browser agents and keep one rule that never bent. The rule was that the agent does all the finding and none of the finishing. No money moves, no forms submit, no account logins, without me doing that step by hand.
I picked the three that most people will actually reach for. Perplexity Comet, which went free for everyone in October 2025 (per Perplexity's announcement) and is still listed for Mac, Windows, iOS and Android as of August 11, 2026. Claude in Chrome, Anthropic's browser extension, which its own help center lists in beta for all paid Claude plans on that same date. And ChatGPT's agent features, which still had a browser of their own in Atlas at the time and have since moved into the ChatGPT desktop app (as of August 2026, per OpenAI's help center).
I gave each one the same standing instruction at the top of every task.
Research this errand end to end and get everything ready for me to approve, but stop before any step that spends money, submits a form, or logs into an account. When you hit that step, pause and show me exactly what you were about to do next.
On the research half of an errand, the agent earned its keep fast. It read messy local business pages, reconciled conflicting opening hours, and put three options side by side with prices and links before I'd finished my coffee. For comparing shipping quotes it opened the tabs, read each one, and handed me a plain-language summary of which was cheapest and which was fastest.
This tracks with the benchmark data. On WebVoyager, a standard test for browser agents, reported success rates run from 69% to 97% depending on the task type. Reading and comparing sits at the high end. That's the work you want to give it: the kind where a wrong answer costs you a re-read, not a refund. If you've ever used a tool that removes the searching-and-summarizing work rather than the deciding, this is the same feeling I described in my Notion AI review, just pointed at the open web instead of your own notes.
The mistake wasn't the agent's. It was mine, and it happened at a cart. One agent had done everything right: found the item, matched the spec, filled the cart, and queued the final "place order" step. My hand was already moving to let it finish, because it had been correct all morning and correctness builds trust. That trust is the trap.
Here's what stopped me. The reason you don't let a browser agent complete a purchase isn't that it fumbles the click. It's prompt injection: hidden instructions buried in a web page that the agent reads as if they came from you. OpenAI calls it one of the most significant risks for browser agents. The UK's National Cyber Security Centre put it more bluntly, saying such attacks may never be fully mitigated. A poisoned page could tell the agent to change the shipping address, add an item, or send your session somewhere it shouldn't go, and the agent has no built-in way to tell your command from the page's.
The agentic-buying problem isn't just security, either. It's contested ground. Amazon sent Perplexity a legal threat in November 2025, as reported by Reuters, over Comet disguising automated purchases as human sessions. When the shopping site itself is fighting the robot at the register, that's your cue to be the one who clicks buy.
What surprised me is that the better agents are built to stop at the money line on their own. ChatGPT's Atlas agent mode was trained to ask before taking important actions, per OpenAI's help center, and it couldn't use your saved passwords or autofill data at all. It also couldn't run code, download files, or install extensions, and tasks couldn't repeat more than once an hour. Those aren't missing features. They're guardrails, and they point at the same boundary I'd drawn by hand.
That browser is the one you can't get anymore. OpenAI's help center still writes about it in the future tense: Atlas "is scheduled to stop working on August 9, 2026", and after that date it "may no longer open, browse, or support browser-based agentic workflows". As of August 11 there's no copy left to grab either, since the product page returns a 404 and chatgpt.com/atlas lands on a download page offering the ChatGPT desktop app, the Chrome extension, iOS, Android and ChatGPT Classic. I went through that migration and what's left of the alternatives in what's left after Atlas.
The replacement draws the line harder, and in writing. The ChatGPT desktop app now has two browsers, and the cloud one that runs a task in the background "does not accept credentials, use autofill or password managers, sign in to websites, or complete payments", per OpenAI's help center. If a site requires one of those steps, the task stops, and a CAPTCHA may stop it too.
The built-in browser goes the other way and handles "richer sign-in, autofill, password management, extensions, downloads, and navigation", with OpenAI's own instruction attached: "Enter credentials only in the browser, never in the chat." One browser can't cross the line and the other one waits for you at it.
So the honest answer to "can it handle my daily tasks" is that the tool and the safe practice agree with each other. The agent wants to hand the money step back to you, and you want to take it.
By the end of the day the agent had teed up every errand and I had finished each paying step myself in ten seconds. That split is the real product. The unbillable hour of finding, comparing, and drafting got mostly absorbed by the agent. The clicks that actually carried risk stayed with me, which is exactly the part you want to own.
Let's do the math on that. If the finding layer eats an hour of your week and the agent, say, clears 80% of it, you've bought back most of a workday a month for the price of a tool you might already pay for. The paying clicks you kept were never the slow part, and this comes with none of the downside of a robot at your checkout.
I'd do it again tomorrow, with the same rule bolted on. An AI browser agent is a genuinely good research assistant for daily tasks and a genuinely bad idea as an autonomous shopper, and the gap between those two isn't going to close soon. The finding half is safe and fast, and worth the setup. The paying half stays yours, not because the agent can't do it, but because the cost of it being wrong is real money and the cost of you doing it is ten seconds.
Bottom line: let the agent run the errand up to the register, and be the one who pays.
| Time | A day of normal errands, agent-assisted |
| Cost | $0 to start with Comet; $20/mo if you use Claude in Chrome on Pro |
| Tools | Perplexity Comet (free), Claude in Chrome (beta, paid plans), ChatGPT agent |
| Difficulty | 2 of 5 to run safely; the discipline is the hard part, not the setup |
I tried to set up a freelance CRM with AI in a day. The dedicated tool wasn't the answer, the $0 spreadsheet nearly was, and here's where it broke.
ChatGPT Atlas was scheduled to stop on August 9. As of August 11 the download is gone and OpenAI's page still says 'scheduled'. Your real alternatives.
What AI cannot do in a day is anything you can't check yourself by tonight. I quit a price audit 13 tools in. Here's the line, and how to spot it.