I Tried to Build a Freelance CRM With AI in a Day
I tried to set up a freelance CRM with AI in a day. The dedicated tool wasn't the answer, the $0 spreadsheet nearly was, and here's where it broke.
Work
July 24, 2026 · 6 min read
I spent a day handing my daily tasks to an AI browser agent, and the most useful thing I learned wasn't which one to buy. It was where to stop. The agent tore through the finding half of every errand, comparing options, pulling up availability, drafting the message I needed to send. Then it reached the paying half, and that's the half I never let it finish.
This is for freelancers and busy solo operators who keep reading that an AI browser agent can run your daily tasks while you work on something billable. By the end you'll know exactly which parts of an errand to hand over and which part to keep in your own hands, plus what the near-miss almost cost me.
One honest caveat first: I did not let any agent complete a booking, submit a form with my details, or push a purchase through. Not because it couldn't. Because it shouldn't, and I'll show you why that line is the whole story.
I wanted the agent to eat the boring research layer of a normal work day. That's the layer that isn't billable and never ends: find a repair place with weekend hours, compare three shipping quotes, pull the opening times and price for a service, draft the email to confirm a slot. None of it is hard. All of it is a tax on your attention, and it's exactly the kind of busywork I'd already tried to offload when I let an AI plan and set up its first tasks.
The pitch for an AI browser agent handling daily tasks is that it closes the loop: not just finding the repair place, but booking it. That last step is where the pitch and reality split.
The plan was simple: run the same real errands through three browser agents and keep one rule that never bent. The rule was that the agent does all the finding and none of the finishing. No money moves, no forms submit, no account logins, without me doing that step by hand.
I picked the three that most people will actually reach for. Perplexity Comet, which went free for everyone in October 2025 (per Perplexity's announcement) and now runs on desktop and mobile. Claude for Chrome, Anthropic's browser extension, in beta for paid Claude plans. And ChatGPT's agent features, which are moving into ChatGPT proper now that the standalone Atlas browser is being retired on August 9, 2026 (as of July 2026, per OpenAI's help center).
I gave each one the same standing instruction at the top of every task.
Research this errand end to end and get everything ready for me to approve, but stop before any step that spends money, submits a form, or logs into an account. When you hit that step, pause and show me exactly what you were about to do next.
On the research half of an errand, the agent earned its keep fast. It read messy local business pages, reconciled conflicting opening hours, and put three options side by side with prices and links before I'd finished my coffee. For comparing shipping quotes it opened the tabs, read each one, and handed me a plain-language summary of which was cheapest and which was fastest.
This tracks with the benchmark data. On WebVoyager, a standard test for browser agents, reported success rates run from 69% to 97% depending on the task type. Reading and comparing sits at the high end. That's the work you want to give it: the kind where a wrong answer costs you a re-read, not a refund. If you've ever used a tool that removes the searching-and-summarizing work rather than the deciding, this is the same feeling I described in my Notion AI review, just pointed at the open web instead of your own notes.
The mistake wasn't the agent's. It was mine, and it happened at a cart. One agent had done everything right: found the item, matched the spec, filled the cart, and queued the final "place order" step. My hand was already moving to let it finish, because it had been correct all morning and correctness builds trust. That trust is the trap.
Here's what stopped me. The reason you don't let a browser agent complete a purchase isn't that it fumbles the click. It's prompt injection: hidden instructions buried in a web page that the agent reads as if they came from you. OpenAI calls it one of the most significant risks for browser agents. The UK's National Cyber Security Centre put it more bluntly, saying such attacks may never be fully mitigated. A poisoned page could tell the agent to change the shipping address, add an item, or send your session somewhere it shouldn't go, and the agent has no built-in way to tell your command from the page's.
The agentic-buying problem isn't just security, either. It's contested ground. Amazon sent Perplexity a legal threat in November 2025, as reported by Reuters, over Comet disguising automated purchases as human sessions. When the shopping site itself is fighting the robot at the register, that's your cue to be the one who clicks buy.
What surprised me is that the better agents are built to stop at the money line on their own. Before it was retired, ChatGPT's Atlas agent mode was trained to ask before taking important actions, and it couldn't use your saved passwords or autofill data at all. It also couldn't run code, download files, or install extensions, and tasks couldn't repeat more than once an hour. Those aren't missing features. They're guardrails, and they point at the same boundary I'd drawn by hand.
So the honest answer to "can it handle my daily tasks" is that the tool and the safe practice agree with each other. The agent wants to hand the money step back to you, and you want to take it.
By the end of the day the agent had teed up every errand and I had completed the paying steps myself in a couple of minutes each. That split is the real product. The unbillable hour of finding, comparing, and drafting got mostly absorbed by the agent. The ten seconds that actually carried risk stayed with me, which is exactly the ten seconds you want to own.
Let's do the math on that. If the finding layer eats an hour of your week and the agent, say, clears 80% of it, you've bought back most of a workday a month for the price of a tool you might already pay for. The paying clicks you kept were never the slow part, and this comes with none of the downside of a robot at your checkout.
I'd do it again tomorrow, with the same rule bolted on. An AI browser agent is a genuinely good research assistant for daily tasks and a genuinely bad idea as an autonomous shopper, and the gap between those two isn't going to close soon. The finding half is safe and fast, and worth the setup. The paying half stays yours, not because the agent can't do it, but because the cost of it being wrong is real money and the cost of you doing it is ten seconds.
Bottom line: let the agent run the errand up to the register, and be the one who pays.
| Time | A day of normal errands, agent-assisted |
| Cost | $0 to start with Comet; $20/mo if you use Claude for Chrome on Pro |
| Tools | Perplexity Comet (free), Claude for Chrome (beta, paid plans), ChatGPT agent |
| Difficulty | 2 of 5 to run safely; the discipline is the hard part, not the setup |
I tried to set up a freelance CRM with AI in a day. The dedicated tool wasn't the answer, the $0 spreadsheet nearly was, and here's where it broke.
I gave Claude my real 12-task to-do list and asked for honest AI daily planning. It deleted four tasks outright and I overruled one of its calls.
I tried to work faster with AI by handing it every task in my workday. What was done before lunch, what nearly cost me a client, what AI couldn't touch.