Ayman Zahran

AI and freelance work

AI agents on freelance platforms

Ayman Zahran · 28 September 2026 · about 9 minutes

I am in the middle of a Doctor of Business Administration at SSBM Geneva. The dissertation asks a narrow question: when freelance platforms move from human freelancers toward autonomous AI agents, how should anyone measure efficiency, return, and the change in the business. This essay is not that paper. There is no survey result here, and I will not pretend a framework is validated before the research is. It is the practitioner version of the question, written from the side of someone who sells skilled work and who also ships AI systems.

The platforms I have in mind are the ones a client already knows: Upwork, Fiverr, Toptal, and the specialist marketplaces around them. They are not job boards with a search box. They are markets for trust. A profile, a review history, a dispute process, and a payment rail sit between a stranger with a budget and a stranger with a skill. Any story about “AI agents replacing freelancers” that ignores that machinery is a demo, not an analysis.

The supply side changes first

Agents show up on the supply side before they show up as a line item in a platform’s annual report. A freelancer uses a model to draft a proposal, to produce a first version of a deliverable, to translate a brief, or to watch a repo and open a pull request. A buyer pastes a ticket into a chat window and expects a finished artifact by morning. Neither of those behaviors requires the platform to “add AI”. They only require that the platform’s rules still assume a person is the one being paid and the one being reviewed.

That assumption is the crack. If the review is of the person, and the work was substantially done by a system the person does not control, the reputation graph is about the wrong object. If the platform then offers its own agent as a cheaper supplier, it is competing with the humans who made the marketplace valuable, using the data those humans produced. Both moves can be legitimate businesses. Neither is measured well by “number of tasks automated”.

What I refuse to count

Hours saved is a slippery number. It is easy to claim and hard to audit, and it treats the freelance product as typing speed. Clients do not buy typing. They buy a result they can accept, on a date, with someone to talk to when the result is wrong. An agent that produces ten drafts nobody accepts has not saved ten drafts’ worth of time. It has added review load.

“Jobs automated” is worse. A job on these platforms is a bundle: understanding the brief, doing the work, handling the revision, showing up when the scope moves, and carrying the reputation if the delivery fails. An agent can take the middle of that bundle. Counting the middle as the whole job flatters the demo and surprises the client at acceptance time.

Generic digital-transformation scorecards have the same hole. They were built for firms adopting software inside a hierarchy. A freelance platform is a multi-sided market. Efficiency for the buyer can be a worse marketplace for the seller, and a higher take rate for the platform can look like ROI while the supply of trusted humans thins out. The dissertation exists because I do not think those generic models survive contact with this market. I do not yet have the field evidence to say what replaces them. I do have a list of measures I am willing to defend as a practitioner.

Measures I will actually use

Time to a first draft that a human is willing to send. Not time to any draft. Revision count before the client accepts. Acceptance rate, not generation rate. Cost per accepted deliverable, including model spend, human review, and the rework when the agent was confident and wrong. Incident rate: how often a delivery has to be recalled, refunded, or explained because the system invented a fact, a credential, or a scope change.

On the platform side I would want the same honesty. What share of proposals in a category are now indistinguishable, and does that change who gets hired? What happens to dispute rates when the supplier is an agent the buyer cannot cross-examine? What happens to repeat hire, which is the only reputation signal I trust more than a star average? If a platform cannot show those numbers, its AI announcement is a feature launch, not an operating result.

What I have shipped, and what I have not

I build with these models in production, on systems I own. merge.news is a daily briefing across cloud, DevOps, and AI: scheduled generation, a multi-provider router across Claude, OpenAI, and Gemini, and a web and mobile client. outbox is a scheduler for X and LinkedIn with AI-assisted drafts and a human still deciding what goes out. I use the Model Context Protocol to give agents tools, and I use agents to draft infrastructure. None of that is a freelance-platform agent that bids, delivers, and cashes out under my name while I sleep.

The boundary is accountability. I will let a model draft. I will not let a model be the party a client argues with. On a marketplace, the profile is a promise that a specific person stands behind the work. An agent can be a tool on that profile. The moment the agent is the profile, the platform owes the buyer a different contract: who pays when it is wrong, who can be banned, and whether the buyer was told.

A hybrid that does not insult either side

The arrangement I think survives is dull. The agent does the repetitive production: first drafts, transformations, watching a pipeline, proposing a patch. The human owns the relationship, the acceptance criteria, and the liability. The platform’s policy says so in language a buyer can find, not in a footnote. Disclosure is part of the product. A client who wanted a person and received an undisclosed model has a complaint even if the file looks fine.

There is a real efficiency story inside that hybrid. I feel it in my own practice when a draft that used to take a morning takes an hour of editing. The return shows up only after the editing, the rejection of the bad third of the draft, and the decision not to send the part I cannot defend. ROI that ignores that hour is marketing.

If you are a platform operator, the useful question is not whether agents are allowed. They are already in the market. The question is which object you are rating, what you tell the buyer, and which number would embarrass you if it moved the wrong way. If you are hiring a freelancer for platform, SRE, or AI work and you want a person who will say what the model did, start at ah.zahran@outlook.com. I wrote about the engineering standard I hold the human side to in Everything as Code.