Close Menu
Technology Mag

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot
    Sky Sports killed off its female-focused Halo brand after just three days

    Sky Sports killed off its female-focused Halo brand after just three days

    November 16, 2025
    The best gifts for dads that have everything (but deserve more)

    The best gifts for dads that have everything (but deserve more)

    November 16, 2025
    The Asus Falcata is an ambitious split ergo gaming keyboard that falls short

    The Asus Falcata is an ambitious split ergo gaming keyboard that falls short

    November 16, 2025
    Facebook X (Twitter) Instagram
    Subscribe
    Technology Mag
    Facebook X (Twitter) Instagram YouTube
    • Home
    • News
    • Business
    • Games
    • Gear
    • Reviews
    • Science
    • Security
    • Trending
    • Press Release
    Technology Mag
    Home » AI Agents Are Terrible Freelance Workers
    Business

    AI Agents Are Terrible Freelance Workers

    News RoomBy News RoomNovember 5, 20253 Mins Read
    Facebook Twitter Pinterest LinkedIn Reddit WhatsApp Email
    AI Agents Are Terrible Freelance Workers

    Even the best artificial intelligence agents are fairly hopeless at online freelance work, according to an experiment that challenges the idea of AI replacing office workers en masse.

    The Remote Labor Index, a new benchmark developed by researchers at data annotation company Scale AI and the Center for AI Safety (CAIS), a nonprofit, measures the ability of frontier AI models to automate economically valuable work.

    The researchers gave several leading AI agents a range of simulated freelance work and found that even the best could perform less than 3 percent of the work, earning $1,810 out of a possible $143,991. The researchers looked at several tools and found the most capable to be Manus from a Chinese startup of the same name, followed by Grok from xAI, Claude from Anthropic, ChatGPT from OpenAI, and Gemini from Google.

    “I should hope this gives much more accurate impressions as to what’s going on with AI capabilities,” says Dan Hendrycks, director of CAIS. He adds that while some agents have improved significantly over the past year or so, that does not mean that this will continue at the same rate.

    Spectacular AI advances have led to speculation about AI soon surpassing human intelligence and replacing vast numbers of workers. In March, Dario Amodei, CEO of Anthropic, suggested that 90 percent of coding work would be automated within a matter of months.

    Previous waves of AI have inspired misplaced predictions about job displacement, for example concerning the imminent replacement of radiologists with AI algorithms.

    The researchers generated a range of freelance tasks through verified Upwork workers. The tasks span a range of work including graphic design, video editing, game development, and administrative chores like scraping data. They combined a description of each job with a directory of files needed to perform the work and an example of a finished project produced by a human.

    Hendrycks says that while AI models have gotten better at coding, math, and logical reasoning in recent years, they still struggle to use different tools and to perform complex tasks that involve numerous steps. “They don’t have long-term memory storage and can’t do continual learning from experiences. They can’t pick up skills on the job like humans,” he says.

    The analysis offers a counterpoint to a benchmark of economic work offered in September by OpenAI called GDPval, which purports to measure economically valuable work. According to GDPval, frontier AI models such as GPT-5 are approaching human abilities on 220 tasks across a range of office jobs. OpenAI did not provide a comment.

    Share. Facebook Twitter Pinterest LinkedIn WhatsApp Reddit Email
    Previous ArticleBaseus’ solar-powered dash cam is down to its best price yet
    Next Article Baseball broke containment during the World Series

    Related Posts

    Meta, Google, and Microsoft Triple Down on AI Spending

    Meta, Google, and Microsoft Triple Down on AI Spending

    November 14, 2025
    Alex Karp Goes to War

    Alex Karp Goes to War

    November 14, 2025
    The AI Data Center Boom Is Warping the US Economy

    The AI Data Center Boom Is Warping the US Economy

    November 14, 2025
    Meet the Chinese Startup Using AI—and a Team of Human Workers—to Train Robots

    Meet the Chinese Startup Using AI—and a Team of Human Workers—to Train Robots

    November 13, 2025
    OpenAI Signs  Billion Deal With Amazon

    OpenAI Signs $38 Billion Deal With Amazon

    November 12, 2025
    TikTok Shop Is Now the Size of eBay

    TikTok Shop Is Now the Size of eBay

    November 10, 2025
    Our Picks
    The best gifts for dads that have everything (but deserve more)

    The best gifts for dads that have everything (but deserve more)

    November 16, 2025
    The Asus Falcata is an ambitious split ergo gaming keyboard that falls short

    The Asus Falcata is an ambitious split ergo gaming keyboard that falls short

    November 16, 2025
    The Mysterious Math Behind the Brazilian Butt Lift

    The Mysterious Math Behind the Brazilian Butt Lift

    November 16, 2025
    Tim Cook could step down as Apple CEO next year

    Tim Cook could step down as Apple CEO next year

    November 15, 2025
    • Facebook
    • Twitter
    • Pinterest
    • Instagram
    • YouTube
    • Vimeo
    Don't Miss
    The Razer Blade 14 Is Still One of the Best Compact Gaming Laptops Games

    The Razer Blade 14 Is Still One of the Best Compact Gaming Laptops

    By News RoomNovember 15, 2025

    The OLED looks great, but one of the benefits of OLED is HDR in gaming,…

    The Steam Machine feels like the TV gaming PC I’ve always wanted

    The Steam Machine feels like the TV gaming PC I’ve always wanted

    November 15, 2025
    Framework’s franken-laptop is back with big chip upgrades and familiar frustrations

    Framework’s franken-laptop is back with big chip upgrades and familiar frustrations

    November 15, 2025
    Pluribus’ third episode throws a bomb into things

    Pluribus’ third episode throws a bomb into things

    November 15, 2025
    Facebook X (Twitter) Instagram Pinterest
    • Privacy Policy
    • Terms of use
    • Advertise
    • Contact
    © 2025 Technology Mag. All Rights Reserved.

    Type above and press Enter to search. Press Esc to cancel.