Image

Microsoft AI Researchers Just Discovered Something That’s Going to Make Their Bosses Extremely Mad

AI automation is typically exactly what it sounds like: automating tasks — many of which were previously carried out by humans — in an attempt to boost productivity and efficiency, often in a prelude to laying off workers wholesale.

However, a new yet-to-be-peer-reviewed paper conducted by a group of Microsoft researchers and spotted by IT Pro found that today’s top AI systems remain eyebrow-raisingly weak at real-world workplace tasks. In fact, they often screw them up badly: the team studied frontier models including OpenAI’s GPT 5.4, Anthropic’s Claude Opus 4.6 and Google’s Gemini 3.1 Pro, and found that during complex assignments, those cutting edge bots corrupted an average of 25 percent of the content in documents. (Older models failed even more severely.)

The researchers concluded that, overall, these “models are not ready for delegated workflows in the vast majority of domains” — which is a very striking finding from Microsoft in particular, which has made massive investments in AI and is actively trying to jam the tech into nearly every aspect of its Windows 11 operating system, often with disastrous results. (Curiously, the paper didn’t evaluate the company’s own Copilot AI.)

In other words, the Redmond giant’s researchers had every incentive to find something positive about AI in the workplace, but instead found that blindly trusting LLMs to handle internal documents will almost certainly result in everything from errors to data deletion.

As bosses everywhere push to replace human labor with AI, the Microsoft paper builds on a growing body of scholarship about “workslop“: AI-powered mush that lazy or clueless workers push onto their colleagues, but which ultimately just needs to be fixed by a careful human laborer.

On AI workslop: Companies Are Being Torn Apart by AI “Workslop,” Stanford Research Finds

The post Microsoft AI Researchers Just Discovered Something That’s Going to Make Their Bosses Extremely Mad appeared first on Futurism.

Releated Posts

China Recalls Nearly Three Million Teslas Over Unsafe Door Handles

For many years now, Tesla’s retracting door handles have proven disastrous in emergencies, trapping occupants within the vehicles…

Aug 25, 2026 3 min read

Feds Investigating Wall Street Bro After His AI Hedge Fund Imploded Spectacularly

The spectacular saga around Situational Awareness isn’t over yet. The once mega-hyped AI hedge fund, founded by 24-year-old Leopold…

Aug 25, 2026 3 min read

Even Babies Are Still Way Better at Learning Than AI Models

Compared to the lofty benchmark of human-level intelligence, AI chatbots still have a long way to go. Sure…

Aug 25, 2026 2 min read

Nobody Wants Anthropic’s Best AI Model Anymore Now That There Are Way Cheaper Alternatives

One thing that could put a dent in the hype around Anthropic’s upcoming, multitrillion dollar IPO? People preferring…

Aug 25, 2026 3 min read