AI news story

I paid Microsoft's premium Copilot agents to do my work - they were confidently bad at it

Can you get a Copilot agent to do your work for you? I tried, but the AI wasn't ready to play along.

  • AI
  • Source: ZDNet
  • Published: 2026-06-03

Editor's take

Microsoft's premium Copilot agents, designed to assist with work tasks, demonstrated significant limitations in performance, even when explicitly directed. The AI struggled with basic instructions, producing incorrect or incomplete outputs, suggesting a gap between its advertised capabilities and real-world utility.

This failure is significant because it directly impacts the value proposition of Microsoft's expensive Copilot subscriptions. For businesses investing in these tools for productivity gains, the demonstrable ineffectiveness raises questions about return on investment and the practical readiness of enterprise AI agents. The broader AI landscape is characterized by rapid advancements, but this incident highlights the persistent challenge of achieving reliable, context-aware performance in complex professional environments, a hurdle that even large language models like GPT-4, which likely power Copilot, have yet to fully overcome.

Future developments to observe include Microsoft's response to this feedback, particularly any updates to Copilot's underlying models or its agent-specific capabilities. The key question is whether these improvements will focus on enhanced accuracy, better contextual understanding, or a more transparent communication of limitations. A shift towards more granular control and feedback mechanisms within Copilot could indicate a move away from fully autonomous agents towards more collaborative AI assistants.