Benioff Dismisses Microsoft Copilot as 'Clippy 2.0,' Touts Salesforce's Slackbot

Copilot is more like Clippy 2.0
Benioff's comparison of Microsoft's AI assistant to the infamous 1990s Office helper, citing poor accuracy and data security failures.
Mark

So Benioff is calling Copilot "Clippy 2.0." That's a pretty harsh comparison. What's actually wrong with Copilot, according to him?

Mimi

He's saying it doesn't work reliably. Poor accuracy, data security problems—he cites Gartner research showing it leaks sensitive information. And then Microsoft asks customers to build their own custom language models to patch the holes.

Luke

Wait—is that Gartner's official assessment, or is that Benioff's interpretation of Gartner's work? There's a difference. And "doesn't work" is pretty broad. Does he mean it fails on specific tasks, or across the board?

Mimi

He's claiming he hasn't found anyone who's had a transformational experience with it. That's anecdotal, though. He's not citing a survey or systematic study.

Mark

So he's positioning Slackbot as the alternative. What makes Slackbot different?

Mimi

It's built on Anthropic's Claude models and lives inside Slack. Benioff says it can access CRM records, edit data, and deploy AI agents for sales work—all without forcing users to leave their existing communication platform.

Luke

That's a design philosophy difference, not necessarily a performance difference. We don't have independent benchmarks comparing Copilot and Slackbot on the same tasks. Benioff is making a business argument, not a technical one.

Mark

He also ditched ChatGPT for Google's Gemini 3.0 last year. Does that tell us anything about his AI strategy?

Mimi

It suggests he's willing to switch vendors based on performance. He praised Gemini 3.0 for reasoning, speed, image quality, and video processing. He said it felt like a fundamental shift.

Luke

That was his personal experience after two hours. It's not a controlled comparison. But it does show Salesforce isn't betting everything on one model provider—they're evaluating options and building on top of what works best for their use case.

  • Benioff's 'Clippy 2.0' label lands as a precise wound — invoking a universally mocked failure to suggest Microsoft is repeating a decades-old mistake with far higher stakes.
  • The accusations are specific and serious: Gartner-cited data leaks, poor accuracy, and a cost burden that pushes customers to build their own AI models just to compensate for Copilot's shortcomings.
  • Salesforce is not merely criticizing from the sidelines — it has launched a Slackbot powered by Anthropic's Claude, designed to operate inside existing workflows without demanding costly platform migrations.
  • Benioff's track record of switching allegiances — from ChatGPT to Google's Gemini 3.0, now championing Claude — signals a deliberate strategy of model agnosticism, shopping for performance over loyalty.
  • The enterprise AI market is fracturing into competing claims of transformation, with customers caught between vendors' promises and the messy reality of tools that don't yet deliver on them.

In the ongoing contest to define the future of enterprise software, Salesforce CEO Marc Benioff has invoked one of technology's most enduring symbols of failure — the 1990s paperclip assistant Clippy — to argue that Microsoft's Copilot repeats history's mistakes rather than transcending them. His critique, grounded in claims of data insecurity and poor accuracy, is also a declaration of intent: Salesforce is building its own AI ecosystem around Anthropic's Claude models, betting that integration and usability will outlast spectacle. The rivalry illuminates a deeper truth about this moment in enterprise AI — that the gap between promise and performance remains wide, and that trust, not novelty, may ultimately determine who wins.

Marc Benioff has made a pointed habit of targeting Microsoft's Copilot, and his latest critique arrived with historical sharpness: he called it 'Clippy 2.0,' invoking the infamous 1990s paperclip assistant that became a symbol of software that misread what users actually needed. The comparison was deliberate and unflattering, suggesting that despite its enterprise prominence, Copilot has failed to evolve beyond its predecessor's core flaw.

Benioff's case was specific. He cited Gartner research indicating that Copilot leaks sensitive data, leaving customers to absorb the consequences. Accuracy, he argued, is poor. And rather than solving these problems, Microsoft reportedly asks customers to build and maintain their own custom language models — an expensive burden that compounds the original failure. Benioff claimed he had yet to meet anyone who found the experience genuinely transformational.

The criticism, however, is inseparable from strategy. Salesforce is aggressively building its own AI ecosystem, with a new Slackbot powered by Anthropic's Claude models at its center. Benioff frames it as the enterprise AI tool that actually works — capable of accessing customer records, editing data, and deploying agents to drive sales outcomes, all within Slack's existing environment. The pitch is integration over disruption: AI that fits the workflow rather than demanding a new one.

This is consistent with Benioff's broader pattern. Last year he publicly abandoned ChatGPT for Google's Gemini 3.0, praising its reasoning and speed after just two hours of use. The move signaled Salesforce's willingness to remain model-agnostic, chasing performance wherever it leads.

What Benioff's commentary ultimately maps is an enterprise AI landscape still defined by significant gaps between ambition and delivery — and a market where trust, accuracy, and seamless integration may matter far more than the scale of any vendor's announcement.

Marc Benioff, the chief executive of Salesforce, has made a habit of publicly criticizing Microsoft's Copilot, and his latest volley arrived with a pointed historical reference. He called the AI assistant "Clippy 2.0"—a jab at the infamous paperclip-shaped Office helper that became synonymous with unwanted intrusion in the 1990s. The comparison was not flattering. Benioff's message was clear: Microsoft's flagship AI tool, despite its prominence in enterprise software, fails to deliver on its promises.

In a post on social media, Benioff laid out his case with specificity. Copilot, he argued, simply does not work well. The accuracy is poor. Security is compromised—he cited Gartner research indicating that Copilot leaks sensitive data, leaving customers to manage the fallout. To make matters worse, he said, Microsoft then asks those same customers to build their own custom language models to fill the gaps. Benioff claimed he had yet to encounter anyone who experienced a genuinely transformational moment using Copilot or who found value in the costly process of training and retraining custom AI systems. The whole enterprise, in his view, amounted to a modern echo of Clippy's notorious failure to understand what users actually needed.

This is not Benioff's first public critique of Microsoft's AI direction. His comments reflect a pattern of skepticism about Copilot's real-world performance, but they also serve a strategic purpose. Salesforce is making an aggressive move into its own AI ecosystem, and Benioff is positioning the company's alternative as the smarter choice for enterprise customers. The centerpiece of that strategy is a new Slackbot, built on Anthropic's Claude models, designed to operate within Slack—the workplace communication platform Salesforce owns. Benioff describes this tool as the top AI CRM sidekick, framing it as every sales leader's best friend. Unlike Copilot, he suggests, Slackbot actually gets work done. It can tap into customer records, edit data objects, and deploy AI agents to execute sales strategies and generate insights—all without forcing users to abandon their existing workflow.

Benioff's public positioning reflects a broader competitive calculus in enterprise AI. Last year, he abandoned ChatGPT in favor of Google's Gemini 3.0, which he praised effusively for its reasoning capabilities, speed, image quality, and video processing. He described the experience as transformational, claiming that after three years of daily ChatGPT use, two hours with Gemini 3.0 convinced him the technology landscape had fundamentally shifted. That endorsement signaled Salesforce's willingness to shop around for the best underlying AI models rather than lock into any single vendor's ecosystem.

What emerges from Benioff's commentary is a portrait of the current enterprise AI market as fragmented and competitive, with significant performance gaps between offerings. His criticism of Copilot carries weight because it is grounded in specific failures: data leaks, poor accuracy, and the burden placed on customers to fix what the vendor could not. His promotion of Slackbot, meanwhile, emphasizes integration and usability—the idea that AI should work within the tools people already use, not force them to adopt new platforms or spend resources on custom model training. Whether Slackbot lives up to that promise remains to be seen, but Benioff's message to enterprise customers is unmistakable: there are better options than Copilot, and Salesforce has built one.

When you look at how Copilot has been delivered to customers, it's disappointing. It just doesn't work, and it doesn't deliver any level of accuracy.
— Marc Benioff, Salesforce CEO
I've used ChatGPT every day for 3 years. Just spent 2 hours on Gemini 3. I'm not going back. The leap is insane.
— Marc Benioff, on switching to Google Gemini 3.0
Envie de l'histoire complète ? Lire l'original sur Times of India ↗
Nous contacter FAQ