It’s undeniable that hallucinations are one of the main sticking points when it comes to the mass rollout of AI technology.
This can be a roadblock for enterprises looking to implement AI, especially when there are constant reports about these hallucinations hitting the headlines.
While some are on the comical side, like when a Virgin Money customer was warned against hateful speech for using the word “virgin,” some are more serious.
Take Cursor, for example. The AI-powered software coding assistant from AI startup Anysphere went viral due to a chatbot hallucination in response to a recent customer service query.
When customers got logged out of their accounts and asked customer support for assistance, an AI chatbot named ‘Sam’ told them that this was "expected behavior” under a new policy — a policy the chatbot had simply invented all on its own.
This led to confusion and distrust among the company’s customers, with some even cancelling their accounts.
While no one is denying the benefits of AI-powered tools, the simple truth is Large Language Models (LLMs), which can be used to develop chatbots, are powerful, but they can also produce these types of frustrating hallucinations if proper guardrails are not put in place.
Speaking to CX Today, Nikola Mrkšić, CEO and Co-Founder of PolyAI, explained how his company has “been able to constrain the behavior of LLMs and use them in places where it makes the most sense to drive customer service conversations.”
The team at PolyAI helps enterprises speak with customers through voice AI agents, and it’s important that these agents not only say, but also do the right things, rather than hallucinate responses or tell you they’ve taken an action when they really haven’t.
Some of the guardrails in place with PolyAI’s agents are powered by retrieval-augmented generation, or RAG, a technique that enables AI agents to cross-reference knowledge from a generative model with a knowledge base.
This ensures that an AI agent checks its generated responses against information the enterprise has confirmed as factual.
In doing so, it prevents inaccurate, irrelevant, and inappropriate responses, and keeps customer conversations within established limits.
Where Enterprises Can Make a Difference Today with AI
In pursuing seamless CX, businesses must evaluate how AI and automation support accuracy, trust, transparency, operational costs, and efficiency.
According to PolyAI’s Mrkšić, enterprises considering where to start implementing AI for CX should consider sophisticated voice AI agents among their first real-world deployments.
Mrkšić said, "AI is such a big and monumental thing, and many people can't resist mounting these large offensives. What they need is a lot of probing attacks on different front lines.




