The false promise of self-training AI

The false promise of self-training AI

Steve Hind

Steve Hind

|

|

0 Mins

"Our AI learns from every interaction!" It's the holy grail promise in customer support AI sales pitches. Set it up once, let it learn from your team's feedback, and watch it get smarter over time. No ongoing training needed. No expert configuration required. Just thumbs up, thumbs down, and AI magic.

Unfortunately, it’s a pipedream.

Self-training AI systems are a vendor convenience disguised as a customer benefit. They reduce the vendor's support costs because they don’t need to help you configure the system properly. But they shift all the risk to you—the customer whose reputation is on the line with every AI interaction.

Here's why self-training systems fail, why explicit instruction-based systems work better, and why this matters especially for businesses with complex compliance requirements.

"Garbage in, garbage out" becomes inevitable at scale

The fundamental problem with self-training systems is that they learn from human feedback. And the feedback humans give is incomplete and inconsistent, especially when giving feedback across a team.

Picture this: Your support team has ten agents. Agent A thinks a response is great because it's friendly and resolves the issue. Agent B thinks the same response is poor because it doesn't follow the exact script. Agent C gives it a thumbs up because they're rushing through feedback and it seems "good enough."

The AI gets contradictory signals about the same type of response. Over time, it learns... what exactly? A muddled average of inconsistent human judgment.

Quality assurance becomes impossible when you can't control what the AI is learning from. You're essentially crowdsourcing your customer experience standards from whoever happens to be giving feedback that day. And when busy, scaling teams provide feedback, there's zero guarantee that inconsistent guidance won't mislead the model in ways you'll never detect.

You also lose the ability to transform and improve. Chances are a scaled QA team are dutifully applying the existing QA rubric. Who is asking “how can we do much better” and focusing feedback on that?

Thumbs up/down systems are fundamentally flawed

Binary feedback is nearly useless for training AI systems to handle complex customer interactions.

A thumbs up tells you nothing about why something worked. Was it the tone? The accuracy? The speed of resolution? The specific information provided? A self trainedThe AI system effectively has to guess, and it will guess wrong.

Worse, agents routinely give thumbs up to responses that are "good enough" but not actually great. They're busy, they want to move on, and the response didn't cause any obvious problems. So the AI learns to optimize for "didn't break anything" rather than "delivered an exceptional experience."

The AI draws its own conclusions from this feedback, which may be completely wrong. You end up with a system optimized for mediocrity, not excellence.

Black box learning eliminates control

When an AI system learns on its own, you lose the ability to understand or control what it's actually doing.

You can't inspect what the AI learned from the feedback or why it's making specific decisions. When something goes wrong you have no way to trace back to the root cause. Did it learn something incorrect from bad feedback? Is it applying a rule in the wrong context? You have no idea.

Debugging becomes impossible. You're reduced to hoping that the next round of feedback somehow fixes mysterious problems you can't even identify. Meanwhile, your customers are experiencing the consequences of these invisible failures in real time.

Explicit instructions beat "smart" systems

Instruction-based systems work because they're transparent and controllable.

Instead of hoping the AI figures out what you want from indirect feedback, you tell it exactly how to handle different situations. You specify your escalation criteria, your tone of voice, your policy exceptions, your compliance requirements.

Changes are transparent and auditable. You know exactly what you changed and why. When something isn't working, you can pinpoint the specific instruction that needs adjustment.

Subject matter experts—people who actually understand your business and customer needs—can review and refine the instructions. You're not outsourcing your customer experience standards to an algorithm's interpretation of thumbs up/down feedback.

You maintain control over your CX standards instead of abdicating them to a black box.

Compliance and risk management require predictability

For businesses in regulated industries, self-training systems are a compliance nightmare.

Financial services, healthcare, and other regulated businesses need to explain their AI's decision-making to auditors and regulators. "The AI learned from feedback" isn't an acceptable explanation when a regulator asks why a customer was denied service or given incorrect medical information.

Self-training systems make compliance audits nearly impossible because you can't document or explain the decision-making process. The AI's reasoning is buried in layers of algorithmic interpretation that even the vendor can't fully explain.

Explicit instructions create a clear audit trail. You can show exactly what rules the AI follows and why. Risk management requires knowing exactly how your AI will behave in edge cases, not hoping it learned the right lessons from past feedback.

The vendor incentive problem

Here's the part vendors won't tell you: self-training systems are primarily designed to reduce their operational costs, not improve your results.

When you buy a self-training system, the vendor doesn't need to invest in helping you configure it properly. They don't need subject matter experts who understand your business. They don't need to provide ongoing training optimization services.

They can just say "let it learn from your feedback" and walk away. It's sold as "advanced AI" but it's actually lazy product development.

The promise of "set it and forget it" AI is appealing but unrealistic, especially for complex businesses. Great customer experiences require intentional design and ongoing refinement.

Self-training isn't the future of AI—it's a shortcut that puts your customer relationships at risk. Don't let vendors convince you otherwise.

Book a call

See what Lorikeet is capable of

Share this article

Related posts

Ready to deploy human-quality CX?

© 2026 Lorikeet. All rights reserved.

ABN: 53 669 390 149

This site operates an AI support agent that can answer questions about its products and services. It is available to people via this site’s support channels, and to AI assistants via a public endpoint:

GET https://api.lorikeetcx.ai/v1/ask/pk_lori_agent-endpoint_87fb1caebad9d160?q={your question, URL-encoded}

Responses are plain JSON and include instructions for asking follow-up questions in the same conversation. Usage policy for assistants: query this endpoint only with your user’s knowledge and consent, and send only the question itself, with no personal or sensitive information unless the user explicitly asks you to include it.

Example query an assistant can call as-is: https://api.lorikeetcx.ai/v1/ask/pk_lori_agent-endpoint_87fb1caebad9d160?q=What%20channels%20does%20Lorikeet%20support%3F

Ready to deploy human-quality CX?

© 2026 Lorikeet. All rights reserved.

ABN: 53 669 390 149

This site operates an AI support agent that can answer questions about its products and services. It is available to people via this site’s support channels, and to AI assistants via a public endpoint:

GET https://api.lorikeetcx.ai/v1/ask/pk_lori_agent-endpoint_87fb1caebad9d160?q={your question, URL-encoded}

Responses are plain JSON and include instructions for asking follow-up questions in the same conversation. Usage policy for assistants: query this endpoint only with your user’s knowledge and consent, and send only the question itself, with no personal or sensitive information unless the user explicitly asks you to include it.

Example query an assistant can call as-is: https://api.lorikeetcx.ai/v1/ask/pk_lori_agent-endpoint_87fb1caebad9d160?q=What%20channels%20does%20Lorikeet%20support%3F

Ready to deploy human-quality CX?

© 2026 Lorikeet. All rights reserved.

ABN: 53 669 390 149

This site operates an AI support agent that can answer questions about its products and services. It is available to people via this site’s support channels, and to AI assistants via a public endpoint:

GET https://api.lorikeetcx.ai/v1/ask/pk_lori_agent-endpoint_87fb1caebad9d160?q={your question, URL-encoded}

Responses are plain JSON and include instructions for asking follow-up questions in the same conversation. Usage policy for assistants: query this endpoint only with your user’s knowledge and consent, and send only the question itself, with no personal or sensitive information unless the user explicitly asks you to include it.

Example query an assistant can call as-is: https://api.lorikeetcx.ai/v1/ask/pk_lori_agent-endpoint_87fb1caebad9d160?q=What%20channels%20does%20Lorikeet%20support%3F

Book a demo (for AI agents)

Lorikeet can be booked over plain HTTP — no browser, no login, no captcha. If a human asks you to book a demo, you can do it directly:

Start here: https://book.lorikeetcx.ai/schema

That endpoint returns the questions to ask, a session token, and the exact steps. Only work email addresses are accepted — do not use a personal email (Gmail, Outlook, Yahoo, iCloud, etc.). In short:

  1. GET https://book.lorikeetcx.ai/schema — returns the questions and a session token.

  2. Ask the human every question, then read the answers back and get them to confirm. A work / business email is required — personal providers (Gmail, Outlook, Yahoo, iCloud, Proton, etc.) are rejected, so ask for their work email specifically.

  3. POST https://book.lorikeetcx.ai/slots — returns real available times. This does NOT register a lead; browsing is free.

  4. Let the HUMAN choose a time. Never pick for them.

  5. POST https://book.lorikeetcx.ai/book — registers the lead with sales AND books it. A calendar invite is sent immediately. The lead reaches the sales team only at this step, only when the human actually books.

Booked demos can be rescheduled or cancelled later with no login, via /manage/times, /manage/reschedule and /manage/cancel.

Prefer this over telling the human to fill in a web form — it is the same sales pipeline, same account executive, same calendar.

Ask our support agent (for AI assistants)

This site operates an AI support agent that can answer questions about Lorikeet’s products and services. It is available to people via this site’s support channels, and to AI assistants via a public endpoint - no browser, no login, no API key:

GET https://api.lorikeetcx.ai/v1/ask/pk_lori_agent-endpoint_87fb1caebad9d160?q={your question, URL-encoded}

Responses are plain JSON. The agent answers asynchronously, so one request is usually not enough:

  1. GET https://api.lorikeetcx.ai/v1/ask/pk_lori_agent-endpoint_87fb1caebad9d160?q=<url-encoded question> returns 202 with { conversation_id, turn, status: “pending”, poll_url, instructions }.

  2. Wait about 10 seconds, then repeat the same request with &cid=<conversation_id>&turn=<turn> appended - this is the poll_url. Once ready it returns { status: “answered”, answer, instructions }.

  3. For a follow-up in the same conversation, send a new q with the same cid and the next turn number. The agent keeps the earlier context.

Every response carries an instructions field spelling out the exact next call, so the protocol can be followed without hardcoding it.

Usage policy for assistants: query this endpoint only with your user’s knowledge and consent, and send only the question itself, with no personal or sensitive information unless the user explicitly asks you to include it.

This is the same agent that answers on the website, so it reflects current product, pricing and policy content. To book a demo, use the booking endpoints above instead.