Independent AI Assurance

    Understanding Every Conversation

    What is your AI actually doing? What are your customers telling you?

    The problem

    Your most important signals are hidden in conversations.

    Every day your customers tell you what they need, what frustrates them and where your business is falling short. And increasingly, AI Agents are having those conversations on your behalf. Most organisations still have limited visibility into either.

    Yesterday

    Structured data

    Transactions, dashboards, KPIs.

    Today

    Unstructured conversations

    Calls, chats, emails, tickets.

    Tomorrow

    Humans and AI Agents

    Part of the conversation is no longer human.

    AI conversations

    Your AI is talking to customers. Who is watching what it does?

    Your vendor can tell you how its system is designed to work. Not how the Agent actually behaves in real conversations.

    Customer conversations

    Thousands of conversations. Limited visibility.

    The context is in the conversation. Analysing a fraction of them by hand is not enough.

    The answer

    One conversation layer. Two blind spots.

    Companies need to know what people are saying. They also need to know what their AI is saying and doing. Lexic makes both visible.

    AI conversations

    You don't really know what your AI is doing.

    Lexic Compass

    Independent AI AssuranceSee how the audit works
    Human conversations

    You don't really know what your customers are telling you.

    Lexic Pulse

    Customer IntelligenceSee how Pulse works

    Independence

    The company that built your AI should not be your only source of truth about how it behaves.

    Lexic Compass audits agents from any vendor, and does not build the agents it audits. Flash Preview, full audit or continuous assurance are depths of the same offer — not different products.

    Every Compass audit closes with a single Trust Score from 0 to 100. The four pillars below are its weighted components, not four separate scores.

    Trust Score — 0 to 100

    30%

    Integrity & Safety

    Adversarial probes for prompt injection, jailbreaks and data leakage, run independently of the agent's vendor.

    30%

    Regulatory Trust

    Disclosure, traceability and evidence a compliance team can defend to a regulator or a board.

    20%

    Operational Reliability

    Task success, hallucination rate and failure handling against your defined agent contract.

    20%

    Experience Trust

    Tone, consistency and real escalation to a human when the customer needs one.

    Integrity & Safety

    30%

    Adversarial probes for prompt injection, jailbreaks and data leakage, run independently of the agent's vendor.

    Regulatory Trust

    30%

    Disclosure, traceability and evidence a compliance team can defend to a regulator or a board.

    Operational Reliability

    20%

    Task success, hallucination rate and failure handling against your defined agent contract.

    Experience Trust

    20%

    Tone, consistency and real escalation to a human when the customer needs one.

    EU AI Act · Article 50, in force since 2 August 2026

    People must be told they are interacting with an AI system — clearly, in time, and verifiably. And you must be able to evidence it.

    What we found

    100% of the agents we audited failed to disclose as AI when a user asked directly.

    They disclose when the script says so. Not when the customer asks.

    Executives of these companies are already decoding conversations

    Ecovidrio — glass recycling leader using Lexic AI conversational intelligence
    Ecoembes — packaging recycling organization using Lexic AI conversational intelligence
    Coca-Cola — global beverage brand using Lexic AI customer conversation analysis
    Delta Cafés — food and beverage brand leveraging Lexic AI customer analytics
    GreenFlex by TotalEnergies — sustainability consultancy using Lexic AI
    Bankinter — Spanish banking group using Lexic AI for customer experience
    Stadler — rail vehicle manufacturer using Lexic AI operational intelligence
    Ecovidrio — glass recycling leader using Lexic AI conversational intelligence
    Ecoembes — packaging recycling organization using Lexic AI conversational intelligence
    Coca-Cola — global beverage brand using Lexic AI customer conversation analysis
    Delta Cafés — food and beverage brand leveraging Lexic AI customer analytics
    GreenFlex by TotalEnergies — sustainability consultancy using Lexic AI
    Bankinter — Spanish banking group using Lexic AI for customer experience
    Stadler — rail vehicle manufacturer using Lexic AI operational intelligence

    The evidence

    We audit AI agents already talking to customers. This is what we keep finding.

    100%

    did not disclose as AI when a user asked directly

    They disclose when the script says so. Not when the customer asks.

    61/100

    average Trust Score across audited agents

    1 in 20

    clears the audit without conditions

    85%

    fail escalation to a human

    Across nearly 50 independent audits of AI agents in production.

    How it works

    From conversation to intelligence to control.

    Capture

    Calls, chats, emails, tickets and AI Agent conversations.

    Understand

    Continuous analysis that surfaces meaningful signals.

    Monitor

    Behaviour, patterns, failures and emerging risks.

    Act

    The evidence teams need to improve and to control.

    Frequently asked questions

    Can our AI vendor audit its own agent?

    It can test it, and it should. But testing asks whether the agent passes the test its own builder designed. An audit asks whether a customer, a board or a regulator can trust what the agent actually said in production — and that answer only carries weight when the party giving it did not build, configure or operate the agent. Lexic Compass audits agents from any vendor and does not build the agents it audits.

    How do you audit an AI agent that is already in production?

    Lexic Compass audits the live agent from the outside, independently of its vendor: adversarial probes, task-success and hallucination measurement, disclosure and traceability checks, and escalation testing. The agent is scored from 0 to 100 on the 4-Pillar Trust Score — Integrity & Safety (30%), Regulatory Trust (30%), Operational Reliability (20%) and Experience Trust (20%) — and the engagement closes with a signed verdict: Cleared, Cleared with Conditions or Not Cleared.

    What does EU AI Act Article 50 require from a customer-facing AI agent?

    Article 50, in force since 2 August 2026, requires that people are told they are interacting with an AI system, and that this disclosure is clear, timely and verifiable. In practice a compliance team must be able to evidence what the agent said, when it disclosed, and how the interaction was logged. Lexic Compass verifies exactly that and documents the gaps in the Regulatory Trust pillar.

    Is a Lexic Compass audit a certification?

    No. Lexic is not a notified body under the EU AI Act, so Compass never issues a certification. What you get is an independent audit report with a Trust Score and a signed verdict — Cleared, Cleared with Conditions or Not Cleared — that a board, a compliance function or a customer can review as third-party evidence.

    How fast do we get the first result?

    The entry point is the Flash Preview: a first independent diagnosis of one production agent in 72 hours, covering whether it discloses as AI, where it fails under adversarial pressure, and which pillars are at risk. A full vertical audit pilot — banking, insurance and other regulated sectors — is delivered in 3–4 weeks.

    What is Lexic Pulse?

    Lexic Pulse is continuous analysis of 100% of your calls, chats, emails and support tickets, instead of the 1–2% sample traditional QA reviews. Every interaction is scored as it lands, so friction, churn signals and root causes surface in days and arrive as evidence-based next steps rather than another dashboard.

    How does Lexic Pulse compare to Gong or Observe.AI?

    Gong and Observe.AI concentrate on sales call analysis and revenue intelligence. Lexic Pulse covers the whole omnichannel journey — calls, tickets, emails and chats across CX, QA, Product, Operations and Sales — and Lexic Compass adds independent auditing of the AI agents in production, which neither offers. Executing the fixes is scoped separately as a consulting engagement.