Elixir

Sample dashboard. Example data — connect your AI traffic to see your own findings.

View your own insights →

Insights

AI traffic audit results for your workspace.

Xybrid maps your AI traffic into workflows, then finds what is expensive, unsafe, or broken.

sample-workspace · production
Workflows discovered 7 6 with findings
Open findings 6
Critical / high 3 1 critical · 2 high
Estimated monthly waste $54.00

Workflows discovered

Workflows explain what keeps happening.

rag-agent 12,400 calls · gpt-4o · $96.40/7d · 94% confidence · inferred
1 critical 1 finding
support-reply-generation 8,240 calls · gpt-4o · $81.20/7d · 91% confidence · callsite
1 high 1 finding
content-moderation 31,050 calls · gpt-4o-mini · $42.70/7d · 88% confidence · inferred
1 high 1 finding
faq-bot 18,400 calls · gpt-4o-mini · $28.90/7d · 90% confidence · fingerprint
1 finding
nightly-report 210 calls · claude-3-5-sonnet · $12.10/7d · 66% confidence · inferred
1 finding
voice-assistant 1,120 calls · whisper-large-v3 · $9.80/7d · 58% confidence · metadata
1 finding
email-classification 24,800 calls · gpt-4o-mini · $18.30/7d · 86% confidence · inferred
no findings

Prioritized findings

Findings explain what should change.

  • critical Security Possible secret sent to AI provider

    rag-agent · 94% confidence

    An API-key-shaped string was detected in 3 prompts forwarded to OpenAI in the last 24 hours.

    → Redact credentials before they reach the provider, or route the workflow through the gateway’s redaction filter.

  • high Cost Expensive model used for a short-output workflow ~$42.80/mo

    support-reply-generation · 88% confidence

    gpt-4o handles ~8.4k calls/day that average 40 output tokens; a smaller model matches quality on this shape.

    → Switch support-reply-generation to gpt-4o-mini — projected ~$42.80/mo saved at current volume.

  • high Quality Refusal rate climbing on a production workflow

    content-moderation · 71% confidence

    12% of moderation calls returned a refusal in the last 7 days, up from 3% the week prior.

    → Review the system prompt — refusals spiked right after the prompt change on the 3rd.

  • medium Cost Duplicate prompts not being cached ~$11.20/mo

    faq-bot · 90% confidence

    26% of faq-bot prompts are exact duplicates within a 5-minute window.

    → Enable prompt caching on faq-bot to cut the redundant spend.

  • medium Latency Workflow has high p95 latency

    nightly-report · 66% confidence

    p95 is 8.9s against a 2.1s median — long-context calls dominate the tail.

    → Cap context to the last 20 messages, or split the summarize step.

  • low Routing Cloud fallback firing more than expected

    voice-assistant · 58% confidence

    9% of calls fell back to cloud last week against a 2% target.

    → Investigate on-device failures for the whisper stage.

Evidence and triage are available once you sign in.