Observability Overview
Taimoe Enterprise AI Gateway provides comprehensive observability into all LLM API calls, agent reasoning steps, and conversation streams routed through your organization. This suite ensures full transparency for cost control, performance optimization, and issue debugging.
The Observability suite comprises five specialized modules accessible under OBSERVABILITY in the Aegis Console:
Observability Modules
1. Usage & Cost
Macro-level analytics on API request throughput, token consumption, failure rates, and spend attribution sliced across Teams, Agents, Virtual Keys, and Runtimes.
2. Requests
Real-time data plane log stream recording every individual LLM API call, including HTTP status codes, latency, model usage, and calculated cost.
3. Conversations
Full transcript viewer and audit history for multi-turn chat sessions across API, Playground, and Teams channels, with Guardrail blocked-turn indicators and CSV/JSON export options.
4. Traces
OpenTelemetry-compliant distributed tracing for multi-step agent reasoning, tool calls, and vector retrievals, complete with interactive span waterfall timelines.
5. Metrics
Real-time visual dashboards charting request throughput, latency trends, token ratios (Prompt vs Completion), and top active agents.
Core Performance Indicators (KPIs)
Across Observability dashboards, key performance metrics provide immediate operational visibility:
- Total Requests / Traces: Total volume of API calls or agent executions handled by the gateway.
- Fail Rate (%): Percentage of failed operations. Healthy systems maintain this near zero.
- Avg Latency (ms): Average end-to-end processing time in milliseconds.
- Total Tokens & Cost: Aggregate prompt and completion tokens processed and their estimated USD cost.