Oodle.ai Launches Agent Observability at $10/Million Traces

✍️ OpenClawRadar📅 Published: July 15, 2026🔗 Source
Oodle.ai Launches Agent Observability at $10/Million Traces
Ad

Oodle.ai (by Kiran and Vijay) launched an Agent Observability product priced at $10 per million spans, with sub-second P99 query latency and zero sampling. The service stores traces in S3 using a custom Parquet-like columnar format, queried via AWS Lambda.

Key details from the HN launch:

  • Custom columnar storage engine built over two years for logs, metrics, and traces, now applied to LLM agent traces.
  • Traces can be MBs to GBs; stored in S3 in a proprietary parquet-like format, queried serverlessly with Lambda.
  • Deterministic analysis per span before LLM evals: detects tool failures, retries, loops, abnormal token usage, latency regressions, schema violations, sentiment, and other production signals.
  • Pricing: $10 per million spans. Ingestion $0.30/GB, retention $0.001/GB/mo (example: 1200 GB/mo + 90 days retention = $362/mo).
  • Out-of-the-box insights: Error Recovery Failures, High Duration, Low User Satisfaction, Model Cost Optimization, Caching Inefficiency, Excessive LLM Turns.
  • Authors previously used Langfuse (6x more expensive). Currently processing 3M+ agent traces/day with zero sampling.

How it works: Oodle stores all traces without sampling, analyzes each span deterministically for failure signals, then optionally runs LLM-based evals on flagged traces. Queries across retained data are uniformly fast — no warm/cold tiers.

Ad

Who it's for: Engineering teams shipping AI agents in production who need affordable, full-fidelity tracing without sampling.

📖 Read the full source: HN AI Agents

Ad

👀 See Also

Dual DGX Sparks vs Mac Studio M3 Ultra: Practical Comparison for Running Qwen3.5 397B Locally
Tools

Dual DGX Sparks vs Mac Studio M3 Ultra: Practical Comparison for Running Qwen3.5 397B Locally

A developer compared running Qwen3.5 397B locally on a $10K Mac Studio M3 Ultra 512GB and a $10K dual DGX Spark setup. The Mac Studio achieved 30-40 tok/s with 800 GB/s bandwidth but slow prefill, while the Sparks delivered 27-28 tok/s with faster compute but complex setup.

OpenClawRadar
Session Siphon: Open Source Tool Consolidates AI Coding Agent Conversations
Tools

Session Siphon: Open Source Tool Consolidates AI Coding Agent Conversations

Session Siphon is a free, open source tool that consolidates and indexes conversation history from multiple AI coding agents across different providers and machines. The developer created it using Claude to solve the problem of tracking conversations across different platforms.

OpenClawRadar
Sonicker: Voice Cloning Web App Built with Claude Code in 4 Days
Tools

Sonicker: Voice Cloning Web App Built with Claude Code in 4 Days

Sonicker is a voice cloning web app that requires only 3 seconds of audio input and supports 10 languages. The developer built it solo in 4 days using Claude Code for the entire frontend, API integration, and deployment.

OpenClawRadar
Fine-tuned Qwen3-0.6B model outperforms 120B teacher on structured function calling
Tools

Fine-tuned Qwen3-0.6B model outperforms 120B teacher on structured function calling

Distil Labs published an end-to-end pipeline that fine-tunes a Qwen3-0.6B model to achieve 79.5% exact match on IoT smart home function calling, outperforming a 120B teacher model by 29 points. The pipeline uses production traces to generate synthetic training data without manual annotation.

OpenClawRadar