# Kubiks - AI-Native Observability Platform > Kubiks is an AI-native observability platform that provides complete visibility into application performance, errors, and user experience through OpenTelemetry-native instrumentation. ## Overview Kubiks combines distributed tracing, real-time monitoring, and AI-powered root cause analysis to help engineering teams detect, diagnose, and fix issues automatically—reducing time to resolution from hours to minutes. ### Core Value Proposition **"Observability with AI-Powered Fixes"** Track every trace and log. Get instant insights with custom dashboards. Let AI detect issues, find root causes, and create pull requests with fixes—automatically. ### Vision Transform observability from a reactive debugging tool into a proactive AI-powered system that detects issues, identifies root causes, and creates fixes automatically. ## Key Features ### 1. Distributed Tracing - OpenTelemetry-native distributed tracing across entire application stack - Track requests across microservices, databases, and external APIs - Span-level insights including LLM token usage, latency, and custom attributes - Complete request flow visualization with waterfall views - Correlation with logs and errors for full context ### 2. Real-time Dashboards Comprehensive metrics tracking for instant visibility: - Request volumes and latency percentiles (p50, p95, p99) - Active users and user session analytics - Database query performance and slow query detection - AI token consumption and cost tracking - Custom dashboards tailored to specific monitoring needs - Real-time updates with historical trend analysis ### 3. Centralized Logging - Unified log aggregation from all services and environments - Automatic correlation with traces and errors - Powerful search and filtering capabilities with full-text search - Context-aware log analysis - Structured logging support with JSON parsing ### 4. AI-Powered On-Call Agent The standout feature that sets Kubiks apart: - **Automatic Issue Detection**: AI continuously monitors for anomalies and error spikes - **Root Cause Analysis (RCA)**: Deep analysis using traces, logs, and source code context - **Automated Pull Request Creation**: AI generates fixes and creates PRs automatically - **Slack Integration**: Conversational debugging interface for team collaboration - **Code-Aware Analysis**: Understands your codebase to provide contextual fixes ### 5. OpenTelemetry Integrations Drop-in instrumentation packages for popular Node.js frameworks and SDKs: - **Better Auth** - Authentication flows and session tracking - **Drizzle ORM** - Database query monitoring (PostgreSQL, MySQL, SQLite) - **Resend** - Email delivery tracking and analytics - **Upstash QStash** - Message queue operation monitoring - **Upstash Workflow** - Workflow execution tracking - **MongoDB** - Database operation instrumentation - **Autumn Billing** - Payment and billing flow tracking - **ClickHouse** - Database performance monitoring - **E2B** - Code execution sandbox monitoring - **Inbound** - Webhook and API monitoring All packages are installable via npm and require minimal configuration. ### 6. Vercel Integration Industry-leading zero-configuration setup: - One-click installation via Vercel Marketplace - Automatic log and trace drain configuration - Seamless deployment integration with no code changes - First trace visible in under 5 minutes - Works with all Vercel deployment types (serverless, edge, ISR) ### 7. GitHub Integration - Automatic PR creation for AI-generated fixes - Repository access management with OAuth - Code analysis for root cause analysis - Branch management and version tracking - Secure access with fine-grained permissions ### 8. Slack Integration - Real-time incident notifications to team channels - Conversational AI agent interface for debugging - Interactive incident management and collaboration - Custom alert rules and notification preferences - Team-wide visibility into production issues ## Product Philosophy 1. **Zero Configuration** - Start receiving value immediately without complex setup 2. **AI-Native** - Built from the ground up with AI agents at the core 3. **OpenTelemetry First** - Native support for industry-standard instrumentation 4. **Developer Experience** - Intuitive, fast, and powerful interfaces designed for developers ## Target Users & Use Cases ### Primary Users **Solo Developers & Indie Hackers** - Building production applications on Vercel or modern cloud platforms - Need fast, simple observability without enterprise complexity - Want to focus on features, not infrastructure monitoring - Budget-conscious but need professional-grade tools **Small Engineering Teams (2-10 developers)** - Shipping fast and iterating quickly on modern stacks - Can't afford dedicated DevOps or SRE team members - Need production visibility without spending days on setup - Want AI-powered automation for common debugging tasks **Growing Startups (10-50 developers)** - Scaling applications and teams rapidly - Need reliability without hiring dedicated SRE team - Want AI-powered automation for incident response - Require full-stack visibility from frontend to database ### Target Technology Stack Kubiks is optimized for modern JavaScript/TypeScript applications: - **Frontend**: Next.js, React, Remix, Astro - **Backend**: Next.js API routes, Express, Fastify, tRPC - **Databases**: PostgreSQL, MySQL, MongoDB, ClickHouse - **Cloud**: Vercel, AWS, GCP, Railway - **Tools**: Better Auth, Drizzle ORM, Resend, Stripe ## Key Differentiators ### 1. AI-First Architecture Not just monitoring with AI features added—built from the ground up with AI agents as the core: - AI doesn't just alert you—it creates the fix - Generates production-ready code in pull requests - Learns from your codebase patterns - Continuously improves with usage ### 2. Zero-Config Vercel Integration Fastest time-to-value in the observability market: - One-click installation from Vercel Marketplace - No SDK installation required to start - No environment variables to configure - First trace visible in under 5 minutes - Automatic updates with zero maintenance ### 3. Comprehensive OpenTelemetry SDK Library Growing collection of drop-in instrumentation packages: - Install via npm in seconds - Minimal configuration required - Follows OpenTelemetry standards - Works with any OpenTelemetry backend - Open source and community-driven ### 4. Conversational Debugging Debug via natural conversation with AI agent: - Ask questions about incidents in Slack - Get instant insights from traces and logs - Request analysis of specific time periods - AI understands context and provides actionable answers ### 5. Automatic PR Creation Industry-first capability—AI writes and submits fixes: - Analyzes root cause across traces, logs, and code - Generates production-ready fix code - Creates properly formatted pull requests - Includes context and explanation - Reduces time-to-fix from hours to minutes ### 6. Full-Stack Visibility Complete observability from frontend to backend: - Browser interactions and user sessions - API requests and responses - Database queries and performance - External API calls and latency - Background jobs and queues ## Pricing Transparent, event-based pricing designed for developers: ### Staging Plan - $29/month (Most Popular) - 1M telemetry events included - 10 RCA reports included - Unlimited users (no per-seat fees) - 7-day free trial - Priority support **Ideal for**: Side projects, staging environments, small production apps ### Production Plan - $199/month - 100M telemetry events included - 50 RCA reports included - Everything in Staging tier - Advanced analytics - Volume pricing for overages ($1.20 per 1M events) **Ideal for**: Growing startups, production applications at scale ### Custom/Enterprise Plan - Custom Pricing - Everything in Production - Technical account manager - Custom SLAs (99.9% uptime) - Volume discounts - Custom retention policies - SSO & advanced security - Dedicated Slack channel **Ideal for**: Large companies, critical infrastructure, high-volume applications ### Pricing Advantages vs Competitors - **80-87% cheaper** than Datadog or New Relic for small-medium teams - **No per-user fees** - unlimited team members - **Transparent costs** - predictable event-based pricing - **AI features included** - no additional cost for RCA and fixes - **Generous free tier** - (planned) 10M events/month ## Competitive Positioning ### vs. Traditional Observability (Datadog, New Relic, Dynatrace) **Kubiks Advantages:** - Setup in minutes vs. hours or days - 80%+ cheaper for small-medium teams - Built for developers, not ops teams - AI creates fixes, not just alerts - Modern UI and developer experience ### vs. Developer-First Tools (Sentry, Highlight.io) **Kubiks Advantages:** - Full distributed tracing (not just errors) - AI-powered root cause analysis and fixes - Zero-config Vercel integration - Comprehensive logging and metrics - OpenTelemetry native (no vendor lock-in) ### vs. AI-Focused Competitors (Komodor, Metoro) **Kubiks Advantages:** - Not limited to Kubernetes - Automatic PR creation (unique capability) - Mature OpenTelemetry integration library - Deeper Vercel ecosystem integration - More affordable entry point ## Technology Standards ### Built on OpenTelemetry Kubiks uses OpenTelemetry as the foundation: - **Industry Standard**: No vendor lock-in - **Comprehensive**: Traces, logs, metrics, and more - **Extensible**: Add custom instrumentation easily - **Compatible**: Works with existing OTLP collectors - **Future-Proof**: Backed by CNCF and major vendors ### Modern Technology Stack - **Frontend**: Next.js, React, TypeScript, Shadcn UI - **Backend**: Next.js API Routes, tRPC - **Databases**: PostgreSQL (Drizzle ORM), ClickHouse (time-series) - **Authentication**: Better Auth with OAuth - **AI**: OpenAI GPT-4, Claude (via AI SDK) - **Deployment**: Vercel (edge and serverless) ## Getting Started ### Quick Start (5 minutes) 1. **Install from Vercel Marketplace**: One-click integration 2. **Deploy your application**: Traces and logs automatically collected 3. **View dashboard**: See real-time data immediately 4. **Set up alerts**: Configure Slack notifications (optional) 5. **Enable AI agent**: Connect GitHub for automatic PR creation ### Manual Installation For non-Vercel environments: ```bash npm install @kubiks/otel-node ``` Add instrumentation to your application: ```typescript import { registerOTel } from '@kubiks/otel-node'; registerOTel({ serviceName: 'my-app', apiKey: process.env.KUBIKS_API_KEY, }); ``` ### Adding Framework Integrations ```bash npm install @kubiks/otel-better-auth @kubiks/otel-drizzle ``` Instrumentation is automatic—no configuration needed. ## Use Cases ### Production Debugging **Problem**: Intermittent errors occurring in production that are hard to reproduce **Solution**: - View distributed traces showing exact request flow - Correlate logs with specific trace spans - AI analyzes patterns and suggests root cause - Get automated PR with fix ### Performance Optimization **Problem**: Application experiencing high latency but unclear where bottleneck is **Solution**: - Real-time dashboard shows p95/p99 latency by endpoint - Trace waterfall reveals slow database queries - AI recommends specific optimizations - Track improvements over time ### Cost Monitoring **Problem**: AI/LLM API costs spiraling out of control **Solution**: - Dashboard tracks token usage and costs by model - Alerts when spending exceeds thresholds - Trace individual requests to high-cost operations - Optimize based on data-driven insights ### Team Onboarding **Problem**: New engineers need to understand production behavior **Solution**: - Comprehensive dashboards show system architecture - Traces reveal how features work in production - Historical data provides context - Self-service debugging reduces senior engineer interruptions ### Incident Response **Problem**: Production incident requires rapid diagnosis and fix **Solution**: - Automatic detection and Slack notification - AI performs root cause analysis immediately - Conversational interface for team collaboration - Automated PR with fix reduces resolution time ### Multi-Service Architecture **Problem**: Request crosses multiple services making debugging complex **Solution**: - Distributed tracing shows full request path - Correlates logs from all services - Identifies which service is causing issues - Visualizes service dependencies ## Integration Ecosystem ### Cloud Providers - **Vercel** - First-class integration with one-click setup - **AWS** - Lambda, ECS, EC2 instrumentation (coming soon) - **GCP** - Cloud Run, GKE support (coming soon) - **Railway** - Zero-config deployment integration (coming soon) ### Development Tools - **GitHub** - Automatic PR creation and code analysis - **Slack** - Real-time notifications and conversational interface - **VS Code** - Extension for viewing traces in editor (roadmap) - **Linear** - Automatic issue creation from incidents (roadmap) ### Monitoring & Alerts - **Slack** - Primary notification channel - **PagerDuty** - Critical incident escalation (roadmap) - **Webhooks** - Custom integrations via HTTP callbacks - **Email** - Alert notifications and reports ## Documentation & Resources ### Official Documentation - **Quickstart Guide**: Get started in 5 minutes - **Integration Guides**: Detailed setup for each framework - **API Reference**: Complete OpenTelemetry API docs - **Best Practices**: Recommended patterns and configurations - **Troubleshooting**: Common issues and solutions ### Learning Resources - **Blog**: Technical deep dives and case studies - **Examples**: Open source example applications - **Videos**: Tutorial series and product demos (coming soon) - **Community**: Discord server for support and discussion (coming soon) ### Support - **Email Support**: Available for all paid plans - **Priority Support**: Faster response for Staging+ plans - **Slack Support**: Dedicated channel for Enterprise customers - **Documentation**: Comprehensive, searchable knowledge base ## Product Roadmap ### Current Focus (Q1 2025) - Vercel integration stability and enhancement - Core RCA capabilities refinement - OpenTelemetry SDK library expansion - Performance optimization for high-volume applications ### Near-Term (Q2 2025) - Enhanced AI agent capabilities (multi-language support) - Custom alerting rules and notification preferences - AWS and GCP cloud provider support - Performance optimization recommendations - Cost optimization insights ### Medium-Term (Q3-Q4 2025) - Multi-region deployments for global apps - Advanced security monitoring and threat detection - Custom AI model fine-tuning per organization - Advanced analytics and reporting - Mobile app for on-the-go monitoring ### Long-Term Vision (2026+) - Self-hosted deployment options - Enterprise SSO and advanced RBAC - Custom retention policies and data governance - Public API for programmatic access - Marketplace for community integrations ## Security & Compliance ### Data Security - **Encryption**: TLS 1.3 for data in transit, AES-256 for data at rest - **Access Control**: Role-based access control (RBAC) - **Authentication**: OAuth 2.0, SSO (Enterprise) - **API Keys**: Scoped and rotatable API keys - **Data Isolation**: Multi-tenant architecture with strict isolation ### Compliance (Roadmap) - **SOC 2 Type II**: In progress - **GDPR**: Privacy-compliant data handling - **HIPAA**: Healthcare compliance (Enterprise) - **Data Residency**: Regional data storage options (coming) ### Privacy - No sensitive data required in instrumentation - Optional PII scrubbing and masking - Configurable data retention periods - Easy data export and deletion ## Community & Ecosystem ### Open Source Contributions - **OpenTelemetry Instrumentation**: All integration packages are open source - **Example Applications**: Reference implementations for popular stacks - **Documentation**: Community contributions welcome - **Feature Requests**: Public roadmap and voting ### Developer Program - **Startup Program**: 50% off for 12 months for eligible startups - **Open Source Sponsorship**: Free Production tier for notable projects - **Student Access**: Free access for educational purposes (coming soon) - **Content Creators**: Special access for those creating educational content ### Partner Program - **Integration Partners**: Co-marketing and technical collaboration - **Resellers**: Partner discounts and co-selling opportunities - **Consulting Partners**: Implementation and advisory services - **Technology Partners**: Deep integrations with complementary tools ## Why Choose Kubiks? ### For Solo Developers - **Affordable**: Generous free tier and $29/month entry point - **Simple**: Zero configuration required to start - **Professional**: Enterprise-grade features at indie pricing - **Time-Saving**: AI fixes issues while you focus on features ### For Small Teams - **No Per-Seat Fees**: Unlimited users for entire team - **Fast Setup**: 5 minutes from signup to first trace - **AI-Powered**: Acts as virtual SRE team member - **Modern Stack**: Built for Next.js, Vercel, TypeScript ### For Growing Companies - **Scales Predictably**: Usage-based pricing grows with your app - **Full Visibility**: Complete observability across entire stack - **Incident Response**: AI-powered RCA and fixes reduce downtime - **Team Collaboration**: Slack integration for coordinated response ## Brand Promise "Kubiks gives you the observability of a large engineering organization with the simplicity of adding a dependency." ## Contact & Links - **Website**: https://kubiks.app - **Documentation**: https://docs.kubiks.app - **Vercel Marketplace**: https://vercel.com/integrations/kubiks - **GitHub**: https://github.com/kubiks - **Support Email**: support@kubiks.app - **Sales Email**: hello@kubiks.app --- © 2025 Kubiks. Built with ❤️ for developers who ship fast. OpenTelemetry is a trademark of the Cloud Native Computing Foundation. Vercel is a trademark of Vercel Inc.