HoneyHive is an AI observability and evaluation platform designed for teams building LLM applications. It provides tools for AI evaluation, testing, and observability, enabling engineers, PMs, and domain experts to collaborate within a unified LLMOps platform. HoneyHive helps teams test and evaluate their applications, monitor and debug LLM failures in production, and manage prompts within a collaborative workspace.
AI observability and evaluation platform for LLM applications.
- Systematically measure AI quality with evals.
- Debug and improve agents with traces.
- Monitor cost, latency, and quality at every step.
- Collaborate with your team in UI or code for artifact management.
- Use HoneyHive to test
- debug
- monitor
- and optimize AI agents. Start by integrating the platform with your AI application using OpenTelemetry or REST APIs. Then
- use the platform's features to evaluate AI quality
- debug issues with distributed tracing
- monitor performance metrics
- and manage prompts and datasets collaboratively.

Rerun is an SDK and visualizer for computer vision and robotics data.


Gen-AI powered AIOps platform for advanced operations management and predictive analytics.


Log4U helps developers log work with AI for interview prep.


Llog: Collaborative analytics and insights tool for LLM interactions and monitoring.


Security gateway for LLMs that detects and masks sensitive data and risky code.


Full-stack cloud observability platform for monitoring infrastructure, logs, and application performance.


AI-powered log assistant for real-time log interaction and analysis.


AI-powered log management and analytics for mobile developers.


AI-powered meal logging and calorie tracking iPhone app.







