Honeyhive Ai
HoneyHive is an AI observability and evaluation platform designed for teams building LLM applications. It provides tools for AI evaluation, testing, and observability, enabling engineers, PMs, and domain experts to collaborate within a unified LLMOps platform. HoneyHive helps teams test and evaluate their applications, monitor and debug LLM failures in production, and manage prompts within a collaborative workspace.
AI observability and evaluation platform for LLM applications.
- Systematically measure AI quality with evals.
- Debug and improve agents with traces.
- Monitor cost, latency, and quality at every step.
- Collaborate with your team in UI or code for artifact management.
- Use HoneyHive to test
- debug
- monitor
- and optimize AI agents. Start by integrating the platform with your AI application using OpenTelemetry or REST APIs. Then
- use the platform's features to evaluate AI quality
- debug issues with distributed tracing
- monitor performance metrics
- and manage prompts and datasets collaboratively.

Open-source AI platform for LLM chatbot management, observability, and evaluation.


Rerun is an SDK and visualizer for computer vision and robotics data.


Full-stack cloud observability platform for monitoring infrastructure, logs, and application performance.


Invisible guardrails for Rails consoles with AI data masking and passwordless authentication.


Security gateway for LLMs that detects and masks sensitive data and risky code.


Llog: Collaborative analytics and insights tool for LLM interactions and monitoring.


Open-source autonomous observability tool that instruments code and automatically fixes bugs.


AI-powered dependency management tool for streamlined updates, licenses, and security.


No-code backend platform with AI for easy system logic creation and management.




