← Back to all projects

Helicone

Logs every model call by sitting in the request path, so nothing in your code has to be instrumented

Apache-2.0
Stars
6.1k
Forks
661
Open issues
155
Last commit
26 Aug 2026

What is Helicone?

Most tracing tools ask you to wrap your calls, which is fine until the code you need to see is inside a library, a background job or a service somebody else owns. Helicone takes the other route: point the base URL at it and every request is logged on the way through, with no SDK in the path. Because it is already in that position it does gateway work too — one key for many providers, automatic fallback when one is down, caching for repeated prompts. The trade is the position itself: a proxy is now between your application and the model, which is a component to run and a hop to account for. Agent runs group into sessions so a multi-step trace reads as one thing rather than thirty unrelated calls.

What can you do with Helicone?

  • Log without changing the calling code — Changing the base URL is the whole integration, so calls made deep inside a library or another team's service are captured too.
  • Read a multi-step run as one thing — Related calls group into a session, so an agent that made thirty requests to finish a task appears as one trace instead of thirty rows.
  • Keep working when a provider does not — Because it is already in the request path, it can route across providers and fall back automatically when one starts failing.
  • Stop paying twice for the same prompt — Repeated requests can be served from a cache, which matters in agent loops where the same context is resent on every turn.
  • Turn logged traffic into a dataset — Recorded requests can be collected for evaluation or fine-tuning, so real usage becomes the test set rather than invented examples.
  • Run it yourself if the data cannot leave — The platform is open source and self-hostable, which is the deciding factor when prompts contain material that must stay inside.

Before you choose Helicone

  • A proxy in the request path is a component you now operate and a hop in every call — the reason to accept that is coverage you cannot get by instrumenting code you do not control.
  • Self-hosting the full platform is a heavier proposition than adding a tracing library: it is a web application with its own datastores, not a package you import.

Frequently asked questions

Is Helicone free for commercial use?

Helicone is released under the Apache-2.0 licence — OSI-approved open source, which permits commercial use.

How can Helicone be deployed?

Helicone is available as Self-hosted / Managed cloud / Runs locally.

Documentation

Reproduced from the Helicone/helicone README, published under Apache-2.0. Read the original ↗

🔍 Observability🕸️ Agent Tracing🚂 LLM Routing
💰 Cost & Latency Tracking📚 Datasets & Fine-tuning🎛️ Automatic Fallbacks

Helicone is an AI Gateway & LLM Observability Platform for AI Engineers

  • 🌐 AI Gateway: Access 100+ AI models with 1 API key through the OpenAI API with intelligent routing and automatic fallbacks. Get started in 2 minutes.
  • 🔌 Quick integration: One-line of code to log all your requests from OpenAI, Anthropic, LangChain, Gemini, Vercel AI SDK, and more.
  • 📊 Observe: Inspect and debug traces & sessions for agents, chatbots, document processing pipelines, and more
  • 📈 Analyze: Track metrics like cost, latency, quality, and more. Export to PostHog in one-line for custom dashboards
  • 🎮 Playground: Rapidly test and iterate on prompts, sessions and traces in our UI.
  • 🧠 Prompt Management: Version prompts using production data. Deploy prompts through the AI Gateway without code changes. Your prompts remain under your control, always accessible.
  • 🎛️ Fine-tune: Fine-tune with one of our fine-tuning partners: OpenPipe or Autonomi (more coming soon)
  • 🛡️ Enterprise Ready: SOC 2 and GDPR compliant

🎁 Generous monthly free tier (10k requests/month) - No credit card required!

Quick Start ⚡️

  1. Get your API key by signing up here and add credits at helicone.ai/credits

  2. Update the baseURL in your code and add your API key.

    import OpenAI from "openai";
    
    const client = new OpenAI({
      baseURL: "https://ai-gateway.helicone.ai",
      apiKey: process.env.HELICONE_API_KEY,
    });
    
    const response = await client.chat.completions.create({
      model: "gpt-4o-mini",  // claude-sonnet-4, gemini-2.0-flash or any model from https://www.helicone.ai/models
      messages: [{ role: "user", content: "Hello!" }]
    });
  3. 🎉 You’re all set! View your logs at Helicone and access 100+ models through one API.

Self-Hosting Open Source LLM Observability

Docker

Helicone is simple to self-host and update. To get started locally, just use our docker-compose file.

# Clone the repository
git clone https://github.com/Helicone/helicone.git
cd docker
cp .env.example .env

# Start the services
./helicone-compose.sh helicone up

Helm

For Enterprise workloads, we also have a production-ready Helm chart available. To access, contact us at enterprise@helicone.ai.

Manual deployment is not recommended. Please use Docker or Helm. If you must, follow the instructions here.

Architecture

Helicone is comprised of five services:

  • Web: Frontend Platform (NextJS)
  • Worker: Proxy Logging (Cloudflare Workers)
  • Jawn: Dedicated Server for serving collecting logs (Express + Tsoa)
  • Supabase: Application Database and Auth
  • ClickHouse: Analytics Database
  • Minio: Object Storage for logs.

Integrations 🔌

Inference Providers

IntegrationSupportsDescription
AI GatewayJS/TS, Python, cURLUnified API for 100+ providers with intelligent routing, automatic fallbacks, and unified observability
Async Logging (OpenLLMetry)JS/TS, PythonAsynchronous logging for multiple LLM platforms
OpenAIJS/TS, PythonInference provider
Azure OpenAIJS/TS, PythonInference provider
AnthropicJS/TS, PythonInference provider
OllamaJS/TSRun and use large language models locally
AWS BedrockJS/TSInference provider
Gemini APIJS/TSInference provider
Gemini Vertex AIJS/TSGemini models on Google Cloud’s Vertex AI
Vercel AIJS/TSAI SDK for building AI-powered applications
AnyscaleJS/TS, PythonInference provider
TogetherAIJS/TS, PythonInference provider
HyperbolicJS/TS, PythonInference provider
GroqJS/TS, PythonHigh-performance models
DeepInfraJS/TS, PythonServerless AI inference for various models
Fireworks AIJS/TS, PythonFast inference API for open-source LLMs

Frameworks

FrameworkSupportsDescription
LangChainJS/TS, PythonUse AI Gateway with LangChain for unified provider access
LlamaIndexPythonFramework for building LLM-powered data applications
LangGraphPythonBuild stateful, multi-actor applications with LLMs
Vercel AI SDKJS/TSAI SDK for building AI-powered applications
Semantic KernelC#, PythonMicrosoft’s AI orchestration framework
CrewAIPythonFramework for orchestrating role-playing AI agents
ModelFusionJS/TSAbstraction layer for integrating AI models into JavaScript and TypeScript applications
PostHogJS/TS, Python, cURLProduct analytics platform. Build custom dashboards.
RAGASPythonEvaluation framework for retrieval-augmented generation
Open WebUIJS/TSWeb interface for interacting with local LLMs
MetaGPTYAMLMulti-agent framework
Open DevinDockerAI software engineer
Mem0 EmbedChainPythonFramework for building RAG applications
DifyNo code requiredLLMOps platform for AI-native application development

This list may be out of date. Don’t see your provider or framework? Check out the latest integrations in our docs. If not found there, request a new integration by contacting help@helicone.ai.

Additional Resources

For more information, visit our documentation.