Helicone
リクエストの経路上に入ってモデル呼び出しを記録するため、コード側に計測を仕込まなくてよい
Heliconeとは
多くの記録ツールは呼び出しを包むことを求めますが、見たいコードがライブラリの中や別のジョブ、他人が持つサービスの中にあると、その方法は使えません。Helicone は別の道を取ります。接続先のURLを向けるだけで、通過するリクエストがすべて記録され、経路上にSDKは入りません。すでにその位置にいるため、ゲートウェイの仕事も兼ねます。複数の提供元を1つの鍵で扱い、片方が落ちたら自動で切り替え、繰り返しのプロンプトはキャッシュから返します。引き換えはその位置そのものです。アプリケーションとモデルの間にプロキシが入るため、運用する部品と経路が1つ増えます。エージェントの実行はセッションとしてまとまるので、多段の記録が無関係な30件の呼び出しではなく1つのまとまりとして読めます。
Heliconeで何ができますか?
- 呼び出し側のコードを変えずに記録する — 接続先のURLを変えるだけで導入が済むため、ライブラリの奥や他チームのサービスの中で行われる呼び出しも記録されます。
- 多段の実行を1つのまとまりとして読む — 関連する呼び出しがセッションにまとまるので、30回要求を出して作業を終えたエージェントが30行ではなく1つの記録として見えます。
- 提供元が止まっても動き続ける — すでに経路上にいるため、複数の提供元へ振り分け、片方が失敗し始めたら自動的に切り替えられます。
- 同じプロンプトに二重で払わない — 繰り返しの要求をキャッシュから返せます。毎ターン同じ文脈を送り直すエージェントのループで効いてきます。
- 記録した通信を評価データにする — 記録した要求を評価や追加学習のために集められます。作り物の例ではなく実際の利用が試験項目になります。
- データを外に出せないなら自分で動かす — 基盤はオープンソースで自社運用できます。プロンプトに社外へ出せない内容が含まれる場合、ここが決め手になります。
Heliconeを選ぶ前に
- 経路上のプロキシは運用対象が1つ増え、すべての呼び出しに1経路加わることを意味します。それを受け入れる理由は、手を入れられないコードまで記録できる点にあります。
- 基盤一式を自社で動かすのは、記録用ライブラリを足すより重い作業です。読み込むパッケージではなく、自前のデータ保存先を伴うWebアプリケーションだからです。
よくある質問
Heliconeは商用利用できますか?
HeliconeはApache-2.0ライセンスで公開されています。OSI承認のオープンソースライセンスで、商用利用が認められています。
Heliconeはどの形で使えますか?
Heliconeはセルフホスト・マネージドクラウド・ローカル実行の形で利用できます。
ドキュメント
Helicone/helicone のREADMEより転載(Apache-2.0)。 原文を読む ↗
| 🔍 Observability | 🕸️ Agent Tracing | 🚂 LLM Routing |
|---|---|---|
| 💰 Cost & Latency Tracking | 📚 Datasets & Fine-tuning | 🎛️ Automatic Fallbacks |
Helicone is an AI Gateway & LLM Observability Platform for AI Engineers
- 🌐 AI Gateway: Access 100+ AI models with 1 API key through the OpenAI API with intelligent routing and automatic fallbacks. Get started in 2 minutes.
- 🔌 Quick integration: One-line of code to log all your requests from OpenAI, Anthropic, LangChain, Gemini, Vercel AI SDK, and more.
- 📊 Observe: Inspect and debug traces & sessions for agents, chatbots, document processing pipelines, and more
- 📈 Analyze: Track metrics like cost, latency, quality, and more. Export to PostHog in one-line for custom dashboards
- 🎮 Playground: Rapidly test and iterate on prompts, sessions and traces in our UI.
- 🧠 Prompt Management: Version prompts using production data. Deploy prompts through the AI Gateway without code changes. Your prompts remain under your control, always accessible.
- 🎛️ Fine-tune: Fine-tune with one of our fine-tuning partners: OpenPipe or Autonomi (more coming soon)
- 🛡️ Enterprise Ready: SOC 2 and GDPR compliant
🎁 Generous monthly free tier (10k requests/month) - No credit card required!
Quick Start ⚡️
-
Get your API key by signing up here and add credits at helicone.ai/credits
-
Update the
baseURLin your code and add your API key.import OpenAI from "openai"; const client = new OpenAI({ baseURL: "https://ai-gateway.helicone.ai", apiKey: process.env.HELICONE_API_KEY, }); const response = await client.chat.completions.create({ model: "gpt-4o-mini", // claude-sonnet-4, gemini-2.0-flash or any model from https://www.helicone.ai/models messages: [{ role: "user", content: "Hello!" }] }); -
🎉 You’re all set! View your logs at Helicone and access 100+ models through one API.
Self-Hosting Open Source LLM Observability
Docker
Helicone is simple to self-host and update. To get started locally, just use our docker-compose file.
# Clone the repository
git clone https://github.com/Helicone/helicone.git
cd docker
cp .env.example .env
# Start the services
./helicone-compose.sh helicone up
Helm
For Enterprise workloads, we also have a production-ready Helm chart available. To access, contact us at enterprise@helicone.ai.
Manual (Not Recommended)
Manual deployment is not recommended. Please use Docker or Helm. If you must, follow the instructions here.
Architecture
Helicone is comprised of five services:
- Web: Frontend Platform (NextJS)
- Worker: Proxy Logging (Cloudflare Workers)
- Jawn: Dedicated Server for serving collecting logs (Express + Tsoa)
- Supabase: Application Database and Auth
- ClickHouse: Analytics Database
- Minio: Object Storage for logs.
Integrations 🔌
Inference Providers
| Integration | Supports | Description |
|---|---|---|
| AI Gateway | JS/TS, Python, cURL | Unified API for 100+ providers with intelligent routing, automatic fallbacks, and unified observability |
| Async Logging (OpenLLMetry) | JS/TS, Python | Asynchronous logging for multiple LLM platforms |
| OpenAI | JS/TS, Python | Inference provider |
| Azure OpenAI | JS/TS, Python | Inference provider |
| Anthropic | JS/TS, Python | Inference provider |
| Ollama | JS/TS | Run and use large language models locally |
| AWS Bedrock | JS/TS | Inference provider |
| Gemini API | JS/TS | Inference provider |
| Gemini Vertex AI | JS/TS | Gemini models on Google Cloud’s Vertex AI |
| Vercel AI | JS/TS | AI SDK for building AI-powered applications |
| Anyscale | JS/TS, Python | Inference provider |
| TogetherAI | JS/TS, Python | Inference provider |
| Hyperbolic | JS/TS, Python | Inference provider |
| Groq | JS/TS, Python | High-performance models |
| DeepInfra | JS/TS, Python | Serverless AI inference for various models |
| Fireworks AI | JS/TS, Python | Fast inference API for open-source LLMs |
Frameworks
| Framework | Supports | Description |
|---|---|---|
| LangChain | JS/TS, Python | Use AI Gateway with LangChain for unified provider access |
| LlamaIndex | Python | Framework for building LLM-powered data applications |
| LangGraph | Python | Build stateful, multi-actor applications with LLMs |
| Vercel AI SDK | JS/TS | AI SDK for building AI-powered applications |
| Semantic Kernel | C#, Python | Microsoft’s AI orchestration framework |
| CrewAI | Python | Framework for orchestrating role-playing AI agents |
| ModelFusion | JS/TS | Abstraction layer for integrating AI models into JavaScript and TypeScript applications |
| PostHog | JS/TS, Python, cURL | Product analytics platform. Build custom dashboards. |
| RAGAS | Python | Evaluation framework for retrieval-augmented generation |
| Open WebUI | JS/TS | Web interface for interacting with local LLMs |
| MetaGPT | YAML | Multi-agent framework |
| Open Devin | Docker | AI software engineer |
| Mem0 EmbedChain | Python | Framework for building RAG applications |
| Dify | No code required | LLMOps platform for AI-native application development |
This list may be out of date. Don’t see your provider or framework? Check out the latest integrations in our docs. If not found there, request a new integration by contacting help@helicone.ai.
Additional Resources
-
LLM Cost API: We have the largest open-source API pricing database with 300+ models and providers such as OpenAI, Anthropic and more. Start querying here.
-
Data Management: Manage and export your Helicone data with our API or access it with our MCP server.
- Guides: ETL, Request Exporting
-
Data Ownership: Learn about Data Ownership and Autonomy
For more information, visit our documentation.