About Langfuse
Langfuse is an open-source LLM engineering platform that helps teams develop, monitor, evaluate and debug AI applications together. You instrument your app, and Langfuse records traces of LLM calls and surrounding logic such as retrieval, embeddings and agent actions, so complex runs and user sessions can be inspected and debugged.
Prompt management lets you centrally version and iterate on prompts, with caching on server and client so that changes do not add latency to your app. Evaluation features cover LLM-as-a-judge, code evaluators, user feedback, manual labeling and custom pipelines through the API and SDKs. Datasets provide test sets and benchmarks, there is an LLM playground for trying prompts, and integrations exist for OpenAI, LangChain and LlamaIndex. It is built on the ClickHouse database, and the Langfuse team has been part of ClickHouse since January 2026.
You can use Langfuse Cloud or self-host it, which the maintainers say takes minutes. The repository license is listed as 'Other' because it combines open-source and enterprise components, so check the terms. It suits teams shipping production LLM features who need visibility into quality, latency and cost.
Key features
- Tracing of LLM calls, retrieval and agent actions
- Prompt versioning with caching
- LLM-as-a-judge and custom evaluations
- Datasets for tests and benchmarks
- Interactive LLM playground
- Integrations with OpenAI, LangChain and LlamaIndex
Good fit for
- →Debugging production LLM applications
- →Tracking prompt versions and quality over time
- →Evaluating agents against test datasets
- Built with
- TypeScript
- Tags
- llmops
- observability
- llm
- tracing
- evaluation
- prompt-management
- self-hosted
- ai
- analytics
Langfuse: questions and answers
- What is Langfuse used for?
- Langfuse is an open-source LLM engineering platform for tracing, evaluating and improving AI applications, with prompt management, datasets and a playground, self-hosted or cloud. It is a good fit for debugging production LLM applications, tracking prompt versions and quality over time, and evaluating agents against test datasets.
- Is Langfuse open source?
- Yes. Langfuse is open source under a custom licence. Its source code is on GitHub at langfuse/langfuse and is written mainly in TypeScript.
- Is Langfuse free?
- Yes. Langfuse is open source, so the software itself is free to use under the terms of its own licence. A managed cloud version is also available.
- Can I self-host Langfuse?
- Yes. Langfuse can be self-hosted on your own server or infrastructure.
- What is Langfuse an alternative to?
- Langfuse is an open-source alternative to Weights & Biases, LangSmith, Arize AI and Braintrust. Other open-source alternatives to Weights & Biases include Opik, Arize Phoenix and Laminar.
- Is Langfuse actively maintained?
- Yes. The most recent commit to Langfuse was on 2 October 2026, and the latest release is v4.50.0, published on 2 October 2026. The project has 35k stars on GitHub.
Open-source alternatives to Langfuse
See all
Opik
AI Infrastructure
Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows wit
Apache-2.0vs Weights & Biases★ 22k
Arize Phoenix
AI Infrastructure
AI Observability & Evaluation
OSSvs Weights & Biases★ 12k
Laminar
AI Infrastructure
Laminar - open-source observability platform purpose-built for AI agents. YC S24.
Apache-2.0vs Weights & Biases★ 3.3k
MLflow
AI Infrastructure
The open source AI engineering platform for agents, LLMs, and ML models. MLflow enables te
Apache-2.0vs Weights & Biases★ 28k
LangWatch
AI Infrastructure
The platform for LLM evaluations and AI agent testing
Apache-2.0vs LangSmith★ 4.9k
TraceRoot
AI Infrastructure
TraceRoot - open source self improving layer for ai agents YC S25
OSSvs Weights & Biases★ 790
SaaS alternatives to Langfuse
See all
Weights & Biases
AI Infrastructure
Experiment tracking, model registry and LLM evaluation for machine learning teams
SaaS
LangSmith
AI Infrastructure
Observability, tracing and evaluation platform for LLM applications
SaaS
Arize AI
AI Infrastructure
Observability and evaluation platform for machine learning models and LLM applications
SaaS
Braintrust
AI Infrastructure
Evaluation and observability platform for building and testing AI applications
SaaS
Galileo
AI Infrastructure
Evaluation and monitoring platform for generative AI applications
SaaS
Fiddler AI
AI Infrastructure
AI observability and model monitoring platform with explainability
SaaS

