LLM observability
Track prompts, traces, errors, runs, and related metadata so teams can inspect how LLM apps behave in production.
Lunary is an observability and prompt management platform for LLM-based applications. It is built to help teams monitor, improve, and secure AI chatbots and other LLM workflows as they move from development into production.
The product combines logs, traces, prompt templates, human review, analytics, and guardrails in one place. The site also emphasizes self-hosting options, enterprise controls, and SDK-based integration so teams can add monitoring without placing Lunary directly in the request path.
Track prompts, traces, errors, runs, and related metadata so teams can inspect how LLM apps behave in production.
Create prompt templates, collaborate with teammates, and use versioning and A/B testing to iterate on prompt changes.
Review responses, label data, and export logs to JSONL for downstream fine-tuning workflows.
Search, filter, and analyze usage, costs, topics, and satisfaction with analytics and custom dashboards.
Mask personal information, manage roles and access, and support SSO/SAML for controlled data access.
Connect through SDKs, an HTTP API, and integrations across OpenAI, LangChain, LiteLLM, Flowise, and other tools.
Use Lunary to inspect production LLM traffic, review traces and errors, and understand how real users interact with a chatbot or agent.
Create templates, compare prompt versions, and run A/B tests when iterating on system prompts or chat flows.
Label logs, review responses, and export data to JSONL when preparing datasets for model fine-tuning or analysis.
Apply PII masking, access controls, and self-hosting when handling sensitive user data or compliance-sensitive workloads.
Share projects, invite teammates, and use human reviews and dashboards to coordinate feedback across technical and non-technical collaborators.
Lunary is a platform for monitoring, improving, and securing AI chatbots. Its FAQ describes it as covering observability, prompt management, evaluations, and LLM guardrails.
According to the FAQ, a run can be types such as llm, trace, chain, embedding, or chat. For LLMs, a run is typically one API call to an inference API or a chat message in a thread.
No. The FAQ says Lunary SDKs run asynchronously rather than sitting between your code and the inference API, so they should not affect request latency.
The FAQ says free-plan data is available for 30 days, paid-plan data is kept indefinitely, and self-hosted data stays with you indefinitely.
Yes. The FAQ says you can self-host for free with the Community Edition, while the Enterprise Edition supports Docker, Kubernetes, and more and is paid.
Prompt Genie 是面向 AI 工作流的提示管理与优化工具,并提供 Claude Code 记忆功能,可复用会话上下文中的重复读取,帮助个人和团队复用提示、对比模型输出并减少不必要的 token 消耗。
Orca 是面向编码代理的 Agent 开发环境,支持在隔离的 git worktree 中并行运行多个 CLI agent,并提供桌面端与移动端协作流程。
Firebase Studio 是一款基于浏览器的全栈应用开发工作区,支持 Gemini 辅助编写代码、应用预览、云模拟器、导入仓库、原型开发、协作与从浏览器部署。
EZsite AI 是一款 AI 网站构建器,可将 URL 转为完整的 React 或 Vue.js 应用。支持托管、自定义域名、代码导出及面向团队的后端功能,帮助快速交付可部署结果。
AI Magicx 是一体化 AI 工作区,集聊天、图片、视频、语音、音乐、邮件和开发任务于一处,帮助创作者、团队和开发者集中使用多种模型,无需在多个工具和订阅间切换。
Paper 是一款设计工具,连接画布、代码和 AI agent,让团队在一个工作流中完成创建、协作与交付。支持桌面应用、基于 MCP 的 agent 访问,以及真实内容和设计 token 工作流。