# comet.com > AI-optimized mirror of comet.com containing 50 pages totalling 54,028 words of clean markdown content, structured data, and semantic HTML. Original source: https://comet.com. Last updated: 2026-07-20T14:34:13.489Z. Each page is available as HTML (with JSON-LD structured data) and Markdown (text-only, ideal for LLMs and RAG). ## Articles & Blog Posts - [Overview | Opik Documentation | Opik Documentation](/content/docs/opik/integrations/overview/index.html): Explore how to log, visualize, and evaluate LLM calls and RAG pipelines using Opik's robust integrations with popular frameworks. (244 words) - [docs/opik/llms-full-txt.html](/content/docs/opik/llms-full-txt.html) (66 words) - [March 3, 2026 | Opik Documentation](/content/docs/opik/changelog/2026/3/3/index.html): Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards. (1,041 words) - [December 9, 2025 | Opik Documentation](/content/docs/opik/changelog/2025/12/9/index.html): Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards. (685 words) - [Log traces | Opik Documentation | Opik Documentation](/content/docs/opik/tracing/advanced/log_traces/index.html): Monitor the flow of your LLM applications with tracing to identify issues and optimize performance using Opik's powerful tools. (1,819 words) - [Custom conversation metric | Opik Documentation | Opik Documentation](/content/docs/opik/evaluation/metrics/custom_conversation_metric/index.html): Learn how to create custom metrics for evaluating multi-turn conversations (909 words) - [Opik's MCP server | Opik Documentation | Opik Documentation](/content/docs/opik/mcp-server/index.html): Configure Opik's Python MCP server with Claude Code, Cursor, and VS Code Copilot to read traces, log scores, and manage prompts from your AI host. (1,293 words) - [May 12, 2026 | Opik Documentation](/content/docs/opik/changelog/2026/5/12/index.html): Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards. (643 words) - [FAQ | Opik Documentation | Opik Documentation](/content/docs/opik/faq/index.html): Explore common questions about Opik, learn how to get started, and find assistance to enhance your experience with our powerful tool. (1,395 words) - [December 18, 2025 | Opik Documentation](/content/docs/opik/changelog/2025/12/18/index.html): Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards. (585 words) - [January 27, 2026 | Opik Documentation](/content/docs/opik/changelog/2026/1/27/index.html): Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards. (720 words) - [February 10, 2026 | Opik Documentation](/content/docs/opik/changelog/2026/2/10/index.html): Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards. (454 words) - [May 5, 2026 | Opik Documentation](/content/docs/opik/changelog/2026/5/5/index.html): Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards. (615 words) - [Optimize agents | Opik Documentation | Opik Documentation](/content/docs/opik/development/optimization-runs/optimization/optimize_agents/index.html): Connect the Agent Optimizer to Agent Frameworks. (1,066 words) - [ML Experiment Tracking | Comet](/content/site/products/ml-experiment-tracking/index.html): Get automatic experiment tracking for ML with tools to version datasets, debug and reproduce models, visualize performance, and collaborate with teammates. (353 words) - [Our Blogs | Opik by Comet](/content/site/blog/page/34/index.html): Comet's blog tracks industry updates, customer stories, product use cases, and more. AI leaders offer advice to optimize Comet's capabilities. (332 words) - [Log media & attachments | Opik Documentation | Opik Documentation](/content/docs/opik/tracing/advanced/log_multimodal_traces/index.html): Capture multimodal traces in Opik by logging images, videos, and audio files using the Python SDK's Attachment type. (1,052 words) - [Conversational metrics | Opik Documentation | Opik Documentation](/content/docs/opik/evaluation/metrics/conversation_threads_metrics/index.html): Describes metrics related to scoring the conversational threads (870 words) - [Log conversations | Opik Documentation | Opik Documentation](/content/docs/opik/tracing/advanced/log_chat_conversations/index.html): Capture and analyze chat conversations on Opik, enabling better tracking of user interactions and improving chatbot performance. (873 words) - [Opik Documentation - Open-Source LLM Observability & Optimization | Opik Documentation](/content/docs/opik/index.html): Build, test, and optimize GenAI apps from prototype to production. Comprehensive tracing, evaluation, and prompt optimization for RAG, agents, and more. (498 words) - [Building Test Suites | Opik Documentation | Opik Documentation](/content/docs/opik/evaluation/advanced/building-test-suites/index.html): Create and manage Test Suites using the SDK, UI, or Ollie to build regression tests from real production failures (600 words) - [July 6, 2026 | Opik Documentation](/content/docs/opik/changelog/2026/7/6/index.html): Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards. (767 words) - [Ollie | Opik Documentation | Opik Documentation](/content/docs/opik/ollie/index.html): Meet Ollie — the AI assistant built into Opik that doesn't just analyze your traces, it helps you improve your agent's code directly from the Opik platform. (614 words) - [June 2, 2026 | Opik Documentation](/content/docs/opik/changelog/2026/6/2/index.html): Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards. (622 words) - [May 26, 2026 | Opik Documentation](/content/docs/opik/changelog/2026/5/26/index.html): Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards. (298 words) - [June 16, 2026 | Opik Documentation](/content/docs/opik/changelog/2026/6/16/index.html): Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards. (564 words) - [Changelog | Opik Documentation](/content/docs/opik/changelog/index.html): Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards. (714 words) - [Vercel AI SDK | Opik Documentation | Opik Documentation](/content/docs/opik/integrations/vercel-ai-sdk/index.html): Start here to integrate Opik into your Vercel AI SDK-based genai application for end-to-end LLM observability, unit testing, and optimization. (695 words) - [Comet - The AI Developer Platform](/content/site/index.html): Comet is the creator of Opik, an end-to-end AI observability platform for developers with best-in-class agent testing, optimization, and monitoring. (639 words) - [Upgrading to Opik 2.0 | Opik Documentation | Opik Documentation](/content/docs/opik/opik-v2-upgrade/index.html): What changed in Opik 2.0 — everything is now organized around projects — how to update your SDK code, where your existing data was moved, and how to relocate it if needed. (727 words) - [Comet at NeurIPS 2025](/content/site/events/neurips/index.html): Join the Comet team for a series of fun offsite events around the NeurIPS 2025 conference in San Diego! (401 words) - [Bounty Program | Opik Documentation | Opik Documentation](/content/docs/opik/contributing/developer-programs/bounties/index.html): Participate in the Opik Bounty Program to enhance the platform, showcase your skills, and earn rewards for your contributions. (452 words) - [Observability Overview | Opik Documentation | Opik Documentation](/content/docs/opik/tracing/overview/index.html): Monitor, debug, and optimize your LLM applications with Opik's observability platform for traces, spans, and conversations (436 words) - [Quickstart | Opik Documentation | Opik Documentation](/content/docs/opik/quickstart/index.html): Integrate Opik with your LLM application to log calls and chains efficiently. Get started with our step-by-step guide. (345 words) - [site/blog/sam-stable-diffusion-for-text-to-image-inpainting/index.html](/content/site/blog/sam-stable-diffusion-for-text-to-image-inpainting/index.html) (2,569 words) - [Opik Agent Optimizer | Opik Documentation | Opik Documentation](/content/docs/opik/development/optimization-runs/overview/index.html): Learn about Opik's automated LLM prompt and agent optimization SDK. Discover MetaPrompt, Few-shot Bayesian, and Evolutionary optimization algorithms for enhanced performance. (663 words) - [Our Blogs | Opik by Comet](/content/site/blog/index.html): Comet's blog tracks industry updates, customer stories, product use cases, and more. AI leaders offer advice to optimize Comet's capabilities. (332 words) - [Batching and update operations | Opik Documentation | Opik Documentation](/content/docs/opik/reference/python-sdk/troubleshooting/batching-and-updates/index.html): Understand how batching affects update operations and how to avoid data loss when using the Opik Python SDK. (517 words) - [What Is an Agent Harness? The Layer That Makes AI Agents Actually Work - Comet](/content/site/blog/agent-harness/index.html): If you’ve shipped an LLM-powered feature beyond a simple chat interface, you’ve already built parts of an agent harness. The retry logic around a failed tool call, the code assembling a system prompt at runtime — these are harness components. Most teams build these pieces incrementally as problems surface in production, which means the agent […] (2,365 words, Jul 17, 2026) - [Agent Tracing: End-to-End Observability for Complex AI Systems](/content/site/blog/ai-agent-tracing/index.html): Learn how to properly instrument agent tracing to identify and debug silent errors in complex agentic AI systems. (2,633 words, Jun 3, 2026) - [Best AI Observability Tools for Agentic Systems](/content/site/blog/ai-observability-tools/index.html): Compare the best AI observability tools for agentic systems in 2026. Learn which platforms are best for tracing, evaluation, debugging, testing, and production monitoring. (4,492 words, May 27, 2026) - [LLM Cost Tracking Solution: How to Monitor AI Spend](/content/site/blog/llm-cost-tracking-solution/index.html): Learn how an LLM cost tracking solution helps monitor token usage, trace AI spend across agentic systems, and optimize costs without sacrificing quality. (1,814 words, May 15, 2026) - [Regression Testing for AI Agents: Introducing Opik Test Suites](/content/site/blog/ai-agent-regression-testing/index.html): We should test AI agents the way we test software — Opik brings straightforward regression testing to your agent development workflow. (1,172 words, Apr 21, 2026) - [RAG Evaluation Guide: Metrics, Methods, and Key Quality Signals](/content/site/blog/rag-evaluation/index.html): Learn how to evaluate RAG systems with proven evaluation metrics for retrieval, generation, and end-to-end quality. (3,560 words, Feb 24, 2026) - [LLM-as-a-Judge: The Ultimate Guide for AI Developers](/content/site/blog/llm-as-a-judge/index.html): How to build reliable, scalable evaluation for LLM apps and agents with llm-as-a-judge, including rubrics, judge architectures, bias mitigation. (1,202 words, Feb 18, 2026) - [Best LLM Observability Tools of 2025 | Compare Top Platforms](/content/site/blog/llm-observability-tools/index.html): Compare the best LLM observability tools, such as Opik, Langfuse, and Datadog to monitor, evaluate, and optimize LLM performance. (3,597 words, Nov 11, 2025) - [Human-in-the-Loop Feedback for Agent Validation](/content/site/blog/thread-level-human-feedback/index.html): Learn how to automate feedback loops, eliminate manual data labeling, and use Opik to capture insights to improve your GenAI apps. (1,413 words, Oct 10, 2025) - [AI Agent Design: How to Build Reliable AI Agent Architecture](/content/site/blog/ai-agent-design/index.html): Discover the breakdown of tried and true design principles for AI agent architecture that enable you to ship and scale real-world AI agents. (1,496 words, Aug 8, 2025) - [LLM Evaluation Complexities for Non-Latin Languages - Comet](/content/site/blog/complexities-for-non-latin-languages-llm-evaluations.html): Large language models (LLMs) have revolutionized natural language processing, yet most development and evaluation efforts have historically centered around Latin-script languages. When these models are extended to non-Latin alphabets, especially Chinese, Japanese, and Korean (CJK), challenges emerge that span linguistic structure, cultural context, and technical implementation. This article offers a deep dive into these challenges […] (2,524 words, Mar 27, 2025) ## Listings & Categories - [Siddharth Mehta, Author at Comet](/content/site/blog/author/siddharthmcomet-com/index.html) (292 words) ## Resources - [Full Page Index](/index.html): Browse all cached pages with rich metadata - [About This Cache](/content/about.html): Methodology, technical details, and usage guidelines - [XML Sitemap](/sitemap.xml): Machine-readable sitemap for crawler discovery - [Robots.txt](/robots.txt): Crawler directives