> For clean Markdown of any page, append .md to the page URL. > For a complete documentation index, see https://www.comet.com/docs/opik/evaluation/advanced/llms.txt. # Advanced > Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards. ## Docs - [Building Test Suites](https://www.comet.com/docs/opik/evaluation/advanced/building-test-suites.md): Create and manage Test Suites using the SDK, UI, or Ollie to build regression tests from real production failures - [Evaluate your agent](https://www.comet.com/docs/opik/evaluation/advanced/evaluate_your_llm.md): Evaluate your LLM applications confidently. Learn the five steps to assess complex LLM chains or agents effectively. - [Resume an interrupted evaluation](https://www.comet.com/docs/opik/evaluation/advanced/resume_evaluations.md): Continue a long-running Opik evaluation after a crash, network blip, or Ctrl-C — replaying only the runs that didn't finish. - [Manage datasets](https://www.comet.com/docs/opik/evaluation/advanced/manage_datasets.md): Evaluate your LLM using datasets. Learn to create and manage them via Python SDK, TypeScript SDK, or the Traces table. - [Evaluate agent trajectories](https://www.comet.com/docs/opik/evaluation/advanced/evaluate_agent_trajectory.md): Evaluate agent trajectories to optimize tool selection and reasoning paths, ensuring efficient agent behavior before production with Opik. - [Evaluate multi-turn agents](https://www.comet.com/docs/opik/evaluation/advanced/evaluate_multi_turn_agents.md): Learn to evaluate multi-turn agents using simulation techniques to enhance chatbot performance and improve user interactions. - [Annotation Queues](https://www.comet.com/docs/opik/evaluation/advanced/annotation_queues.md): Optimize your AI projects by enabling SMEs to efficiently review and annotate outputs using Opik's intuitive Annotation Queues feature. - [Manually logging experiments](https://www.comet.com/docs/opik/evaluation/advanced/log_experiments_with_rest_api.md): Evaluate your LLM application by logging pre-computed experiments and boosting confidence in performance with this detailed guide. - [Exporting experiment results](https://www.comet.com/docs/opik/evaluation/advanced/export_experiment_results.md): Export the full results of an experiment to CSV with the Opik Python SDK, matching the columns of the comparison page without its row cap.