# Agenta Blog

Articles and engineering posts from the Agenta team on building, evaluating, and shipping AI agents.

Canonical HTML page: <https://agenta.ai/blog>

## All posts

- [7 Best Manus Alternatives in 2026](https://agenta.ai/blog/best-manus-alternatives) — Article, Sep 1, 2026. Compare the best Manus alternatives for one-off tasks, ongoing AI coworkers, business automation, self-hosting, and team workflows.
- [Introducing Agenta 2.0](https://agenta.ai/blog/introducing-agenta-2-0) — Article, Jul 22, 2026. Agenta 2.0 is the open-source workspace for building and running agents: build them through chat, improve them with feedback, share them with your team.
- [CI/CD for LLM Prompts: How to Build a Prompt Deployment Pipeline](https://agenta.ai/blog/cicd-for-llm-prompts) — Engineering, Feb 11, 2026. How to build a CI/CD pipeline for LLM prompts. Covers webhook integration, automated evaluation gates, and three deployment paths.
- [Git vs. Prompt Management Tools: Which Should You Use?](https://agenta.ai/blog/git-vs-prompt-management-tools) — Engineering, Feb 11, 2026. Should you use Git or a dedicated tool for prompt versioning? Honest comparison with decision framework and the hybrid approach.
- [Prompt Drift: What It Is and How to Detect It](https://agenta.ai/blog/prompt-drift) — Engineering, Feb 11, 2026. What is prompt drift and why do LLM outputs change without prompt edits? Learn the three causes, how to detect drift, and how to prevent it.
- [Prompt Management for Non-Engineers: How Product Teams Can Own Their AI Prompts](https://agenta.ai/blog/prompt-management-for-non-engineers) — Engineering, Feb 11, 2026. How product managers and domain experts can contribute to AI prompt quality without writing code. A practical guide to collaborative prompt management.
- [Prompt Versioning: The Complete Guide](https://agenta.ai/blog/prompt-versioning-guide) — Engineering, Feb 11, 2026. Learn how to version LLM prompts for teams. Covers Git-based approaches, dedicated systems, three integration paths, and step-by-step setup.
- [Building the Data Flywheel: How to Use Production Data to Improve Your LLM Application](https://agenta.ai/blog/building-the-data-flywheel-how-to-use-production-data-to-improve-your-llm-application) — Article, Dec 19, 2025. Learn how to build a data flywheel for your LLM application using production data. Discover the 5-step process: from error analysis and clustering to creating golden test sets and improving prompts with Agenta's LLMOps platform. Master continuous improvement for reliable AI.
- [Top Open-Source Prompt Management Platforms 2026](https://agenta.ai/blog/top-open-source-prompt-management-platforms) — Engineering, Dec 17, 2025. Discover which open-source prompt management platform is right for your team. In-depth comparison of Agenta, Langfuse, Phoenix, Latitude, and Pezzo with features, licensing, and code examples.
- [Launch Week #2 Day 5: Jinja2 Prompt Templates](https://agenta.ai/blog/launch-week-2-day-5-jinja2-prompt-templates) — Article, Nov 14, 2025. Agenta prompt playground now supports Jinja2 prompt templates. Create dynamic LLM prompts with conditional logic. Prompt management with Jinja2 templating support. 
- [Commercial Open Source Is Hard: Our Journey](https://agenta.ai/blog/commercial-open-source-is-hard-our-journey) — Article, Nov 13, 2025. Today we're open-sourcing most of our product. Here's what we learned after three failed attempts.
- [Launch Week #2 Day 4: Open Sourcing Evaluation](https://agenta.ai/blog/launch-week-2-day-4-open-sourcing-evaluation) — Article, Nov 13, 2025. We're open-sourcing all functional features of Agenta under the MIT license.
- [Launch Week #2 Day 3: Evaluation SDK](https://agenta.ai/blog/launch-week-2-day-3-evaluation-sdk) — Article, Nov 12, 2025. We're launching the evaluation SDK today. The evaluation SDK allows you to evaluate complex agents and LLM workflows using built-in or custom evaluators.
- [Launch Week #2 Day 2: Online Evaluation](https://agenta.ai/blog/launch-week-2-day-2-online-evaluation) — Article, Nov 11, 2025. Today we're launching online evaluation. With online evaluation each request is automatically evaluated. This allows you to monitor things in production.
- [Launch Week #2 Day 1: New Evaluation Dashboard ](https://agenta.ai/blog/launch-week-2-day-1) — Article, Nov 10, 2025. Launch Week Day 1: Redesigned evaluation dashboard with side-by-side comparison, detailed debugging, and customizable LLM-as-a-judge evaluators.
- [LLM as a Judge: Guide to LLM Evaluation & Best Practices](https://agenta.ai/blog/llm-as-a-judge-guide-to-llm-evaluation-best-practices) — Article, Sep 30, 2025. A practical guide to LLM as a judge: design, implement, and automate LLM evaluation and RAG evaluation for your AI projects.
- [Top LLM Gateways 2025](https://agenta.ai/blog/top-llm-gateways) — Engineering, Sep 30, 2025. We compare and test the top LLM gateways in 2025. These includes Litellm, Helicone, BricksLLM and Kong AI Gateway.
- [Top LLM Observability platforms 2025](https://agenta.ai/blog/top-llm-observability-platforms) — Engineering, Sep 29, 2025. Explore the best LLM Observability platforms of 2025. Compare open-source and enterprise tools like Agenta, Langfuse, Langsmith and more. 
- [The guide to structured outputs and function calling with LLMs](https://agenta.ai/blog/the-guide-to-structured-outputs-and-function-calling-with-llms) — Article, Sep 10, 2025. Get reliable JSON from any LLM using structured outputs, JSON mode, Pydantic, Instructor, and Outlines. Complete production guide with OpenAI, Claude, and Gemini code examples for consistent data extraction.
- [The AI Engineer's Guide to LLM Observability with OpenTelemetry](https://agenta.ai/blog/the-ai-engineer-s-guide-to-llm-observability-with-opentelemetry) — Article, Aug 27, 2025. Learn why LLM observability is critical for production AI. This guide covers traces, OpenTelemetry (OTel), and the LLMOps workflows you need to build reliable apps
- [The Ultimate Guide to RAG Chunking Strategies](https://agenta.ai/blog/the-ultimate-guide-for-chunking-strategies) — Article, Aug 15, 2025. Learn 4 essential chunking strategies for RAG systems: syntactic, recursive, semantic, and cluster-based. Compare performance with code examples and evaluation metrics.
- [Building in Public: Why We're Publishing Our Roadmap](https://agenta.ai/blog/building-in-public-why-we-re-publishing-our-roadmap) — Article, Aug 12, 2025. Why we're open-sourcing our product roadmap in Agenta
- [July 2025 Product Updates](https://agenta.ai/blog/july-2025-product-updates) — Article, Aug 7, 2025. Product updates for July 2025. Adding tool and image support to the LLM playground to improve your prompt engineering flow, new observability integrations, and feedback endpoint to capture evaluations from your end-users
- [Humanloop Sunsetting - Migration and Alternative](https://agenta.ai/blog/humanloop-sunsetting-migration-and-alternative) — Engineering, Jul 22, 2025. Humanloop has been acquired and goes offline on September 8, 2025. Agenta is an ideal alternative that lets you version prompts, evaluate, and monitor LLM apps easily. Migrate your prompts and workflows to Agenta with free white-glove migration support.
- [Top techniques to Manage Context Lengths in LLMs](https://agenta.ai/blog/top-6-techniques-to-manage-context-length-in-llms) — Article, Jul 16, 2025. Overcome LLM token limits with 6 practical techniques. Learn how you can use truncation, RAG, memory buffering, and compression to overcome the token limit and fit the LLM context window.
- [Top 10 Techniques to Improve RAG Applications](https://agenta.ai/blog/top-10-techniques-to-improve-rag-applications) — Article, Jul 9, 2025. A practical guide to RAG architectures, chunking strategies, reranking, and evaluation—improve your RAG system's accuracy and performance.
- [How to Evaluate RAG: Metrics, Evals, and Best Practices](https://agenta.ai/blog/how-to-evaluate-rag-metrics-evals-and-best-practices) — Article, Jul 1, 2025. A practical guide to RAG evaluation, evaluation metrics, RAGAS, and LLM evaluation. Learn how to measure and improve your RAG systems.
- [Launch Week Day 5: SOC2 Type 2 Compliance](https://agenta.ai/blog/soc2-type2) — Article, Apr 18, 2025. It's Official: Agenta Is Now SOC2 Type 2 Compliant
- [Launch Week Day 4 – Structured Output in the Playground](https://agenta.ai/blog/structured-outputs-playground) — Article, Apr 17, 2025. Enforce JSON and schema-validated responses straight from the Agenta playground.
- [Launch Week Day 3: Prompt & Configuration Registry](https://agenta.ai/blog/introducing-prompt-registry) — Article, Apr 16, 2025. Today we're launching the Prompt & Configuration Registry. It's a place to manage all your LLM prompts and configurations.
- [Launch Week Day 2: Custom Workflows](https://agenta.ai/blog/introducing-custom-workflows) — Article, Apr 15, 2025. Agenta helps teams build better LLM applications. Today, we're releasing Custom Workflows - a way to connect your entire application to Agenta's playground and evaluation tools.
- [Launch Week Day 1: AI Model Hub](https://agenta.ai/blog/introducing-ai-model-hub) — Article, Apr 14, 2025. Agenta helps you build LLM applications with our playground for prompt engineering and evaluation tools. Now, we're making it possible to use virtually any model you have access to, no matter where it's hosted.
- [Agenta Launch Week #1: April 15-19](https://agenta.ai/blog/announcing-agenta-launch-week-1-april-15-19) — Article, Apr 9, 2025. We're excited to share that our first-ever Launch Week is happening next week, April 15-19!
- [What We Learned Building a Prompt Management System](https://agenta.ai/blog/what-we-learned-building-a-prompt-management-system) — Article, Mar 18, 2025. For teams building production-grade LLM applications, a systematic approach to prompt management is no longer optional—it's essential infrastructure. Here's why it matters and how to get it right.
- [Introducing prompt Playground 2.0: A New Prompt Engineering IDE](https://agenta.ai/blog/prompt-playground) — Article, Feb 6, 2025. Streamling your prompt engineering with Playground 2.0. An integrated LLM playground for testing and comparing prompts and models.
- [The Definitive Guide to Prompt Management Systems](https://agenta.ai/blog/the-definitive-guide-to-prompt-management-systems) — Article, Jan 22, 2025. Explore why prompt management is crucial for scaling AI applications from pilots to production.
- [Product Updates November 2024 - LLM Observability and Prompt management](https://agenta.ai/blog/product-update-november-2024) — Article, Nov 26, 2024. In this product update we introduce LLM observability, prompt management, and a new interface to configure LLM-as-a-judge evaluators.
- [Introducing Open-Source LLM Observability with Agenta](https://agenta.ai/blog/open-source-llm-observability) — Article, Nov 13, 2024. Agenta introduces open-source LLM observability and LLM monitoring for LLM applications. It allows you to trace inputs, outputs, and meta-data with two-lines of code. It is OpenTelemetry compliant and comes with many integrations out of the box (OpenAI, LiteLLM, LangChain, Instructor and more).
- [Agenta Achieves SOC2 Type I Certification](https://agenta.ai/blog/agenta-achieves-soc2-type-i-certification) — Article, Jan 15, 2024. Agenta secures SOC2 Type 1 certification, ensuring your LLM development data stays protected with enterprise-level security. Build your AI applications with confidence.

---

## Machine-readable

- [Sitemap](https://agenta.ai/sitemap-index.xml)
- [llms.txt](https://agenta.ai/llms.txt)
- [OpenAPI specification](https://agenta.ai/openapi.json)
- [Documentation](https://docs.agenta.ai)

Every page on https://agenta.ai serves this representation when requested with
`Accept: text/markdown`.
