The PromptFoo.tech blog

Editorial pieces on AI workflows, automation, agents, MCP, and prompt engineering — with a bias toward what actually ships.

MCP 9 min

What Is MCP? A Practical Guide to the Model Context Protocol

MCP is quickly becoming the USB-C of AI tools. Here's what it actually is, when to use it, and how to get value from it in a week.

Updated Jun 20, 2025
Automation 11 min

n8n vs Make vs Zapier in 2025: Which to Pick

A practical comparison of the three big automation platforms across pricing, AI-readiness, self-hosting, and real-world usability.

Updated Jun 14, 2025
Prompt Engineering 10 min

Prompt Engineering Fundamentals That Still Matter in 2025

Models got smarter but prompt structure still moves outcomes. Here are the fundamentals that survived every model release.

Updated Jun 1, 2025
Tutorials 9 min

How to Write an AI Workflow That Doesn't Break in Production

Six lessons from workflows that survived contact with real traffic — retries, schema validation, cost caps, and human gates.

Updated May 25, 2025
Guides 12 min

Choosing an LLM in 2025: A Buyer's Guide for Workflow Builders

How to pick between GPT, Claude, Gemini, and open models for real production workflows — not vibes benchmarks.

Updated May 30, 2025
MCP 14 min

Building Your First MCP Server: A Practical Walkthrough

A hands-on tutorial for building a working MCP server in TypeScript in under an hour — with the debugging tips nobody puts in the README.

Updated Jun 12, 2026
Prompt Engineering 12 min

Prompt Engineering Patterns That Actually Work in Production

The prompt patterns we keep reaching for when reliability matters more than cleverness — with the failure modes each one solves.

Updated Jun 10, 2026
AI Agents 11 min

AI Agents vs Workflows: When to Use Which (and When Not To)

Everyone's shipping agents. Most of them should have been workflows. Here's how to tell the difference before you commit.

Updated Jun 8, 2026
Automation 13 min

n8n vs Make vs Zapier in 2026: Which Automation Platform Fits You?

A pragmatic comparison of the three dominant automation platforms — pricing, AI integrations, self-hosting, and where each one actually wins.

Updated Jun 12, 2026
engineering 11 min

Prompt Caching in Production: The Cheat Code Most Teams Miss

How to use Anthropic and OpenAI prompt caching to cut RAG and agent bills by 60-90% without changing model quality.

Updated Jun 28, 2026
engineering 12 min

Evals That Actually Catch Regressions

Most eval suites give teams false confidence. Here's how to build an eval harness that catches real regressions in production LLM workflows.

Updated Jun 27, 2026
engineering 10 min

Structured Outputs and Tool Use: The Practical Guide

How to get 99%+ valid JSON from OpenAI, Anthropic, and open-weights models — with the failure modes each still has.

Updated Jun 26, 2026
engineering 13 min

RAG That Doesn't Suck in 2026

Hybrid retrieval, reranking, evals, and when to just stuff everything into a long-context model instead.

Updated Jun 25, 2026
infrastructure 11 min

Self-Hosting LLMs in 2026: When It Actually Makes Sense

The real economics of running Llama 3.3, Mistral, and DeepSeek yourself — and the cases where it beats API providers.

Updated Jun 24, 2026
engineering 10 min

Agent Observability: Debugging LLM Workflows in Production

The tracing, replay, and evaluation patterns that turn opaque agent runs into debuggable systems.

Updated Jun 23, 2026
guides 9 min

How to Choose an LLM in 2026

A decision tree that maps common workflow requirements to the right model — GPT, Claude, Gemini, Llama, DeepSeek.

Updated Jun 22, 2026
guides 10 min

Building an AI Workflow That Actually Ships

A 6-week playbook from idea to production for internal AI workflows — the traps that stall most teams and how to avoid them.

Updated Jun 21, 2026
mcp 9 min

MCP Servers Actually Worth Installing in 2026

A curated list of Model Context Protocol servers that earn their place in a Claude Desktop or Cursor config — with the ones to skip.

Updated Jun 20, 2026
coding 9 min

Cursor vs GitHub Copilot in 2026: The Honest Comparison

Where each tool wins, where each falls short, and how to pick for your team based on how you actually work.

Updated Jun 19, 2026
engineering 11 min

The 8 Cost-Optimization Patterns Every LLM Team Should Know

Practical techniques to cut your OpenAI and Anthropic bill by 60-90% without hurting quality — with the numbers.

Updated Jun 18, 2026