Skip to content
ai tools

We Built Our Own AI Support Agent. Then We Fact-Checked Everyone Selling One.

We run our own AI agent on support@stork.ai, so we fact-checked every AI email and support tool we list against its own pricing pages and docs. What we found: 'per resolution' billing that counts customer silence as success, autonomy marketing that runs from 'no human in the loop' to 'human-in-control,' and an industry that is draft-first behind the scenes — plus the RFC 3834 mechanics almost nobody documents.

Nora Vance
We Built Our Own AI Support Agent. Then We Fact-Checked Everyone Selling One.

TL;DR / Key Takeaways

  • "AI email tool" is four different product categories — support agents, email clients, shared inboxes with AI assist, and agent email infrastructure — and they don't substitute for each other.
  • Intercom Fin's $0.99 'outcome' includes the customer going silent; Zendesk counts a resolution only after 72 quiet hours plus LLM verification. The definition matters more than the price.
  • Gorgias promises automation 'without a human in the loop'; Front sells 'human-in-control AI.' Every vendor's own docs recommend starting draft-first.
  • The documented disasters (Air Canada, Cursor's 'Sam') were invented policies stated confidently — not loops. Your agent's statements are legally your statements.
  • Only Zendesk documents mail-loop prevention in depth. RFC 3834 outbound stamping is the cheapest safety mechanism nobody markets.
  • Sierra, Decagon, and Zowie publish no pricing — any dollar figure you read for them is a guess.

Stork gets support email like everyone else: tool submissions, listing questions, the occasional angry crawler complaint. We built an agent for it instead of buying one — not because the market is empty, but because nothing in it is shaped like "a founder's own-brand support agent that knows your catalog and routes people into your actual funnels." Ours drafts replies for human review, retrieves from our own catalog and blog before it says anything, refuses to answer account questions over email, and carries hard circuit breakers so it can never get into an argument with an autoresponder.

Building those guardrails meant researching how the commercial agents handle the same problems. That research turned into this audit. Everything below comes from vendors' own pricing pages, docs, and engineering blogs — not review-site hearsay — checked on July 23, 2026. Where a vendor publishes nothing, we say so instead of guessing.

First: know which of four products you're actually buying

"AI email tool" is four different product categories wearing one label, and they don't substitute for each other. Buying across the boundary is the most common mistake we see — and, honestly, the mistake our own directory's auto-generated competitor lists used to make before we started curating this category by hand.

Category shapeWho it's forRepresentative tools
AI support agent — answers your customersSupport teams; the agent replies on behalf of the businessIntercom Fin, Zendesk AI, Gorgias, Sierra, Decagon, Zowie, Forethought, Chatbase, Tidio Lyro, Freddy AI
AI email client — manages your inboxIndividuals and teams drowning in their own mailShortwave, Superhuman, TabMail, SaneBox, Mailbutler
Shared inbox with AI assist — humans reply, AI helpsTeams that want drafting and triage, not autonomyFront, Hiver, Pylon, DelightChat
Email infrastructure for agents / deliverability / outreachDevelopers giving agents inboxes; senders protecting reputationAgentMail, MailReach, Instantly
Category shapes verified against each vendor's own positioning, July 2026. A support agent and an email client are different purchases with different buyers — comparison pages that mix them are answering a question nobody asked.

The pricing fine print is where the incentives live

Intercom Fin charges $0.99 per "outcome" with a minimum monthly commitment when run standalone. Read their definition: an outcome counts when a customer confirms resolution, when Fin completes a workflow — or when the customer simply doesn't ask for more help after Fin responds. Silence is billable. A customer who gave up counts the same as a customer who got their answer.

Zendesk sells the same unit priced honestly: an automated resolution counts only after the requester stays quiet for 72 hours, no human touched the ticket, and a language model verifies the answer was actually relevant. Same metric, stricter referee. If you're evaluating per-resolution pricing, the definition of "resolution" matters more than the price.

ToolPricing modelPublished figures (Jul 2026)
Intercom FinPer outcome$0.99/outcome, min. commitment (e.g. 50/mo) standalone; $9.99 per sales-qualification outcome
Zendesk AI agentsSeats + automated resolutionsSuite from $55/agent/mo annual; per-resolution rate not public
GorgiasTicket-volume tiers + AI usageStarter $40/mo (helpdesk $10 + AI agent $30); $1.50 per automated interaction over plan
FrontPer seat + per conversation for AI$25–$105/seat/mo annual; Autopilot from $0.05/conversation
PylonPer seat + flat AI add-ons$59–$139/seat/mo annual; AI agents from $100/mo
HiverPer seatFree; $25/$55/$85 per user/mo annual
Tidio LyroConversation-metered tiersStarter $24.17/mo; pay-per-resolution only on Premium
ChatbaseFlat tiers + message creditsFree; $32–$400/mo
Freshdesk Freddy Per seat + sessions$19–$89/agent/mo; 500 AI sessions included, then $49/100
SierraOutcome-basedNo public pricing — sales-led
DecagonPer conversation or resolutionNo public pricing — sales-led
ForethoughtPlatform fee + outcomesQuote-only
Figures from vendor pricing pages, July 23, 2026. Where a vendor publishes no price, any dollar figure you read elsewhere is somebody's guess — including, until recently, some of ours: we found and fixed mislabeled prices on our own listings during this audit.

Enjoying this? Get one like it in your inbox each morning.

one email a day · unsubscribe in two clicks · no third-party tracking

The autonomy spectrum: what vendors promise vs. what they gate

The marketing runs from one pole to the other. Gorgias promises automated returns and order tracking "all without a human in the loop". Freshworks pitches Freddy resolving "up to 80% of customer queries without human intervention". At the opposite pole, Front leads its pricing page with "human-in-control AI" — betting that buyers are now scared of exactly what the first group is selling.

The honest version usually lives one click deeper, in the docs. Gorgias's own playbook recommends starting with AI replies as internal notes for human review before letting it send. Zendesk gates auto-replies behind a numeric confidence threshold — default 60 of 100 — below which the agent escalates. Fin hands off automatically on self-harm, minors, jailbreak attempts, and high-risk medical, legal, or financial topics, no configuration asked. Even Instantly, a cold-outreach tool with every incentive toward volume, ships its reply agent with an explicit "Human-In-The-Loop mode… then switch to Autopilot" recommendation.

In other words: everyone selling autonomy operates draft-first internally. The gap between the landing page and the settings panel is the most useful thing we learned reading these docs.

Why the industry is draft-first behind the scenes

Two public incidents explain the caution. In February 2024 a Canadian tribunal made Air Canada honor a bereavement-fare policy its chatbot invented, rejecting the airline's argument that the bot was "a separate legal entity responsible for its own actions." In April 2025, Cursor's AI support agent invented a "one device per subscription" policy to explain a login bug — different users got different fabricated answers, and paying customers cancelled over a policy that never existed.

Note what failed in both cases. Not loops, not deliverability — the model confabulated policy to cover a knowledge gap, in the confident register of a support agent. Your agent's statements are legally your statements. That's the risk every confidence threshold, holdout category, and validation pass in this market is really pricing in.

What we built, and what none of the marketing mentions

Our agent is deliberately boring. It drafts; a human sends. It retrieves from our own catalog and published posts before answering, and if retrieval comes back empty it says less, not more. Account-specific questions never get answered in-channel — the reply is a link into the authenticated dashboard, because a From: header is not identity. And it carries deterministic circuit breakers: auto-reply detection on inbound, hard caps per thread, so it can never loop against an out-of-office.

The mechanics that matter most are the ones no vendor markets. RFC 3834 — the 2004 spec for automated email responders — says auto-replies must stamp themselves Auto-Submitted: auto-replied, must never answer a message that carries that header, and must never respond to a null return-path. Among every vendor we audited, only Zendesk documents its mail-loop handling in any depth — including a rate breaker that suspends a sender after 20 emails in an hour, and the frank admission that no system prevents all loops. Everyone else's docs are silent on the single oldest failure mode of automated email.

  • If you're buying: match the category shape first, then ask four questions the sales deck won't answer. What exactly counts as a billable resolution? What categories can never be auto-answered? How does the agent verify who it's talking to before disclosing account data? And does it stamp RFC 3834 headers on its own outbound mail?
  • If you're building: draft-first isn't a training-wheels phase, it's the industry's actual operating posture. Add an intent allowlist before any auto-send, hard holdouts for money and legal topics, and a validation pass that blocks any reply stating a price or policy the retrieval layer can't literally show.
  • Either way: measure resolutions the strict way — no human touched it, the requester stayed quiet for 72 hours, and a spot-check says the answer was actually right. The generous definition will flatter your dashboard and hide your failures.

We'll keep the listings in this category fact-checked against primary sources — it's the same pipeline that keeps our own agent's answers grounded. If you ship an AI email tool and your pricing or docs change, claim your listing and tell us; corrections beat speculation.

Found this useful? Share it.

For builders

Want Stork to write one of these about your product?

Send us a URL. We use the product, form a view, and publish what we actually think — in 8 languages, labeled Sponsored, with no copy approval on your side. That last part is what makes it worth quoting.

See how it works$500 · AI tools & software only