Every Search API passed. That was the useful result.
All three Search APIs passed 27 of 27 retrieval attempts. The tie exposed why pass rates are guardrails, not product rankings.
Blog
Benchmarks, integration notes, and practical analysis of APIs for coding agents.
All three Search APIs passed 27 of 27 retrieval attempts. The tie exposed why pass rates are guardrails, not product rankings.
A technical walkthrough of the protocol, normalization layer, evaluator, artifacts, and compromises behind our Search API benchmark.
Date coverage varied from 33% to 100% in our Search API run. Here is why that matters for some agent tasks and misleads in others.
Exa, Perplexity Search, and Parallel Search all passed our retrieval test. The useful differences appeared in latency, metadata, onboarding, and integration surface.
An agent-ready API reduces discovery, authentication, and verification work. Here are the signals we inspect before trusting an integration.
A reproducible email API benchmark needs a fixed task, fresh workspaces, one send attempt, provider-native evidence, and honest blocked outcomes.
Mailgun and SendGrid both send transactional email, but differ in authentication, payload shape, sender setup, test mode, and acceptance evidence.
Resend and Postmark both expose straightforward send APIs, but their test addresses, test tokens, sandbox servers, responses, and idempotency differ.