A light estimate, computed — not measured, not edited
Every Syftly answer carries a confidence label. Today it reads light estimate: the verdict is aggregated from public benchmarks — with sources and dates — and computed from structured fields, not produced by first-hand Syftly measurement and not chosen by an editor. Here is exactly how it works, and where it is weak.
- 5categories — Transcription, Text-to-speech, Web search, Scraping & browser, OCR & document extraction
- 31provider offerings, each a callable API + model-tier
- 33published queries, curated, hand-written & indexable
- 33computed decision axes the engine maps questions onto
Light estimate rests on public research and benchmarks (with attribution) plus AI-as-a-judge to synthesise. It is broad and useful, but always labelled as light. Hard tested would rest only on first-hand measured facts — latency, price and uptime probed by Syftly, accuracy scored against a verified ground truth. That tier does not exist yet; nothing here is presented as hard tested. AI-as-a-judge never counts as hard.
Each category is produced by a fixed, version-controlled research recipe: a source hierarchy (Tier-1 provider docs and recognised benchmarks carry the ranking; marketing never does), then a two-phase pipeline — deterministic extraction of the hard fields, then judged synthesis on top of that grounded data. The winner on each axis is then computed from those structured fields — “cheapest” is literally a min() over the prices. So a new recipe run changes the numbers and the winners recompute by themselves. Credibility comes from provenance — a dated source plus a confidence label plus attribution — not from human approval.
A free question is mapped — deterministically, with no LLM call — onto a decision axis; its winner is computed from the ranking. Every category declares its own axes; the seven below (transcription) are the worked example. If nothing matches, the answer falls back to the category default (the top of the ordered ranking). An in-category question never 404s.
| Axis | Computed by |
|---|---|
| Cheapest | lowest directly-comparable per-minute price |
| Most accurate | lowest word error rate (WER) |
| Most multilingual | highest supported-language count |
| Lowest latency | lowest published streaming latency |
| Best price-to-accuracy | lowest price × WER |
| Capability filter | top-ranked offering that has diarization, word-timestamps or custom vocabulary |
| Language | most accurate offering that supports the asked language (e.g. Dutch, Spanish) |
- WER is English-leaning.Transcription accuracy uses the Artificial Analysis aggregate WER, which weights English heavily. A “most accurate for Dutch” verdict is a light estimate, not a Dutch-specific measurement.
- Token-priced models are excluded from “cheapest”. An LLM-based transcriber billed per input-audio token (e.g. Gemini) understates real cost, so it is not directly comparable to per-minute pricing and is kept out of the price axis.
- Latency figures are heterogeneous. Some are independent P50 numbers, others vendor claims; they are not strictly comparable, so a latency verdict carries that caveat.
- Only curated answers are indexable. Hand-written published queries are crawlable; on-demand engine answers for the long tail are served
noindexso the site never fills with thin near-duplicate pages.
Every ranking carries its own dated sources; this is the union across all 5 categories.
- Artificial Analysis — Speech to Text2026-06-19
- Open ASR Leaderboard2026-06-19
- Artificial Analysis — Text to Speech Leaderboard (Speech Arena, blind-vote ELO)2026-06-22
- Coval — Best Text-to-Speech Providers in 2026 (independent TTFA/TTFB benchmark, captured 2026-05-04)2026-06-01
- ElevenLabs — API Pricing2026-06-22
- Cartesia — Pricing2026-06-22
- Google — Gemini Developer API Pricing2026-06-22
- OpenAI — API Pricing2026-06-22
- MiniMax — Product Pricing (API docs)2026-06-22
- Deepgram — Pricing (Aura-2 TTS)2026-06-22
- Brave Search API — Pricing (provider docs)2026-06-22
- Exa — API Pricing (provider docs)2026-06-22
- Tavily — Credits & Pricing (provider docs)2026-06-22
- Linkup — Pricing (provider docs)2026-06-22
- Perplexity — API Pricing (provider docs)2026-06-22
- SerpApi — Plans and Pricing (provider docs)2026-06-22
- AImultiple — Agentic Search: Benchmark 8 Search APIs for Agents (independent eval, updated 2026-05-25)2026-06-22
- Nebius announces agreement to acquire Tavily2026-02-10
- Bright Data — Web Unlocker pricing2026-06-22
- Zyte API pricing — Zyte documentation2026-06-22
- Firecrawl pricing2026-06-22
- ScrapingBee pricing2026-06-22
- ScraperAPI pricing2026-06-22
- Apify pricing2026-06-22
- Browserbase — Plans and Pricing (docs)2026-06-22
- Proxyway — Web Scraping API Report 2025 (independent benchmark)2026-06-22
- Scrapeway — ScrapingBee vs Firecrawl benchmark (run 13-19 June 2026)2026-06-22
- ScrapeCreators — Best Web Scraping APIs 2026 (aggregator, updated 31 May 2026)2026-06-22
- Mistral Pricing2026-06-22
- Introducing Mistral OCR 32026-06-22
- CodeSOTA — OmniDocBench Leaderboard2026-06-22
- OmniDocBench (CVPR 2025) — opendatalab2026-06-22
- Google Cloud Document AI Pricing2026-06-22
- AWS Textract Pricing2026-06-22
- Azure Document Intelligence Pricing2026-06-22
- Reducto Pricing2026-06-22
- LlamaParse / LlamaIndex Pricing2026-06-22