New · Speaker diarization on STT

Audio AI infrastructure
transcription + synthesis in one API.

Speech-to-Text in 60+ languages at 130× real-time, multilingual Text-to-Speech, and tier-1 Voice Cloning. Same cluster, same transparent pricing, one unified balance.

orchard.ts
TypeScript
import OpenAI from "openai";
 
const orchard = new OpenAI({
apiKey: process.env.ORCHARD_API_KEY,
baseURL: "https://api.orchardrun.com/v1",
});
 
// 1. Transcribe · 60+ languages · 130× real-time
const transcript = await orchard.audio.transcriptions.create({ file });
 
// 2. Synthesize · 17 languages · latency under 2 s
const audio = await orchard.audio.speech.create({
input: "Welcome back, your order is on its way.",
voice: "claribel",
});
 
// 3. Clone a voice from 10 s of reference audio
const voice = await fetch("https://api.orchardrun.com/v1/voices", { method: "POST", body });

Drop into any TypeScript / Node app · Python SDK identical

Your voice. Any language. In milliseconds.

Speech to Text
AUDIO → SPEAKERS
AUDIOIN
diarize
SEGMENTS · SPEAKERSOUT
Hit play — each speaker labelled live
Text to Speech
TEXT → VOICE
TEXTIN

Voice synthesis at production scale. One API, one balance, seventeen languages.

synthesize
VOICE · ClaribelOUT
Clone Voice
VOICE → SAME VOICE
ORIGINAL VOICEIN
Mateo Bustamante, Orchard co-founder, voice reference for the cloning demo
Mateo
co-founder · 8s ref
clone
CLONED VOICEOUT
Mateo Bustamante, Orchard co-founder, voice reference for the cloning demo
Esta es mi voz. Y ahora puedo decir cualquier cosa, en mi …

Audio rendered with Orchard · pre-cached, instant playback

THREE PRODUCTS · ONE STACK

Audio AI infrastructure ready for production.

Transcription + synthesis + voice cloning. Same API, same billing, same owned cluster.

STT · IN YOUR EDITOR

Orchard Dictate

VS CODE

Press Cmd+Shift+8, speak, paste at cursor. Works in Cursor, Claude Code, Copilot — and any editor.

Install on Marketplace
8
USE CASES · WHAT DEVS BUILD ON ORCHARD

Where Orchard fits.

Six real workflows where Orchard replaces expensive transcription and synthesis APIs, or fragmented service stacks — all on one API.

Workflow 01

Conversational AI agents

Transcribe WhatsApp, Telegram or live call audio and feed the context to your LLM. Low latency, controlled cost per minute.

Workflow 02

Bulk audio processing

Transcribe podcasts, meetings, interviews or thousands of files a day. No rate limits on paid plans, no per-file caps.

Workflow 03

Call analysis & support

Turn customer calls into text + automatic insights for your team. Speaker diarization, ideal for call centers and QA.

Workflow 04

Voice cloning for marketing

Clone real voices for videos, ads and automations. Same consistent voice across hundreds of assets.

Workflow 05

Automated pipelines

Audio → Transcript → Summary → Action with your LLM of choice. Webhooks, retries and batching native to the API.

Workflow 06

Custom voice assistants

Build your own conversational assistant or branded voice agent. Text-to-speech in the voice you define, in the language you need.

LIVE ARCHITECTURE
STT · TTS · Clone · LLM
Loading workflow…

Voice notes in seconds. Podcasts in minutes.

60× real-time average sustained. 1 hour of audio in under 1 minute.

10× cheaper than the competition

$0.00042/min on Pro plan. Simple plans, no surprise costs.

Drop-in replacement

Industry-standard API. Existing SDKs work without changes — migrate in minutes.

Global community
+1,584users
+29countries

Builders, indie hackers and audio-first teams across Latin America, the US and Europe ship faster on Orchard's pay-as-you-go speech stack.

🇦🇷AR🇮🇳IN🇺🇸US🇩🇪DE🇬🇧GB🇨🇦CA🇵🇱PL🇺🇦UA🇪🇸ES🇧🇷BR🇦🇪AE🇿🇦ZA🇮🇱IL🇷🇴RO🇳🇱NL🇰🇷KR🇮🇹IT🇫🇷FR🇲🇽MX🇨🇴CO🇨🇱CL🇵🇪PE🇹🇷TR🇯🇵JP🇦🇺AU
60+STT languages
broadest speech-recognition coverage
17TTS · clone langs
multilingual synthesis + voice clone
4products · 1 bill
STT · TTS · Dictate + free Clone beta
60×real-time STT
1 hour of audio in under a minute
<5%WER
competitive with best-in-segment
0daily caps
monthly quota, roll your own pace
1API key
OpenAI Whisper drop-in
60+STT languages
broadest speech-recognition coverage
17TTS · clone langs
multilingual synthesis + voice clone
4products · 1 bill
STT · TTS · Dictate + free Clone beta
60×real-time STT
1 hour of audio in under a minute
<5%WER
competitive with best-in-segment
0daily caps
monthly quota, roll your own pace
1API key
OpenAI Whisper drop-in
Pricing

Simple plans, no surprises.

Start free · Upgrade as your volume grows

Free
$0

500 min on signup

≈ 500K chars TTS · 1 cloned voice

Try all 3 products. No card required.

  • 500 minutes on signup
  • Automatic monthly refill
  • 1 concurrent request
  • Community support
Hobby
$1/month

1,500 min/month

≈ 1.5M chars TTS · 3 cloned voices

Coffee-money tier. STT + TTS + Clone Voice share the balance.

  • 500 minutes on signup
  • Webhooks + SRT/VTT
  • 1 concurrent request
  • Billed annually
Popular
Starter
$10/month

15,000 min/month

≈ 15M chars TTS · 10 cloned voices

Bots and small SaaS. All 3 products on shared balance.

  • Webhooks + SRT/VTT
  • 3 concurrent requests
  • Email support
  • Speaker diarization
Pro
$25/month

60,000 min/month

≈ 60M chars TTS · 50 cloned voices

Production volume. All 3 products on shared balance.

  • Priority queue
  • 10 concurrent requests
  • Python + Node SDK
  • Speaker diarization

Optional diarization · Custom SLA · Dedicated capacity

FAQ

Questions builders ask before integrating.

Our API mirrors the OpenAI Whisper request/response format, so most SDKs work without code changes. Point them at our endpoint and swap the key. Median migration is under an hour.

Stop stitching three vendors.
Ship voice on one API.

500 free minutes on signup. No card. Same balance across STT and TTS. Pay once, use everything.

About us

Audio infrastructure at scale.

We founded Orchard with one purpose: to reshape the Voice Infrastructure industry. As heavy consumers ourselves, we kept hitting the same gaps in the market — exactly where we decided to differentiate: price, volume and concurrency. That's why we built three core verticals: STT, TTS and Voice Cloning.

Our strongest surface today is STT batch — and we're going for the global #1 spot. We back it up with three hard numbers: the cheapest minute on the market, a WER competitive with the best engines in the segment, and an RTF that sustains high volume and massive concurrency without throttling. That combination of quality, speed and price isn't on offer anywhere else.

In parallel, our TTS is consolidating as the default base for voice agents, voice assistants and conversational products — a segment growing double digits as every product turns voice-first.

Voice Cloning is the bet we're most excited about for what's next. It already works great for the current use cases, and where we're investing heavily is the pipeline: capturing prosody, rhythm and emotion with a precision that generic voice cloning will never reach. The goal: when a customer uploads 30 seconds of audio, the model doesn't just reproduce the timbre — it replicates the way they speak, not just the voice they have.

Mateo Bustamante, Co-founder at Orchard
Mateo Bustamante
Co-founder

Processing audio at scale breaks you: pipelines that don't scale, latencies that kill the product, costs that crush you exactly when you grow. The next wave of software is going to be voice-first and agentic. And we want to be the infrastructure it stands on, not the bottleneck.

Ramiro Alvarez, Co-founder at Orchard
Ramiro Alvarez
Co-founder

We're aiming to be the world #1 in STT batch. Today we combine the lowest cost on the market with quality on par with the leaders, and we're investing heavily in voice cloning for LATAM, where capturing real accents is still an unsolved problem.

Security & Compliance

Built for enterprise from day one

Encryption, audit logs, and multi-region replication are table stakes. Watermarking and consent for cloned voices — that's our line.

GDPR Compliant

DPA available, full data deletion on request

CCPA Compliant

California consumer rights honored end-to-end

AES-256 at Rest

All stored audio and transcripts encrypted

TLS 1.2+ in Transit

Modern cipher suites only, no downgrade

Multi-region Replication

Database tier replicated with point-in-time recovery

RBAC for Org Seats

Role-based access control across organization seats

unique

EU AI Act Ready

Consent framework + watermarking on roadmap

in progress

SOC 2 Type II

In progress — targeting Q4 2026

DPA·Privacy·Terms·Request security report