Orukeet: a new shape for speech recognition
An open speech model for 25 languages, with fitted Gabor filters.
Speech understanding models for transcription and vocal delivery. Choose a model for recorded audio or live speech, with language support and outputs that vary by model.
Let your agent pick up on hesitation, frustration, or excitement, then use that context to shape its response.
“I’m not sure what to do next.”
Illustrative workflow
An open speech model for 25 languages, with fitted Gabor filters.
Small differences in speech acquire meaning within communities.
A faster Orukeet offshoot, with the accuracy tradeoff measured.
Past questions help a quantum memory choose which qubits to connect.
A speech model built from 499 fruit fly neurons.
Six speech systems learn the same filter shape, predicted by signal theory.
Speech models organize voices into shapes you can explore in 3-D.
Emotion structure transfers across languages, with a measurable gap.
Pitch, texture and timing contribute to overlapping impressions of a voice.
Testing spoken questions alongside handwritten math.
Prices in USD. Included minutes reset monthly.
$9/ month
Billed monthly
7-day free trial
Start 7-day free trial of Hobby$0 today. Then $9/month. Cancel before the trial ends.
$49/ month
Billed monthly
7-day free trial
Start 7-day free trial of Builder$0 today. Then $49/month. Cancel before the trial ends.
$199/ month
Billed monthly
7-day free trial
Start 7-day free trial of Production$0 today. Then $199/month. Cancel before the trial ends.
Contact us
Tailored to your team
Resonance 1, Fourier, Realtime, and Proficiency share the speech understanding allowance. Orukeet transcription has a separate allowance; optional tasks and request rounding reduce its estimated minutes. Allowances reset monthly. Trials include 25%; extra usage shares one spending cap.
Now serving: Resonance 2 — emotion and style on existing plans; new organizations need separate approval.
Oruk at the top of the evaluations below.
Emotion accuracy ↑
91%
1stby scoreEmotion accuracy ↑
48%
1stby score7-class accuracy ↑
77.8%
1stby scoreMacro-F1 × 100 ↑
56.2
1stby scoreMUStARD · standard 5-fold
7-class accuracy ↑
49.3%
1stby scoreOruk trained on all 280 clips; not held out.
Bring us your audio and your use case. We’ll help you find the right model and integration.
Send audio. Get a transcript, speaker turns, and vocal context.
Oruk turns speech into transcripts, speaker turns, and labels for emotion and delivery. Choose a model for recorded audio or live conversations.
Speech-to-text tells you what was said. Oruk also analyzes how it sounded: frustrated, excited, hesitant, sarcastic, and more.
Voice agents with more context, searchable call reviews, expressive captions, research tools, and other products that work with speech.
Yes. Open the full demo to use your microphone or upload a short recording. You can try it without an account.
Try your own audioUse Resonance 1 for a complete English recording: it returns a transcript, emotion and speaking-style labels, and timed segments. Use Realtime for live multilingual transcription with phrase-level emotion. Realtime is in preview and does not return the full file-analysis label set. Orukeet is an option for short English recordings or streamed utterances, with final text after commit.
Spectra-2 transcribes files in 25 languages; Realtime supports 32 locales. Original Resonance 1 and Fourier support English in WAV, FLAC, MP3, M4A, OGG, or WebM. Formats and limits vary by model. Transcription coverage does not establish emotion accuracy.
Resonance 2 buffers audio in memory until reset or session close and pending work finishes. Stored results exclude audio; 24-hour authorized replay is not a deletion deadline. Optional diarization keeps audio up to 48h and results up to 24h. Metadata is kept for billing, security and support. No training, fine-tuning or evaluation without explicit written agreement.
Data handlingPerformance varies with the model, language, recording, and task. The benchmark report includes dated results, confidence intervals, and evaluation conditions. Emotion scores describe vocal expression, not someone’s inner state.
Read the evaluationsPlans start at $9/month with a 7-day free trial. New organizations need separate approval for Resonance 2.
See pricing