Skip to content
@Oruk-AI

Oruk AI

Speech models and APIs for transcription and vocal expression. Resonance-2 Preview, open Orukeet speech recognition, and reproducible research.

oruk

Oruk is a speech lab. We build models for transcription, emotion and speaking-style analysis, with a hosted API and open research.

Speech API

Resonance analyzes prerecorded English audio. One request can return the transcript, 15 emotion scores, 16 speaking-style scores and timed segments. Scores are independent model outputs, not calibrated probabilities of a person's inner state.

curl https://speech-api.oruk.ai/v1/audio/analysis \
  -H "Authorization: Bearer $ORUK_API_KEY" \
  -F file=@call.wav -F model=oruk-resonance

The file API also has separate transcription, emotion and speaking-style endpoints. The Realtime preview streams transcription in 32 locales and phrase-level emotion over WebSocket. Original Resonance’s full emotion and speaking-style analysis uses the English file endpoints.

Resonance-2 Preview is a separate clip-level emotion and speaking-style API. It returns all 31 continuous scores, six signed axes and selected labels that can be empty. Send 0.1–120 seconds of audio, up to 30 MiB, to /v1/audio/resonance-2 using your existing API key and shared speech allowance. This route does not transcribe or diarize audio. SDK 0.2.10 users call it through ordinary HTTP.

Listen to six recorded examples, including mistakes, and inspect the model/calibration revisions and all 546 predictions from the September 17 acted-emotion diagnostic. Training overlap has not been audited; the diagnostic is not an independent held-out evaluation.

Plans include audio minutes, measured by the second. Standard self-serve plans have a seven-day trial that requires a card, charges $0 today and can be canceled before the trial ends. Promotional offers have their own terms. See current plans and allowances.

Orukeet: local speech recognition

Orukeet is a 25-language speech recognizer built from NVIDIA Parakeet TDT 0.6B v3. It has NeMo, ONNX and native inference exports.

Start with the tested local-transcription tutorial, including native Q8 and sherpa-onnx CPU examples, pinned artifacts and recorded outputs. The model card documents the weights, licenses and evaluation limitations. OpenWhispr 1.10.0 includes Orukeet as its recommended local model.

Benchmarks and research

oruk-bench contains the published speech-emotion evaluation toolkit and results. The benchmark's seven-class mapping and multilingual dataset describe its evaluation protocol, not the hosted API's native outputs or language support. Oruk's historical model was evaluated in-distribution; those results do not establish the current API's accuracy on independent audio. Read the results and methodology together.

Research articles · RSS · API scope · Contact

Popular repositories Loading

  1. orukeet orukeet Public

    Orukeet: multilingual ASR with fitted, frozen Gabor kernels and native inference

    Python 132 77

  2. oruk-bench oruk-bench Public

    Speech-emotion benchmark: 64k held-out clips, 7 classes, ~20 languages, one scoring protocol for every model

    Python 1

  3. .github .github Public

    oruk organization profile

  4. oruk-ai.github.io oruk-ai.github.io Public

    oruk — speech AI landing (GitHub Pages)

    HTML

  5. oruk-mcp oruk-mcp Public

    Hosted MCP tools for Oruk transcription, emotion and speaking-style analysis. Python/TypeScript SDKs and the separate Resonance-2 REST route are documented in the README.

  6. openwhispr openwhispr Public

    Forked from OpenWhispr/openwhispr

    Voice-to-text dictation app with local (Nvidia Parakeet/Whisper) and cloud models (BYOK). Privacy-first and available cross-platform.

    JavaScript 1

Repositories

Showing 9 of 9 repositories
  • FluidAudio Public Forked from FluidInference/FluidAudio

    Frontier CoreML audio models in your apps — text-to-speech, speech-to-text, voice activity detection, and speaker diarization. In Swift, powered by SOTA open source.

    Oruk-AI/FluidAudio's past year of commit activity
    Swift 0 Apache-2.0 453 0 1 Updated Sep 27, 2026
  • openwhispr Public Forked from OpenWhispr/openwhispr

    Voice-to-text dictation app with local (Nvidia Parakeet/Whisper) and cloud models (BYOK). Privacy-first and available cross-platform.

    Oruk-AI/openwhispr's past year of commit activity
    JavaScript 0 MIT 1,093 0 0 Updated Sep 25, 2026
  • orukeet Public

    Orukeet: multilingual ASR with fitted, frozen Gabor kernels and native inference

    Oruk-AI/orukeet's past year of commit activity
    Python 132 MIT 77 0 2 Updated Sep 24, 2026
  • oruk-mcp Public

    Hosted MCP tools for Oruk transcription, emotion and speaking-style analysis. Python/TypeScript SDKs and the separate Resonance-2 REST route are documented in the README.

    Oruk-AI/oruk-mcp's past year of commit activity
    0 0 0 0 Updated Sep 19, 2026
  • .github Public

    oruk organization profile

    Oruk-AI/.github's past year of commit activity
    0 0 0 0 Updated Sep 19, 2026
  • pipecat-oruk Public

    Oruk-maintained streaming speech and phrase-emotion integration for Pipecat, with a verified WebRTC demo.

    Oruk-AI/pipecat-oruk's past year of commit activity
    Python 0 MIT 0 0 0 Updated Sep 18, 2026
  • pipecat-docs Public Forked from pipecat-ai/docs

    Official Pipecat docs repo

    Oruk-AI/pipecat-docs's past year of commit activity
    MDX 0 BSD-2-Clause 120 0 0 Updated Sep 12, 2026
  • oruk-ai.github.io Public

    oruk — speech AI landing (GitHub Pages)

    Oruk-AI/oruk-ai.github.io's past year of commit activity
    HTML 0 0 0 0 Updated Sep 11, 2026
  • oruk-bench Public

    Speech-emotion benchmark: 64k held-out clips, 7 classes, ~20 languages, one scoring protocol for every model

    Oruk-AI/oruk-bench's past year of commit activity
    Python 0 Apache-2.0 1 0 1 Updated Sep 6, 2026

People

This organization has no public members. You must be a member to see who’s a part of this organization.