FasterWhisper

Speech transcription engine used by the Speaches runtime.

Agentic Friendly

Component Category

Inference / speech-to-text

Component Description

FasterWhisper is an optimized implementation of Whisper for efficient speech transcription workloads. BullSequana AI runs it inside the backend-owned Speaches service rather than as an independently deployed component.

Why It Is Used

In BullSequana AI Foundation, FasterWhisper supports audio transcription use cases with lower latency and lower resource consumption than a baseline Whisper deployment.

Learn More

Deployment notes

FasterWhisper runs inside the Speaches service. Speech-to-text is exposed through the authenticated BSQAI API, which supports batch transcription and real-time WebSocket sessions. Operators select CPU or GPU execution and the corresponding Whisper model through the AI service configuration.

Interacts With

  • BSQAI API, which exposes batch and real-time speech-to-text operations.
  • Speaches, which hosts FasterWhisper and the platform's speech-generation engines.
  • AI Web Portal, which uses speech services for voice dictation.

On this page