Blog | VoicePing Skip to main content

Blog

Insights, tips, and updates from VoicePing

Offline Speech Transcription Benchmark: 16 Models Across Android, iOS, macOS, and Windows
Speech Recognition On-Device AI

Offline Speech Transcription Benchmark: 16 Models Across Android, iOS, macOS, and Windows

Comprehensive benchmark of 16 on-device speech-to-text models across 9 inference engines on Android, iOS, macOS, and Windows

Akinori Nakajima - VoicePing
9 min
Offline Text-to-Speech Benchmark: 18 Models Across Android and iOS
Text-to-Speech On-Device AI

Offline Text-to-Speech Benchmark: 18 Models Across Android and iOS

Comprehensive benchmark of 18 on-device text-to-speech models including Kokoro, Piper, Matcha, Kitten, and VITS on Android and iOS

Akinori Nakajima - VoicePing
6 min
VoicePing Product Update — February 2026: 14 Languages, Bilingual Mode, and Searchable Transcripts
AI Translation

VoicePing Product Update — February 2026: 14 Languages, Bilingual Mode, and Searchable Transcripts

VoicePing now supports 14 languages with bilingual display, rebuilt translation engine, searchable transcripts, custom phrase training, and downloadable subtitles and dubbed audio.

VoicePing Editorial
6 min
First Real-Time AI Translation Implementation at Mint and Print International Conference 2025
Events AI Translation

First Real-Time AI Translation Implementation at Mint and Print International Conference 2025

The Mint and Print International Conference is a biennial international conference where central banks from around the world gather. At the 2025 conference, VoicePing's real-time AI translation service was adopted to overcome language barriers, enabling approximately 200 participants from over 50 central banks to understand presentations in their native languages.

VoicePing Editorial
8 min
Whisper in Production: Real-Time Dual-Language Switching, the Failures We Hit, and the Architecture That Works
ASR Whisper

Whisper in Production: Real-Time Dual-Language Switching, the Failures We Hit, and the Architecture That Works

How VoicePing engineered Bilingual Mode for automatic, low-latency language switching inside a single WebSocket stream powered by customized Whisper V2 models.

Akira Noda - VoicePing
9 min
Evaluating Speaker Diarization Models: A Practical Comparison
Speaker Diarization NeMo

Evaluating Speaker Diarization Models: A Practical Comparison

Technical comparison of NeMo MSDD and Pyannote 3.1 across 6 real-world test scenarios

Ashar Mirza - VoicePing
4 min

Try VoicePing for Free

Experience communication beyond language barriers with real-time voice translation

Get Started Free