One-for-All Multimodal Evaluation Toolkit Across Text, Image, Video, and Audio Tasks
-
Updated
Jul 23, 2026 - Python
One-for-All Multimodal Evaluation Toolkit Across Text, Image, Video, and Audio Tasks
Versatile Evaluation of Speech and Audio
Benten is an audio evaluation and analytics platform built specifically for Voice AI agents. It connects directly to voice AI platforms to discover your agents, process call audio, and measure conversation quality, response latency, and speech dynamics.
Extract grounded evidence from video files to enable automated review and visual understanding for coding agents.
Comparative analysis and benchmarking of open-source TTS models for Malay/Indonesian/English, with WER, speaker-similarity, VRAM, and prosody evaluation.
A standalone tool for evaluating Automatic Speech Recognition (ASR) models, particularly optimized for medical/clinical speech recognition, using Word Error Rate (WER) metric
Professional portfolio for AI Evaluation, Audio Analysis, and UX Research. Specialized in Human-in-the-Loop (HITL) data integrity and design-focused research.
Add a description, image, and links to the audio-evaluation topic page so that developers can more easily learn about it.
To associate your repository with the audio-evaluation topic, visit your repo's landing page and select "manage topics."