chat_with_audio
Interactive audio analysis using GPT-4o audio models:
How to use it
chat_with_audio is exposed by the MCP Server Whisper MCP server. Add the server to your MCP client (Claude Desktop, Cursor, Windsurf and others), and the chat_with_audio tool becomes available to the model automatically. See the full listing for setup details and every tool this server provides.
Install MCP Server Whisper
bunx dotenv-cli -- claudeOther tools in MCP Server Whisper (12)
Adds analysis of speech patterns and key points
Compresses audio files that exceed size limits
Converts audio files to supported formats (mp3 or wav)
Generate text-to-speech audio using OpenAI's TTS API:
Includes tone, emotion, and background details
Gets the most recently modified audio file with model support info
Lists audio files with comprehensive filtering and sorting options:
Creates formal, business-appropriate transcriptions
Transforms the transcript into a narrative form
Advanced transcription using OpenAI's models:
Enhanced transcription with specialized templates: