Media
Detect likely speech-range activity and export speech/non-speech ranges as JSON, CSV, SRT markers, or an FFmpeg cut list.
Call this tool from your code in three languages.
# 1) Request a presigned URL → returns { uploadUrl, storageKey }
curl -X POST 'https://api.elysiatools.com/api/upload/presign/audio-voice-activity-segmentation' \
-H 'Content-Type: application/json' \
-d '{"filename":"audioFile.ext","contentType":"application/octet-stream","size":12345}'
# 2) PUT the file bytes directly to the presigned uploadUrl
curl -X PUT '<presigned uploadUrl>' \
--data-binary @/path/to/file.ext
# 3) Call the tool, passing the returned storageKey for each file field
curl -X POST 'https://api.elysiatools.com/en/api/tools/audio-voice-activity-segmentation' \
-F 'audioFile=uploads/2026/01/01/your-tool-1700000000000-abc123.ext' \
-F 'sensitivity=balanced' \
-F 'minimumSpeechSeconds=0.3' \
-F 'minimumSilenceSeconds=0.25' \
-F 'exportFormat=json'Send a POST request with your inputs as JSON. File parameters require a separate upload first.
POST https://api.elysiatools.com/en/api/tools/audio-voice-activity-segmentation| Name | Type | Required | Description |
|---|---|---|---|
| audioFile | fileupload required | Yes | — |
| sensitivity | select | Yes | — |
| minimumSpeechSeconds | number | No | — |
| minimumSilenceSeconds | number | No | — |
| exportFormat | select | Yes | — |
File result
{
"filePath": "/public/processing/randomid.ext",
"fileName": "output.ext",
"contentType": "application/octet-stream",
"size": 1024,
"metadata": {
"key": "value"
},
"error": "Error message (optional)",
"message": "Notification message (optional)"
}Add this tool to your Model Context Protocol server so AI agents can list and call it.
Add this block to your MCP client configuration:
{
"mcpServers": {
"elysiatools-audio-voice-activity-segmentation": {
"name": "audio-voice-activity-segmentation",
"description": "Detect likely speech-range activity and export speech/non-speech ranges as JSON, CSV, SRT markers, or an FFmpeg cut list.",
"baseUrl": "https://api.elysiatools.com/mcp/sse?toolId=audio-voice-activity-segmentation",
"command": "",
"args": [],
"env": {},
"isActive": true,
"type": "sse"
}
}
}After connecting to the SSE endpoint, list the exposed tools:
{
"jsonrpc": "2.0",
"id": 1,
"method": "tools/list"
}Invoke the tool by its id, passing arguments built from its parameters:
{
"jsonrpc": "2.0",
"id": 2,
"method": "tools/call",
"params": {
"name": "audio-voice-activity-segmentation",
"arguments": {
"audioFile": "https://example.com/file.ext",
"sensitivity": "balanced",
"minimumSpeechSeconds": 0.3,
"minimumSilenceSeconds": 0.25,
"exportFormat": "json"
}
}
}Questions or issues? Contact [email protected]