MOSI · ASR
MOSS Transcribe 1.0
General speech transcription optimized for clear, structured text output.
moss-transcribe-1.0Current model-specific price from our catalog. Requests require your project API key and sufficient balance. Availability is checked at execution time.
Input and capabilities
Upload a short audio or video file using multipart/form-data: at most 30 seconds and 20 MiB. Use WAV for the examples below. Container compatibility depends on the model.
- transcription
Automatic language detection: supported.
No language-specific selection is declared for this model.
API endpoint
POST https://api.speechinfra.com/v1/inference/moss-transcribe-1.0
Authenticate with your Speechinfra project key. Read the quickstart and billing guide.
cURL
curl --fail-with-body "https://api.speechinfra.com/v1/inference/moss-transcribe-1.0" \
-H "Authorization: Bearer $SPEECH_API_KEY" \
-F "file=@clip.wav" \
--form-string "response_format=json"Python
# pip install httpx
import os
import httpx
headers = {"Authorization": "Bearer " + os.environ["SPEECH_API_KEY"]}
with open("clip.wav", "rb") as audio:
response = httpx.post(
"https://api.speechinfra.com/v1/inference/moss-transcribe-1.0",
headers=headers, files={"file": ("clip.wav", audio, "audio/wav")},
data={
"response_format": "json"
}, timeout=90,
)
response.raise_for_status()
print(response.json())Node.js / fetch
// Node.js 22+, no SDK required.
import { readFile } from 'node:fs/promises';
const form = new FormData();
form.set('file', new Blob([await readFile('clip.wav')]), 'clip.wav');
for (const [key, value] of Object.entries({
"response_format": "json"
})) form.set(key, value);
const response = await fetch("https://api.speechinfra.com/v1/inference/moss-transcribe-1.0", {
method: 'POST', body: form, signal: AbortSignal.timeout(90000),
headers: { Authorization: 'Bearer ' + process.env.SPEECH_API_KEY },
});
if (!response.ok) throw new Error(await response.text());
console.log(await response.json());Request parameters
Generated from the API schema for this deployment. Conditional parameters apply only when their controlling setting is enabled.
| Parameter | Type | Details |
|---|---|---|
fileRequired | string | Short audio or video file. Direct API limit: 30 seconds, 20 MiB. |
response_format | string | Return a JSON transcript or plain text. For speakers and keyterms, choose MOSS Diarize Pro. enum: json, text · default: "json" All accepted values[ "json", "text" ] |
Complete request schema
{
"type": "object",
"additionalProperties": false,
"required": [
"file"
],
"properties": {
"file": {
"type": "string",
"format": "binary",
"description": "Short audio or video file. Direct API limit: 30 seconds, 20 MiB."
},
"response_format": {
"type": "string",
"enum": [
"json",
"text"
],
"default": "json",
"description": "Return a JSON transcript or plain text. For speakers and keyterms, choose MOSS Diarize Pro."
}
}
}Response schema
The schema below describes the response; optional fields depend on model settings.
View complete response schema
{
"anyOf": [
{
"type": "object",
"required": [
"text"
],
"properties": {
"text": {
"type": "string"
}
}
},
{
"type": "string",
"description": "Plain transcript when response_format=text."
}
]
}