Analyze voice/audio for safety concerns with transcription
Upload an audio file (mp3, wav, m4a, ogg, flac, webm, mp4 — max 25MB) for automatic transcription followed by full safety analysis on the transcript. Returns timestamped segments, per-category analysis (bullying, grooming, unsafe, emotions), and an overall risk score. Ideal for monitoring voice chats, audio messages, and recorded conversations.
Authorizations
API key as Bearer token
Response
Default Response
none, low, medium, high, critical A concern was observed at ANY severity, including monitor-only cases. For production branching prefer recommended_action.
Confidence in the overall verdict (0-1).
Routing key, weakest to strongest: none, monitor, flag_for_review, block, immediate_intervention. Safe to switch on.
none, monitor, flag_for_review, block, immediate_intervention Human-readable expansion of recommended_action for a moderator UI. Free text: do not branch on it.
Why this verdict was reached. Built from category and severity labels only; never contains the transcript or media description.
Detected or specified language code. 27 supported languages: en is stable, the other 26 are beta (es, pt, uk, sv, no, da, fi, de, fr, nl, pl, it, tr, ro, el, cs, hu, bg, hr, sk, lt, lv, et, sl, mt, ga).
Language support maturity: "stable" (en) or "beta" (all others)
stable, beta Number of credits consumed by this request
Per-call override of account incident logging: "true" forces persistence, "false" suppresses it. Omit for account default.
Echo of the metadata you provided in the request