Rev AI connects to a Nagent workspace with an API key. Once it is connected, agents can call 11 Rev AI actions, such as "Get Account", "Get Captions" and "Get Custom Vocabulary Details". Nothing is enabled on connect: each action is allowed one at a time, and an action with side effects runs or waits for a person according to the agent's level.
Every operation an agent can call against Rev AI, with input parameters and output schema.
REV_AI_DELETE_CUSTOM_VOCABULARYTool to delete a completed custom vocabulary and its data. Use when you need to remove an unused vocabulary after confirming it's no longer needed.
Input parameters
Unique identifier of the custom vocabulary to delete
Output
Data from the action execution
Error if any occurred during the execution of the action
Whether or not the action execution was successful or not
REV_AI_DELETE_JOB_BY_IDTool to delete a completed transcription job and its data. Use when you need to permanently remove a finished job after confirming it's no longer needed.
Input parameters
The unique identifier of the job to delete
Output
Data from the action execution
Error if any occurred during the execution of the action
Whether or not the action execution was successful or not
REV_AI_GET_ACCOUNTTool to retrieve developer account details. Use after authenticating with Rev AI.
Output
Data from the action execution
Error if any occurred during the execution of the action
Whether or not the action execution was successful or not
REV_AI_GET_CAPTIONSTool to retrieve captions (SRT or VTT) for a completed Rev.ai transcription job. Use after confirming the job status is 'completed'.
Input parameters
Caption format: SRT (application/x-subrip) or WebVTT (text/vtt).
The ID of the completed transcription job.
Optional audio channel number for multi-channel jobs.
Output
Data from the action execution
Error if any occurred during the execution of the action
Whether or not the action execution was successful or not
REV_AI_GET_CUSTOM_VOCABULARY_DETAILSTool to retrieve custom vocabulary processing details. Use when needing to fetch the status and submitted phrases for a specific custom vocabulary after creation.
Input parameters
Unique identifier of the custom vocabulary to retrieve.
Output
Data from the action execution
Error if any occurred during the execution of the action
Whether or not the action execution was successful or not
REV_AI_GET_JOB_BY_IDTool to fetch details of a transcription job by its ID. Use when confirming job status and metadata are accurate.
Input parameters
Unique identifier of the transcription job to retrieve.
Output
Data from the action execution
Error if any occurred during the execution of the action
Whether or not the action execution was successful or not
REV_AI_GET_LIST_OF_JOBSTool to get list of transcription jobs from the past 30 days. Use when you need to retrieve and paginate through recent transcription tasks.
Input parameters
Maximum number of jobs to return (default 100, max 1000)
Only return jobs created before the given job ID.
Only return jobs created after the given job ID.
Output
Data from the action execution
Error if any occurred during the execution of the action
Whether or not the action execution was successful or not
REV_AI_GET_TRANSCRIPT_BY_IDTool to retrieve the transcript of a completed Rev.ai job. Use after confirming job is complete. Supports JSON format (with timestamps and speaker info) or plain text format.
Input parameters
Identifier for the transcription job.
Output format. Supported values: application/vnd.rev.transcript.v1.0+json (default) for JSON with timestamps and speaker info, or text/plain for plain text.
Output
Data from the action execution
Error if any occurred during the execution of the action
Whether or not the action execution was successful or not
REV_AI_START_STREAM_TRANSCRIPTIONTool to start a WebSocket transcription stream. Use when you need real-time speech-to-text streaming via Rev.ai.
Input parameters
Language code in ISO 639-1 format for transcription, default 'en'.
Optional user metadata string to associate with the stream.
Priority of transcription processing: 'speed' or 'quality'.
Initial timestamp (in seconds) from which to start transcription.
Select a transcriber model, if multiple are available.
Audio MIME type and parameters (e.g., 'audio/x-raw;layout=interleaved;rate=16000;format=S16LE;channels=1', 'audio/x-wav', 'audio/x-flac').
Whether to filter profanity from transcripts.
Whether to send detailed partial results.
Whether to remove disfluencies (um, uh) from transcripts.
Whether to skip post-processing phase.
ID of a custom vocabulary to apply, if available.
If set, server will delete the stream results after given seconds.
Whether to enable speaker-switch detection.
Maximum seconds to wait for WebSocket connection establishment.
Maximum duration in seconds of each transcription segment.
Output
Data from the action execution
Error if any occurred during the execution of the action
Whether or not the action execution was successful or not
REV_AI_SUBMIT_CUSTOM_VOCABULARYTool to submit a custom vocabulary for improved speech recognition. Use when you want to process domain-specific terms asynchronously.
Input parameters
List of phrases or words to include in the custom vocabulary.
Optional user-defined metadata for the custom vocabulary.
Optional unique identifier for the custom vocabulary.
Output
Data from the action execution
Error if any occurred during the execution of the action
Whether or not the action execution was successful or not
REV_AI_SUBMIT_TRANSCRIPTION_JOBTool to submit a new transcription job. Use when you have a media URL or file bytes ready for async processing.
Input parameters
Binary audio content to upload if media_url is not used
ISO 639-1 or BCP-47 code to override default language
Arbitrary string returned verbatim with job status
HTTP(S) URL of the media file to transcribe
Webhook URL called on job completion
If true, filter profanity from the transcript
If true, do NOT perform speaker diarization
If true, do NOT auto-punctuate transcript
List of custom vocabularies to apply
Automatically delete job this many seconds after completion
Settings for advanced transcription options.
Number of audio channels to treat as separate speakers
Output
Data from the action execution
Error if any occurred during the execution of the action
Whether or not the action execution was successful or not
Agents can call 11 Rev AI actions on Nagent, including "Get Account", "Get Captions" and "Get Custom Vocabulary Details". Get Account: Retrieve developer account details. Each action is listed on this page with its input parameters and its output.
Rev AI connects with an API key, under your workspace's own connection. Nothing is enabled on connect: each action is allowed one at a time and can be scoped to the agents that need it.
It returns data, error and successful.