GroqCloud connects to a Nagent workspace with an API key. Once it is connected, agents can call 7 GroqCloud actions, such as "Create Audio Transcription", "Create Response" and "Create Audio Translation". Nothing is enabled on connect: each action is allowed one at a time, and an action with side effects runs or waits for a person according to the agent's level.
Every operation an agent can call against GroqCloud, with input parameters and output schema.
GROQCLOUD_CREATE_AUDIO_TRANSCRIPTIONTool to transcribe audio into text in the same language as the audio. Use when you need to convert speech to text while preserving the original language. Supports multiple formats including mp3, mp4, wav, and webm.
Input parameters
The audio file to transcribe. Supported formats: flac, mp3, mp4, mpeg, mpga, m4a, ogg, wav, webm.
Model ID for transcription. whisper-large-v3-turbo is faster, whisper-large-v3 may be more accurate.
Optional text to guide the model's style or continue a previous audio segment. The prompt should match the audio language.
Language of the input audio in ISO-639-1 format (e.g., 'en', 'es', 'fr'). Supplying this will improve accuracy and latency.
Sampling temperature between 0 and 1. Higher values (e.g., 0.8) make output more random, lower values (e.g., 0.2) make it more focused. If set to 0, model uses log probability to auto-adjust temperature.
Output format. Use 'verbose_json' for timestamp information, 'json' for basic text, or 'text' for plain text output.
Timestamp granularities to populate. Requires response_format='verbose_json'. Options: 'word' (adds latency), 'segment' (no additional latency). Can specify both.
Output
Data from the action execution
Error if any occurred during the execution of the action
Whether or not the action execution was successful or not
GROQCLOUD_CREATE_RESPONSETool to create a model response for the given input. Beta endpoint with simplified interface compared to chat completions. Use when you need a streamlined API for generating model responses.
Input parameters
Response format configuration.
Optional identifier for tracking end-user requests for monitoring and compliance
Text input to the model or an array of input items
ID of the model to use. See available models at https://console.groq.com/docs/models
Response storage flag. Currently only supports false or null
List of tools available to the model. Maximum of 128 functions
Nucleus sampling parameter controlling cumulative probability cutoff. Range 0-1
Enable streaming mode to receive response data as server-sent events
Custom key-value pairs for storing additional information. Maximum of 16 pairs
Configuration for reasoning capabilities.
Context truncation strategy.
Controls randomness. Range 0-2. Lower is more deterministic, higher is more creative
Controls which tool is called. Values: 'none', 'auto', 'required', or specific function
System message inserted as the first item in the model's context
Service tier for processing the request.
Upper bound for tokens in the response, including visible and reasoning tokens
Enable parallel execution of multiple tool calls
Output
Data from the action execution
Error if any occurred during the execution of the action
Whether or not the action execution was successful or not
GROQCLOUD_GROQ_CREATE_AUDIO_TRANSLATIONTool to translate an audio file into English text. Use when you have a non-English recording and need an accurate English transcript. Use after confirming the file path.
Input parameters
The audio file to translate to English. Supported formats: mp3, wav, etc.
Model ID for translation (e.g., 'whisper-large-v3'). whisper-large-v3-turbo may not support translations.
Optional prompt to guide the translation output.
Sampling temperature between 0.0 and 1.0 to control randomness.
Output format: 'json', 'verbose_json', or 'text'.
Output
Data from the action execution
Error if any occurred during the execution of the action
Whether or not the action execution was successful or not
GROQCLOUD_GROQ_CREATE_CHAT_COMPLETIONTool to generate a chat-based completion for a conversation. Use when you have a list of prior messages and need the model's next reply. Response completion text is at choices\[0\].message.content in the returned envelope.
Input parameters
Number of chat completion choices to generate (must be 1)
Up to 4 stop sequences where the model will stop generating further tokens
Unique identifier for the end user for monitoring/abuse detection
ID of the model to use Verify valid IDs via GROQCLOUD_LIST_MODELS before use; hard-coded IDs may be deprecated. Different models have different token limits and rate quotas — check model metadata before large-scale completions.
Nucleus sampling parameter (0 to 1)
Ordered list of messages comprising the conversation
Sampling temperature between 0 and 2
Maximum number of tokens to generate in the completion
Output
Data from the action execution
Error if any occurred during the execution of the action
Whether or not the action execution was successful or not
GROQCLOUD_GROQ_RETRIEVE_MODELTool to retrieve detailed information about a specific model. Use after listing models when you need metadata for a chosen model. Returned metadata may change as models update; do not cache.
Input parameters
Identifier of the model to retrieve Must be an exact ID from GROQCLOUD_LIST_MODELS; approximated or guessed IDs will fail.
Output
Data from the action execution
Error if any occurred during the execution of the action
Whether or not the action execution was successful or not
GROQCLOUD_LIST_MODELSTool to list all available models and their metadata. Always call this to retrieve current model IDs rather than using hard-coded or cached identifiers, as deprecated names cause failures in GROQCLOUD_GROQ_RETRIEVE_MODEL and GROQCLOUD_GROQ_CREATE_CHAT_COMPLETION. Returns availability and metadata only — excludes usage stats, latency metrics, and pricing. Response may include many models; filter client-side by provider, family, modality, or context length. Frequent polling combined with high-volume requests risks HTTP 429 rate_limit_exceeded; use backoff and minimize call frequency.
Output
Data from the action execution
Error if any occurred during the execution of the action
Whether or not the action execution was successful or not
GROQCLOUD_LIST_VOICESTool to retrieve available TTS voices for Groq PlayAI models. Use when you need to discover voice options before calling text-to-speech. Note: static list maintained manually; no live endpoint exists.
Output
Data from the action execution
Error if any occurred during the execution of the action
Whether or not the action execution was successful or not
Agents can call 7 GroqCloud actions on Nagent, including "Create Audio Transcription", "Create Response" and "Create Audio Translation". Create Audio Transcription: Transcribe audio into text in the same language as the audio. Each action is listed on this page with its input parameters and its output.
GroqCloud connects with an API key, under your workspace's own connection. Nothing is enabled on connect: each action is allowed one at a time and can be scoped to the agents that need it.
Create Audio Transcription takes 1 required input: file. It also takes 6 optional inputs: model, prompt, language, temperature, response_format and timestamp_granularities. It returns data, error and successful.