Skip to main content
POST
Generate content (Gemini)
This page uses the same generateContent operation as Generate content (Gemini), with the playground above pre-filled for plain text chat. The notes below describe the Gemini-native fields you can add to generationConfig to request audio understanding or generation with structured parts.
Set generationConfig.responseModalities to ["AUDIO"] to request audio output, and configure generationConfig.speechConfig.voiceConfig.prebuiltVoiceConfig.voiceName to choose a prebuilt voice for generated speech.

Gemini-native request fields

Use a text-to-speech-capable model such as gemini-2.5-flash-preview-tts in the model path parameter when requesting audio output.

Example: requesting speech audio

Response fields

The response follows the standard generateContent shape. When audio output is requested, the returned parts contain inline audio data instead of text:
array
Candidate responses returned by the model.
object
Token accounting, including promptTokenCount, candidatesTokenCount, and totalTokenCount.
object
Prompt blocking feedback when applicable.

Example response

200

Authorizations

Authorization
string
header
required

Your DGrid API key. All endpoints use Authorization: Bearer <DGRID_API_KEY>.

Path Parameters

model
string
required

Target model ID, such as gemini-1.5-pro.

Body

application/json
contents
object[]

Input content array with role and parts.

generationConfig
object

Generation configuration.

Response

Generated content candidates.

candidates
object[]

Candidate responses returned by the model.

usageMetadata
object

Token accounting metadata.