Skip to content

Commit 9d52581

Browse files
openrouter-docs-sync[bot]OpenRouter SDK Bot
andauthored
chore: update OpenAPI spec from monorepo (#1022)
Co-authored-by: OpenRouter SDK Bot <sdk-bot@openrouter.ai>
1 parent 0748fd2 commit 9d52581

1 file changed

Lines changed: 32 additions & 3 deletions

File tree

‎.speakeasy/in.openapi.yaml‎

Lines changed: 32 additions & 3 deletions
Original file line numberDiff line numberDiff line change
@@ -28606,6 +28606,15 @@ components:
2860628606
oneOf:
2860728607
- $ref: '#/components/schemas/ContainerAutoEnvironment'
2860828608
- $ref: '#/components/schemas/ContainerReferenceEnvironment'
28609+
SpeechInput:
28610+
anyOf:
28611+
- type: 'string'
28612+
- items:
28613+
$ref: '#/components/schemas/SpeechTurn'
28614+
minItems: 1
28615+
type: 'array'
28616+
description: 'Text to synthesize, or a list of turns for multi-speaker input. Each turn has its own text, voice, and instructions. Multi-speaker input is currently supported by Gemini TTS models only.'
28617+
example: 'Hello world'
2860928618
SpeechInputReference:
2861028619
description: 'Reference content part for stateless voice cloning or voice design'
2861128620
discriminator:
@@ -28713,9 +28722,7 @@ components:
2871328722
voice: 'en_paul_neutral'
2871428723
properties:
2871528724
input:
28716-
description: 'Text to synthesize'
28717-
example: 'Hello world'
28718-
type: 'string'
28725+
$ref: '#/components/schemas/SpeechInput'
2871928726
input_references:
2872028727
description: 'Reference content for stateless voice cloning or voice design. Audio mode: one to three `input_audio` parts, each optionally paired with a `text` part carrying its transcript (a single clip accepts its transcript before or after it; with multiple clips each transcript immediately follows its clip); only routed to endpoints that support voice cloning (and multiple references when more than one part is sent). Image mode: exactly one `image_url` part; only routed to endpoints that support image references. The two modes cannot be mixed. An empty array is treated as no reference.'
2872128728
example:
@@ -28727,6 +28734,10 @@ components:
2872728734
items:
2872828735
$ref: '#/components/schemas/SpeechInputReference'
2872928736
type: 'array'
28737+
instructions:
28738+
description: 'Delivery instructions for the whole request, such as tone, pacing, or emotion. Supported by OpenAI gpt-4o-mini-tts and Gemini TTS models. Ignored by other providers.'
28739+
example: 'Speak in a warm and friendly tone.'
28740+
type: 'string'
2873028741
model:
2873128742
description: 'TTS model identifier'
2873228743
example: 'mistralai/voxtral-mini-tts-2603'
@@ -28788,6 +28799,24 @@ components:
2878828799
- 'model'
2878928800
- 'input'
2879028801
type: 'object'
28802+
SpeechTurn:
28803+
properties:
28804+
instructions:
28805+
description: 'Delivery instructions for this turn, such as tone, pacing, or emotion. Overrides the top-level `instructions`.'
28806+
example: 'whispering'
28807+
type: 'string'
28808+
text:
28809+
description: 'Text to synthesize'
28810+
example: 'Hi Jane.'
28811+
type: 'string'
28812+
voice:
28813+
description: 'Voice for this turn. Defaults to the top-level `voice`.'
28814+
example: 'Kore'
28815+
minLength: 1
28816+
type: 'string'
28817+
required:
28818+
- 'text'
28819+
type: 'object'
2879128820
StopServerToolsWhen:
2879228821
description: 'Stop conditions for the server-tool agent loop. Any condition firing halts the loop (OR logic). When set, this overrides `max_tool_calls`. When a condition fires while the model is still emitting tool calls, the pending tool calls are executed and one final turn is made with tool calls disabled so the response ends with a natural-language answer instead of an unfinished tool call.'
2879328822
example:

0 commit comments

Comments
 (0)