> ## Documentation Index
> Fetch the complete documentation index at: https://dragonwingdocs.qualcomm.com/llms.txt
> Use this file to discover all available pages before exploring further.

# List available ASR models

> Returns available ASR models with capabilities and supported parameters.



## OpenAPI

````yaml /Ubuntu/ai-workflows/api-specs/audio-analytics/openapi.yaml get /transcriptions/models
openapi: 3.0.3
info:
  title: Audio Analytics API
  version: 1.0.0
  description: >
    REST API providing AI-powered audio processing: speech recognition (ASR),

    text translation (T2T), and speech synthesis (TTS).


    ## Services


    | Service | Endpoint prefix | Description |

    |---------|----------------|-------------|

    | ASR | `/transcriptions/` | Whisper-based speech-to-text (file or live
    WebSocket streaming) |

    | T2T | `/translations/` | Opus/NLLB text translation (batch supported) |

    | TTS | `/tts/` | MeloTTS speech synthesis; returns raw PCM audio |


    ## Audio formats (ASR input)

    - WAV, 16-bit PCM; 16 kHz mono recommended; max 25 MB


    ## TTS output

    - Supported sample rates: 44100, 22050, 16000 Hz (auto-resampled)


    ## WebSocket streaming sequence


    ```

    App                                           Server
     |===POST /transcriptions/create================>|
     |<==200 {session_id, state: asr_initialized}====|
     |
     |===WS /stream=================================>|
     |<==  {state: connection_established}===========|
     |
     |===WS {transcriptions_session_audio} (loop)===>|
     |<===WS {transcript.event,     speech_start}====|
     |<===WS {transcript.text.delta, ...}============|  (zero or more)
     |<===WS {transcript.event,     speech_end}======|
     |<===WS {transcript.text.done,  ...}============|
     |
     |===POST /transcriptions/close=================>|  (after transcript.text.done)
    ```


    ## Authentication

    None required.
  contact:
    name: Qualcomm Support Forums
    url: https://mysupport.qualcomm.com/supportforums/s/
  license:
    name: BSD-3-Clause-Clear
    url: https://opensource.org/licenses/BSD-3-Clause-Clear
servers:
  - url: http://{device-ip}:{device-port}/audio-analytics/v1/api
    description: Audio Analytics API Server
    variables:
      device-ip:
        description: IP address and port of the host device running the Audio Analytics API
        default: localhost
      device-port:
        default: '8085'
security: []
tags:
  - name: Health
    description: Service health and status endpoints
  - name: Transcriptions
    description: Speech-to-text audio transcription (ASR)
  - name: Translations
    description: Text-to-text translation between languages
  - name: TTS
    description: Text-to-speech synthesis
  - name: WebSocket
    description: Real-time streaming communication
paths:
  /transcriptions/models:
    get:
      tags:
        - Transcriptions
      summary: List available ASR models
      description: Returns available ASR models with capabilities and supported parameters.
      operationId: listTranscriptionModels
      responses:
        '200':
          description: List of available ASR models
          content:
            application/json:
              schema:
                type: array
                items:
                  $ref: '#/components/schemas/ASRModel'
        '500':
          description: Internal server error
          content:
            application/json:
              schema:
                $ref: '#/components/schemas/ErrorResponse'
components:
  schemas:
    ASRModel:
      type: object
      description: Automatic Speech Recognition model information
      properties:
        name:
          type: string
          description: Unique model identifier
          example: whisper-small-quantized
        display_name:
          type: string
          description: Human-readable model name
          example: Whisper Small (Quantized)
        version:
          type: string
          description: Model version
          example: 1.0.0
        description:
          type: string
          description: Model description
          example: OpenAI Whisper small model for general speech recognition
        capabilities:
          type: object
          description: Model capabilities
          properties:
            streaming:
              type: boolean
              description: Supports streaming transcription
            file_based:
              type: boolean
              description: Supports file-based transcription
            real_time:
              type: boolean
              description: Supports real-time processing
            language_detection:
              type: boolean
              description: Can automatically detect language
            confidence_scores:
              type: boolean
              description: Provides confidence scores
        parameters:
          type: object
          description: Model parameters and constraints
          properties:
            max_file_size_mb:
              type: integer
              description: Maximum file size in megabytes
            supported_formats:
              type: array
              items:
                type: string
              description: Supported audio formats
            sample_rates:
              type: array
              items:
                type: integer
              description: Supported sample rates in Hz
            chunk_duration_ms:
              type: integer
              description: Chunk duration for streaming in milliseconds
      required:
        - name
    ErrorResponse:
      type: object
      description: Standard error response structure
      properties:
        error:
          type: object
          properties:
            message:
              type: string
              description: Human-readable error message
              example: Invalid request, missing 'file' parameter
            type:
              type: string
              description: Error type or category
              example: invalid_request_error
            param:
              type: string
              nullable: true
              description: Parameter related to the error (if applicable)
              example: file
            code:
              type: string
              nullable: true
              description: Error code (if applicable)
              example: null
          required:
            - message
            - type
      required:
        - error

````