> ## Documentation Index
> Fetch the complete documentation index at: https://veniceai-feat-models-redesign.mintlify.site/llms.txt
> Use this file to discover all available pages before exploring further.

# xAI Speech to Text v1 API

> xAI Speech to Text v1 API on Venice: speech to text at $0.11 per audio hour ($1.86 per 1,000 minutes). Anonymized access without your identity.

export const HubMount = ({view, children, ...props}) => {
  const [hub, setHub] = useState(null);
  const [failed, setFailed] = useState(false);
  useEffect(() => {
    let alive = true;
    const w = window;
    if (!w.__veniceModelHub) {
      const urls = w.location.hostname === 'localhost' ? ['http://localhost:3333/data/model-hub.bundle.json', '/data/model-hub.bundle.json'] : ['/data/model-hub.bundle.json'];
      const load = i => fetch(urls[i], {
        cache: 'no-cache'
      }).then(res => {
        if (!res.ok) throw new Error(`bundle ${res.status}`);
        return res.json();
      }).catch(err => i + 1 < urls.length ? load(i + 1) : Promise.reject(err));
      const Frag = <></>.type;
      const h = (type, props, ...kids) => {
        const T = type;
        const {key, ...rest} = props || ({});
        if (!kids.length) return <T key={key} {...rest} />;
        if (kids.length === 1) return <T key={key} {...rest}>{kids[0]}</T>;
        return <T key={key} {...rest}>{kids.map((kid, i) => <Frag key={i}>{kid}</Frag>)}</T>;
      };
      w.__veniceModelHub = load(0).then(bundle => new Function(`return (${bundle.code})`)()({
        h,
        Fragment: Frag,
        useState,
        useEffect,
        useRef,
        useMemo,
        useCallback
      }));
    }
    w.__veniceModelHub.then(instance => {
      if (alive) setHub(instance);
    }).catch(() => {
      w.__veniceModelHub = null;
      if (alive) setFailed(true);
    });
    return () => {
      alive = false;
    };
  }, []);
  const View = hub ? hub[view] : null;
  if (View) return <View {...props}>{children}</View>;
  if (failed) {
    return <div className="vx-mount is-failed">
        <p className="vx-mount-note">The interactive model catalog could not load. The full data is below.</p>
        {children}
      </div>;
  }
  return <div className="vx-mount" aria-busy="true">
      <div className="vx-mount-skeleton" aria-hidden="true"><span /><span /><span /></div>
      <div className="vx-mount-source">{children}</div>
    </div>;
};

<HubMount view="ModelPage" data={{"family":{"slug":"xai-speech-to-text-v1","name":"xAI Speech to Text v1","modality":"audio","task":"stt","provider":"xai","primary":"stt-xai-v1","variants":["stt-xai-v1"],"created":1776470400,"updated":1776470400,"privacy":["anonymized"]},"models":[{"id":"stt-xai-v1","name":"xAI Speech to Text v1","type":"asr","modality":"audio","task":"stt","variant":"standard","provider":"xai","created":1776470400,"privacy":"anonymized","pricing":{"perSecond":0.000031,"perMinute":0.00186,"perHour":0.1116},"headline":{"value":0.1116,"unit":"per audio hour","basis":"hour"},"endpoints":[{"id":"transcriptions","method":"POST","path":"/audio/transcriptions","name":"Transcribe audio","status":"stable","recommended":true}],"family":"xai-speech-to-text-v1"}],"related":{"similar":[{"slug":"wizper","name":"Wizper (Whisper v3)","provider":"openai","modality":"audio","privacy":["private"],"created":1776384000,"headline":{"value":0.36,"unit":"per audio hour","basis":"hour"},"variants":1},{"slug":"parakeet-asr","name":"Parakeet ASR","provider":"nvidia","modality":"audio","privacy":["private"],"created":1760136444,"headline":{"value":0.36,"unit":"per audio hour","basis":"hour"},"variants":1},{"slug":"elevenlabs-scribe-v2","name":"ElevenLabs Scribe V2","provider":"elevenlabs","modality":"audio","privacy":["anonymized"],"created":1776384000,"headline":{"value":0.6012,"unit":"per audio hour","basis":"hour"},"variants":1},{"slug":"whisper-large-v3","name":"Whisper Large V3","provider":"openai","modality":"audio","privacy":["private"],"created":1736899200,"headline":{"value":0.36,"unit":"per audio hour","basis":"hour"},"variants":1}],"versions":[]},"providers":{"xai":{"slug":"xai","name":"xAI","logo":"/images/icons/models/grok.svg"},"openai":{"slug":"openai","name":"OpenAI","logo":"/images/icons/models/openai.svg"},"nvidia":{"slug":"nvidia","name":"NVIDIA","logo":"/images/icons/models/nvidia.svg"},"elevenlabs":{"slug":"elevenlabs","name":"ElevenLabs","logo":"/images/icons/models/elevenlabs.svg"}},"faq":[{"q":"How much does the xAI Speech to Text v1 API cost?","a":"$0.000031 per second of audio, which is $0.11 per hour. Prices are in USD and can be paid in DIEM at parity."},{"q":"What is the xAI Speech to Text v1 model ID?","a":"Use `stt-xai-v1` as the `model` parameter."},{"q":"Is the xAI Speech to Text v1 API private?","a":"xAI Speech to Text v1 is anonymized: Venice forwards requests to the provider without your identity, but the provider may retain prompt data, so use a private model for sensitive work."},{"q":"Which endpoint does the xAI Speech to Text v1 API use?","a":"Call `POST /audio/transcriptions`."}]}} />

<div className="vx-static">
  <Accordion title="Plain-text specification">
    # xAI Speech to Text v1 API

    xAI Speech to Text v1 is a speech-to-text model by xAI, available on the Venice API as `stt-xai-v1`. Requests are anonymized, so the provider never sees your identity.

    ## xAI Speech to Text v1 API pricing

    | Model ID | Variant | Privacy | Price |
    | - | - | - | - |
    | `stt-xai-v1` | Standard | Anonymized | \$0.11 per audio hour |

    ## xAI Speech to Text v1 specifications

    | Spec | Value |
    | - | - |
    | Provider | xAI |
    | Released | Apr 18, 2026 |
    | Privacy | Anonymized |

    ## How to use the xAI Speech to Text v1 API

    Send requests to `POST https://api.venice.ai/api/v1/audio/transcriptions` with `"model": "stt-xai-v1"` and your API key.

    ```bash theme={null}
    curl https://api.venice.ai/api/v1/audio/transcriptions \
      -H "Authorization: Bearer $VENICE_API_KEY" \
      -F model=stt-xai-v1 \
      -F file=@meeting.mp3
    ```

    ## xAI Speech to Text v1 API FAQ

    ### How much does the xAI Speech to Text v1 API cost?

    $0.000031 per second of audio, which is $0.11 per hour. Prices are in USD and can be paid in DIEM at parity.

    ### What is the xAI Speech to Text v1 model ID?

    Use `stt-xai-v1` as the `model` parameter.

    ### Is the xAI Speech to Text v1 API private?

    xAI Speech to Text v1 is anonymized: Venice forwards requests to the provider without your identity, but the provider may retain prompt data, so use a private model for sensitive work.

    ### Which endpoint does the xAI Speech to Text v1 API use?

    Call `POST /audio/transcriptions`.

    ## Related models

    * [Wizper (Whisper v3) API](/models/wizper): \$0.36 per audio hour
    * [Parakeet ASR API](/models/parakeet-asr): \$0.36 per audio hour
    * [ElevenLabs Scribe V2 API](/models/elevenlabs-scribe-v2): \$0.60 per audio hour
    * [Whisper Large V3 API](/models/whisper-large-v3): \$0.36 per audio hour
  </Accordion>
</div>


This documentation is built and hosted on [Mintlify](https://mintlify.com), a developer documentation platform.