Bonus: Top up now and we'll double your first deposit — get x2 credits instantly.

wav2vec2-xls-r-300m-hebrew API: Pricing & Specs

wav2vec2-xls-r-300m-hebrew is an open-weight speech recognition model that transcribes audio to text.

What is wav2vec2-xls-r-300m-hebrew?

wav2vec2-xls-r-300m-hebrew is an open-weight speech recognition model that transcribes audio to text, published by imvladikon with open weights at imvladikon/wav2vec2-xls-r-300m-hebrew. LLM API does not route this model today, so there is no endpoint or code sample for it here. What this page does give you is the self-hosting picture: published weight sizes, the VRAM each quantisation needs and the commands to serve it yourself. Every figure is read from the public repository and refreshed weekly. At import it had 1,179,294 downloads on Hugging Face in the previous 30 days.

Can you self-host wav2vec2-xls-r-300m-hebrew?

No. We found no public weights for wav2vec2-xls-r-300m-hebrew on Hugging Face, so there is nothing to download, quantize or serve on your own GPUs: no VRAM budget, no vLLM or GGUF build to plan for. The only way to run it is over an API.

Not on LLM.API yet

We do not route this model through our API at the moment. Browse the models you can call today — most workloads have a close match already live.

Not available on LLM API yet

We do not route this model through our API at the moment, so there is no endpoint or code snippet for it yet. Browse the models you can call today — most workloads have a close match already live.

Browse available models

When to Use — When NOT to Use

Use it if...

  • You need transcripts of recordings or calls produced on your own hardware (wav2vec2-xls-r-300m-hebrew)
  • Audio must stay inside your network for privacy or compliance reasons
  • The weights are published openly, so you can fine-tune on your own voices or vocabulary
  • You are comparing self-hosting cost against a per-minute or per-character API

Avoid if...

  • LLM API does not route this model today, so there is no endpoint here to call
  • You want zero operations — serving, scaling and upgrades are yours to run
  • You need a chat or reasoning model — this is an audio model
  • You need a guaranteed real-time SLA before testing on your own audio

Get one key to every model

Swap your API key. Keep your code.