Silero Language Classifier 95 β€” GGUF

GGUF conversion of Silero's 95-language classifier for use with CrispASR.

Model Details

  • Architecture: Learned STFT frontend + 8-stage MobileNet-style depthwise-separable conv encoder with interleaved post-norm transformers + attention-weighted pooling + 95-language / 58-group classifiers
  • Parameters: 507 tensors, ~4M parameters
  • Input: Raw 16 kHz mono PCM audio (best results on clips < 20 seconds)
  • Output: 95-language log-probabilities + 58-language-group log-probabilities
  • License: MIT (same as upstream Silero)

Files

File Type Size Notes
silero-lid-lang95-f32.gguf F32 16 MB Full precision, recommended

Quantized versions (Q8_0, Q5_0) were tested but break accuracy β€” the model is dominated by small Conv1d kernels where block quantization is destructive. At 16 MB F32, the model is already very small.

Usage with CrispASR

# Language detection pre-step for backends without native LID
crispasr --backend cohere -m cohere-transcribe-q5_0.gguf \
         -f audio.wav -l auto \
         --lid-backend silero --lid-model silero-lid-lang95-f32.gguf

# Standalone detection (via the C API)
silero_lid_context * ctx = silero_lid_init("silero-lid-lang95-f32.gguf", 4);
float conf;
const char * lang = silero_lid_detect(ctx, samples, n_samples, &conf);
// lang = "en, English"
silero_lid_free(ctx);

Supported Languages (95)

The model supports 95 languages across 58 language groups. Language detection works best on audio clips between 3-20 seconds of speech.

Conversion

Converted from the ONNX export (lang_classifier_95.onnx) using models/convert-silero-lid-to-gguf.py from the CrispASR repo.

Acknowledgements

Provenance and EU AI Act Art. 53 note

  • Upstream model: snakers4/silero-lang-classifier-95.
  • Upstream licence: mit. This repository redistributes under the same terms; it grants no rights the upstream licence does not.
  • What was done here: format conversion and/or quantisation only (GGUF). No training, no fine-tuning, no merging, no distillation, no change to architecture, vocabulary or capability. Only the numeric representation of the upstream weights differs.
  • Training data: documented β€” where it is documented at all β€” by the upstream provider; see the upstream model card. No training data was used, added or selected by this repository.
  • Provider status: under Regulation (EU) 2024/1689 the upstream authors remain the provider of this model. Converting the serialisation format does not make this repository the provider of a new general-purpose AI model, and no such claim is made. Questions about training content, copyright policy or model capability belong upstream.
Downloads last month
352
GGUF
Model size
4.22M params
Architecture
silero_lid
Hardware compatibility
Log In to add your hardware

32-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. πŸ™‹ Ask for provider support