nemo_curator.models.asr.indic_canary
nemo_curator.models.asr.indic_canary
Adapter for an Indic Canary model exported as a TensorRT-LLM engine.
Module Contents
Classes
Data
API
Run static-batch Indic Canary inference from a prebuilt engine directory.
cross_kv_cache_fraction
kv_cache_free_gpu_memory_fraction
max_new_tokens
max_samples
min_duration_sec
min_samples
num_beams
pnc
Run engine-sized sub-batches without exposing a second batch control.
Validate local engine artifacts without allocating GPU state.
Load the TensorRT-LLM runtime on its one required GPU.
Transcribe supported rows and preserve their original positions.
Release the engine and its CUDA allocations.