nemo_curator.models.audio.sed.tensorrt
nemo_curator.models.audio.sed.tensorrt
TensorRT execution for the CNN14 SED neural core.
Module Contents
Classes
Functions
Data
API
Bases: Module
CNN14 neural core; the checkpoint’s spectrogram frontend stays in PyTorch.
Bases: PANNsSEDAdapter
Run the PANNs CNN14 neural core with a target-specific TensorRT engine.
Audio preprocessing, checkpoint resolution, batch padding, and the
canonical SEDResult contract match PANNsSEDAdapter. The checkpoint’s
spectrogram and log-mel frontend remains in PyTorch; only the CNN14 neural
core runs in TensorRT. Engines are valid only for the GPU compute capability
and TensorRT version recorded in their adjacent JSON sidecar.
Run one TensorRT call and preserve the PANNs adapter result schema.
Load the PyTorch frontend and TensorRT runtime on one CUDA device.
Release the TensorRT context, PyTorch frontend, and CUDA cache.
Persistent TensorRT runner with shape-specific context memory.
Reusable SED adapter preserving the checkpoint’s PyTorch frontend.
Maximum log-mel frame count accepted by the engine profile.
Return framewise probabilities for padded [batch, samples] input.
Reject an engine whose immutable build contract differs from this adapter.
Run the checkpoint’s exact spectrogram and log-mel frontend.
Restore PANNs framewise output geometry from CNN14 segment outputs.