nemo_curator.models.client
nemo_curator.models.client
Submodules
Package Contents
Classes
API
Interface representing a client connecting to an LLM inference server and making requests asynchronously
Run an async request with the client’s concurrency and retry policy.
Internal implementation of query_model without retry/concurrency logic. Subclasses should implement this method instead of query_model.
Query the model with automatic retry and concurrency control.
Setup the client.
Bases: AsyncLLMClient
A wrapper around OpenAI’s Python async client for querying models
Internal implementation of query_model without retry/concurrency logic.
Query a model and return its raw response with retry and concurrency control.
Setup the client.
Interface representing a client connecting to an LLM inference server and making requests synchronously
Setup the client.