> For clean Markdown of any page, append .md to the page URL.
> For a complete documentation index, see https://docs.nvidia.com/nemo/automodel/llms.txt.
> For AI client integration (Claude Code, Cursor, etc.), connect to the MCP server at https://docs.nvidia.com/nemo/automodel/_mcp/server.

# Bamba-9B

> Reference Bamba-9B checkpoint and architecture details for NeMo AutoModel, with model-family information, setup guidance, and upstream resources.

[Bamba](https://huggingface.co/ibm-ai-platform/Bamba-9B) is a hybrid SSM-attention language model from IBM, combining Mamba-2 selective state space layers with standard transformer attention for efficient long-context processing.

Use this page as a checkpoint and architecture reference. Set up NeMo AutoModel with the
[latest container](https://catalog.ngc.nvidia.com/orgs/nvidia/containers/nemo-automodel) or follow the
[installation instructions](/get-started/installation).

## Model Reference

### Model Architecture

| Property                  | Value                                                     |
| ------------------------- | --------------------------------------------------------- |
| Task                      | Text Generation                                           |
| Architecture              | `BambaForCausalLM`                                        |
| Parameters                | 9B                                                        |
| Hugging Face Organization | [ibm-ai-platform](https://huggingface.co/ibm-ai-platform) |

### Available Models

| Model    | HF ID                                                                         |
| -------- | ----------------------------------------------------------------------------- |
| Bamba 9B | [`ibm-ai-platform/Bamba-9B`](https://huggingface.co/ibm-ai-platform/Bamba-9B) |

## Related Resources

* [LLM Fine-Tuning Guide](/recipes-e2e-examples/sft-peft)