Skip to main content

← All model recipe tasks

m2m_100 model family

Tasks

Supported architectures and task heads

The Hugging Face values below are copied from checkpoint metadata at the recorded revision. The TRTMC task contract comes from the exact E2E recipe. Architecture identity selects or describes a source graph; it does not imply that TRTMC reproduces every Hugging Face head with the same model_type. For example, an encoder recipe may consume the base model and intentionally return hidden states instead of the checkpoint's pretraining or classification logits.

HF model_typeHF architecture / pipeline classTRTMC task contractExact recipe profiles
m2m_100M2M100ForConditionalGeneration
Source: config.architectures
Metadata: config.json at f8d333a098d1
text_generation_causal
nllb-200-distilled-600m

Declared recipes

Each row comes from a model-owned E2E manifest. It declares a test recipe; it is not a live pass receipt for every hardware target.

RecipeExact Hugging Face checkpointTaskBuild configurationDeclared E2E casesManifest
nllb-200-distilled-600mfacebook/nllb-200-distilled-600Mfp16; single devicenllb-200-distilled-600mtests/e2e/models/m2m_100/manifests/nllb-200.json

Family-owned configuration

This family does not declare a family-owned --set namespace. Use the explicit CLI options shown below and the shared configuration namespaces documented in Configure Runtime Behavior.

Family-specific CLI contracts

Inputs and options below are filtered by the declared E2E task and the methods and configuration fields used by that family's native runtime implementation. The global CLI parser accepts a wider union of flags; flags absent here are not declared for this family.

trtmc run

Generate text from a text prompt.

Declared recipes: nllb-200-distilled-600m

trtmc run <bundle.bundle> --prompt "<text>" [generation options]
Supported input or optionRequirementRuntime behavior
--prompt <TEXT>RequiredText input for this causal language-model recipe.
--max-new-tokens <N>OptionalLimit the number of generated tokens or audio frames.
--source-language-token-id <N>OptionalSet the source-language token used by a multilingual encoder-decoder runtime.
--forced-bos-token-id <N>OptionalForce the first decoder token for a multilingual encoder-decoder runtime.
--num-samples <N>OptionalRun N independent text generations.
--output <PATH>OptionalWrite generated samples as JSON Lines.
--benchmark <N>OptionalRun N timed generation iterations.
--warmup <N>Optional with --benchmarkWarm-up iterations before generation timing.
Code basis

Runtime provider: m2m_100

  • src/cli/main.cpp
  • include/trtmc/pipeline.h
  • tests/e2e/models/m2m_100/manifests/nllb-200.json
  • src/runtime/models/m2m_100/plugin.cpp

Bundle lifecycle commands

These commands apply to every declared recipe in this family; they do not add task inputs or model capabilities.

trtmc build

Build one exact checkpoint into a TensorRT-Model-Connect bundle.

trtmc build <hf-id> --output <bundle.bundle>

trtmc inspect

Inspect bundle metadata, runtime identity, and packaged sections.

trtmc inspect <bundle.bundle>

See the CLI Reference for all options and limitations.