Model Optimizer

Getting Started

  • Overview
  • Installation
  • Quick Start: PTQ - PyTorch
  • Quick Start: PTQ - ONNX
  • Quick Start: PTQ - PyTorch to ONNX
  • Quick Start: PTQ - Windows
  • Quick Start: QAT
  • Quick Start: Pruning
  • Quick Start: Distillation
  • Quick Start: Speculative Decoding
  • Quick Start: Sparsity

Guides

  • Support Matrix
  • Recipes
  • ModelOpt Config System
  • Quantization
  • Saving & Restoring
  • Pruning
  • Distillation
  • Speculative Decoding
  • Sparsity
  • NAS
  • AutoCast (ONNX)
  • Autotune (ONNX)

Deployment

  • TensorRT-LLM
  • Onnxruntime
  • Unified HuggingFace Checkpoint

Examples

  • All GitHub Examples

Reference

  • Changelog
  • modelopt API
    • deploy
    • onnx
    • torch
      • distill
      • export
      • fastgen
      • kernels
      • nas
      • opt
      • peft
      • prune
      • puzzletron
        • activation_scoring
        • anymodel
        • artifact_coverage
        • artifact_import
        • artifact_import_contract
        • artifact_inventory
        • benchmarks
        • block_config
        • bypass_distillation
        • campaigns
        • candidates
        • checkpoint_transactions
        • dataset
        • depth
        • diagnostics
        • distillation
        • distributed_eval
        • evaluation
        • execution_record
        • export
        • granularity
        • identity
        • manifest
        • mip
        • orchestration
        • pipeline_config
        • plugins
        • post_mip
        • pruning
        • replacement_library
        • rpc_eval
        • sampling
        • scenarios
        • scoring
        • scoring_parent
        • search_space
        • security_policy
        • solution_registry
        • stage_runner
        • stages
        • subblock_stats
        • tools
        • utils
      • quantization
      • sparsity
      • speculative
      • trace
      • utils

Support

  • Contact us
  • FAQs
Model Optimizer
  • modelopt API
  • torch
  • puzzletron
  • anymodel
  • models
  • qwen3
  • View page source

qwen3

Modules

modelopt.torch.puzzletron.anymodel.models.qwen3.qwen3_converter

modelopt.torch.puzzletron.anymodel.models.qwen3.qwen3_model_descriptor

Previous Next

© Copyright 2023-2025, NVIDIA Corporation.

Built with Sphinx using a theme provided by Read the Docs.