infrastructure
ModelOpt 0.47.0 introduces FP8 Vision Encoder recipes for qwen3_vl and qwen3_5 models.
ModelOpt 0.47.0 introduces quantization with Autotune for ONNX models, benchmarking placements in requested runtime precision.
1 sources · 1 verified claims