[6078291][OMNIML-3716] Add ViT FP8 + Torch-TRT example, wire softmax_quantizer in _QuantAttention - #1569
Merged
Commits
Commits on Jul 1, 2026
- andcommitted
- committed
- committed
- committed
- andcommitted
- andcommitted
- andcommitted
- andcommitted
- andcommitted
- andcommitted
- andcommitted
- andcommitted
- andcommitted
- andcommitted
- andcommitted
- andcommitted
[6078291] torch_trt: quantize non-causal (ViT) attention softmax-P via eager p_bmm_quantizer wrapper
andcommitted