载入中...
搜索中...
未找到
#include "tensor/AffineQuant.h"#include "tensor/AffineQuantInternal.h"#include "tensor/OnnxGpuKernels.h"#include <algorithm>#include <cmath>#include <limits>命名空间 | |
| namespace | eve |
| Build metadata (engine git commit, build time, third-party version). | |
| namespace | eve::tensor |
| namespace | eve::tensor::affine |
函数 | |
| Result< std::vector< uint8_t > > | eve::tensor::affine::quantize (std::span< const float > input, float scale, int zeroPoint, bool signedValues) |
| Affine quantization to int8/uint8 bytes, saturating and rounding ties to even. | |
| Result< QuantizedActivation > | eve::tensor::affine::dynamicQuantize (std::span< const float > input) |
| Quantize finite FP32 activations with ONNX DynamicQuantizeLinear semantics. | |
| Result< std::vector< float > > | eve::tensor::affine::dequantize (ByteView input, std::span< const float > scales, std::span< const int32_t > zeros, size_t inner=1) |
| Affine dequantization using scalar or per-axis scale/zero point. | |
| Result< std::vector< int32_t > > | eve::tensor::affine::matmul (ByteView a, ByteView b, size_t m, size_t k, size_t n, int aZero=0, int bZero=0, OnnxCompute *compute=nullptr) |
| Integer row-major [M,K] x [K,N], subtracting scalar zero points. | |
| Result< std::vector< int32_t > > | eve::tensor::affine::conv (ByteView x, ByteView w, const ConvShape &shape, int xZero, std::span< const int32_t > wZeros, OnnxCompute *compute=nullptr) |
| Integer Conv, supporting groups, asymmetric padding and dilation. | |