载入中...
搜索中...
未找到
OnnxModel.h
浏览该文件的文档.
18enum class OnnxElement : int { Float32 = 1, UInt8 = 2, Int8 = 3, Int32 = 6, Int64 = 7, Bool = 9 };
96 [[nodiscard]] Result<OnnxGpuResult> runGpu(std::span<const OnnxNamedTensor> feeds, OnnxCompute& compute,
Move-only, checked operation results for the common layer.
Native ONNX import and CPU/GPU execution using tensor kernels, without ONNX Runtime.
Definition OnnxModel.h:63
~OnnxModel()
Destroy model-owned graph and packed initializers; outputs remain valid.
Definition AffineQuant.cpp:9
OnnxElement
ONNX wire element types; distinct from block-quantized Tensor storage.
Definition OnnxModel.h:18
@ Float32
Owning GPU execution output and completed dispatch count.
Definition OnnxModel.h:42
Owning admission report; unsupported nodes remain inspectable but cannot execute.
Definition OnnxModel.h:48
std::vector< std::string > unsupportedNodes
Definition OnnxModel.h:52
Per-call deterministic RNG and optional strict finite-output diagnostic.
Definition OnnxModel.h:37
Owning ONNX boundary tensor: row-major little-endian bytes and exact integer shape.
Definition OnnxModel.h:25
Actual transfers and submissions made during one GPU run.
Definition OnnxStorage.h:26