Runtime, Backend Test, and Kernel Examples#

Examples showing how to retrieve, inspect or run ONNX backend test cases exposed by onnx_light.onnx.backend.

Benchmark Abs: onnxruntime vs onnx-light

Benchmark Abs: onnxruntime vs onnx-light

Benchmark a subset of backend test cases against onnxruntime

Benchmark a subset of backend test cases against onnxruntime

Benchmark the initialization steps: onnxruntime vs onnx-light

Benchmark the initialization steps: onnxruntime vs onnx-light

Calibrates quantization with graph kernels

Calibrates quantization with graph kernels

Extend ReferenceEvaluator with a custom kernel

Extend ReferenceEvaluator with a custom kernel

Profile the runtime memory of a model with the event log

Profile the runtime memory of a model with the event log

Quantizes and dequantizes selected pages of a KV cache

Quantizes and dequantizes selected pages of a KV cache

Replace a built-in kernel with a Python one and prove it ran

Replace a built-in kernel with a Python one and prove it ran

Retrieve a backend test case and display its model and data

Retrieve a backend test case and display its model and data

Run a model with the runtime and inspect intermediate results

Run a model with the runtime and inspect intermediate results

Run an ONNX model casting a float tensor into an int2 tensor

Run an ONNX model casting a float tensor into an int2 tensor

Run the reference evaluator with tensor, sequence and dictionary inputs/outputs

Run the reference evaluator with tensor, sequence and dictionary inputs/outputs

Uses every portable quantization profile from Python

Uses every portable quantization profile from Python

Gallery generated by Sphinx-Gallery