Fast-loading implementation sequence#
- Date:
2026-08
Large-model startup follows one explicit four-document sequence:
fix existing parser, external-data, and initializer-materialization defects in Model-loading bug fixes;
implement Prepared and asynchronous execution;
connect adaptive I/O, model resolution, prepared tensors, first-token overlap, eviction, and device variants in Completing native fast model loading;
only after the native path is complete, consume its stable ownership and payload contracts in onnxruntime through Using onnx-light fast loading in onnxruntime.
Parallel-for profiling may proceed alongside step 1, but its executor instrumentation must be stable before step 2 begins.
Assignable issue sequence#
Step |
Issues |
Order |
|---|---|---|
|
#4608–#4610 |
|
|
#4613–#4617 |
|
|
#4618–#4623 |
|
|
#4611–#4612 |
#4611 is the
completed onnx-light payload-contract prerequisite. #4612 is the final
implementation PR in |