Skip to content

Deploying Models

Models run on the Xisom runtime in ONNX format. You upload a model, then choose it in a flow together with an input datasource and outputs. Enabling the flow loads the model.

Model Manager screen listing deployed ONNX models and their status Model Manager screen listing deployed ONNX models and their status
Model Manager — upload an ONNX bundle and check its validation status.
  • ONNX — the runtime accepts ONNX models and executes them with the execution provider that matches your hardware. On the Jetson TX2 that is CUDA, with CPU as the fallback. See Hardware Setup for the modes.
  1. Go to Models and upload your .onnx file. The platform validates the file (extension, type, and schema) before it is available.

  2. Create a flow and pick this model for an input datasource. If the model is not in the list, its window size or feature count disagrees with the input. Adjust the datasource or regenerate the model to match.

  3. Enable the flow, then press Start. Enabling loads the model into the inference service; Start begins producing predictions. See Your First Inference for the full loop.

A fresh install ships a demo model with random weights — predictions are meaningless until you replace it. Upload your trained model as above, or drop it into the model-data directory and restart the inference service. The offline-bundle install page covers the file-drop path: Replace the demo model (step 5).

  • Upload rejected, the model is missing from a flow’s list, or enabling the flow fails to load it — see the Troubleshooting runbook.