XReduce
XR LabsDocs
SearchOpen menu

Custom model loaders

XReduce works out of the box with HuggingFace transformers models. For anything else - a custom PyTorch checkpoint, a quantized variant, a model served from your own inference endpoint - custom loaders let you wire up whatever you're running.

This feature is available in private beta. and we'll help you set up a loader for your model.