Skip to main content

Integrations

MatrixHub is built to integrate seamlessly with standard ML systems and high-performance inference frameworks.


🚀 GPU Inference Engines​

MatrixHub acts as a private, high-speed cache endpoint for your serving nodes. Set the HF_ENDPOINT redirect when starting an inference engine. For detailed integration instructions, see:


Model Distribution and P2P Acceleration​