Deploy and optimise AI inference on constrained devices with model conversion, quantisation, accelerators, latency, power, privacy, and fallback.