About this role
Job title: AI Compiler & Inference Engineer
What you’ll work on:
- AI model intake, conversion and deployment
- K3 NPU compiler/runtime and supported operators
- ONNX, PyTorch export & model runtimes
- Quantisation, calibration & numerical validation
- Edge/accelerator NPU and heterogeneous inference
- Python/C++ and Linux deployment pipelines
- Model equivalence testing and production acceptance
Ideal candidates: Have strong hands-on experience with ML inference systems, model compilers/runtimes, ONNX/PyTorch, quantisation and production deployment.