Airon.ai builds and operates AI infrastructure and dedicated bare-metal NVIDIA GPU compute for enterprise customers. The Systems Engineer, Inference will own the production serving path for open-weight large language models, from GPU kernels through gateways, while improving performance, scalability, reliability, observability, and cost efficiency.