Nexera is an AI-native consulting firm that builds intelligent, custom solutions for organizations navigating the age of AI. The AI Infrastructure & LLM Inference Engineer will build, deploy, optimize, and operate secure production platforms for self-hosted language and multimodal model inference across on-premises, private-cloud, hybrid, and restricted-network environments. The role also includes reliability engineering, automation, observability, client collaboration, and knowledge transfer.