New Relic is a global team dedicated to shaping the future of observability through their intelligent platform. They are seeking a Senior Software Engineer to join their Alerts teams, focusing on backend services that handle high-volume telemetry and alerting workloads, ensuring reliability and customer impact.
Responsibilities:
- Work collaboratively on a team using agile practices to ship software incrementally with frequent customer feedback
- Design, develop, and deploy backend services in Java/Kotlin that process high-volume telemetry and alerting workloads, with reliability and customer impact top of mind
- Collaborate with product managers and engineers who specialize in high-throughput data streaming systems, computing infrastructure, design, UIs, and customer-facing APIs
- Implement exciting new Alerting features that affect our entire pipeline, and also help reduce tech debt and retire old architecture
- Advocate for architecture improvements, provide future direction, and clearly articulate reasons why while assessing tradeoffs
- Develop and deploy your code to customers multiple times per day
- Be part of a small team of engineers collectively accountable for the reliability and security of the team's software
- Write clean, well-tested, and maintainable code; participate in peer code reviews and contribute to internal technical documentation
- Maintain a work-life balance that enables you to thrive by leveraging flex time-off, ten weeks of paid parental leave, and our DataNerds4Good volunteer time off program
- Submit PRs to other teams' codebases with low friction by providing the context the team needs to understand and approve the changes
Requirements:
- 5+ years of professional backend software engineering experience, preferably in a SaaS or product-based environment
- Strong proficiency in Java (Kotlin actually, but willingness is fine). You should have a solid grasp of OOP principles, RESTful APIs, and multi-threaded programming
- Experience building multi-threaded Java services and shipping reliable high-throughput services to customers in a production environment
- Experience with relational databases: complex SQL, optimization, pagination, partitioning, and scaling
- Experience working with distributed systems and an understanding of how to write code and queries that perform at scale
- Experience delivering APIs consumed by internal and/or external customers
- Demonstrated empathy for the end user — you understand that backend data logistics, persistence, and retrieval at scale directly affect what customers experience, and you make engineering decisions with that impact in mind
- Experience working in an agile environment characterized by rapid change
- Strong interpersonal skills, including the ability to seek consensus, lead by example, and exhibit persistence and tenacity
- AI & LLM Development: Hands-on experience building with LLMs and AI agents — designing prompts, integrating LLM APIs, building retrieval-augmented workflows, evaluating model output quality, or developing/maintaining MCP (Model Context Protocol) servers to expose telemetry data to AI agents. A genuine curiosity about how agentic systems are reshaping observability is a strong plus
- Familiarity with message queuing systems and streaming patterns like Kafka (preferred), Flink, Spark Streaming, AMQP (RabbitMQ), or gRPC
- Familiarity with Kubernetes, Docker, and Terraform
- Cloud computing experience (compute, storage, and analytics with AWS, GCP, or Azure)
- Frontend awareness or working knowledge (React, TypeScript, GraphQL, CSS) — enough to collaborate effectively with frontend partners