About the company
Nexthink is the leader in digital employee experience management software. The company provides IT leaders with unprecedented insight allowing them to see, diagnose and fix issues at scale impacting employees anywhere, with any application or network, before employees notice the issue. As the first solution to allow IT to progress from reactive problem solving to proactive optimization, Nexthink enables its more than 1,500 customers to provide better digital experiences to more than 25+ million employees. Dual headquartered in Lausanne, Switzerland and Boston, Massachusetts, Nexthink has 9 offices worldwide.
Responsibilities
- Design, build, and scale distributed backend systems that serve millions of users.
- Drive end-to-end delivery of features: from design, implementation, deployment, and monitoring.
- Architect real-time, low-latency APIs and backend services for AI-driven agent pipelines that aggregate platform metrics into executive reporting.
- Develop cloud-native services in Java, ensuring reliability, performance, and scalability.
- Work on infrastructure components, including AWS-based deployments, containerization (Docker), and orchestration (Kubernetes/ECS).
- Build robust APIs and backend services that power AI-driven products and real-time interactions.
- Ensure resiliency, fault tolerance, and compliance in high-throughput environments.
- Collaborate with product managers, infrastructure engineers, and other teams to deliver production-ready systems.
- Contribute to technical discussions, share knowledge, and help evolve best practices in backend and infra engineering.
Requirements
- 7+ years of professional backend engineering experience in Java.
- Proficient in modern Java (17+).
- Java frameworks knowledge (Micronaut, Spring).
- Proven experience designing and scaling backend systems for millions of users or high-throughput real-time services.
- Solid understanding of cloud-native architectures (AWS or similar), containerization (Docker), orchestration (Kubernetes/ECS), and GitOps, CI/CD.
- Experience with distributed data systems (e.g. Apache Kafka).
- Performance tuning and scaling experience in high-throughput environments.
- Strong background in high availability systems, failover, and resiliency design.
- Exposure to LLM integration (prompting, evaluation, observability).
- Proficient in using AI tools for software development.
- Curious mindset with an eagerness to learn and adapt.
- Excellent knowledge of observability practices: logging, monitoring, tracing, metrics.
- Strong communication skills in English, able to explain complex technical topics clearly.