Design, build, and maintain production features with a strong focus on product security, influencing secure engineering practices in a cloud-native environment.
Senior Software Engineer, Observability Delivery
📜 Description
- Design, build, and operate high-scale observability pipelines for logs, metrics, traces, and exceptions.
- Lead cross-functional technical initiatives to identify and resolve scaling bottlenecks.
- Develop and maintain production software primarily using Go and Ruby.
- Work across cloud infrastructure, Kubernetes, and networking to improve distributed systems.
- Partner with engineering teams to enhance observability tools and platforms.
- Take ownership of system health through monitoring and incident response.
🛠️ Requirements
- 6+ years of experience in Software Engineering or related discipline.
- Proven experience maintaining and delivering production software in languages such as Go, Ruby, or others.
- 3+ years of experience building and operating production services in cloud environments.
- 2+ years of experience deploying and troubleshooting workloads using Kubernetes.
- Experience with logging, metrics, distributed tracing, and telemetry platforms.
✨ Benefits
- Certain roles may be eligible for an annual bonus and stock-based rewards based on individual impact.
- Some positions may also offer sales incentives depending on the role and applicable plan terms.
- Remote work opportunity within the United States.
- Generous learning and professional growth opportunities.
- Benefits designed to support employees and their individual working preferences.
- Inclusive work environment that welcomes people from diverse backgrounds and provides reasonable accommodations throughout the hiring
- Applications are accepted on an ongoing basis until the position is filled, with the posting open for a minimum of three days.
- Jobgether - Senior Software Engineer, Observability Delivery
Full job description
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior Software Engineer, Observability Delivery based in United States.
This role sits at the intersection of software engineering, cloud infrastructure, and large-scale observability. You will build and operate critical pipelines for logs, metrics, traces, and exceptions that engineering teams rely on to monitor and diagnose production services. Working across distributed systems, Kubernetes, networking, and cloud infrastructure, you will help ensure telemetry platforms remain reliable, scalable, efficient, and easy to operate. You will lead initiatives to remove scaling bottlenecks and evolve production infrastructure safely as demand grows. The role also involves close collaboration with service teams to improve observability capabilities and operational practices. It is an opportunity to make a direct impact on the reliability of critical infrastructure while providing technical leadership across teams.
Accountabilities:
- Design, build, and operate high-scale observability pipelines for logs, metrics, traces, and exceptions, balancing capacity, reliability, performance, and cost.
- Lead cross-functional technical initiatives to identify and resolve scaling bottlenecks, improve telemetry collection and processing, and safely evolve critical production infrastructure.
- Develop and maintain production software, primarily using Go and Ruby, while configuring and integrating open-source and commercial observability technologies.
- Work across cloud infrastructure, Kubernetes, virtual machines, networking, and service connectivity to improve the dependability and operability of distributed systems.
- Partner with engineering teams to understand their observability requirements and enhance the tools and platforms they use to monitor, diagnose, and maintain their services.
- Take ownership of system health through monitoring, incident response, on-call participation, and continuous improvements driven by operational experience.
- Provide technical leadership through architecture and design proposals, code and design reviews, mentoring, and collaboration across engineering teams.
- 6+ years of experience in Software Engineering, Computer Science, or a related technical discipline, with proven experience maintaining and delivering production software in languages such as C, C++, C#, Java, JavaScript, Go, Ruby, Rust, or Python; equivalent combinations of education and experience are also considered.
- Alternatively, an Associate's Degree with 5+ years of relevant experience, a Bachelor's Degree with 4+ years, a Master's Degree with 2+ years, or a Doctorate in Computer Science, Electrical Engineering, Electronics Engineering, Mathematics, Physics, Computer Engineering, or a related field.
- 3+ years of experience building and operating production services or infrastructure within cloud or large-scale distributed environments.
- 2+ years of experience deploying, configuring, or troubleshooting production workloads using Kubernetes or comparable container orchestration technologies.
- 2+ years of experience developing production software in Go, Ruby, or a comparable general-purpose programming language.
- Experience building or operating shared platforms used by other engineering teams is highly valued.
- Familiarity with logging, metrics, distributed tracing, and OpenTelemetry concepts or tooling, as well as experience with telemetry platforms such as Datadog or comparable solutions, is preferred.
- Experience with Azure or another major cloud platform, including hybrid or self-managed infrastructure, along with knowledge of networking, service connectivity, and distributed-systems operations.
- Strong technical leadership, communication, collaboration, and mentoring skills, with the ability to lead ambiguous technical initiatives across teams.
- Base salary range of USD $124,000–$329,200 per year, depending on factors such as geographic location, experience, knowledge, skills, and abilities.
- Certain roles may be eligible for an annual bonus and stock-based rewards based on individual impact.
- Some positions may also offer sales incentives depending on the role and applicable plan terms.
- Remote work opportunity within the United States.
- Generous learning and professional growth opportunities.
- Benefits designed to support employees and their individual working preferences.
- Inclusive work environment that welcomes people from diverse backgrounds and provides reasonable accommodations throughout the hiring process.
- Applications are accepted on an ongoing basis until the position is filled, with the posting open for a minimum of three days.
Requirements
Benefits
Similar jobs
Search more Software Engineer jobsAs the first dedicated IT hire, you will design, implement, and manage the corporate technology environment, focusing on Microsoft 365 and identity management.
The Senior ServiceNow Developer will design, develop, and support scalable ServiceNow solutions, enhancing asset management and operational efficiency in a large enterprise environment.
Sr. Data Center Operations Engineer
As a Sr. Data Center Operations Engineer, you will deploy, maintain, and scale critical server and network infrastructure in global production data centers.
Senior Systems Software Engineer- EDA Infrastructure
NVIDIA is seeking a Senior Systems Software Engineer to build and manage large-scale infrastructure platform services, focusing on automation and reliability.
