Senior Software Engineer, Observability Delivery
GoRubyKubernetesCloud Infrastructuredistributed systemstelemetryOpenTelemetryDatadogCJava
About the Role
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior Software Engineer, Observability Delivery based in United States. This role sits at the intersection of software engineering, cloud infrastructure, and large-scale observability. You will build and operate critical pipelines for logs, metrics, traces, and exceptions that engineering teams rely on to monitor and diagnose production services. Working across distributed systems, Kubernetes, networking, and cloud infrastructure, you will help ensure telemetry platforms remain reliable, scalable, efficient, and easy to operate. You will lead initiatives to remove scaling bottlenecks and evolve production infrastructure safely as demand grows. The role also involves close collaboration with service teams to improve observability capabilities and operational practices. It is an opportunity to make a direct impact on the reliability of critical infrastructure while providing technical leadership across teams. This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior Software Engineer, Observability Delivery based in United States. This role sits at the intersection of software engineering, cloud infrastructure, and large-scale observability. You will build and operate critical pipelines for logs, metrics, traces, and exceptions that engineering teams rely on to monitor and diagnose production services. Working across distributed systems, Kubernetes, networking, and cloud infrastructure, you will help ensure telemetry platforms remain reliable, scalable, efficient, and easy to operate. You will lead initiatives to remove scaling bottlenecks and evolve production infrastructure safely as demand grows. The role also involves close collaboration with service teams to improve observability capabilities and operational practices. It is an opportunity to make a direct impact on the reliability of critical infrastructure while providing technical leadership across teams. Accountabilities: Design, build, and operate high-scale observability pipelines for logs, metrics, traces, and exceptions, balancing capacity, reliability, performance, and cost. Lead cross-functional technical initiatives to identify and resolve scaling bottlenecks, improve telemetry collection and processing, and safely evolve critical production infrastructure. Develop and maintain production software, primarily using Go and Ruby, while configuring and integrating open-source and commercial observability technologies. Work across cloud infrastructure, Kubernetes, virtual machines, networking, and service connectivity to improve the dependability and operability of distributed systems. Partner with engineering teams to understand their observability requirements and enhance the tools and platforms they use to monitor, diagnose, and maintain their services. Take ownership of system health through monitoring, incident response, on-call participation, and continuous improvements driven by operational experience. Provide technical leadership through architecture and design proposals, code and design reviews, mentoring, and collaboration across engineering teams. Requirements 6+ years of experience in Software Engineering, Computer Science, or a related technical discipline, with proven experience maintaining and delivering production software in languages such as C, C++, C#, Java, JavaScript, Go, Ruby, Rust, or Python; equivalent combinations of education and experience are also considered. Alternatively, an Associate's Degree with 5+ years of relevant experience, a Bachelor's Degree with 4+ years, a Master's Degree with 2+ years, or a Doctorate in Computer Science, Electrical Engineering, Electronics Engineering, Mathematics, Physics, Computer Engineering, or a related field. 3+ years of experience building and operating production services or infrastructure within cloud or large-scale distributed environments. 2+ years of experience deploying, configuring, or troubleshooting production workloads using Kubernetes or comparable container orchestration technologies. 2+ years of experience developing production software in Go, Ruby, or a comparable general-purpose programming language. Experience building or operating shared platforms used by other engineering teams is highly valued. Familiarity with logging, metrics, distributed tracing, and OpenTelemetry concepts or tooling, as well as experience with telemetry platforms such as Datadog or comparable solutions, is preferred. Experience with Azure or another major cloud platform, including hybrid or self-managed infrastructure, along with knowledge of networking, service connectivity, and distributed-systems operations. Strong technical leadership, communication, collaboration, and mentoring skills, with the ability to lead ambiguous technical initiatives across teams. Benefits Base salary range of USD $124,000–$329,200 per year , depending on factors such as geographic location, experience, knowledge, skills, and abilities. Certain roles may be eligible for an annual bonus and stock-based rewards based on individual impact. Some positions may also offer sales incentives depending on the role and applicable plan terms. Remote work opportunity within the United States. Generous learning and professional growth opportunities. Benefits designed to support employees and their individual working preferences. Inclusive work environment that welcomes people from diverse backgrounds and provides reasonable accommodations throughout the hiring process. Applications are accepted on an ongoing basis until the position is filled, with the posting open for a minimum of three days. How Jobgether works: We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team. We appreciate your interest and wish you the best! Why Apply Through Jobgether? Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time. #LI-CL1
You'll be redirected to Jobgether's application page
Job Details
Salary
$124K–$329K
Location
United States
Job type
Full-time
Category
Cloud Engineering
Experience
6+ years
Posted
Today
Job Highlights
- $124K–$329K salary
- 6+ years level role
- 100% Remote — open to candidates in United States
About Jobgether
This job is hosted by Jobgether. Clicking Apply opens their site.
Remote Work Style
Mixed
Mix of flexible and scheduled meetings
Your Match
See how well your skills line up with this role, and what you're missing.
AI Cover Letter
Generate a cover letter tailored to this job from your profile.