Senior DevOps Engineer, Observability
AWSTerraformKubernetesGitHub ActionsPrometheusGrafanaPythonGoHelmCI/CD
About the Role
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior DevOps Engineer, Observability based in United States . This is a senior infrastructure-focused role supporting reliable, scalable SaaS and on-premise environments. You will design and operate cloud infrastructure, automation, CI/CD pipelines, and observability systems that enable engineering teams to deliver software efficiently. The position combines hands-on DevOps, SRE, and platform engineering with a strong focus on reliability, performance, security, and automation. You will help build internal platforms and self-service capabilities while improving monitoring, alerting, incident response, and service-level objectives. The role offers the opportunity to work across complex infrastructure challenges in a fast-moving B2B software environment. You will collaborate closely with product and engineering teams and contribute to systems that operate at scale. It is well suited to an experienced engineer who treats infrastructure as code and internal platforms as products. This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior DevOps Engineer, Observability based in United States . This is a senior infrastructure-focused role supporting reliable, scalable SaaS and on-premise environments. You will design and operate cloud infrastructure, automation, CI/CD pipelines, and observability systems that enable engineering teams to deliver software efficiently. The position combines hands-on DevOps, SRE, and platform engineering with a strong focus on reliability, performance, security, and automation. You will help build internal platforms and self-service capabilities while improving monitoring, alerting, incident response, and service-level objectives. The role offers the opportunity to work across complex infrastructure challenges in a fast-moving B2B software environment. You will collaborate closely with product and engineering teams and contribute to systems that operate at scale. It is well suited to an experienced engineer who treats infrastructure as code and internal platforms as products. Accountabilities: Design, build, and maintain infrastructure supporting SaaS and on-premise engineering environments. Operate and optimize AWS infrastructure with a focus on cost efficiency, security, scalability, reliability, and performance. Manage cloud services including EC2, VPC, IAM, RDS, EKS, and Kubernetes-based environments. Develop and improve infrastructure-as-code using Terraform and Helm. Build and maintain CI/CD pipelines and developer automation, particularly using GitHub Actions. Contribute to internal platform tooling and developer self-service capabilities. Enhance observability and incident response systems through effective monitoring, alerting, instrumentation, and SLO management. Collaborate with stream-aligned product teams to understand infrastructure needs and continuously improve internal platforms. Support security and compliance initiatives, including implementation and maintenance of SOC 2 controls. Create and maintain technical documentation, onboarding resources, and internal support processes. Participate in an on-call rotation and contribute to reliable incident response and operational practices. Contribute to infrastructure initiatives involving event streaming, Change Data Capture (CDC), or large-scale, multi-tenant observability systems. Requirements: 5+ years of professional experience in DevOps, Site Reliability Engineering (SRE), platform engineering, or a closely related field. At least 2 years of experience working within a B2B software startup or similarly fast-paced technology environment. Strong hands-on AWS experience, including services such as EC2, VPC, IAM, RDS, and EKS. Solid Kubernetes experience combined with infrastructure-as-code expertise using Terraform, Helm, or comparable tools. Experience designing and maintaining CI/CD pipelines and automation tooling, ideally with GitHub Actions. Familiarity with observability platforms and technologies such as Prometheus, Grafana, Mimir, Loki, or similar solutions. Proficiency in at least one scripting or programming language, such as Python, Go, or shell scripting. Experience with either Change Data Capture (CDC) and event-streaming systems or scaling large, multi-tenant observability platforms covering ingestion, analysis, and alerting. Strong communication, collaboration, troubleshooting, and technical documentation skills. Ability to work effectively in a fast-paced environment with ambiguity and evolving priorities. Experience in security-focused or SOC 2-compliant environments is a plus. Familiarity with MQTT, AMQP, or other messaging technologies is advantageous. Experience with network automation or related infrastructure ecosystems is a plus. Open-source contributions or experience working with open-source technologies is valued. Familiarity with AI-assisted development tools such as Copilot, ChatGPT, or Cursor is beneficial. Benefits: Fully remote working arrangement for candidates based in Brazil/LATAM. Full-time position within an engineering-focused environment. Competitive LATAM compensation of $75,000–$85,000 USD annually. Opportunity to work on cloud infrastructure, Kubernetes, infrastructure-as-code, CI/CD, and observability at scale. Significant ownership over infrastructure reliability, performance, automation, and developer experience. Exposure to SaaS and on-premise environments and complex multi-tenant systems. Collaboration with experienced engineering and product teams in a fast-paced B2B software environment. Opportunity to contribute to open-source technologies and infrastructure communities. Professional exposure to modern observability, platform engineering, security, and automation practices. How Jobgether works: We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team. We appreciate your interest and wish you the best! Why Apply Through Jobgether? Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time. #LI-CL1
You'll be redirected to Jobgether's application page
Job Details
Salary
$75K–$85K
Location
United States
Job type
Full-time
Category
DevOps / SRE
Experience
5+ years
Posted
Today
Job Highlights
- $75K–$85K salary
- 5+ years level role
- 100% Remote — open to candidates in Brazil, United States
About Jobgether
This job is hosted by Jobgether. Clicking Apply opens their site.
Remote Work Style
Mixed
Mix of flexible and scheduled meetings
Your Match
See how well your skills line up with this role, and what you're missing.
AI Cover Letter
Generate a cover letter tailored to this job from your profile.