Senior Site Reliability Engineer, NetBox Delivery
AWSKubernetesTerraformPostgreSQLGitHub ActionsDjangoPrometheusGrafanaArgoCDFluxCD
About the Role
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior Site Reliability Engineer, NetBox Delivery based in Brazil. This is a senior engineering opportunity focused on the reliability, performance, and delivery of a widely used network automation platform. You will help own the path from core software releases through dependable production deployments across cloud and self-managed environments. The role combines software engineering, SRE, platform engineering, observability, and supply chain security. You will work hands-on with technologies including AWS, Kubernetes, Terraform, PostgreSQL, GitHub Actions, and monitoring platforms. As an early member of a new team, you will have meaningful influence over engineering practices, release processes, and operational standards. You will collaborate across engineering teams, lead reliability initiatives, and address production issues at their technical source. The environment values technical ownership, simplicity, open-source collaboration, clear communication, and continuous improvement. This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior Site Reliability Engineer, NetBox Delivery based in Brazil. This is a senior engineering opportunity focused on the reliability, performance, and delivery of a widely used network automation platform. You will help own the path from core software releases through dependable production deployments across cloud and self-managed environments. The role combines software engineering, SRE, platform engineering, observability, and supply chain security. You will work hands-on with technologies including AWS, Kubernetes, Terraform, PostgreSQL, GitHub Actions, and monitoring platforms. As an early member of a new team, you will have meaningful influence over engineering practices, release processes, and operational standards. You will collaborate across engineering teams, lead reliability initiatives, and address production issues at their technical source. The environment values technical ownership, simplicity, open-source collaboration, clear communication, and continuous improvement. Accountabilities Own and improve the software build and release pipeline, from container base images through availability across cloud and enterprise deployments. Establish reliable release handoffs between core engineering and downstream cloud and enterprise teams so releases reach customers efficiently and predictably. Improve application performance and production reliability, including startup behavior, PostgreSQL performance, and other critical system components. Build and maintain comprehensive observability for both the application and release pipeline, including monitoring, alerting, and service-level objectives (SLOs). Serve as an escalation point for performance and reliability incidents, identifying root causes and contributing fixes to the underlying application when appropriate. Strengthen software supply chain security and contribute to controls supporting SOC 2 compliance for the build and delivery pipeline. Participate in on-call rotations, lead incident response for relevant issues, and facilitate effective postmortems and follow-up improvements. Drive cross-team engineering initiatives from technical proposals and RFCs through implementation, migration, and adoption. Requirements 5+ years of experience in software engineering, platform engineering, SRE, or a closely related discipline, with a strong track record of producing robust and maintainable software. Production experience with Django and PostgreSQL at scale, including schema design, migration planning, and query-performance optimization under real-world workloads. Strong container-building expertise, including base image design, Python dependency management, vulnerability scanning, image signing, and other software supply chain security practices. Hands-on experience with technologies such as AWS EC2, VPC, IAM, and RDS; Kubernetes and Helm; GitHub Actions; ArgoCD or FluxCD; Terraform; Prometheus; and Grafana, or comparable tooling. Experience working within AI-augmented software development environments, including tools such as Claude Code and practices for making agentic development workflows reliable. Demonstrated ability to drive initiatives across team boundaries, from writing technical proposals through completing complex migrations or operational changes. Familiarity with network automation or the NetBox ecosystem is a plus. Open-source contributions or experience participating in open-source projects is advantageous. Experience in B2B software, startups, or high-growth technology organizations is beneficial. Deep knowledge of supply chain security technologies such as cosign, Sigstore, or SLSA is a plus. Experience operating high-throughput or performance-sensitive systems for enterprise customers is advantageous. Strong communication, ownership, problem-solving, and collaboration skills, with an ability to balance reliability, simplicity, and delivery speed. Benefits Remote position based in Brazil as part of a distributed team. Competitive compensation aligned with the applicable LATAM compensation structure. Opportunity to work on open-source and commercial software used by organizations managing complex networks. Meaningful ownership and influence as an early engineer on a newly established delivery and reliability team. Exposure to modern cloud infrastructure, Kubernetes, observability, infrastructure-as-code, and software supply chain security practices. Collaborative environment focused on clear communication, technical ownership, simplicity, and community. Opportunity to contribute to cross-functional engineering initiatives and open-source projects. Equal opportunity workplace with support for reasonable accommodations throughout the hiring process. How Jobgether works: We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team. We appreciate your interest and wish you the best! Why Apply Through Jobgether? Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time. #LI-CL1
You'll be redirected to Jobgether's application page
Job Details
Salary
Not disclosed
Location
Brazil
Job type
Full-time
Category
Site Reliability Engineering
Experience
5+ years
Posted
Yesterday
Job Highlights
- 5+ years level role
- 100% Remote — open to candidates in Brazil
- Full-time position
About Jobgether
This job is hosted by Jobgether. Clicking Apply opens their site.
Remote Work Style
Mixed
Mix of flexible and scheduled meetings
Your Match
See how well your skills line up with this role, and what you're missing.
AI Cover Letter
Generate a cover letter tailored to this job from your profile.