About the Role
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior DevOps Engineer based in the United States. This is a senior-level opportunity to help operate and evolve a globally distributed cloud infrastructure supporting millions of connected devices and users. You will play a key role in improving reliability, scalability, performance, automation, and cloud cost efficiency across complex production environments. The role spans cloud infrastructure, networking, databases, messaging, observability, CI/CD, and infrastructure as code. You will collaborate with engineering leaders and technical teams across countries and time zones while contributing to both immediate operational priorities and long-term infrastructure strategy. The position offers significant influence over architecture, automation, capacity planning, and cloud optimization. This is a fully remote role with occasional travel, supporting a global technology platform. This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior DevOps Engineer based in the United States. This is a senior-level opportunity to help operate and evolve a globally distributed cloud infrastructure supporting millions of connected devices and users. You will play a key role in improving reliability, scalability, performance, automation, and cloud cost efficiency across complex production environments. The role spans cloud infrastructure, networking, databases, messaging, observability, CI/CD, and infrastructure as code. You will collaborate with engineering leaders and technical teams across countries and time zones while contributing to both immediate operational priorities and long-term infrastructure strategy. The position offers significant influence over architecture, automation, capacity planning, and cloud optimization. This is a fully remote role with occasional travel, supporting a global technology platform. Accountabilities: Manage and continuously improve highly scalable and reliable systems supporting millions of devices and users across geographically distributed platform instances. Drive infrastructure reliability, scalability, performance, automation, and cloud cost-efficiency initiatives. Collaborate with engineering leaders and peers across multiple countries to establish repeatable, scalable infrastructure objectives and practices. Partner with DevOps leadership on short- and long-term infrastructure initiatives, including architecture, budgeting, implementation, capacity planning, and cloud cost optimization. Identify opportunities to automate operational processes and improve engineering workflows. Design, deploy, and operate cloud applications and infrastructure across AWS and GCP environments. Support networking, messaging, databases, observability, and high-availability services within production environments. Participate in a scheduled 24/7 on-call rotation to support production systems and maintain service reliability. Contribute to the long-term technical direction of infrastructure while balancing operational requirements, scalability, resilience, and efficiency. Requirements: 8+ years of professional experience supporting Docker-based microservices, PaaS environments, and cloud technologies. 5+ years of networking experience supporting high-availability applications on AWS and GCP, with experience across both platforms preferred. Experience with HTTP- and MQTT-based ingestion and messaging solutions. 5+ years of hands-on Linux administration experience, including system, user, and machine administration, package management, and scripting with Python and Bash. 5+ years of experience deploying and operating cloud applications, particularly across AWS and GCP; Azure experience is a plus. 3+ years of experience managing high-availability Kafka clusters, with strong knowledge of brokers, producers, consumers, partitions, and Kafka internals. 3+ years of database administration experience with technologies such as MySQL, Cassandra, or other NoSQL databases. 3+ years of experience with log collection and analysis, performance monitoring, and tuning using platforms such as Coralogix, Prometheus, OpenTelemetry, New Relic, Datadog, or similar tools. 3+ years of experience with infrastructure automation, particularly Terraform; Ansible experience is valuable. Strong knowledge of modern CI/CD tools, automation practices, and deployment workflows. Hands-on experience with AI-assisted development and automation tools such as Claude, Cursor, Gemini, or similar technologies. Strong understanding of DevOps architecture and operations, including infrastructure strategy, capacity planning, budgeting, cost optimization, implementation, reliability, and scalability. Excellent communication and collaboration skills, with the ability to work effectively across teams, functions, countries, and time zones. Ability to work effectively in a collaborative engineering environment and solve complex infrastructure and distributed-systems challenges. Benefits: 100% remote position. Candidates should be located in the Eastern Time Zone of the United States or Canada, or in Belfast, Northern Ireland, to support collaboration with the broader engineering team. Occasional travel may be required. Opportunity to work on large-scale distributed systems supporting millions of connected devices and users. Exposure to complex challenges across cloud infrastructure, networking, messaging, databases, observability, security, automation, and high availability. Significant opportunity to influence infrastructure architecture, technical direction, scalability, resilience, automation, and cloud efficiency. Collaboration with an experienced international engineering team. How Jobgether works: We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team. We appreciate your interest and wish you the best! Why Apply Through Jobgether? Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time. #LI-CL1
You'll be redirected to Jobgether's application page
Job Details
Salary
Not disclosed
Location
United States
Job type
Full-time
Category
DevOps / SRE
Experience
8+ years
Posted
Today
Job Highlights
- 8+ years level role
- 100% Remote — open to candidates in United States
- Full-time position
About Jobgether
This job is hosted by Jobgether. Clicking Apply opens their site.
Remote Work Style
Mixed
Mix of flexible and scheduled meetings
Your Match
See how well your skills line up with this role, and what you're missing.
AI Cover Letter
Generate a cover letter tailored to this job from your profile.