[Instituição de Pagamento] Analista SRE Pleno
Site reliability engineeringCloud InfrastructureSQLIncident ManagementChange Managementmonitoring alertsdatabase performanceOperational ReliabilityProduction SupportTechnical Documentation
About the Role
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Instituição de Pagamento - Analista SRE Pleno based in United States . This is a mid-level Site Reliability Engineering role focused on ensuring the availability, performance, and reliability of financial and payment technology solutions. You will proactively monitor critical payment flows, investigate incidents, and help maintain a resilient and efficient production environment. The role combines operational excellence with cloud infrastructure, database performance, incident management, change management, and continuous improvement. You will work closely with infrastructure, product, technology, and leadership teams to keep services stable and meet contracted SLA and uptime expectations. Your work will directly support secure and reliable payment experiences for businesses and individuals within a regulated financial environment. The position offers an opportunity to contribute to the evolution of a growing payment institution while strengthening operational processes, documentation, and technical practices. You will thrive in this role if you are analytical, proactive, collaborative, and motivated by solving operational challenges in a technology-driven financial ecosystem. This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Instituição de Pagamento - Analista SRE Pleno based in United States . This is a mid-level Site Reliability Engineering role focused on ensuring the availability, performance, and reliability of financial and payment technology solutions. You will proactively monitor critical payment flows, investigate incidents, and help maintain a resilient and efficient production environment. The role combines operational excellence with cloud infrastructure, database performance, incident management, change management, and continuous improvement. You will work closely with infrastructure, product, technology, and leadership teams to keep services stable and meet contracted SLA and uptime expectations. Your work will directly support secure and reliable payment experiences for businesses and individuals within a regulated financial environment. The position offers an opportunity to contribute to the evolution of a growing payment institution while strengthening operational processes, documentation, and technical practices. You will thrive in this role if you are analytical, proactive, collaborative, and motivated by solving operational challenges in a technology-driven financial ecosystem. Accountabilities Create, adjust, and remove monitoring alerts across email and monitoring platforms, evaluating their effectiveness and maintaining the related knowledge base. Monitor payment processing flows and overall platform health to proactively identify potential operational risks. Analyze application server and database performance, including SQL maintenance routines, to support system reliability and efficiency. Monitor compliance with contracted SLA and uptime targets, escalating risks and breaches to coordination and management when necessary. Analyze, identify, and resolve incidents and service requests according to established operational processes. Immediately escalate systemic unavailability and critical incidents to the appropriate coordination and management teams. Collect technical evidence for the creation of issues and work requests related to critical incidents that have not yet been classified as bugs. Document incident resolutions and publish relevant solutions in the knowledge base to strengthen operational continuity. Create and submit formal change requests for system and infrastructure maintenance activities. Plan and execute cloud environment updates and resource-sizing activities in collaboration with infrastructure teams. Support version validation and testing before releases are promoted to production. Maintain the release and patch calendar and communicate planned changes to relevant stakeholders and operational teams. Recommend product and operational improvements based on incident patterns, demand trends, and recurring challenges. Participate in daily team meetings, sharing operational insights and supporting the resolution of issues raised by internal teams. Stay current with emerging technologies, operational procedures, technical manuals, and product releases. Execute activities requiring formal approval in accordance with established governance processes. Manage access controls for authorized users who provide operational and technical support. Requirements: Professional experience in Site Reliability Engineering, production support, infrastructure, systems operations, or a related technology discipline. Strong understanding of system availability, performance, monitoring, incident management, and operational reliability. Experience creating, configuring, and evaluating monitoring alerts and maintaining technical knowledge bases. Ability to analyze application and database performance, with practical knowledge of SQL maintenance routines. Experience supporting cloud environments and coordinating infrastructure resource sizing and system updates. Familiarity with change management processes, formal change requests, version releases, patches, and production deployments. Ability to investigate incidents systematically, collect technical evidence, identify root causes, and escalate critical issues appropriately. Strong organizational skills for managing operational calendars, documentation, procedures, and multiple concurrent priorities. Excellent communication and collaboration skills, with the ability to interact effectively with infrastructure, product, technology, and leadership teams. Proactive mindset focused on identifying risks before they become incidents and continuously improving operational processes. Strong sense of responsibility, attention to detail, and commitment to security, reliability, governance, and regulatory compliance. Comfortable working in a regulated financial and payments environment where operational controls and formal approvals are essential. Benefits: Annual bonus of up to 3.3 salaries. Career development plan supported by 360° performance evaluations, continuous feedback, internal recognition, and regular 1:1 meetings with management. Financial education, psychological guidance, legal counseling, and social support programs. Unimed health insurance and life insurance. Support for flu vaccination and employee wellness initiatives. Meal/food allowance. Coffee and fresh fruit available at the office. Recreation space with leisure areas, barbecue facilities, and games. SESC partnership providing access to cultural, educational, and tourism activities. Education allowance and support for learning new languages. Incentives to attend industry events and pursue professional certifications. Transportation allowance, private parking, and bicycle parking. Birthday voucher. 40-hour workweek with flexible working hours. How Jobgether works: We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team. We appreciate your interest and wish you the best! Why Apply Through Jobgether? Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time. #LI-CL1
You'll be redirected to Jobgether's application page
Job Details
Salary
Not disclosed
Location
United States
Job type
Full-time
Category
Site Reliability Engineering
Experience
Mid
Posted
Today
Job Highlights
- Mid level role
- 100% Remote — open to candidates in Brazil, United States
- Full-time position
About Jobgether
This job is hosted by Jobgether. Clicking Apply opens their site.
Remote Work Style
Mixed
Mix of flexible and scheduled meetings
Your Match
See how well your skills line up with this role, and what you're missing.
AI Cover Letter
Generate a cover letter tailored to this job from your profile.