Regístrese para acceder a todas las funciones de nuestro servicio
  • Búsqueda de ofertas de trabajo
  • Favoritos
  • Crear CV
    Nuevo
  • Alertas de empleo

Senior Site Reliability Engineer (LATAM)

Reap

About Reap

Reap is a global financial technology company headquartered in Hong Kong with employees across multiple countries. We enable financial connectivity and access for businesses worldwide by combining traditional finance with stablecoins for efficient money movement.

Through our stablecoin-powered corporate cards, payments, and expense management tools, we streamline financial operations and help businesses scale. Our APIs enable businesses to integrate stablecoin-enabled finance into their own products and services — from issuing Visa cards to facilitating cross-border payments.

Backed by leading investors including Index Ventures and HashKey Capital, Reap is building the future of borderless, stablecoin-enabled finance.

About the Role

Reap is building a Site Reliability Engineering practice, and this role is central to it.

We run card issuing, payouts, FX and stablecoin settlement across multiple AWS regions, under PCI DSS and financial regulation. The platform is growing quickly — into new markets, new products, and now agent-initiated payments — and the infrastructure underneath it needs to grow up with it. That means real service ownership, reliability measured in SLIs and SLOs rather than intuition, and a platform that engineering teams can serve themselves from instead of queueing for.

That work is largely still ahead of us, which is the appeal. You will help decide what reliability means at Reap, what the platform looks like, and what good engineering practice is in this domain — rather than inheriting someone else's answers.

This is a deeply hands-on senior individual contributor role. We expect you to lead by example and drive technical excellence through the systems you build, the standards you set, and the way you help the engineers around you level up. The team is deliberately flat and distributed across time zones.

Technologies You'll Use

  • Cloud: AWS, multi-account across multiple regions — Transit Gateway, PrivateLink, site-to-site VPN, WAF, KMS

  • Compute: ECS/Fargate, Lambda and EKS

  • Infrastructure as Code: Terraform and CloudFormation

  • CI/CD: GitHub Actions, Argo CD

  • Databases: Aurora PostgreSQL, ElastiCache/Redis

  • Messaging and streaming: SQS, EventBridge, Kafka

  • Observability: New Relic, CloudWatch

  • Languages and scripting: Python, Go and Bash

What You'll Do

As a Senior Site Reliability Engineer at Reap, you will join a team of experienced engineers transitioning a DevOps organisation into Site Reliability Engineering. You will own the reliability of a payments platform while rebuilding the foundation it runs on — the lights stay on while the platform gets replaced underneath them. We treat repetitive manual work as a bug in the platform, not a chore for a human, and your job is to delete whole categories of it rather than absorb them faster.

  • Define what reliability means here: set SLIs and SLOs with product and engineering teams, introduce error budgets, and make "ship or stabilise?" a matter of arithmetic rather than argument

  • Take part in our on-call rotation, and help build the incident response and blameless postmortem practice around it

  • Drive Reap to full Infrastructure as Code coverage: bring the remaining legacy infrastructure under IaC, and get to no manual provisioning, no drift, and every resource defined, versioned and reproducible

  • Consolidate our Terraform estate into a coherent, well-structured codebase with module standards, governance and automated drift detection

  • Automate account provisioning and environment setup so that new regions and services can be stood up repeatably and consistently

  • Build the golden paths and self-service interfaces that let product and engineering teams provision, deploy and observe without filing a ticket, with escape hatches for the cases they do not cover

  • Design and implement an ephemeral environment platform so any developer, and any coding agent, can get an isolated production-like environment on demand and have it cleaned up automatically

  • Implement industry-standard observability across logging, metrics and distributed tracing, with alerting that is actionable and trusted

  • Own cloud operations for PCI DSS and regulated financial systems — uptime, failover, capacity, disaster recovery and incident response

  • Embed security into the infrastructure layer: secrets management, least-privilege IAM, network segmentation and compliance controls

  • Build infrastructure that lets AI assistants and agents operate safely: sane blast radius, strong isolation, auditable actions

  • Partner with product and engineering teams so that reliability work is negotiated rather than imposed — we treat the platform as a product and our engineering teams as its customers

Skills We're Looking For

  • Real SRE practice, not just the vocabulary: you have defined SLOs, run error budgets, or built an incident and postmortem process that people actually used

  • Strong Linux and computer-systems fundamentals: internals, networking, and administration. This work rests on them

  • Deep Terraform: state management, module design, and enforcing standards across a team. You have owned a full IaC migration or a greenfield buildout at scale

  • Strong AWS across multi-account, multi-region estates: Control Tower, IAM, networking, RDS, and cost management

  • Containers and orchestration: comfortable operating ECS and Fargate, and able to build and run production Kubernetes at scale — cluster lifecycle and upgrades, autoscaling, resource limits and multi-tenant workload isolation

  • Serverless and event-driven systems: Lambda, SQS and EventBridge in production, including the failure modes that only show up at scale — retries, poison messages, ordering and idempotency

  • GitOps and delivery: Argo CD or equivalent, and CI/CD pipelines with GitHub Actions that engineers actually enjoy using

  • Observability in practice: logging, metrics and tracing with tools such as New Relic, CloudWatch, Datadog or Prometheus — and getting standards adopted, not just published

  • Python, Go or Bash for automation and tooling

  • Full-lifecycle ownership: you own a service across its whole life — from understanding the need, through design and delivery, to observability, incident response, capacity, upgrades and security patching

  • Experience operating in a regulated environment, and comfort with the constraints that come with PCI DSS and financial compliance

  • Self-management and cross-team influence: you can run your own projects and build alignment without authority. Nobody will sequence your work for you

  • Communication as a first-class skill, weighted equally with technical depth. You can take a problem, explain your approach in plain language, and decompose it into work — for a person or for an agent

  • AI-assisted engineering: you already use AI tools seriously for code, review, automation, incident analysis and documentation, and you have opinions about where they work and where they do not

  • Genuine comfort with production incidents. In payments, high blast-radius incidents preempt everything

Bonus Skills

  • Significant experience in SRE, DevOps or infrastructure engineering, including time in an organisation with a mature SRE practice — defined SLOs, self-service deploys, and real on-call

  • Experience in fintech, payments, card issuing or another regulated environment

  • Hands-on PCI DSS work: designing cardholder-data environments, or reducing scope

  • Kubernetes or EKS inside a PCI DSS environment: isolating the cardholder-data environment through namespace and network-policy segmentation, dedicated node groups, admission control and pod security standards — and the audit logging that makes it provable

  • Having designed and operated ephemeral or on-demand environment platforms, including the hard parts — database seeding, service dependencies and secrets in short-lived environments

  • Having built an internal developer platform, account factory or landing zone from scratch

  • Configuration management with Ansible or similar

  • AWS certifications

  • Curiosity about stablecoins and the Web2 / Web3 intersection

After submitting your application, please check your inbox for a confirmation email. If you don't see it, kindly check your spam or junk folder and adjust your settings to ensure future communication reaches your inbox. You can follow the steps here .

Vacante publicada el 4 días atrás
Empleos similares que podrían interesarleBasado en la vacante Senior Site Reliability Engineer (LATAM) en Argentina
  •  ...AI Native Executive Assistant LatAm Remote About Our Client Our client is a fast-moving, early-stage AI health technology company operating...  ...replace manual processes with scalable systems. No traditional engineering background is required, but genuine comfort building with AI... 

    Hire Overseas

    Teletrabajo
    10 días atrás
  •  ...operating capability that helps us safely and reliably manage a growing fleet of customer cloud...  ...cloud changes while partnering with engineering teams to continually improve the...  ...Cloud Operations, Production Operations, Site Reliability Engineering, Platform Operations... 

    Domino Data Lab

    Teletrabajo
    8 días atrás
  • The Mission We aren’t looking for an agency e-commerce media buyer whose entire playbook relies on setting up basic Facebook catalog ads or broad Google Search campaigns. We need a battle-tested Media Buyer who understands the underground mechanics of iGaming, sweepstakes...
    Trabajo remoto
    Argentina
    6 horas atrás
  •  ...Importante empresa de tecnología, con gran presencia en LATAM, ofrece servicios de consultoría y herramientas de software para todas las plataformas. Es una compañía dedicada a la tecnología de la información, especializada en soluciones, seguridad, capacitación y optimización... 
    Buenos Aires
    13 horas atrás
  •  ...recommend the appropriate engagement. Prepare proposals, negotiate scope, and close new business opportunities. Collaborate with senior consultants on larger engagements. Maintain an organized sales pipeline and follow-up process. Use AI-powered sales tools to... 

    VeloceTalent

    Argentina
    3 días atrás
  •  ...information online Ability to prioritize multiple client requests Comfortable learning online platforms and systems Reliable and professional Previous customer service, reservations, hospitality, or coordination experience is helpful... 
    Trabajo remoto

    Sweetepictravels

    Rosario, Santa Fe
    19 horas atrás
  •  ...-focused mindset Comfortable researching information online Able to learn web-based platforms Strong follow-up skills Reliable and professional Customer service, reservations, hospitality, or administrative experience is helpful Benefits Fully... 
    Trabajo remoto

    The Hiring Platform

    Buenos Aires
    19 horas atrás
  •  ...attention to detail Customer-service mindset Strong organizational abilities Comfortable with online systems and research Reliable and responsive Previous reservation or customer service experience is helpful but not required Benefits Remote work... 
    Trabajo remoto

    The Hiring Platform

    Cordoba, Córdoba
    19 horas atrás
  •  ...craft or quality. About the Role We’re looking for a Senior Infrastructure Engineer to improve reliability and stability of Webflow’s customer-facing,...  ...Engineer if you: - Have a background as an infrastructure, site reliability engineer or cloud engineer with an enthusiasm... 

    Webflow

    America, Buenos Aires
    hace un mes
  •  ...looking for a backend-leaning Senior Software Engineer who is energized by distributed systems, cares about reliability as a feature, and wants to...  ...what to fix on their site, and then the agentic workflows...  ...impact, and help the team build reliable systems on top of non-... 

    Webflow

    America, Buenos Aires
    27 días atrás
  •  ...- Ability to identify trends and conversion opportunities within the loan origination process - Partner with product managers and engineers to design data requirements to ensure proper success measurement of new features - Create anomaly detection that identifies shifts... 

    Lendbuzz

    Teletrabajo
    8 días atrás
  •  ...comunicación verbal y escrita. - Empatía y orientación al cliente. Ofrecemos - Misión a nivel regional: lograr que los negocios de LATAM se despreocupen de su logística. - Horario: Lunes a Viernes de 9:00 a 19:00 GMT-6 (CDMX). Con flexibilidad de horario en fechas de lanzamientos... 

    Skydropx

    Teletrabajo
    5 días atrás
  •  ...Data Analyst / Data Engineer – SQL, BI & Data Analytics | Remote Position Type: Full-Time, Remote Working Hours: U. S. Business Hours...  ...a Data Analyst / Data Engineer to transform business data into reliable reporting, actionable insights, and scalable data systems. Depending... 

    Pavago

    Teletrabajo
    4 días atrás
  • 150 AR$ por semana

    QuickTeam is looking for a driven Sales Development Representative to own our sales pipeline end to end — researching leads, getting them on the phone, and closing deals. Location: (Remote) | Schedule: Full-time, 40 hours/week, EST hours | Compensation: $150/week base ...

    QuickTeam

    Teletrabajo
    2 días atrás
  •  ...EY Global Delivery Services (GDS) in Argentina seeks a RMS Global Security Senior Associate to develop expertise across Global Security disciplines and support the STIC with intelligence monitoring, threat assessments and data analytics insights. You will work under Global... 

    Ernst & Young Advisory Services Sdn Bhd

    Buenos Aires
    23 horas atrás
  •  ...Fluent Trade Technologies seeks an experienced DevOps engineer to join a team dedicated to automated financial trading systems. You will integrate our software onto customer servers and Fluent servers, and become a technical expert on all Fluent Trade Technologies.... 

    OpenTalent

    Argentina
    23 horas atrás
  •  ...RMS Global Security Senior Associate - EY GDS Location: CABA Other locations: Primary Location Only Date: Sep 11, 2026 Requisition ID: 1742088 The Senior Associate within RMS-Global Security team would be required to develop subject matter expertise across... 

    Ernst & Young Advisory Services Sdn Bhd

    Buenos Aires
    23 horas atrás
  •  ...Engineer and automate platform solutions to meet and exceed expectations of Product Manager. Proactively evolve and apply DevSecOps methodologies, standards and leading practices. Ensure re-use through consumption and expansion of shared platform technology assets... 

    Pyramid Consulting, Inc

    Argentina
    23 horas atrás
  •  ...Estamos buscando un Semi Senior Backend Developer con con al menos 3 años de experiencia para sumarse a nuestro equipo de desarrollo...  ...Nivel mínimo de educación: Universitario (En Curso) SOMOS LATAM , OPERAMOS EN 15 PAÍSES, DESARROLLAMOS MEDIOS DE PAGO | PAGOS DIGITALES... 

    Credencial Argentina S.A

    Buenos Aires
    23 horas atrás
  •  ...development. This is the job In Argentina, we are seeking a Senior Node.js Engineer to join our engineering team and lead backend development...  ...systems. You will design, build, and maintain scalable, reliable services, collaborate with product and frontend teams, and... 

    Avenga

    Argentina
    23 horas atrás
  •  ...EPAM Systems is seeking a Senior DevOps Engineer who leads outcomes, designs scalable cloud architectures, and improves reliability. You will own automation, CI/CD pipelines, and production stability across enterprise environments. Responsibilities include building... 

    EPAM Systems

    Argentina
    23 horas atrás
  •  ...Scale Up Recruiting Partners is seeking a Backend Engineer to design, build, and maintain scalable backend systems supporting AI and GenAI solutions for enterprise customers. This role focuses on Python backend development, building services that connect AI models... 

    Scale Up Recruiting Partners

    Rosario, Santa Fe
    23 horas atrás
  •  ...EPAM Systems is seeking a Lead Data Software Engineer (Java+GCP) to drive design and implementation of scalable data solutions on Google Cloud Platform. You will mentor engineers, lead data pipelines, and collaborate with stakeholders to deliver high-quality data products... 

    EPAM Systems

    Argentina
    23 horas atrás
  •  ...MSD LATAM is seeking a dedicated finance professional to join the regional team for Argentina, Uruguay, and Paraguay. You will ensure accurate recording, reconciliation, and month-end close while maintaining internal controls and standards. You will interact with internal... 

    MSD LATAM

    Buenos Aires
    23 horas atrás
  • Despegar busca talento en el equipo de crecimiento para liderar la agenda comercial y coordinar acciones entre áreas, con autonomía y foco en resultados medibles. El perfil ideal genera impacto real y propone soluciones de alta calidad. Se valora estudiante avanzado...

    Despegar

    Buenos Aires
    23 horas atrás
  • United Imaging Healthcare - Latin America is seeking an experienced Regional Sales Manager to drive market share for UIH diagnostic imaging products in Argentina. You will lead cross‑functional efforts with marketing, logistics and service to execute projects and achieve...

    United Imaging Healthcare - Latin America

    Buenos Aires
    23 horas atrás
  •  ...Argentina is seeking a Quality Assurance leader to drive Food Safety and Product Quality at the Pacheco and Villa Mercedes production sites. You will oversee teams, ensure compliance with HACCP and global standards, and lead improvement initiatives to protect the company... 

    Mondelēz International

    General Pacheco, Buenos Aires
    23 horas atrás
  •  ...Build What Matters at RevStar Role Title: Data & AI Engineer (Databricks Specialist) Reports To: Data & AI Practice Lead Location: Remote (US-Based / Eastern or Central Time Zone Preferred) Employment Type: Contract Ready to build greenfield Lakehouse solutions... 

    RevStar Consulting Inc

    Argentina
    23 horas atrás
  •  ...EPAM Systems is seeking a DevOps Engineer to strengthen platform reliability through automation, troubleshooting, and resilient cloud operations. You will act as an escalation point for complex incidents, run Root Cause Analysis, and deliver high-risk changes like cluster... 

    EPAM Systems

    Argentina
    23 horas atrás
  • $ 25.000

     ...We are seeking a Senior Python Engineer who brings strong backend development expertise along with a solid grasp of cloud-native solutions...  ...making Maintain code quality, security, performance and reliability Assist with deployment, monitoring and troubleshooting... 

    EPAM Systems

    Argentina
    23 horas atrás

¿Desea recibir más vacantes?

Suscríbase y reciba vacantes similares a Senior Site Reliability Engineer (LATAM). ¡Sea el primero en aplicar!