Talent.com
Capital One
Principal Associate SRECapital One • Mexico City, Mexico City
Principal Associate SRE

Principal Associate SRE

Capital One • Mexico City, Mexico City
Hace más de 30 días
Descripción del trabajo

Overview

WeWork Reforma Latino (97001), Mexico, Ciudad de Mexico, Ciudad de MexicoPrincipal Associate SRE

We're building a Site Reliability Engineering center in Mexico City and hiring Principal Associate SREs to join one of our founding teams. You'll work on payment-critical systems across the Discover Network, Diners Club International, and PULSE - contributing to settlement reliability, alert quality, observability, and automation that directly impacts millions of transactions daily.

This is a ground-floor opportunity. You'll be part of the first cohort of engineers in CDMX, working alongside experienced SRE leaders to build the operational muscle that allows Mexico City to own reliability outcomes independently. Depending on team placement, you'll focus on one of the following areas:

  • Settlement - ensuring batch settlement cycles complete accurately, on time, and in compliance with regulatory requirements across domestic credit/debit and international cross-border networks

  • Alert Signal & Observability - reducing alert noise, building automated severity classification, and creating customer impact dashboards that make incident response faster and more decisive

  • Reliability Automation & Platform Convergence - building automated runbooks, driving Capital One platform adoption, and developing AI-powered remediation workflows

What You'll Do

  • Build and maintain reliability tooling - observability dashboards, automated alerts, runbooks, and remediation scripts that reduce toil and improve mean time to recovery

  • Develop automation solutions - using Python, Java, and shell scripting to eliminate manual operational processes, from certificate rotation to compliance artifact generation

  • Troubleshoot and debug complex production issues - diagnose failures across distributed systems spanning on-prem data centers and AWS, identify root causes, and implement durable fixes

  • Contribute to observability - configure and tune monitoring in Datadog and Observe, build dashboards that surface actionable signals, and reduce unactionable alert volume

  • Support incident response - participate in on-call rotations, respond to production incidents, drive diagnosis, and contribute to blameless postmortems

  • Leverage AI tools to accelerate engineering - use agentic AI automation (Claude Code and others) to develop solutions, generate runbook drafts, and build automation agents

  • Manage secrets and certificates - automate rotation and provisioning, ensuring security posture without manual toil

  • Deliver through CI/CD pipelines - build, test, and deploy automation via continuous integration and API automation frameworks

What Success Looks Like

  • Independently troubleshooting and resolving production issues within your domain without escalation

  • At least one operational process fully automated and running in production

  • Contributing measurably to team OKRs - whether that's alert noise reduction, MTTR improvement, or settlement cycle reliability

  • Producing or improving runbooks and dashboards that your teammates and partner teams actively use

The Environment

You'll work across hybrid on-prem and cloud infrastructure supporting real-time and batch financial transaction systems at global scale. The tech stack includes Python, Java, shell scripting, AWS, Kubernetes, OpenShift, CI/CD pipelines, and API automation frameworks. Observability runs on Datadog and Observe with extensive dashboard configuration. Secret management uses HashiCorp Vault. You'll use agentic AI tools (Claude Code and others) to develop automation solutions and accelerate your engineering output. The systems span three on-prem data centers and AWS, with both modern cloud-native services and legacy payment platforms. Strong troubleshooting and debugging skills are essential.

Basic Qualifications

  • Professional English fluency

  • Bachelor's degree

  • Background in SRE, production operations, or reliability engineering

  • At least 4 years of experience in DevOps Engineering (internship experience does not apply)

  • 4+ years of experience in at least one of the following: Java, Python, Go

  • At least 2 years of experience with Cloud Native technologies (Amazon Web Services, Microsoft Azure, Google Cloud Platform)

  • 2+ years of experience with container orchestration services including Docker or Kubernetes

  • Experience with Shell or Bash scripting

  • At least 2 years of Unix or Linux system administration experience

Preferred Qualifications

  • Experience developing automation solutions using agentic AI tools (Claude Code, Copilot CLI)

  • Troubleshooting and debugging skills across distributed systems

  • Familiarity with payments, financial services, or other regulated high-availability domains

  • Knowledge or experience of Networking concepts (TCP/DNS/TLS)

At Capital One, we respect individual differences in culture, religion, and ethnicity. Likewise, we promote equal opportunities and development for all personnel. In the hiring process, we seek to provide equal employment opportunities to candidates, regardless of race, color, religion, gender, sexual orientation, marital or civil status, national origin, disability, or any other situation protected by federal, state, or local laws.

For technical support or questions about Capital One's recruiting process, please send an email to

Crear una alerta de empleo para esta búsqueda

Principal Associate SRE • Mexico City, Mexico City

Ofertas similares

(Sr) Project Manager, Feasibility Site Activation

ICONMexico City, Mexico

Project Manager, FSA - Homebased - Mexico or Brazil.ICON is a global healthcare intelligence and clinical research organisation united by a mission to bring new medicines and treatments to patients... Mostrar más

CTM & Sr CTM

ICONMexico City, Mexico

Sr CTM and CTM - Mexico - Remote.ICON plc is a world-leading healthcare intelligence and clinical research organization.We’re proud to foster an inclusive environment driving innovation and excelle... Mostrar más

Site Management Associate II

ICONMexico City, Mexico

At ICON Strategic Solutions, you will work in a sponsor-dedicated model, supported by ICON’s global expertise.As the world’s largest FSP organisation, with over 90 sponsor partnerships, we offer st... Mostrar más

Principal, Platform Architecture

MastercardMexico City, Mexico City, MX

Our Purpose Mastercard powers economies and empowers people in 200+ countries and territories worldwide.Together with our customers, we’re helping build a sustainable economy where everyone can pr... Mostrar más

 • Oferta promocionada

Private Equity Associate

T-mappCiudad de México, CDMX, Mexico
Quick Apply

Private Equity en la búsqueda de su próximo.Buscamos un perfil con alta capacidad analítica, curiosidad por los negocios y excelentes habilidades financieras, que quiera desarrollar su carrera dent... Mostrar más

Sr Associate, Investment Banking - M&A

ScotiabankCiudad de México, MX

Contribuye al área de Investment Banking México con la ejecución y cierre de mandatos de fusiones y adquisiciones y mercado de capitales (ECM).Además, cumple son sus metas individuales, y ayuda a l... Mostrar más

 • Oferta promocionada

Principal Engineer II

ICONMexico City, Mexico

Principal Engineer II - Mexico - Hybrid Model (Office with Flex).ICON is a global healthcare intelligence and clinical research organisation united by a mission to bring new medicines and treatment... Mostrar más

Principal SDET

CRH Talento de ITCiudad de México, Ciudad de México, Mexico

Work Shift: On-Site, Monday through Friday 09:00 am to 06:00 pm.BS in Computer science or related field or 13 years of technical experience as an SDE/T or similar role.Experience implementing softw... Mostrar más