5+ years of SRE, DevOps, or production engineering experience operating mission-critical, high-availability systems
Hands-on experience with IT infrastructure platforms, virtualized and hyperconverged environments such as VxRail, and physical production infrastructure
Hands-on experience with Linux systems, including performance tuning, kernel parameters, and security hardening
Track record leading incident response and postmortem processes for customer-impacting services
Solid knowledge of networking fundamentals: TCP/IP and routing
Proficiency in at least one scripting or programming language (Python, Shell scripting) for automation and tooling
Experience with observability stacks (Prometheus, Grafana) and distributed tracing
AWS infrastructure experience across services like EC2, S3, CloudWatch, KMS, EKS/ECS, VPCs, RDS/DynamoDB, and SQS/SNS
IT Regulatory audit experience, including evidence gathering, control validation, audit prep, and remediation follow-through
Experience with infrastructure-as-code tools (like Puppet) and Git-based workflows
Working knowledge of AI-assisted engineering tools (Claude, Cursor, Claude Code)
Strong written and verbal communication skills in Spanish and English
Experience with Mexican payment rails (SPEI)
SQL knowledge for operational analysis and troubleshooting
Relevant cloud (AWS Certified Cloud Practitioner certification or higher) or IT infrastructure certifications