This job is no longer available
This job expired on 17/08/2026. It no longer accepts applications.
Site Reliability Engineer (SRE) – Regional Multi Project Platform
Qualcomm · Tijuana
Job description
About the role
The Site Reliability Engineer (SRE) will lead the design, implementation and operation of a regional multi‑project cloud platform. The role focuses on building highly available, scalable and cost‑efficient infrastructure primarily on AWS while integrating with OpenStack environments.
Key responsibilities
- Design, build and manage cloud infrastructure on AWS integrated with OpenStack.
- Develop and maintain Infrastructure as Code using Terraform, Ansible and Kubernetes manifests/Helm charts.
- Implement scalability, high‑availability, performance and reliability patterns across services and regions.
- Operate and scale production Kubernetes clusters in large‑scale environments.
- Partner with development and QA teams to improve system reliability, automate scaling and apply SRE principles.
- Manage CI/CD pipelines (e.g., Jenkins) and automate provisioning, configuration and lifecycle management.
- Write and maintain runbooks, automate repetitive tasks and reduce operational toil.
- Operate, tune and scale data and streaming platforms such as Kafka, Zookeeper, NiFi, Elasticsearch, MySQL and Vertica.
- Design AI‑assisted operational bots and workflows using LLM‑based agent frameworks.
- Develop and operate monitoring and observability solutions with Prometheus, Grafana and the ELK stack.
- Lead incident response, root‑cause analysis and post‑incident reviews.
Required profile
- Extensive experience operating large‑scale distributed systems.
- Strong background in cloud infrastructure, especially AWS and OpenStack.
- Proven expertise in Infrastructure as Code and container orchestration.
- Hands‑on experience with CI/CD pipelines and automation tools.
- Ability to design and implement monitoring, alerting and incident‑management processes.
Required skills
- AWS
- OpenStack
- Terraform
- Ansible
- Kubernetes
- Helm
- Jenkins
- Prometheus
- Grafana
- ELK stack
- Kafka
- Zookeeper
- NiFi
- Elasticsearch
- MySQL
- Vertica
- LLM‑based agent frameworks (e.g., Claude Agent SDK)
Questions fréquentes
Why are you reporting this job?
Explore further
Salaries, guides and searches in Mexico.
Salaries by job title
Boost your chances
Upload your CV — we will match you with relevant openings.
Analyzing your CV...
Qualcomm
Tijuana