Jobiglo

No results.

This job is no longer available

This job expired on 17/08/2026. It no longer accepts applications.

Site Reliability Engineer (SRE) – Regional Multi Project Platform

Qualcomm · Tijuana

🇬🇧 English
AWS OpenStack Terraform Ansible Kubernetes Helm Jenkins Prometheus Grafana ELK stack Kafka Zookeeper NiFi Elasticsearch MySQL Vertica LLM‑based agent frameworks

Job description

About the role

The Site Reliability Engineer (SRE) will lead the design, implementation and operation of a regional multi‑project cloud platform. The role focuses on building highly available, scalable and cost‑efficient infrastructure primarily on AWS while integrating with OpenStack environments.

Key responsibilities

  • Design, build and manage cloud infrastructure on AWS integrated with OpenStack.
  • Develop and maintain Infrastructure as Code using Terraform, Ansible and Kubernetes manifests/Helm charts.
  • Implement scalability, high‑availability, performance and reliability patterns across services and regions.
  • Operate and scale production Kubernetes clusters in large‑scale environments.
  • Partner with development and QA teams to improve system reliability, automate scaling and apply SRE principles.
  • Manage CI/CD pipelines (e.g., Jenkins) and automate provisioning, configuration and lifecycle management.
  • Write and maintain runbooks, automate repetitive tasks and reduce operational toil.
  • Operate, tune and scale data and streaming platforms such as Kafka, Zookeeper, NiFi, Elasticsearch, MySQL and Vertica.
  • Design AI‑assisted operational bots and workflows using LLM‑based agent frameworks.
  • Develop and operate monitoring and observability solutions with Prometheus, Grafana and the ELK stack.
  • Lead incident response, root‑cause analysis and post‑incident reviews.

Required profile

  • Extensive experience operating large‑scale distributed systems.
  • Strong background in cloud infrastructure, especially AWS and OpenStack.
  • Proven expertise in Infrastructure as Code and container orchestration.
  • Hands‑on experience with CI/CD pipelines and automation tools.
  • Ability to design and implement monitoring, alerting and incident‑management processes.

Required skills

  • AWS
  • OpenStack
  • Terraform
  • Ansible
  • Kubernetes
  • Helm
  • Jenkins
  • Prometheus
  • Grafana
  • ELK stack
  • Kafka
  • Zookeeper
  • NiFi
  • Elasticsearch
  • MySQL
  • Vertica
  • LLM‑based agent frameworks (e.g., Claude Agent SDK)

Questions fréquentes

Le salaire n'est pas communiqué publiquement par le recruteur. Vous pouvez postuler et négocier directement avec Qualcomm.
Cliquez sur "Postuler maintenant" en haut de la page. Vous pouvez importer votre CV en 1 clic — Jobiglo extrait automatiquement vos informations et postule pour vous.

Why are you reporting this job?

Thank you for your report. We will review this job.

Explore further

Salaries, guides and searches in Mexico.

💬 Chat with us on Telegram Chat on WhatsApp

Published 3 months ago

19 views · 0 interested

Boost your chances

Upload your CV — we will match you with relevant openings.

Analyzing your CV...

Qualcomm

Tijuana