Kairos
Back to jobs

Site Reliability Engineer

On-site
DeepIntentPune, IN3 hours agoWebsite
Fresh
Engineering

Compensation

Salary undisclosed
Apply
Share

Description

What You'll Do:
  • Deploy, configure, and maintain Kubernetes clusters for our microservices architecture.
  • Utilize Git and Helm for version control and deployment management.
  • Implement and manage monitoring solutions using Prometheus and Grafana.
  • Work on continuous integration and continuous deployment (CI/CD) pipelines.
  • Containerize applications using Docker and manage orchestration.
  • Manage and optimize AWS services, including but not limited to EC2, S3, RDS, and AWS CDN.
  • Maintain and optimize MySQL databases, Airflow, and Redis instances.
  • Write automation scripts in Bash or Python for system administration tasks.
  • Perform Linux administration tasks and troubleshoot system issues.
  • Utilize Ansible and Terraform for configuration management and infrastructure as code.
  • Demonstrate knowledge of networking and load-balancing principles.
  • Collaborate with development teams to ensure applications meet reliability and performance standards.
Who you are:
  • Bachelor’s degree in engineering (CS / IT) or equivalent degree from a well-known Institute / University.
  • 2+ years of experience in a Site Reliability Engineer role or similar.
  • Proven experience with Kubernetes, Git, Helm, Prometheus, Grafana, CI/CD, Docker, and microservices architecture.
  • Strong knowledge of AWS services, MySQL, Airflow, Redis, AWS CDN.
  • Proficient in scripting languages such as Bash or Python.
  • Hands-on experience with Linux administration.
  • Familiarity with Ansible and Terraform for infrastructure management.
  • Understanding of networking principles and load balancing.
  • Hands-on experience in setting up, optimizing, and securing analytical distributed data sources such as ClickHouse, Druid, or similar distributed database systems for data storage and analytics.
  • Intermediate DBA skills required (mandatory).
  • Experience with Jenkins for continuous integration (Good to have).
  • Basic understanding of Google Cloud Platform (GCP) and data center operations (Good to have). 

Stack

PythonGCPTerraformCI/CDAirflowRedisAWSKubernetesDocker
Posted
Sep 22, 2026
Last seen
Sep 22, 2026
First seen
Sep 22, 2026

Similar roles

Browse more AI jobs