Search Jobs

Search by job, company or skills

Site Reliability Engineer

Site Reliability Engineer

Tata Consultancy Services
Early Applicant
  • Posted a day ago
  • Be among the first 10 applicants

Job Description

TCS Hiring for Site Reliability Engineer

Experience Range: - 06 To 10 Years (Mandatory)

(Note: Candidates below 6 years of IT experience shall not be considered)

Job Locations : Bengaluru, Chennai, Hyderabad

JOB DESCRIPTION

Must-Have Skills:

  • Strong experience as a Site Reliability Engineer, Cloud Platform Engineer, Platform Engineer, DevOps Engineer, or Infrastructure Engineer.
  • Hands-on experience with Google Cloud Platform (GCP).
  • Strong expertise in Kubernetes, preferably GKE.
  • Experience with Terraform and Infrastructure as Code (IaC).
  • Experience with CI/CD tools and platform automation.
  • Strong understanding of SRE concepts including: SLIs, SLOs, Error Budgets, Automation, Toil Reduction, Operational Readiness
  • Experience managing production services, incident management, monitoring, alerting, capacity planning, and resilience.
  • Strong troubleshooting skills across cloud platforms, Kubernetes, networking, security, and applications.
  • Experience with observability tools such as Dynatrace or OpenTelemetry.
  • Experience supporting large-scale distributed systems and business-critical platforms.

Good-to-Have Skills:

  • Experience with Kafka and Confluent Platform.
  • Experience with PostgreSQL and AlloyDB.
  • Experience with Istio and service mesh technologies.
  • Banking or Core Banking domain experience.
  • Experience with Thought Machine Vault.
  • Experience in GCP operations, cloud cost optimization, and FinOps practices.
  • Understanding of operational resilience, security, risk, and compliance in regulated environments.
  • Experience working with strategic technology suppliers and managed service providers.
  • Experience coaching engineering teams on cloud platform standards and SRE best practices.

Key Responsibilities:

  • Build, operate, and support cloud-native platforms on Google Cloud Platform (GCP).
  • Manage and maintain Kubernetes environments, preferably Google Kubernetes Engine (GKE), including deployments, ingress, networking, scaling, upgrades, troubleshooting, and operational support.
  • Implement and manage Infrastructure as Code (IaC) using Terraform.
  • Develop and enhance CI/CD pipelines to automate platform build, deployment, configuration, and operational processes.
  • Own services throughout their lifecycle, including production readiness, change delivery, incident management, problem management, monitoring, alerting, capacity planning, resilience, and continual improvement.
  • Apply Site Reliability Engineering (SRE) principles including automation, toil reduction, SLIs, SLOs, error budgets, operational readiness, and reliability improvement.
  • Troubleshoot issues across cloud infrastructure, Kubernetes, networking, security, integration services, applications, and platform services.
  • Implement and support observability solutions using Dynatrace, OpenTelemetry, or equivalent tools.
  • Support cloud-native platform and data services including Kafka, PostgreSQL, AlloyDB, and related technologies.
  • Work with Istio and service mesh technologies to support secure and scalable service-to-service communication.
  • Collaborate with engineering teams, technology vendors, and managed service providers to ensure operational excellence.
  • Drive cloud platform standards, SRE best practices, and continuous service improvement.

Minimum Qualification:

15 years of Full Time Education

Interested candidates kindly share your updated resume to [Confidential Information] along with filling the following details to maximize your chances of getting calls from TCS. The details are as follows:

Section 1: Basic Details:

  • Candidate Name:
  • Email ID:
  • Phone Number:
  • Current Location (City):
  • Preferred Location (City):

Section 2: Educational Details:

  • Highest Regular Full‑Time Qualification
  • (Diploma / Graduation / Post Graduation / Others)
  • 10th Completion Year
  • 12th / Intermediate Completion Year
  • Graduation Completion Year
  • Post‑Graduation Completion Year (if applicable)

Section 3: Work Experience:

  • Current Company Name
  • Total Experience (in years)
  • Number of Companies Worked
  • Employment Timeline
  • (Example: Company A – 2021 to Present, Company B – 2019 to 2021)

Section 4: Compensation Details:

  • Current CTC (in LPA)
  • Expected CTC (in LPA)
  • Notice Period
  • (Immediate / 15 Days / 30 Days / 60+ Days)

Section 5: Documentation & Compliance (Yes / No):

  • Valid ID Proof Available (Aadhar & Pan Card): (Yes/No)
  • Educational Documents Available: (Yes/No)
  • Employment Documents Available: (Yes/No)
  • UAN Number Available: (Yes/No)

More Info

Job Type:
Industry:
Employment Type:

Key Skills

OpenTelemetry

SRE concepts

AlloyDB

Similar Jobs

8-10 yrs
Bengaluru, India
Skills:
JavaC#UnixData StructuresIdentity And Access ManagementWindowsDatadogDevopsAlgorithmsGcpLinuxApplication SecurityTerraformAzurePythonAWSEntra IDInfrastructure as CodeGoBigPanda
5-7 yrs
Bengaluru, India
Skills:
PrometheusGrafanaDatadogTerraformPythonJavaRustC++Incident ResponseDynatraceSplunkKubernetesProduction codingGoSLISLOObservabilitynixMonitoringInfrastructure automationAI-assisted reliability workflowsAI-assisted developmentCI/CDError-budget
5-8 yrs
Bengaluru, India
Skills:
Automation ScriptingSecurity ComplianceLoggingInfrastructure ManagementMLopsDockerAutomation FrameworksPythonKubernetesDevOps methodologiesInfrastructure as CodetracingLLMOpsmultimodal AI workflowsAiAgentic AI applicationsCI CDMonitoringalertingContinuous TrainingMachine Learning lifecycle managementLarge Language Modelsobservability
4-8 yrs
Bengaluru, India
Skills:
Cloud networkingElkPrometheusBashGrafanaDatadogTerraformLinuxKubernetesPythonAWSGitOpsGoPagerDutyEKSGitLab CICI/CDArgoCDZenduty
6-8 yrs
Bengaluru, India
Skills:
PrometheusBashGrafanaDatadogTerraformAzureKubernetesPythonAWSGKEGitOpsGoCloud Monitoring