- അനുഭവം
- 7–8 yrs
- ശമ്പളം
- INR 500,000 – INR 1,500,000 / year
- ഓപ്പണിംഗുകൾ
- 1
- പോസ്റ്റ് ചെയ്തു
- 4 മണിക്കൂർ മുമ്പ്
- പ്രവർത്തന രീതി
- ഓഫീസിൽ
- വിദ്യാഭ്യാസം
- B.Tech / B.E.
- യോഗ്യത
- Candidates holding a Bachelor of Technology or Bachelor of Engineering degree in any discipline are eligible to apply.
- പുനരാരംഭിക്കുക
- അപേക്ഷിക്കാൻ നിർബന്ധം
നിങ്ങൾ എവിടെ ജോലി ചെയ്യും
ജോലി വിവരണം
Overview
We are seeking an experienced Site Reliability Engineering (SRE) Lead with 7 to 8 years of expertise, proficient in implementing SRE principles alongside DevOps practices. This role requires leading production support activities including L1/L2 support, 24/7 monitoring, incident management, and detailed root cause analysis.
Key Responsibilities
- Lead and manage production support operations, ensuring incident responsiveness and resolution.
- Utilize hands-on experience with monitoring and observability platforms such as ELK, Prometheus, Grafana, LGTM, Dynatrace, New Relic, and Splunk to maintain system health.
- Apply strong troubleshooting skills to address issues in production, cloud environments, and microservices architectures.
- Implement and maintain automation and scripting solutions, predominantly using Python and other scripting languages.
- Operate and optimize container orchestration platforms using Kubernetes.
- Employ Infrastructure as Code (IaC) and Configuration Management tools including Terraform and Ansible.
- Support Continuous Integration and Continuous Deployment (CI/CD) pipelines and oversee deployment readiness processes.
- Ensure high availability and stability in large-scale production systems.
- Leverage knowledge of database systems, including SQL and NoSQL, to optimize performance.
Required Qualifications
- Bachelor's degree in Engineering or Technology in any specialization.
- 7 to 8 years of relevant experience in Site Reliability Engineering and production support.
- Familiarity with cloud service providers like AWS, Azure, or Google Cloud Platform.
- Strong foundation in Linux operating system, networking protocols, and security concepts.
കഴിവുകൾ
നെറ്റ്വർക്കിംഗ്
കുബേർനെറ്റസ്
ടെറാഫോം
മൂലകാരണ വിശകലനം
ഉൽപ്പാദന പിന്തുണ
സംഭവ മാനേജ്മെന്റ്
CI/CD pipelines
DevOps practices
Python Scripting
Ansible
Site Reliability Engineering
Linux Fundamentals
Cloud platforms (AWS, Azure, GCP)
Monitoring Tools (ELK, Prometheus, Grafana, Splunk)
Database Optimization (SQL & NoSQL)