InOrg Global Logo

InOrg Global

Site Reliability Engineer (SRE)

Reposted 13 Days Ago
Remote
Hiring Remotely in India
Mid level
Remote
Hiring Remotely in India
Mid level
The Site Reliability Engineer will monitor and ensure the reliability and performance of platforms using ELK and TICK stacks, bridging development and operations.
The summary above was generated by AI

About VivaOps :

VivaOps is a leading DevSecOps platform company specializing in GitLab - The comprehensive DevOps platform, to transform and secure software development processes. We help organizations to streamline their DevSecOps journey by offering a complete range of GitLab services, from advisory, to implementation and managed services, to accelerate deployment, optimize security, and improve collaboration. Our deep expertise in GitLab enables clients to unlock the full potential of their development pipelines, driving efficiency, innovation, and competitive advantage.

Job Title: Site Reliability Engineer (SRE)

Location: Remote

Shift Timings: 5:30 PM to 3:00 AM IST to ensure support for global operations.

Job Description:

We are seeking a skilled Site Reliability Engineer (SRE) to join our dynamic team. The ideal candidate will have a strong background in both log and metrics monitoring stacks, specifically ELK (Elasticsearch, Logstash, Kibana) and TICK (Telegraf, InfluxDB, Chronograf, Kapacitor). As an SRE, you will be responsible for ensuring the reliability, availability, and performance of our customer’s platforms and services, bridging the gap between development and operations.

Required Skills and Qualifications:

● Minimum 3+ years of experience in Site Reliability Engineering, DevOps, or a related role.

● Proficiency in the ELK stack (Elasticsearch, Logstash, Kibana) for log monitoring.

● Experience with the TICK stack (Telegraf, InfluxDB, Chronograf, Kapacitor) for metrics monitoring.

● Strong scripting skills in languages such as Python, Bash, or Ruby.

● Understanding of Operating System:

Ubuntu(OpenStack) - Must have

Debian and Redhat etc.,

● DevOps Platforms

Gitlab - Good to have

Or similar

● Solid understanding of Grafana and Prometheus.

● Having worked with ServiceNow or something similar.

● Experience with configuration management tools like Ansible, Puppet, or Chef.

● Familiarity with containerization and orchestration tools like Docker and Kubernetes.

● Understanding of cloud platforms (Any of AWS, Azure, or GCP) and their services.

● Bachelor’s degree in computer science, Information Technology, or a related field.

● Excellent problem-solving skills and attention to detail.

● Strong communication and collaboration abilities.

About Inorg :

InOrg handles India Operations for VivaOps and is dedicated to empowering organizations to achieve global growth at scale. InOrg is establishing a Global Capability Center (GCC) for VivaOps.

Similar Jobs

11 Hours Ago
Remote
India
Senior level
Senior level
Information Technology • Marketing Tech • Social Media
Build and manage observability platforms, cloud infrastructure, Kubernetes and ArgoCD environments, and reliability tooling across AWS and on-premises systems. Own patching compliance, vulnerability remediation, incident support, cloud migration, and service performance. Develop AI-assisted automation, consult development teams on monitoring standards, and mentor engineers while improving operational practices and reducing toil.
Top Skills: AksAnsibleArgocdAWSAws CdkEksElasticFargateGkeGoGradleGrafana LgtmJenkinsKubernetesLinuxMavenMimirMySQLOpentelemetryPostgresPrometheusPythonRubyServicenowSite24X7Terraform
2 Days Ago
Remote or Hybrid
India
Entry level
Entry level
Artificial Intelligence • Information Technology • Machine Learning • Professional Services • Software • Analytics • Consulting
Respond to P1/P2 production incidents in a 24/7 rotation, triage infrastructure failures using Kubernetes, RabbitMQ, and PostgreSQL logs, stabilize systems, implement infrastructure fixes, escalate application issues, communicate incident status, and support system hardening and scaling improvements at senior levels.
Top Skills: AWSAzureGCPKubernetesPostgresRabbitMQ
2 Days Ago
In-Office or Remote
India
Senior level
Senior level
Cloud • Security • Software • Cybersecurity
The Senior Site Reliability Engineer improves the reliability, scalability, availability, and performance of distributed content delivery systems. Responsibilities include defining SLOs and SLIs, monitoring platforms, debugging incidents, implementing corrective actions, automating operational processes, participating in design reviews, and guiding scalable infrastructure design. The role collaborates with Product and Engineering teams and applies software engineering, systems administration, cloud, DevOps, and SRE practices.
Top Skills: AdbmsBashCloud ComputingDatadogDevOpsGrafanaJavaScriptOracle SqlPrometheusPythonUnix/Linux

What you need to know about the Delhi Tech Scene

Delhi, India's capital city, is a place where tradition and progress co-exist. While Old Delhi is known for its rich history and bustling markets, New Delhi is defined by its modern architecture. It's clear the region places a strong emphasis on preserving its cultural heritage while embracing technological advancements, particularly in artificial intelligence, which plays a central role in shaping the city's tech landscape, fueled by investments in research and development.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account