Social Discovery Group Logo

Social Discovery Group

Site Reliability Engineer (SRE)

Posted One Month Ago
In-Office or Remote
Hiring Remotely in India
Mid level
In-Office or Remote
Hiring Remotely in India
Mid level
Own and improve production infrastructure reliability, deployments, Infrastructure-as-Code, Kubernetes environments, automation, CI/CD, monitoring, alerting, and observability. Investigate incidents, optimize system performance, maintain documentation and runbooks, and support DNS, WAF, CDN, and caching infrastructure. The role requires strong Linux administration, Bash scripting, networking, Git, and containerization skills, with independent ownership and collaboration across development and operations teams.
The summary above was generated by AI

Social Discovery Group (SDG) is a group of social discovery companies. SDG solves the problems of loneliness, isolation, and disconnection - transforming virtual intimacy into the new normal. SDG’s products redefine the way people interact and connect with one another.

Our portfolio includes social entertainment platforms designed to connect people online across different cultures and regions of the world.

We bring together a team of like-minded people and IT professionals who specialize in creating and developing globally impactful social discovery products. Our international team of digital nomads works remotely from all over the world.

We’re proud to be a two-time “Great Place to Work” winner (USA & Japan, 2024–2025) and a Top-5 Company for Work-From-Anywhere Jobs (FlexJobs, 2025).

We are looking for a Site Reliability Engineer (SRE) passionate about infrastructure reliability, automation, and the development of scalable production systems.

Your main tasks will be:

  • Own and improve production infrastructure reliability and stability
  • Prepare, execute, and support deployments and infrastructure changes
  • Build and maintain Infrastructure-as-Code solutions using Ansible and Terraform
  • Support and optimize Kubernetes-based and containerized environments
  • Develop automation scripts and internal operational tooling
  • Monitor system health, investigate incidents, and proactively improve observability
  • Participate in CI/CD improvements together with Development, QA, DevOps, and SRE teams
  • Work with monitoring and alerting systems to reduce downtime and improve system performance
  • Maintain technical documentation, runbooks, and operational procedures
  • Support DNS, WAF, CDN, and caching infrastructure where required

We expect from you:

  • 3+ years of experience in SRE, DevOps, System Administration, or Build/Release Engineering
  • Strong Linux administration and troubleshooting skills
  • Hands-on experience with Kubernetes and containerization technologies (Docker/Podman)
  • Experience with CI/CD pipelines, preferably GitLab CI
  • Practical experience with Infrastructure-as-Code and configuration management tools (Ansible and/or Terraform)
  • Experience with observability and monitoring tools such as Prometheus, Grafana, Zabbix, or VictoriaMetrics
  • Good understanding of networking fundamentals, DNS, HTTP/HTTPS, load balancing, and troubleshooting
  • Experience with Git and modern software delivery workflows
  • Ability to work independently, take ownership, and proactively improve infrastructure
  • Fluent Russian level for technical documentation and team communication

Nice to have:

  • AWS or GCP experience
  • RabbitMQ / AMQP experience
  • Cloudflare, Akamai, WAF, CDN experience
  • Experience with tracing and advanced observability tooling

What do we offer:

  • REMOTE OPPORTUNITY to work full-time;
  • Vacation 28 calendar days per year;
  • 7 wellness days per year (time off) that can be used to deal with household issues, to lie down and recover without taking sick leave;
  • Bonuses up to $5000 for recommending successful applicants for positions in the company;
  • 50% payment for professional training, international conferences, and meetings;
  • Corporate discount for English lessons;
  • ​Health benefits. According to the paychecks, if you are not eligible for corporate medical insurance, the company will compensate you with up to $1,000 gross per year per employee. This can be spent on self-purchase of health insurance or on doctor’s fees for yourself and close relatives (spouse, children);
  • ​Workplace organization. The company provides all employees with an equipped workplace and all the necessary equipment (table, armchair, wifi, etc.) in our offices or co-working locations. In the other locations, the company provides reimbursement of workplace costs up to $1000 gross once every 3 years, according to the paychecks. This money can be spent on the rent of the co-working room, on equipping the working place at home (desk, chair, Internet, etc.) during those 3 years;
  • Internal gamified gratitude system: receive bonuses from colleagues and exchange them for our merchandise, team building activities, massage certificates, etc.

Sounds good? Join us now!
The initial pay level or pay range for this role will be shared with candidates during the recruitment process and before the commencement of employment.

Similar Jobs

Yesterday
Easy Apply
Remote
Easy Apply
Expert/Leader
Expert/Leader
Cloud • Security • Software • Cybersecurity • Automation
Provide technical direction for GitLab Dedicated, a managed single-tenant SaaS platform. Lead architecture and transformation across resilience, failover, tenant orchestration, change management, automation, and platform integrations. Identify systemic reliability and scalability risks, establish reusable platform patterns, strengthen service ownership, and guide cross-team technical decisions. Mentor senior engineers and advance engineering excellence across the organization.
Top Skills: Cloud InfrastructureDevsecopsDistributed SystemsGoInfrastructure As CodeObservabilityPythonRuby
Yesterday
Remote
Entry level
Entry level
Artificial Intelligence • Blockchain • Information Technology • Consulting
Build and scale automated infrastructure provisioning across physical data centers and AWS/GCP environments. Develop continuous delivery pipelines for software, servers, and network switches. Manage vendor relationships, procurement, hardware purchasing, and remote workforce setup. Collaborate within a multidisciplinary infrastructure team with shared responsibility for systems and networks. The role requires automation, scripting, systems and network engineering, vendor management, and continuous learning. Familiarity with authentication protocols, Azure Active Directory, and highly available PostgreSQL architectures is desirable.
Top Skills: Active DirectoryAnsibleAWSGCPKerberosMicrosoft Azure Active DirectoryOauth 2.0PostgresSAML
2 Days Ago
Remote
Senior level
Senior level
Software
Build and operate scalable, resilient, distributed services across AWS and Azure. The Senior Site Reliability Developer will automate infrastructure, develop platforms and frameworks, establish monitoring and incident response practices, define runbooks, reduce operational toil, lead delivery projects, mentor engineers, and participate in technical interviews and on-call rotations. The role requires expertise in cloud infrastructure, infrastructure as code, CI/CD, Kubernetes, observability, telemetry, and SRE practices including SLOs, SLIs, and error budgets.
Top Skills: Amazon CloudwatchAmazon EcrAmazon EcsAmazon EksAmazon KinesisAmazon RdsAmazon RedshiftAmazon SqsAnsibleArtifactoryAuth0AWSAzureAzure Container AppsAzure DevopsAzure MonitorCloudamqpDockerDropwizardElasticsearchGitGithub ActionsGitlabHibernateJavaJavaScriptJenkinsKubernetesMongoDBMySQLObserve Inc.OpentelemetryPrometheusPythonRabbitMQReactRedshift SpectrumReduxSpring BootTemporalTerraformTypescript

What you need to know about the Delhi Tech Scene

Delhi, India's capital city, is a place where tradition and progress co-exist. While Old Delhi is known for its rich history and bustling markets, New Delhi is defined by its modern architecture. It's clear the region places a strong emphasis on preserving its cultural heritage while embracing technological advancements, particularly in artificial intelligence, which plays a central role in shaping the city's tech landscape, fueled by investments in research and development.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account