Proofpoint Logo

Proofpoint

SRE Engineer II

Reposted One Month Ago
Remote
Hiring Remotely in India
Mid level
Remote
Hiring Remotely in India
Mid level
Build, deploy, scale, and operate highly available distributed systems across regions. Maintain Kubernetes infrastructure, Helm charts, CI/CD pipelines, and automation. Implement observability, alerting, incident response, and runbooks. Troubleshoot production issues, perform root cause analysis, and collaborate with development teams to ensure production readiness and scalability.
The summary above was generated by AI

About Us:

 

Proofpoint is a global leader in human- and agent-centric cybersecurity. We protect how people, data, and AI agents connect across email, cloud, and collaboration tools. Over 80 of the Fortune 100, 10,000 large enterprises, and millions of smaller organizations trust Proofpoint to stop threats, prevent data loss, and build resilience across their people and AI workflows. Our mission is simple: safeguard the digital world and empower people to work securely and confidently. Join us in our pursuit to defend data and protect people.

How We Work:

At Proofpoint you’ll be part of a global team that breaks barriers to redefine cybersecurity guided by our BRAVE core values: 

Bold in how we dream and innovate

Responsive to feedback, challenges and opportunities

Accountable for results and best in class outcomes

Visionary in future focused problem-solving

Exceptional in execution and impact

About the role:

We are looking for a skilled Site Reliability Engineer (SRE) to join our team and help build, deploy, scale, and operate highly reliable distributed systems. You will play a key role in deploying and managing platforms across multiple regions, ensuring high availability, scalability, and performance.

Key Responsibilities:

• Deploy, manage, and scale distributed platforms across multiple geographic regions

• Design and maintain Kubernetes-based infrastructure for large-scale applications

• Build and manage Helm charts for efficient and repeatable deployments

• Monitor system health using Grafana dashboards and metrics; proactively identify and resolve issues

• Improve system reliability, performance, and scalability through automation and best practices

 • Handle large-scale deployments and improve infrastructure for growth

• Collaborate with development teams to ensure smooth CI/CD and production readiness

• Implement observability, alerting, and incident response processes

• Troubleshoot production issues and perform root cause analysis

• Write and maintain run books for incident response

Required Qualifications

• 4–5 years of experience in Site Reliability Engineering, DevOps, or similar roles

• Strong hands-on experience with Kubernetes in production environments

 • Strong experience with infrastructure as code (Terraform, git...)

• Strong experience with AWS (eks, vpc, s3, ecr, iam role etc)

• Solid experience with Helm charts for application deployment

• Strong experience in bash scripting and tooling

• Experience with large-scale distributed systems and high-availability architectures

• Strong understanding of containerization, micro-services, and cloud-native ecosystems

• Experience with CI/CD pipelines and automation tools

• Good debugging and problem-solving skills in production environments Preferred Skills

• Proficiency in at least one of the following: Golang, Python

• Experience building and managing Grafana dashboards and metrics

• Knowledge of monitoring and observability stacks (Prometheus, Loki, etc.)

• Experience with multi-region deployments and global infrastructure

Why Proofpoint?

At Proofpoint, we believe that an exceptional career experience includes a comprehensive compensation and benefits package. Here are just a few reasons you’ll love working with us:

  • Competitive compensation

  • Comprehensive benefits

  • Career success on your terms

  • Flexible work environment

  • Annual wellness and community outreach days

  • Always on recognition for your contributions

  • Global collaboration and networking opportunities

 

Our Culture:

Our culture is rooted in values that inspire belonging, empower purpose and drive success-every day, for everyone.

We encourage applications from individuals of all backgrounds, experiences, and perspectives. If you need accommodation during the application or interview process, please reach out to [email protected].


How to Apply

Interested? Submit your application along with any supporting information- we can’t wait to hear from you!

Similar Jobs

6 Days Ago
Remote
India
Mid level
Mid level
Information Technology • Productivity • Software • Manufacturing
Build and operate reliable AWS-based SaaS platforms through observability, infrastructure automation, CI/CD, container and serverless operations, incident response, security controls, and self-healing systems. The role owns services end to end, improves SLOs and incident metrics, maintains Terraform and GitHub Actions automation, participates in 24/7 on-call, and mentors engineers while applying governed AI-assisted engineering practices.
Top Skills: Amazon CloudwatchAmazon Ec2Amazon Ecs FargateAmazon EksAmazon Rds PostgresqlAmazon S3Amazon VpcAWSAws IamAws LambdaBashCi/CdDnsGitGithub ActionsGrafanaOpenobserveOpentelemetryPagerdutyPrometheusPythonTerraform
7 Days Ago
Remote
India
Senior level
Senior level
Software
The Site Reliability Engineer will design scalable infrastructure, automated deployment pipelines, monitoring, and alerting systems. Responsibilities include troubleshooting production incidents, participating in on-call rotations, maintaining systems, improving security and compliance, optimizing reliability and efficiency, and mentoring junior engineers. The role requires collaboration with development and cross-functional teams, technical communication, infrastructure automation, and expertise in distributed systems, cloud platforms, and networking.
Top Skills: AnsibleAWSAzureChefDockerGCPGoGrafanaJavaKubernetesNagiosPrometheusPuppetPythonRubyTerraform
7 Days Ago
In-Office or Remote
India
Junior
Junior
Cloud • Security • Software • Cybersecurity
Build and improve reliable, scalable distributed content delivery systems. Define SLIs and SLOs, enhance monitoring and alerting, analyze performance data, resolve complex incidents, automate operational tasks, and participate in architecture reviews. The role requires scripting, Oracle SQL analysis, Unix/Linux expertise, and experience with observability tools including Prometheus, Grafana, ADBMS, and Datadog. Collaboration with product and cross-functional teams is central to ensuring high availability, performance, and resilience.
Top Skills: AdbmsBashCloud ComputingDatadogDevOpsGrafanaJavaScriptOracle SqlPrometheusPythonUnix/Linux

What you need to know about the Delhi Tech Scene

Delhi, India's capital city, is a place where tradition and progress co-exist. While Old Delhi is known for its rich history and bustling markets, New Delhi is defined by its modern architecture. It's clear the region places a strong emphasis on preserving its cultural heritage while embracing technological advancements, particularly in artificial intelligence, which plays a central role in shaping the city's tech landscape, fueled by investments in research and development.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account