Prodege

Site Reliability Engineer II (SREII)

Prodege pune, Maharashtra, India

Market Research · 501-1,000 employees

Sep 14
Remote sre Mid (2-5 yrs) Full-time India
Create a free account to apply — email only, no card. You can also save this posting or score it against your profile with AI.

About the role

The SRE II is responsible for deploying, managing, and optimizing cloud infrastructure using Terraform and AWS to ensure stability and scalability. This includes managing CI/CD pipelines, automating operational tasks, and handling incident response to reduce operational toil.

What they look for

Terraform AWS Bash Python PHP MySQL Jenkins GCP CI/CD Infrastructure as Code Monitoring Incident Management Docker Kubernetes Cloud Infrastructure Automation

Requirements

Candidates must have a bachelor's degree or equivalent experience and at least 4 years of IT operations experience focusing on AWS and Terraform. Proficiency in scripting languages like Python, Bash, or PHP and experience with Jenkins is required.

Full description

Job Description:

Read this part first:

This is a build-and-automate role, not a babysit-the-servers one. Prodege is moving fast and changing fast, and we need an SRE who finds that energizing rather than exhausting.

If you want a static environment where nothing changes and the runbook is already written, this isn't your role, and that's okay. But if you're the kind of engineer who sees manual toil and automates it away, who'd rather harden and scale infrastructure than firefight it, and who's comfortable owning production as the systems around you keep evolving, keep reading.

What you'll own:

You'll play a key role in our cloud and data infrastructure, making sure our products run on a stable, scalable, secure foundation. You'll own AWS/Terraform environments, CI/CD pipelines, and MySQL performance, work that directly affects release velocity, site reliability, and customer experience.

Through thoughtful automation and scripting, you'll reduce operational toil, increase consistency, and free engineering teams to focus on shipping features. Paired with strong monitoring, incident response, and documentation, you'll help create predictable, repeatable operations as we grow, and by embedding security best practices and continuously evaluating new tools, you'll help the org run more efficiently while de-risking the infrastructure over time.

What you'll do:

  • Infrastructure management: Use Terraform to define and provision AWS infrastructure. Configure and maintain AWS services (EC2, S3, RDS, Lambda, VPC).
  • Automation and scripting: Build and manage automation scripts and tools in Bash, Python, and PHP to streamline operations and improve efficiency.
  • CI/CD integration: Implement and manage continuous integration and deployment pipelines using Jenkins.
  • Monitoring and optimization: Monitor system performance, availability, and resource usage, and implement optimizations for efficiency and reliability.
  • Incident management: Troubleshoot and resolve infrastructure issues, outages, and performance problems quickly and effectively.
  • Collaboration: Work with cross-functional teams to support application deployments and address infrastructure needs.
  • Documentation: Create and maintain comprehensive documentation for infrastructure configurations, processes, and procedures.
  • Security: Ensure all infrastructure and operations adhere to security best practices and compliance standards.
  • Continuous improvement: Evaluate and adopt new technologies and practices to improve infrastructure performance and operational efficiency.

What success looks like:

You consistently deliver stable, secure, scalable AWS infrastructure that supports our products without surprise outages or performance bottlenecks. Deployments run smoothly through well-maintained CI/CD pipelines and automation, with minimal manual intervention and short lead times for changes.

Systems are actively monitored, incidents are investigated quickly, root causes are documented, and meaningful preventative fixes get implemented. Cross-functional teams feel supported because infrastructure needs are anticipated, clearly communicated, and backed by current documentation. Over time, you're known for reducing operational toil, improving performance and cost efficiency, and thoughtfully introducing new tools and practices that raise the bar on how we run production.

Why this role is worth your time:

  • Real ownership of production: AWS/Terraform environments, CI/CD, and database performance are yours, not a slice of someone else's stack
  • Automation over toil: you're measured on eliminating manual work, not absorbing it
  • Your work is visible in the numbers: release velocity, uptime, and lead time for changes all move because of what you build
  • Room to modernize: you evaluate and introduce new tools, not just maintain what exists
  • Backed for growth: Prodege closed a major Blackstone investment in Q1 2026, which means real momentum and real investment in the infrastructure you'd be running

A bit about Prodege:

Prodege is a marketing and consumer insights platform that helps leading brands, marketers, and agencies answer their business questions, acquire customers, grow revenue, and build brand loyalty. We go the extra mile to "Create Rewarding Moments" for our partners, consumers, and team. We operate with startup speed and a startup appetite for change, backed by the resources of a major investment partner.

The must-haves:

  • Bachelor's degree (or equivalent) in Computer Science, Software Engineering, Information Technology, or a related discipline, or equivalent professional experience in a similar infrastructure/DevOps engineering role
  • 4+ years in IT operations or a related role, with a strong focus on Terraform and AWS
  • Proficiency in Terraform for infrastructure as code (IaC)
  • Hands-on experience with AWS services (EC2, S3, RDS, Lambda)
  • Experience with scripting languages including Bash and Python/PHP
  • Knowledge of Jenkins for CI/CD pipeline management
  • AWS knowledge required; experience with Google Cloud Platform (GCP) is a plus
  • Strong analytical and troubleshooting skills, with the ability to resolve complex infrastructure issues
  • Excellent verbal and written communication, with the ability to convey technical information clearly to technical and non-technical stakeholders

The nice-to-haves:

  • Knowledge across multiple cloud providers
  • Certification in public cloud disciplines
  • Hands-on experience with Docker and Kubernetes
  • Use of AI in deployment and uptime automation

Similar roles