Jobgether

Senior Site Reliability Engineer, NetBox Delivery

Jobgether Brazil

Internet Marketplace Platforms · 11-50 employees

2 d ago
Remote sre Senior (5-10 yrs) Full-time Brazil
Create a free account to apply — email only, no card. You can also save this posting or score it against your profile with AI.

About the role

You will own and improve the software build and release pipeline while ensuring high availability and performance across cloud and enterprise deployments. Additionally, you will serve as an escalation point for reliability incidents and drive cross-team engineering initiatives to improve system standards.

What they look for

Site Reliability Engineering AWS Kubernetes Terraform PostgreSQL GitHub Actions Django Observability Software supply chain security Helm ArgoCD Prometheus Grafana Python Network automation Incident response

Requirements

Candidates must have 5+ years of experience in software or platform engineering with strong production expertise in Django, PostgreSQL, and container technologies. You should also demonstrate the ability to drive complex technical initiatives and possess deep knowledge of software supply chain security practices.

Benefits

Remote work Competitive compensation Open-source project contribution Professional development Collaborative environment Equal opportunity workplace

Full description

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior Site Reliability Engineer, NetBox Delivery based in Brazil.

This is a senior engineering opportunity focused on the reliability, performance, and delivery of a widely used network automation platform. You will help own the path from core software releases through dependable production deployments across cloud and self-managed environments. The role combines software engineering, SRE, platform engineering, observability, and supply chain security. You will work hands-on with technologies including AWS, Kubernetes, Terraform, PostgreSQL, GitHub Actions, and monitoring platforms. As an early member of a new team, you will have meaningful influence over engineering practices, release processes, and operational standards. You will collaborate across engineering teams, lead reliability initiatives, and address production issues at their technical source. The environment values technical ownership, simplicity, open-source collaboration, clear communication, and continuous improvement.

\n

Accountabilities

• Own and improve the software build and release pipeline, from container base images through availability across cloud and enterprise deployments.

• Establish reliable release handoffs between core engineering and downstream cloud and enterprise teams so releases reach customers efficiently and predictably.

• Improve application performance and production reliability, including startup behavior, PostgreSQL performance, and other critical system components.

• Build and maintain comprehensive observability for both the application and release pipeline, including monitoring, alerting, and service-level objectives (SLOs).

• Serve as an escalation point for performance and reliability incidents, identifying root causes and contributing fixes to the underlying application when appropriate.

• Strengthen software supply chain security and contribute to controls supporting SOC 2 compliance for the build and delivery pipeline.

• Participate in on-call rotations, lead incident response for relevant issues, and facilitate effective postmortems and follow-up improvements.

• Drive cross-team engineering initiatives from technical proposals and RFCs through implementation, migration, and adoption.

Requirements

• 5+ years of experience in software engineering, platform engineering, SRE, or a closely related discipline, with a strong track record of producing robust and maintainable software.

• Production experience with Django and PostgreSQL at scale, including schema design, migration planning, and query-performance optimization under real-world workloads.

• Strong container-building expertise, including base image design, Python dependency management, vulnerability scanning, image signing, and other software supply chain security practices.

• Hands-on experience with technologies such as AWS EC2, VPC, IAM, and RDS; Kubernetes and Helm; GitHub Actions; ArgoCD or FluxCD; Terraform; Prometheus; and Grafana, or comparable tooling.

• Experience working within AI-augmented software development environments, including tools such as Claude Code and practices for making agentic development workflows reliable.

• Demonstrated ability to drive initiatives across team boundaries, from writing technical proposals through completing complex migrations or operational changes.

• Familiarity with network automation or the NetBox ecosystem is a plus.

• Open-source contributions or experience participating in open-source projects is advantageous.

• Experience in B2B software, startups, or high-growth technology organizations is beneficial.

• Deep knowledge of supply chain security technologies such as cosign, Sigstore, or SLSA is a plus.

• Experience operating high-throughput or performance-sensitive systems for enterprise customers is advantageous.

• Strong communication, ownership, problem-solving, and collaboration skills, with an ability to balance reliability, simplicity, and delivery speed.

Benefits

• Remote position based in Brazil as part of a distributed team.

• Competitive compensation aligned with the applicable LATAM compensation structure.

• Opportunity to work on open-source and commercial software used by organizations managing complex networks.

• Meaningful ownership and influence as an early engineer on a newly established delivery and reliability team.

• Exposure to modern cloud infrastructure, Kubernetes, observability, infrastructure-as-code, and software supply chain security practices.

• Collaborative environment focused on clear communication, technical ownership, simplicity, and community.

• Opportunity to contribute to cross-functional engineering initiatives and open-source projects.

• Equal opportunity workplace with support for reasonable accommodations throughout the hiring process.

\nHow Jobgether works:

We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team.

We appreciate your interest and wish you the best!

Why Apply Through Jobgether?

Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time.

#LI-CL1

Similar roles