Tight Line

DevOps Engineer

Tight Line Rio de Janeiro, Rio de Janeiro, Brazil

Software Development · 51-200 employees

Sep 08
Remote devops Senior (5-10 yrs) Full-time Brazil
Create a free account to apply — email only, no card. You can also save this posting or score it against your profile with AI.

About the role

The DevOps engineer will maintain and improve Kubernetes workloads and GitLab CI/CD pipelines to ensure operational reliability. They will also collaborate with developers to troubleshoot production issues and implement automation for system efficiency.

What they look for

Kubernetes GitLab CI/CD Docker Linux Automation Troubleshooting Infrastructure Monitoring YAML Containerization Scripting Deployment pipelines Observability Technical documentation System administration

Requirements

Candidates must have 4-6 years of relevant DevOps experience with strong proficiency in Kubernetes, Docker, and Linux fundamentals. The role requires the ability to write and modify CI/CD pipelines and possess sufficient English communication skills for technical teamwork.

Full description

Location: Fully remote LATAM region Working hours: 9:00 am–5:00 pm US Eastern Time Engagement: Full-time, long-term contractor

About the role

We’re looking for a hands-on DevOps engineer to help maintain and improve the Helios platform. You’ll work with developers and experienced engineers to support Kubernetes workloads, improve deployment pipelines, and troubleshoot production issues.

This is a mid-level individual contributor position. You’ll learn how our existing systems fit together and contribute through implementation, automation, and operational support.

What you’ll do

  • Maintain and improve Kubernetes workloads, including deployment configurations, service connectivity, and application configuration.
  • Build, maintain, and troubleshoot GitLab CI/CD pipelines for automated testing, container builds, and deployments.
  • Write and optimize multi-stage Dockerfiles, keeping images secure, efficient, and maintainable.
  • Investigate failed deployments, broken builds, Linux issues, and application runtime errors.
  • Write scripts and automation to reduce repetitive work and improve operational reliability.
  • Collaborate with developers to identify problems across application code, containers, pipelines, and infrastructure.
  • Use logs, metrics, and monitoring tools to investigate production issues and support lasting fixes.
  • Maintain technical documentation and runbooks, review changes, and participate in regular team planning.

During your first three to six months, your main priorities will be maintaining Helios workloads, improving CI/CD pipelines, and supporting day-to-day DevOps operations.

Required experience

  • Around 4–6 years of relevant experience, with substantial hands-on DevOps work.
  • Production Kubernetes experience, including deploying workloads and troubleshooting application or configuration issues.
  • GitLab CI/CD experience, with the ability to write and modify YAML pipelines and investigate pipeline failures.
  • Docker experience, including writing and troubleshooting multi-stage builds.
  • Strong Linux fundamentals, including processes, permissions, logs, networking, and resource troubleshooting.
  • Development experience, with the ability to write code, understand application behavior, and contribute to debugging and automation.
  • Experience supporting production systems and working collaboratively with application teams.
  • English communication skills sufficient for technical discussions, documentation, and daily teamwork.

Useful additional experience

These skills are helpful but are not required to apply:

  • Helm charts.
  • Nginx configuration and reverse proxies.
  • Application-level integration with Kafka or RabbitMQ.
  • OpenTelemetry, Elasticsearch/ELK, Grafana, or equivalent observability tools.

Technologies you can learn on the job

You do not need prior experience with every technology in our stack. Depending on your background, you can develop your skills in:

  • Cloudflare Workers and Wrangler.
  • Lua scripting.
  • Service reliability indicators and objectives (SLIs/SLOs), and distributed tracing.
  • Python, Node.js, and C#.
  • PostgreSQL administration beyond basic operations.

Similar roles