Job Description

At Freshworks, uptime is sacred. As a Senior Site Reliability Engineer (SRE), you'll be the engineer behind the curtain—designing for resilience, automating recovery, and ensuring our systems stay fast, stable, and observable at scale. You’ll partner closely with engineering, platform, and product teams to shift reliability left and set the standard for performance and availability.

If you live for clean telemetry, root cause resolution, and engineering chaos into confidence, this is your playground.

Responsibilities

  • Write scripts for automating DevOps tasks such as configuration management, provisioning, and deployments using Python, Ruby, or Go.
  • Integrate scripts with DevOps tools and pipelines.
  • Manage user accounts, permissions, and file systems.
  • Perform advanced Linux administration and shell scripting tasks.
  • Automate system administration tasks using shell scripts.
  • Design and implement CI/CD pipelines for automating deployments and testing.
  • Utilize popular CI/CD tools such as Jenkins and GitLab CI/CD.
  • Integrate CI/CD pipelines with version control systems and container orchestration platforms.
  • Set up and manage monitoring and logging solutions.
  • Use tools for collecting, analyzing, and visualizing application and infrastructure logs.
  • Troubleshoot issues based on monitoring and logging data.
  • Utilize Git for version control and collaboration.
  • Perform branching, merging, and conflict resolution using Git.
  • Set up and manage Git repositories.
  • Work effectively with developers, operations teams, and other stakeholders.
  • Document DevOps processes and procedures.
  • Troubleshoot complex DevOps issues.
  • Identify root causes of problems and implement solutions.