Enumerate Logo

Enumerate

Senior Site Reliability Engineer

Reposted 19 Days Ago
Remote
Hiring Remotely in Chile
4K-5K Annually
Senior level
Remote
Hiring Remotely in Chile
4K-5K Annually
Senior level
Lead design and ownership of cloud and platform infrastructure, define reference architectures and governance, build CI/CD and automation pipelines, manage observability and cost optimization, lead incident response and reliability practices, mentor engineers and drive platform modernization across distributed systems.
The summary above was generated by AI

We’re looking for a Senior Site Reliability Engineer who can own the architecture, governance, and cost efficiency of our cloud and platform infrastructure. In this role you’ll design and evolve our production environments, define standards and best practices, and partner with engineering and IT teams to build scalable, reliable systems that are easy to operate and cost-effective to run.

You’ll be a hands-on technical leader: designing reference architectures, building CI/CD and automation pipelines, leading incident response practices, and setting guardrails around security, reliability, and cost management across our platforms.

This role is a remote contractor role. We are seeking candidates located in LATAM.

Key Responsibilities

Architecture & Infrastructure Ownership

  • Design, implement, and evolve cloud infrastructure architectures for high availability, reliability, security, and scale.
  • Define and maintain reference architectures and patterns for services, applications, and environments across the organization.
  • Develop workflow processes and standards for building, deploying, and maintaining applications within a distributed architecture.
  • Lead infrastructure modernization initiatives (e.g., containerization, Kubernetes adoption, infrastructure as code, platform consolidation).

Governance, Standards & Cost Management

  • Establish and enforce governance standards for infrastructure, CI/CD, observability, and operational practices.
  • Define and maintain policies for environment management, access control, configuration management, and change management.
  • Implement cost management practices (e.g., tagging, budget alerts, rightsizing, reservations/committed use, auto-scaling policies) to optimize cloud spend.
  • Partner with product and engineering leadership to balance performance, reliability, and cost-efficiency across environments.
  • Use DORA metrics and industry benchmarks to drive continuous improvement in delivery and operational performance.

CI/CD, Automation & Operations

  • Design, implement, and maintain CI/CD pipelines for multiple applications and environments using tools such as Git, Azure DevOps, GitLab, or Jenkins.
  • Develop and manage automation pipelines for deployment, configuration, and infrastructure management.
  • Build and maintain monitoring, alerting, and logging systems to ensure visibility, high availability, and performance of applications and services.
  • Manage cloud infrastructure resources and services to ensure reliability, security, and scalability.

Incident Management & Reliability

  • Lead incident response efforts, including triage, root cause analysis, and post-incident reviews.
  • Contribute to and maintain incident response processes, runbooks, and on-call practices.
  • Partner with engineering teams to design resilient systems and reduce mean time to recovery (MTTR).

Leadership, Mentorship & Cross-Functional Collaboration

  • Collaborate with software engineering, QA, product, and IT teams to determine the best way to tackle complex infrastructure, security, and delivery challenges.
  • Mentor engineers in DevOps and platform practices, tools, and standards across the organization.
  • Lead departmental initiatives related to DevOps, platform engineering, and infrastructure disciplines; present plans and progress to stakeholders.
  • Drive new department initiatives based on organizational needs and your expertise in modern technologies and industry trends.
  • Stay current on emerging technologies, tools, and best practices; evaluate their potential application within our tech stack.
Required Experience
  • BS or MS in Computer Science, Engineering, or a related technical field, or equivalent practical experience.
  • 6+ years of experience with container orchestration services (Kubernetes preferred).
  • 6+ years of experience administering and deploying CI/CD tooling (e.g., Git, Azure DevOps, Jira, GitLab, Jenkins).
  • 6+ years of experience managing scalable applications in one or more major cloud providers.
  • 8+ years of significant experience with both Windows and Linux operating system environments.
  • 7+ years of experience with scripting and automation using tools such as PowerShell, Bash, or Python.
  • 4+ years of experience with infrastructure-as-code and orchestration platforms (e.g., Terraform, ARM/Bicep, CloudFormation, Ansible, etc.).
  • Experience with compliance and security frameworks, including SOC 2, ISO/IEC 27001, FedRAMP, and NIST SP 800-53, with active participation in audits, controls implementation, and ongoing compliance efforts.
  • Demonstrated expertise designing architectures for scalable, reliable, and secure tech stacks in distributed systems.
  • Demonstrated expertise implementing workflow processes for operating and maintaining applications in distributed architectures.
Qualifications & Skills
  • Strong experience working in agile-leaning software development environments and across varying application stacks.
  • Deep understanding of best practices and IT operations in distributed, cloud-native architectures.
  • Experience defining and implementing governance and guardrails around infrastructure, CI/CD, and security.
  • Strong grasp of cloud cost management and optimization techniques (e.g., usage analysis, rightsizing, scaling policies).
  • Excellent problem-solving, troubleshooting, and incident management skills.
  • Excellent oral and written communication skills; capable of presenting complex technical concepts to technical and non-technical audiences.
  • Process-oriented with strong documentation skills and attention to detail.
  • Ability to translate loosely defined product or platform requirements into robust, scalable technical solutions.

Total monthly compensation:
$4,000$5,000 USD

Similar Jobs

12 Days Ago
Remote
Senior level
Senior level
Cloud • Information Technology • Software • Analytics
Lead SRE tasks within client teams: design and implement GCP-based infrastructure, drive IaC and CI/CD automation, improve observability (Grafana, telemetry), ensure reliability and security, collaborate with engineering leadership, and support cloud platform evolution and operational excellence.
Top Skills: Container ClustersContinuous Integration And Continuous Delivery (Ci/Cd)DatabasesGoogle Cloud Platform (Gcp)GrafanaInfrastructure As CodeLinuxObservability Semantic ConventionsScripting LanguagesServerlessWindows
19 Days Ago
Remote
USA
Senior level
Senior level
Fintech • Information Technology
Operate and improve brokerage platform reliability: on-call incident response, define SLIs/SLOs, enhance observability, deploy infrastructure via GitOps, and own PostgreSQL performance, migrations, HA/DR, and mentoring.
Top Skills: AlertingDnsGitopsGoKubernetesLinuxLoad Balancing (L4)Load Balancing (L7)LogsMetricsPostgresPythonTlsTracingVpc
6 Days Ago
Remote
84K-144K Annually
Senior level
84K-144K Annually
Senior level
Payments
Senior SRE responsible for architecting and building scalable Kubernetes/AWS infrastructure, improving reliability and observability, leading capacity planning and AI enablement, mentoring engineers, defining SLAs/alerting, and driving CI/CD and microservices best practices across a distributed architecture.
Top Skills: Ai Developer Tooling (Claude CodeAWSAws Aurora RdsCi/CdCodexCursor)DatadogElasticsearchEvent-Driven ArchitectureGemini CliGitGitlabGoGrafanaKubernetesMySQLNew RelicPostgresPrometheusPythonQueuesRdsReactReact NativeRestful ApisStream ProcessingTypescript

What you need to know about the NYC Tech Scene

As the undisputed financial capital of the world, New York City is an epicenter of startup funding activity. The city has a thriving fintech scene and is a major player in verticals ranging from AI to biotech, cybersecurity and digital media. It also has universities like NYU, Columbia and Cornell Tech attracting students and researchers from across the globe, providing the ecosystem with a constant influx of world-class talent. And its East Coast location and three international airports make it a perfect spot for European companies establishing a foothold in the United States.

Key Facts About NYC Tech

  • Number of Tech Workers: 549,200; 6% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Capgemini, Bloomberg, IBM, Spotify
  • Key Industries: Artificial intelligence, Fintech
  • Funding Landscape: $25.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Greycroft, Thrive Capital, Union Square Ventures, FirstMark Capital, Tiger Global Management, Tribeca Venture Partners, Insight Partners, Two Sigma Ventures
  • Research Centers and Universities: Columbia University, New York University, Fordham University, CUNY, AI Now Institute, Flatiron Institute, C.N. Yang Institute for Theoretical Physics, NASA Space Radiation Laboratory

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account