Knotch Logo

Knotch

DevOps Engineer

Reposted 10 Days Ago
Remote
Hiring Remotely in Canada
100K-150K Annually
Senior level
Remote
Hiring Remotely in Canada
100K-150K Annually
Senior level
As a DevOps Engineer, you'll build and maintain scalable infrastructure, manage CI/CD pipelines, and optimize systems for performance and reliability in a fast-paced environment.
The summary above was generated by AI

About Knotch

We're a growth-stage technology company helping brands optimize content performance and apply AI to modern marketing. Our culture is fast-paced, entrepreneurial, and highly adaptable—we move quickly, test ideas, and evolve with our customers.


Through the Knotch platform, brands can measure the impact of their content, identify what drives results, and continuously improve performance across channels. By combining performance data, strategic insights, and AI-driven capabilities, we help marketing teams make smarter decisions and get more impact from every piece of content they produce.

About the Role

As we evolve into an AI-native platform powered by agentic systems and large-scale data pipelines, the reliability, scalability, and observability of our infrastructure becomes mission-critical. We’re not just deploying services — we’re operating complex, production-grade AI systems that enterprise clients depend on every day.


As a DevOps Engineer, you’ll take part in building and scaling the foundation that powers everything we ship — from our core platform to our AI agents. You’ll work across infrastructure, CI/CD, observability, and security to ensure our systems are fast, resilient, and cost-efficient.


This role goes beyond "keeping the lights on". You’ll help define how Knotch operates as an AI-first company, shaping infrastructure strategy, enabling developer velocity, and ensuring our systems scale alongside rapid product innovation. If you want to own infrastructure in a high-impact environment, work closely with engineering teams across the stack, and directly influence how production systems are built and operated, this is *that* role. If you've made it this far, please kindly input this code with your application: DEVOPS-ENG-2026.


Responsibilities

  • Design, build, and maintain scalable, secure, and highly available infrastructure across pre-production and production environments.
  • Develop and manage CI/CD pipelines to enable fast, reliable, and repeatable deployments across multiple environments.
  • Own infrastructure as code (IaC) practices using tools like Terraform to ensure consistency and reproducibility.
  • Manage environment lifecycle (development, staging, production), including promotion workflows and configuration management.
  • Partner closely with Engineering, Data, and AI teams to support system performance, reliability, and scalability.
  • Implement and maintain monitoring, logging, and alerting systems to ensure high visibility into system health and performance.
  • Optimize infrastructure for cost, performance, and reliability, especially for compute- and data-intensive AI workloads.
  • Support Kubernetes-based deployments and container orchestration for distributed systems.
  • Contribute to security best practices across infrastructure, including IAM, networking, and application-level protections.
  • Create dashboards and reporting systems to provide visibility into system performance, uptime, and operational metrics.
  • Document architecture, operational processes, and infrastructure decisions to support knowledge sharing and onboarding.
  • Act as a DevOps/SRE partner across teams, helping troubleshoot issues and improve system reliability.


Qualifications
You have a minimum 5+ years of experience in DevOps, Cloud, SRE, or Infrastructure Engineering roles within SaaS, PaaS, or cloud-native environments, with at least 3 years of experience working in GCP and AWS cloud environments.
Must Haves

  • Prior experience in growth-stage and/or startup environement scaling from $10M to $20M+ ARR with a lean team.
  • Recent experience with Google Cloud Provider (GCP) is required, including IAM, networking, and data services.
  • Hands-on experience with Infrastructure as Code tools such as Terraform.
  • Experience building and maintaining CI/CD pipelines (GitHub Actions, ArgoCD, or similar).
  • Solid experience with Kubernetes, Docker, and containerized environments.
  • Familiarity with deployment tools such as Helm.
  • Experience with monitoring and observability tools like Prometheus and Grafana.
  • Strong understanding of system reliability, scalability, and performance optimization.
  • Ability to work across multiple systems and priorities in a dynamic environment.
  • Strong documentation and communication skills, with attention to clarity and detail.
Nice-to-Haves (not mandatory)
  • Supplementary experience supporting AI/ML or data-intensive workloads in production environments.
  • Familiarity with workflow orchestration or data pipeline tools.
  • Experience with cost optimization strategies for cloud infrastructure.
  • Exposure to security frameworks and compliance best practices.
  • Experience working with distributed or globally deployed systems.

How to be Successful

  • Infrastructure ownership mindset: You have experience building services in GCP from scratch. You take responsibility for system reliability, performance, and scalability — not just deployments.
  • Strong DevOps fundamentals: You understand CI/CD, IaC, observability, and containerization deeply and apply best practices consistently.
  • Systems thinking: You think holistically about how services interact, scale, and fail — and design accordingly.
  • Collaboration-first approach: You work closely with engineers across disciplines to enable velocity and reliability.
  • Pragmatic decision-making: You balance speed, cost, and reliability without over-engineering.
  • Operational excellence: You prioritize monitoring, alerting, and incident response as core parts of system design
  • Adaptability: You thrive in fast-moving environments and can context switch effectively across priorities.
  • Continuous improvement mindset: You proactively identify gaps and improve systems, processes, and tooling over time.

Why Join Knotch

We offer a unique opportunity to build and scale infrastructure that powers a truly AI-native platform. Our stack includes modern tools like Kubernetes, Terraform, AWS, Snowflake, and cutting-edge AI systems, all supported by a team actively building at the intersection of data, AI, and enterprise SaaS.


You’ll have real ownership and impact, influencing how systems are designed, deployed, and operated across the company. The work is highly consequential: the infrastructure you build will directly support production AI systems used by enterprise clients. Most importantly, you’ll be part of a team where infrastructure is not an afterthought — it’s a core part of how we innovate and scale.

The expected salary range for this opportunity is $100,000–$150,000 CAD plus our other perks and benefits, depending on skills and experience.


Our Benefits and Perks

Knotch is a fully remote company. Candidates may work from anywhere within Canada or the U.S., with a mandatory EST working time zone. Some of our other great benefits include:

  • Comprehensive medical, dental, and vision insurance eligibility
  • 401(k) plan
  • Unlimited PTO
  • 10+ company-paid holidays
  • A daily company-wide break, and more!


Equal Opportunity Employer

Knotch is a US-based equal opportunity employer. We strive to provide equal opportunities in all of our processes, including our hiring and employee experience. We pride ourselves on our three values: transparency, relentlessness, and inclusiveness.

 

We commit to daily work towards leading with empathy, reducing bias through periodic training, and engaging with and uplifting communities of marginalized groups. We condemn all forms of racism and discrimination on the basis of race, religion, ethnicity, nationality, gender identity, sexual orientation, age, marital status, pregnancy or parenthood status, veteran status, disability status, or any other identifier. We encourage all employees, clients, investors, candidates, vendors, and friends of Knotch to deliver honest feedback directly or anonymously so that we may always seek to improve as an organization dedicated to diversity, equity, inclusion, and belonging.


HQ

Knotch New York, New York, USA Office

Knotch is remote 1st and we plan to stay that way. We have personnel in 4 countries, including the US, Canada, India and Romania, and 14 US states and we are open to adding more.

Similar Jobs

Yesterday
In-Office or Remote
Mid level
Mid level
Information Technology • Analytics • Consulting • Pharmaceutical
Build and maintain CI/CD pipelines with GitHub Actions; provision infrastructure with Ansible/Terraform; manage Docker containers and AWX/Semaphore jobs; improve observability, troubleshoot incidents, support on-call rotations, document systems, and follow secure DevOps practices while growing AWS/CloudOps skills.
Top Skills: AnsibleAnsible TowerAWSAws LokiAwxAzure MirageBashDockerDockerfilesGitGithub ActionsGoLinux/UnixPythonSemaphoreTerraform
Yesterday
Remote
Senior level
Senior level
AdTech • Marketing Tech
Own and improve AWS infrastructure and platform services using IaC (Terraform/Terrateam). Support reliable cloud and Kubernetes deployments, enhance observability (metrics, logs, tracing), troubleshoot distributed systems, create standards and documentation, and partner with engineering teams on scalable, resilient design.
Top Skills: AnsibleArgocdAWSConsulGithub ActionsGrafanaHashicorp VaultKubernetesPackerPrometheusRedisTerraformTerrateam
Yesterday
In-Office or Remote
United States
Mid level
Mid level
Digital Media • Fintech • Gaming • Sports
Maintain and improve cloud infrastructure across multiple countries, manage Kubernetes and CI/CD, optimize performance and costs, handle monitoring and security audits, lead deployments, and mentor junior team members.
Top Skills: Apache RocketmqAws CloudwatchAws CodepipelineAws Ec2Aws LambdaAws OpsworksChefCloudflareDockerDroneDruidEbsElastic JobElasticacheElkFeignGrafanaJavaJenkinsKubernetesLog4J2MemcachedMybatisMySQLNetflix EurekaNetflix RibbonNetflix ZuulNettyNginxOraclePrometheusRancherRedisRedissonRsyslogS3SaltSpring BootTypescriptVpcVuejs

What you need to know about the NYC Tech Scene

As the undisputed financial capital of the world, New York City is an epicenter of startup funding activity. The city has a thriving fintech scene and is a major player in verticals ranging from AI to biotech, cybersecurity and digital media. It also has universities like NYU, Columbia and Cornell Tech attracting students and researchers from across the globe, providing the ecosystem with a constant influx of world-class talent. And its East Coast location and three international airports make it a perfect spot for European companies establishing a foothold in the United States.

Key Facts About NYC Tech

  • Number of Tech Workers: 549,200; 6% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Capgemini, Bloomberg, IBM, Spotify
  • Key Industries: Artificial intelligence, Fintech
  • Funding Landscape: $25.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Greycroft, Thrive Capital, Union Square Ventures, FirstMark Capital, Tiger Global Management, Tribeca Venture Partners, Insight Partners, Two Sigma Ventures
  • Research Centers and Universities: Columbia University, New York University, Fordham University, CUNY, AI Now Institute, Flatiron Institute, C.N. Yang Institute for Theoretical Physics, NASA Space Radiation Laboratory

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account