Category Labs Logo

Category Labs

Senior DevOps / Infrastructure Engineer

Posted 3 Days Ago
Remote or Hybrid
Hiring Remotely in New York, NY, USA
180K-250K Annually
Senior level
Remote or Hybrid
Hiring Remotely in New York, NY, USA
180K-250K Annually
Senior level
Operate and maintain Monad's globally distributed validator, full node, and archive fleets; own infrastructure-as-code, observability, and release automation; build AI-driven agentic operations and guardrails; harden services, manage secrets, and codify runbooks and tooling to reduce toil and support safe, automated operation of production node and model workloads.
The summary above was generated by AI

Category Labs (formerly known as Monad Labs) is a team of systems engineers and researchers on a mission to design and build at the frontier of decentralized technology. We strive to deliver significant improvements over existing blockchain solutions. After raising $225M in series A funding, led by Paradigm, we are growing our team.

We’re the team behind Monad, a high-performance, EVM-compatible Layer 1 whose public mainnet is now live. We write the core software that runs it: a parallel-execution EVM, a custom state database, and a BFT consensus client, all developed in the open.

The Role

We're looking for a Senior DevOps / Infrastructure Engineer to operate the infrastructure behind Monad, and to push how much of that operation can be driven by AI. You'll keep our globally-distributed validator, full node, and archive fleet healthy across mainnet and testnet, own our infrastructure-as-code and observability, and build the agentic tooling and guardrails that let a small team safely operate a large fleet. As more of our engineering shifts toward AI, this role is central to designing the workflows, deterministic guardrails, and security perimeters within which autonomous agents operate our infrastructure; you'll also help stand up and operate the infrastructure behind our own growing model workloads.

What You'll Do
  • Operate the Monad node fleet: health, sync, upgrades, and recovery across validators, full nodes, archive/historical, and indexer nodes on mainnet and testnet, including safe, staged rollouts and incident response.

  • Own our infrastructure-as-code: Ansible for fleet configuration, Terraform + Atlantis for cloud and DNS, and Kubernetes/Flux (GitOps) for platform services.

  • Build and operate observability and alerting (Prometheus, Grafana, Loki); create dashboards and alerts that catch problems before they page while minimizing false positives.

  • Automate the release pipeline: node upgrades, canary rollouts, snapshot/restore, and the guardrails that bound blast radius (e.g., protecting validators from automated changes).

  • Design and build agentic operations: develop AI agents, tooling (e.g., MCP servers), and runbooks-as-code that let agents safely investigate, diagnose, and execute routine operations, with deterministic guardrails and human oversight.

  • Codify operational knowledge into tools and automation that the whole team, and its agents, can reuse.

  • Harden nodes and services, manage secrets, and continuously drive down manual toil.

Who You Are
  • You have 5+ years in DevOps, SRE, or Infrastructure Engineering, operating production systems at scale.

  • You have strong Linux, systemd, networking, and shell fundamentals, and you're comfortable debugging live systems over SSH.

  • You have deep, hands-on infrastructure-as-code experience with Ansible and Terraform.

  • You have experience with observability stacks (Prometheus, Grafana, Loki, or equivalents).

  • You have hands-on fluency with AI-assisted engineering: you use coding agents and LLM tooling in your daily workflow and have judgment on where it helps and where it's risky.

  • You have experience designing automation with safe guardrails, and you bring calm, methodical incident response.

  • You have programming and scripting experience (e.g., Python, bash).

  • Experience with Kubernetes and GitOps (Flux or Argo) is a plus.

  • Experience building AI agent tooling, MCP servers, or agent orchestration frameworks is a plus.

  • Experience serving inference, either locally or as a service is a plus.

  • Previous experience with blockchain clients or node operations is a plus.

  • A Bachelor of Science in Computer Science, Engineering, or a related field is a plus.

Why Work with Us
  • Challenging problems: You’ll work on extremely challenging problems with massive impact. See our Blogs and Publications & Talks for a flavor of the problems we are solving in the real world.

  • Huge opportunity: The Ethereum Virtual Machine (EVM) standard is ubiquitous, but existing EVM-compatible chains are very slow. Monad’s core innovations offer developers the best of both worlds (portability and performance) and are a game-changer for mass user adoption in crypto.

  • The right team: You’ll be part of a small, exceptional team (engineers and researchers make up 90% of the team).

  • Open by default: Our core software is public on GitHub. You’ll build in the open, and your work ships where the whole ecosystem can see it.

  • Culture: We’re a lean team working together to achieve very ambitious goals. We are united in our culture of collaboration, low ego, and high-quality output. As an early member of our team, you’ll help to shape our culture.

  • Compensation: You’ll receive a competitive salary and equity package.

  • Resources and growth: We’re well-capitalized, with backing from leading venture funds like Paradigm, Electric Capital, Greenoaks, Dragonfly, and Coinbase Ventures. We keep a lean team, and this is a rare opportunity to join. You’ll learn a lot and grow as our company scales.

 
How We Use AI

We’re an AI-native team, and we expect engineers to use coding agents and keep up as the tooling evolves. A few things we believe:

  • AI is leverage, not a crutch. Review what it generates with the same scrutiny you’d give a teammate’s PR, and own every line you ship.

  • Judgment is what matters, not how long you typed by hand. We won’t ask for “N years of [tool].” The stack turns over every few months, so what matters is that you pick up new tools fast and know where and when they apply.

 
Salary and Benefits

The base salary range for this role is $180,000 – $250,000. This reflects the minimum and maximum range across US locations. It does not include benefits, token, or equity incentives. The final offer may vary based on factors such as relevant skills, experience, domain expertise, and work location. If you are based outside of the US, we have geographic considerations that may impact your final compensation.

 

Benefits for all Full-Time Employees:

  • Private health insurance options

  • Flexible paid time off

  • Monthly wellness reimbursement

  • Paid parental leave

 

Benefits for US employees:

  • World-class benefits package with 100% paid medical, dental, and vision insurance including 75% coverage for dependents and HSA + FSA options

  • 401(k) with company match

  • Lunch and dinner stipend (in-office NYC)

 

Benefits for employees hired through an EOR (outside of the US) will be based on EOR offerings and country-specific requirements.

 

Category Labs is an Equal Employment Opportunity (EEO) employer and welcomes all qualified applicants. Applicants will receive fair and impartial consideration without regard to race, sex, color, religion, national origin, age, disability, veteran status, genetic data, or other legally protected status.

HQ

Category Labs New York, New York, USA Office

New York, New York, United States, 10018

Similar Jobs

Yesterday
Remote
United States
Senior level
Senior level
Blockchain • Software
Lead design, deployment, and operation of Caldera's blockchain infrastructure. Build and expand DevOps pipelines, manage Kubernetes/AWS deployments, implement IaC (Terraform/Helm), enhance observability, optimize performance, and maintain infrastructure security for production rollup chains.
Top Skills: AWSCloudflareDatadogDockerGoGrafanaHelmKubernetesPrometheusTerraformTypescriptVercel
4 Days Ago
Remote
USA
175K-250K Annually
Senior level
175K-250K Annually
Senior level
Information Technology
Lead Infrastructure/DevOps for a DoD IL4 web modernization. Build AWS GovCloud IaC with TypeScript/CDK, enforce NIST compliance, architect high-performance CI/CD, implement Zero Trust, use agentic AI to automate infrastructure, and support full-stack debugging and platform reliability.
Top Skills: Aws CdkAws GovcloudCacCdk-NagClaude CodeCodexEcsEksGithub ActionsIcamMerge QueuesMtlsNistNode.jsOracle 19CPl/SqlPostgresPrismaRemote CachingStigTemporalTest ShardingTypescript
16 Days Ago
Remote
USA
Senior level
Senior level
AdTech • Artificial Intelligence • Marketing Tech
As a Senior DevOps Engineer, you will manage and optimize multi-cloud infrastructure, drive Infrastructure as Code, improve CI/CD processes, and enhance system reliability and performance.
Top Skills: AWSBigQueryCloudflareEksElkGCPHetznerKubernetesOpensearchRdsS3SqsTerraform

What you need to know about the NYC Tech Scene

As the undisputed financial capital of the world, New York City is an epicenter of startup funding activity. The city has a thriving fintech scene and is a major player in verticals ranging from AI to biotech, cybersecurity and digital media. It also has universities like NYU, Columbia and Cornell Tech attracting students and researchers from across the globe, providing the ecosystem with a constant influx of world-class talent. And its East Coast location and three international airports make it a perfect spot for European companies establishing a foothold in the United States.

Key Facts About NYC Tech

  • Number of Tech Workers: 549,200; 6% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Capgemini, Bloomberg, IBM, Spotify
  • Key Industries: Artificial intelligence, Fintech
  • Funding Landscape: $25.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Greycroft, Thrive Capital, Union Square Ventures, FirstMark Capital, Tiger Global Management, Tribeca Venture Partners, Insight Partners, Two Sigma Ventures
  • Research Centers and Universities: Columbia University, New York University, Fordham University, CUNY, AI Now Institute, Flatiron Institute, C.N. Yang Institute for Theoretical Physics, NASA Space Radiation Laboratory

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account