Support and operate OpenShift/Kubernetes platforms and RHEL systems for a trading exchange. Maintain automation (Ansible, Jenkins, ArgoCD, GitOps), cloud workloads (AWS/Azure), authentication/DNS/time services, enterprise hardware/storage, and developer platform integrations. Monitor, respond to incidents, perform patching and changes, and collaborate with developers, traders, and infrastructure teams to ensure security, resilience, and availability.
TXSE is building the next-generation exchange infrastructure to support transparent, efficient, and resilient capital markets. With SEC approval and $275MM in funding, we are currently hiring a Site Reliability Engineer to help with a greenfield infrastructure build out.
Requirements
- Proficiency in automation and scripting tools (Ansible and Terraform are required as they are a large part of our environment).
- Proficiency supporting Linux Systems, general Linux Sys Administration, and Linux command line expertise.
- Monitoring & Incident Response: Monitor platform and infrastructure health, respond to incidents, troubleshoot service issues, and assist with root cause analysis to reduce downtime and operational risk. Exposure to Grafana, Prometheus or similar monitoring systems.
- Cloud Operations: Support workloads in AWS and Azure, including troubleshooting issues related to compute, networking, storage, access, and security in hybrid and multi-cloud environments.
- Enterprise Hardware & Storage Support: Support infrastructure running on Dell PowerEdge servers and assist in maintaining enterprise storage platforms, including performance monitoring, replication support, and capacity tracking.
- Experience with both on-prem and cloud infrastructure
- Experience with CI/CD and developer platform tools such as Jenkins, ArgoCD, GitHub Enterprise, SonarQube, Artifactory, or GitHub Actions runners.
Nice to have
- Red Hat OpenShift and/or Kubernetes in production or non-production environments.
- Multi-tenant cloud experience
- Systems/Software Development experience in Python
- Authentication services, DNS, and time synchronization concepts in distributed environments.
- Experience with enterprise server hardware and storage platforms is a plus
- Experience working in low latency environments, start-ups, prior experience within exchanges/trading firms
- Ability to context switch, and work highly effective in autonomous environments.
Similar Jobs
Fintech • Financial Services
Leads Site Reliability Engineering for enterprise workplace technology platforms, improving stability, availability, performance, observability, and operational resilience. Provides technical leadership, develops Python automation, analyzes telemetry and outages, and designs infrastructure solutions. Drives AI-assisted incident triage, root cause analysis, proactive remediation, workflow automation, and self-healing capabilities. Collaborates with engineers, business teams, and vendors while managing risks, controls, modernization initiatives, and continuous reliability improvements.
Top Skills:
Agentic AiAgileAPIsContainersElasticGenerative AiGrafanaInfrastructure As CodeKubernetesLarge Language ModelsLow-Code/No-Code PlatformsPrometheusPrompt EngineeringPythonRetrieval-Augmented GenerationSplunk
Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Leads teams designing and deploying AI-driven enterprise and cloud security solutions. Responsibilities include developing machine learning systems, integrating data infrastructure and pipelines, performing advanced modeling and analysis, and building AI applications with Python, Java, and C++. The role involves client engagement, strategic problem-solving, stakeholder validation, coaching, process innovation, and operational excellence while addressing complex cybersecurity and privacy challenges.
Top Skills:
Ai SystemsAWSC++Cloud SecurityData EngineeringData PipelinesDatabricksGCPJavaMachine LearningAzurePythonScikit-LearnSnowflakeTensorFlow
Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Leads the design and development of AI-driven enterprise and cloud security solutions. Manages teams delivering data, analytics, and machine learning engineering projects; oversees AI model implementation, data infrastructure, pipelines, and platform deployments. Uses Python and C++ for algorithm development, analyzes data, improves data quality, ensures technical compliance, mentors staff, manages client expectations, and drives innovation across cybersecurity and privacy initiatives.
Top Skills:
AWSC++DatabricksGCPAzurePythonSnowflake
What you need to know about the NYC Tech Scene
As the undisputed financial capital of the world, New York City is an epicenter of startup funding activity. The city has a thriving fintech scene and is a major player in verticals ranging from AI to biotech, cybersecurity and digital media. It also has universities like NYU, Columbia and Cornell Tech attracting students and researchers from across the globe, providing the ecosystem with a constant influx of world-class talent. And its East Coast location and three international airports make it a perfect spot for European companies establishing a foothold in the United States.
Key Facts About NYC Tech
- Number of Tech Workers: 549,200; 6% of overall workforce (2024 CompTIA survey)
- Major Tech Employers: Capgemini, Bloomberg, IBM, Spotify
- Key Industries: Artificial intelligence, Fintech
- Funding Landscape: $25.5 billion in venture capital funding in 2024 (Pitchbook)
- Notable Investors: Greycroft, Thrive Capital, Union Square Ventures, FirstMark Capital, Tiger Global Management, Tribeca Venture Partners, Insight Partners, Two Sigma Ventures
- Research Centers and Universities: Columbia University, New York University, Fordham University, CUNY, AI Now Institute, Flatiron Institute, C.N. Yang Institute for Theoretical Physics, NASA Space Radiation Laboratory


