Top Engineering Jobs in New York City, NY

12 Days AgoSaved
In-Office
2 Locations
143K-210K Annually
Senior level
143K-210K Annually
Senior level
Cloud • Information Technology • Machine Learning
Design and review data center white space infrastructure for high-density computing environments. Coordinate electrical, mechanical, cooling, low-voltage, cabling, and network pathway systems; review drawings, submittals, RFIs, and field conditions; oversee consultants; resolve design and constructability issues; and support construction, commissioning, and deployment. Contribute to design standards, scalability, technical documentation, and mentorship across CoreWeave’s data center portfolio.
Top Skills: AsanaAshraeAutocadBicsiBluebeamBranch CircuitsBuswaysCable TrayFan WallsGrounding SystemsHot Aisle ContainmentLadder RackLiquid CoolingNfpaOverhead ConveyancePdusRack Power DistributionRemote CdusRevitRppsSmartsheetStructured CablingTia-942Uptime Institute
14 Days AgoSaved
In-Office
2 Locations
153K-204K Annually
Senior level
153K-204K Annually
Senior level
Cloud • Information Technology • Machine Learning
Build and operate full-stack applications and AI-facing features for CoreWeave’s internal data platform. Responsibilities include developing TypeScript, React, Next.js, and Python services; designing APIs and relational data models; integrating analytical data platforms; deploying on Kubernetes; and owning production reliability, testing, observability, and on-call support. The role also involves collaborating with non-engineers to define processes, shipping LLM-backed applications, and evaluating AI output quality.
Top Skills: BigQueryCi/CdDelta LakeDockerHelmHudiIcebergKafkaKubernetesLanggraphNext.JsPostgresPythonReactSnowflakeSparkSQLStarrocksTypescript
14 Days AgoSaved
In-Office
2 Locations
101K-134K Annually
Junior
101K-134K Annually
Junior
Cloud • Information Technology • Machine Learning
Build and ship full-stack software for CoreWeave’s internal data infrastructure and operational applications. Develop TypeScript, React, and Next.js frontends; Python services on Kubernetes; SQL queries and data models; and AI-enabled interfaces such as text-to-SQL tools, retrieval systems, and agents. Own scoped features through production, including testing and post-release fixes, while collaborating with senior engineers and internal users.
Top Skills: BigQueryCi/CdDockerGitHelmHttp ApisJavaScriptKubernetesLlmsNext.JsPythonReactSnowflakeSparkSQLStarrocksTypescript
17 Days AgoSaved
In-Office
2 Locations
182K-242K Annually
Senior level
182K-242K Annually
Senior level
Cloud • Information Technology • Machine Learning
Leads systems engineering teams developing high-throughput file, block, and object storage for AI cloud infrastructure. Drives architecture, performance tuning, latency optimization, benchmarking, and 10x scaling across distributed storage systems. Provides technical mentorship, performance management, recruitment, and cross-functional leadership while maintaining a coding-focused culture and maximizing GPU utilization.
Top Skills: CephCloud StorageDistributed SystemsGoGpudirect StorageHigh-Performance ComputingLustreMinioRust
Reposted 18 Days AgoSaved
In-Office
New York, NY, USA
165K-242K Annually
Senior level
165K-242K Annually
Senior level
Cloud • Information Technology • Machine Learning
The role involves leading the Observability Insights initiative at CoreWeave, focusing on developing APIs, agentic tools, and improving telemetry for AI systems. Candidates should possess significant experience in backend engineering and observability systems, especially for multi-tenant environments.
Top Skills: ClickhouseGoGrafanaKubernetesLokiPrometheusPythonVictoria Metrics
New

Cut your apply time in half.

Use ourAI Assistantto automatically fill your job applications.

Use For Free
Application Tracker Preview
23 Days AgoSaved
In-Office
New York, NY, USA
182K-242K Annually
Senior level
182K-242K Annually
Senior level
Cloud • Information Technology • Machine Learning
Lead and develop an infrastructure engineering team responsible for internal Kubernetes platforms and foundational services. Set roadmaps, ownership, operating practices, and success metrics; guide reliable infrastructure and automation; establish testing, observability, SLO, incident response, change management, and on-call practices; resolve cross-team dependencies; and manage hiring, coaching, performance, priorities, and risk communication.
Top Skills: Cloud InfrastructureDistributed SystemsGitopsInfrastructure AutomationKubernetesObservabilityProgressive DeliverySlos
28 Days AgoSaved
In-Office
2 Locations
153K-204K Annually
Senior level
153K-204K Annually
Senior level
Cloud • Information Technology • Machine Learning
Design, develop, and maintain network datapath monitoring and observability infrastructure for GPU cloud services. Build real-time network telemetry and analytics pipelines across host networking, smart NICs, and overlay/underlay networks. Collaborate with DevOps, production, and data platform teams; troubleshoot Kubernetes and cloud networking issues; participate in on-call support, code reviews, architecture decisions, and continuous infrastructure improvements.
Top Skills: BgpCloud InfrastructureGoKernel NetworkingKubernetesKubernetes ControllersKubernetes OperatorsNetwork TelemetryNetwork VirtualizationObservabilityOverlay NetworksPythonSmart NicsSoftware-Defined Networking (Sdn)Tcp/IpUnderlay Networks
29 Days AgoSaved
In-Office
New York, NY, USA
153K-204K Annually
Senior level
153K-204K Annually
Senior level
Cloud • Information Technology • Machine Learning
Design and implement distributed, exabyte-scale storage systems for AI workloads, including S3-compatible object storage and dedicated storage clusters. Improve storage reliability, durability, security, observability, throughput, latency, and resilience using technologies such as RDMA, GPU Direct Storage, and distributed filesystems. Monitor and troubleshoot production systems, develop performance dashboards, analyze telemetry, collaborate across infrastructure teams, and mentor engineers.
Top Skills: CCephClickhouseDaosFuseGoGpu Direct StorageGrafanaKubernetesNfsPrometheusRdmaRustS3
29 Days AgoSaved
In-Office
2 Locations
182K-242K Annually
Senior level
182K-242K Annually
Senior level
Cloud • Information Technology • Machine Learning
Design, build, and operate secure sandboxed runtime environments for Kubernetes and GPU-accelerated workloads. Develop GPU-aware scheduling, isolation, and resource management strategies; optimize container, VM, I/O, and GPU performance; conduct profiling and benchmarking; and contribute to Linux, container runtime, virtualization, and GPU driver architecture. Collaborate with security, platform, and infrastructure teams to establish runtime isolation and performance standards for multi-tenant systems.
Top Skills: BashCC++Container RuntimesGoGpu DriversGvisorKata ContainersKubernetesKubevirtLinuxQemuRustVirtualization
One Month AgoSaved
In-Office
New York, NY, USA
182K-242K Annually
Senior level
182K-242K Annually
Senior level
Cloud • Information Technology • Machine Learning
Lead an engineering team responsible for CI, build systems, artifact management, and release automation. Set technical direction, modernize developer workflows, improve reliability and performance, track delivery metrics, manage technical debt, and drive adoption of platform capabilities. Mentor engineers through regular one-on-ones and career development while partnering with senior technical stakeholders. The role requires hands-on experience with scalable CI and build platforms, Kubernetes, cloud technologies, automation, and infrastructure as code.
Top Skills: BazelCi/CdCloud PlatformsContainerizationGithub ActionsInfrastructure As CodeJfrog ArtifactoryKubernetesTerraform
All Filters
JobType
New Jobs
Job Category
Experience
Industry
Company Name
Company Size

Sign up now Access later

Create Free Account