General Compute Logo

General Compute

Founding Platform Engineer

Posted Yesterday
Be an Early Applicant
Hybrid
New York, NY, USA
Entry level
Hybrid
New York, NY, USA
Entry level
Build the foundational inference cloud platform, including its control plane, API, serving layer, routing, model placement, scheduling, reliability, and observability. The role involves scaling a heterogeneous GPU and ASIC fleet, partnering with data center and model deployment teams, and shaping platform architecture as an early individual contributor at a small startup.
The summary above was generated by AI
About us

General Compute is the neocloud for alternative chips.

Inference is fragmenting: purpose-built silicon from SambaNova, Cerebras, Positron, d-Matrix, and others already beats GPUs on decode, and we productionize that hardware — we buy the racks, find the data center space, and run it for our customers. Each piece of hardware runs the workload it's actually built for: prefill stays on GPUs, decode moves to the chip built for it, and today that means generating tokens 5–7× faster than existing GPU-based competitors. Our customers are frontier labs, fast-growing AI application companies, and asset-light clouds.

We closed a $15M seed round in May 2026, and have since closed a $400M debt facility — $100M funded upfront by Upper90, with the balance available for drawdown — collateralized by our inference chips.

About the role

You will build the inference cloud itself — the control plane, API, and serving layer that turn racks into a sellable product. There's no existing platform team to inherit or manage, no legacy system to work around, and no established playbook to follow — just the platform itself to build, with reliability treated as core infrastructure from day one rather than something bolted on after the first outage.The technical problem is also genuinely unsolved elsewhere. The fleet is heterogeneous by design — GPUs for prefill, multiple ASIC vendors for decode — so there's no single-vendor playbook to lean on; you'll be defining how a mixed-hardware inference cloud gets scheduled, routed, and served reliably, in close partnership with the teams standing up the physical fleet.

What you'll do:
  • Build/own the control plane — routing, model placement, scheduling across a mixed ASIC/GPU pool

  • Build the API and serving layer exposing rack capacity as a sellable product

  • Build in reliability and observability from day one

  • Scale the platform ahead of the demand curve

  • Partner closely with data center deployment and model bring-up teams

  • Be a founding technical voice on platform architecture

What we need from you:
  • Strong systems engineering background on cloud control planes/serving infra at scale

  • Comfort being a high-impact IC rather than a manager

  • Track record building reliability from scratch

  • Comfort with hardware heterogeneity/ambiguity

  • Genuine interest in being an early hire at a ~6–7 person company.

Nice-to-haves:
  • LLM-serving infra experience (vLLM, TGI, Ray Serve, etc.)

  • Experience running non-NVIDIA accelerators (TPUs/ASICs) in production.

Similar Jobs

17 Days Ago
Hybrid
New York, NY, USA
160K-190K Annually
Senior level
160K-190K Annually
Senior level
Artificial Intelligence • Healthtech • Sales • Software
Lead the platform engineering for a voice-first AI product serving life sciences field teams. Own backend services, APIs, real-time systems, multi-agent runtime, cloud infrastructure, integrations (CRM, Slack, Teams), observability, security/compliance, and data/auditability. Ship internal and customer-facing tooling and collaborate closely with co-founders and a small engineering team.
Top Skills: APIsCloud InfrastructureMicrosoft TeamsMulti-Agent SystemsObservabilityReal-Time StreamingSalesforceSlackVeevaVoice Ai
One Month Ago
In-Office
New York, NY, USA
Senior level
Senior level
AdTech • Big Data • Marketing Tech • Analytics
Design and build the platform layer, solve challenging infrastructure and systems problems, define and drive the roadmap, and scale systems to handle trillions of data points to support growth of major consumer brands.
2 Minutes Ago
Hybrid
New York, NY, USA
80K-110K Annually
Senior level
80K-110K Annually
Senior level
AdTech • Big Data • Digital Media • Software
Manages technical escalations for Magnite’s DV+ programmatic advertising platform. Queries large datasets with SQL, troubleshoots issues across systems, APIs, reporting pipelines, and integrations, and applies AI tools to automate workflows and improve efficiency. The role coordinates with Product, Engineering, Account Management, Technical Operations, and external partners while documenting solutions, diagnosing root causes, and driving ambiguous issues to resolution.
Top Skills: Ai ToolsAPIsExcelSQL

What you need to know about the NYC Tech Scene

As the undisputed financial capital of the world, New York City is an epicenter of startup funding activity. The city has a thriving fintech scene and is a major player in verticals ranging from AI to biotech, cybersecurity and digital media. It also has universities like NYU, Columbia and Cornell Tech attracting students and researchers from across the globe, providing the ecosystem with a constant influx of world-class talent. And its East Coast location and three international airports make it a perfect spot for European companies establishing a foothold in the United States.

Key Facts About NYC Tech

  • Number of Tech Workers: 549,200; 6% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Capgemini, Bloomberg, IBM, Spotify
  • Key Industries: Artificial intelligence, Fintech
  • Funding Landscape: $25.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Greycroft, Thrive Capital, Union Square Ventures, FirstMark Capital, Tiger Global Management, Tribeca Venture Partners, Insight Partners, Two Sigma Ventures
  • Research Centers and Universities: Columbia University, New York University, Fordham University, CUNY, AI Now Institute, Flatiron Institute, C.N. Yang Institute for Theoretical Physics, NASA Space Radiation Laboratory

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account