Plot Technologies, Inc. Logo

Plot Technologies, Inc.

Staff Data Engineer

Reposted 9 Days Ago
Hybrid
New York City, NY, USA
Senior level
Hybrid
New York City, NY, USA
Senior level
Own and harden Plot's core data platform: stabilize high-throughput ingestion, design read-optimized schemas and partitioning, separate transactional vs analytical workloads, ensure performant data access for dashboards and AI agents, write RFCs, and raise reliability and observability across teams.
The summary above was generated by AI
Why Plot

Plot is building the next generation of social intelligence — an AI-native platform that helps some of the world's most influential brands understand culture, find emerging trends, and activate their communities at scale. We're a high-growth, venture-backed startup supported by the co-founders of Reddit and Plaid, with customers like Tory Burch, Lululemon, La Roche-Posay, and other household names.

We operate at the intersection of applied AI and large-scale data systems, turning millions of short-form videos every day into structured representations of culture and consumer behavior. Plot makes unstructured social media legible at scale, enabling brands to build a deep, systematic understanding of trends, audiences, and influence.

You'll join a small but mighty team where engineers own systems end-to-end — not tickets. You'll work directly with founders, shape foundational architecture, and make decisions that define how Plot scales for years to come. If you enjoy ambiguous problems, deep systems thinking, and building infrastructure that unlocks entire product surfaces, Plot is the place. It’s also a great fit if you aspire to be a technical founder one day — we operate with a strong founder’s mindset.

 
What You'll Work On
  • Own the core data systems that power Plot. You’ll take responsibility for how millions of enriched videos flow from ingestion → storage → analytics → product surfaces.

  • Stabilize and harden high-throughput write paths. Our pipelines ingest and enrich data continuously; you’ll diagnose failure modes, reduce error rates, and make ingestion reliable, observable, and predictable.

  • Design data access patterns that stay fast under pressure. You’ll ensure dashboards and analytical queries remain responsive even while large volumes of data are being written concurrently.

  • Separate transactional and analytical concerns. You’ll lead decisions around which workloads belong where, and design the systems and contracts that keep them from stepping on each other.

  • Build read-optimized models for insight, not just storage. This includes partitioning strategies, rollups, and schemas that reflect how data is actually queried — by humans and by AI.

  • Enable AI-native analytics. You’ll shape how our internal tools and AI agents query and reason over data, making it possible to answer complex analytical questions without brute-force scans.

  • Make architectural calls that matter. You’ll evaluate tradeoffs, write RFCs, and guide the evolution of Plot’s data platform as scale, product needs, and AI capabilities grow.

  • Raise the bar for reliability and performance across the team. You’ll define best practices, unblock other engineers, and become a go-to resource for data-related decisions.

     
What We're Looking For
  • Deep experience operating Postgres in production under real-world read/write concurrency. You understand partitioning, locking, autovacuum, connection pooling, and why systems fail in subtle ways.

  • Strong systems intuition. You think in bottlenecks, failure modes, and tradeoffs — not just indexes or tooling.

  • Experience designing or owning analytical data architectures, including read-optimized schemas, aggregations, or large-scale querying patterns.

  • Comfort leading ambiguous, high-impact work. You can take a messy system, form a clear mental model, and chart a pragmatic path forward.

  • High ownership mindset. You like digging into logs, metrics, and real data — and you care deeply about correctness, reliability, and long-term maintainability.

  • Curiosity, pragmatism, and a desire to build foundational systems that other engineers and AI systems depend on. No degree requirements.

  • Minimum 5 years of relevant experience

Similar Jobs

3 Days Ago
Remote or Hybrid
USA
195K-290K Annually
Senior level
195K-290K Annually
Senior level
Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Lead design, build, and deploy of large-scale data platforms for LLMs, RAG, and agentic AI systems. Hands-on coding, architecting fault-tolerant pipelines, establishing MLOps/DataOps best practices, mentoring engineers, and operationalizing research into production across Exabyte-scale distributed systems.
Top Skills: AirflowAWSBigQueryDaskDevsecopsDockerFlinkGCPGoJvmKafkaKubeflowKubernetesLangchainLlamaindexLlmsMlflowOciPulsarPythonRetrieval-Augmented Generation (Rag)RustSagemakerSnowflakeSparkVertex Ai
19 Days Ago
In-Office or Remote
US
165K-250K Annually
Senior level
165K-250K Annually
Senior level
Consumer Web • eCommerce • Machine Learning • Software • Sports • Analytics
Lead architecture and reliability of the data platform, build and harden ingestion (batch and CDC), own BigQuery+dbt transformations, establish testing/observability/governance, mentor engineers, develop Python tooling and APIs, run POCs for lakehouse/open table formats, and define AI-assisted development practices.
Top Skills: AWSCi/CdClaude CodeDbtEstuaryGCPGitGoogle BigqueryHevoLakehouseOpen Table FormatsPythonSQL
22 Days Ago
In-Office
2 Locations
207K-275K Annually
Expert/Leader
207K-275K Annually
Expert/Leader
Cloud • Information Technology • Machine Learning
Lead architecture and standards for CoreWeave's enterprise data ecosystem: design lakehouse/streamhouse architectures, modeling and semantic layers, governance, data quality, lineage, observability, and reusable data products. Drive cross-domain solutions, resolve performance and scalability constraints, evaluate core data technologies, conduct architecture reviews, and mentor senior engineers.
Top Skills: Apache FlussApache HudiApache IcebergApache PaimonAutomqClickhouseData VaultDelta LakeFlinkJavaKafkaKubernetesPulsarPythonRustScalaSparkSQLStarrocksTrino

What you need to know about the NYC Tech Scene

As the undisputed financial capital of the world, New York City is an epicenter of startup funding activity. The city has a thriving fintech scene and is a major player in verticals ranging from AI to biotech, cybersecurity and digital media. It also has universities like NYU, Columbia and Cornell Tech attracting students and researchers from across the globe, providing the ecosystem with a constant influx of world-class talent. And its East Coast location and three international airports make it a perfect spot for European companies establishing a foothold in the United States.

Key Facts About NYC Tech

  • Number of Tech Workers: 549,200; 6% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Capgemini, Bloomberg, IBM, Spotify
  • Key Industries: Artificial intelligence, Fintech
  • Funding Landscape: $25.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Greycroft, Thrive Capital, Union Square Ventures, FirstMark Capital, Tiger Global Management, Tribeca Venture Partners, Insight Partners, Two Sigma Ventures
  • Research Centers and Universities: Columbia University, New York University, Fordham University, CUNY, AI Now Institute, Flatiron Institute, C.N. Yang Institute for Theoretical Physics, NASA Space Radiation Laboratory

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account