Peregrine Technologies, Inc. Logo

Peregrine Technologies, Inc.

Staff Software Engineer, Data Infrastructure

Reposted One Month Ago
Be an Early Applicant
In-Office
New York, NY, USA
200K-275K Annually
Senior level
In-Office
New York, NY, USA
200K-275K Annually
Senior level
Design, build, and operate high-throughput real-time data integration and storage systems. Architect open table format layers at petabyte scale, build and optimize Spark-based batch and streaming pipelines, manage stream processing (Kafka/Flink), orchestrate pipelines (Airflow), and ensure performance, reliability, and cost efficiency across the data stack. Collaborate with platform and product teams and own end-to-end data infrastructure.
The summary above was generated by AI

Backed by leading Silicon Valley investors, Peregrine helps public safety organizations, state and local and governments, federal agencies, and private-sector institutions address society’s challenges with unprecedented speed and accuracy. Our AI-enabled platform turns siloed and disconnected data into operational intelligence — instantly surfacing mission-critical information to empower better, faster decisions that improve outcomes at every touchpoint. Today Peregrine supports hundreds of customers across 30+ states and two countries, serving more than 125 million people — and we’re amplifying our impact as we expand into the enterprise and internationally. 

Team

As an engineering team, we believe strongly that empathy improves our solutions. Seeing how people use the product is a priority and the way we get to the right answer. Engineers will have the opportunity to work closely with our team onsite to understand the variety of use cases that Peregrine serves.

We value both ownership and collaboration—you will take full responsibility for major features and work closely with other engineers to drive them to completion. We believe that humility and empathy are essential for building the right solutions—you will collaborate directly with our deployment team and users as we iterate to solve their problems. Perseverance and creativity are crucial to executing our vision.

Role

We are looking for a Staff Data Infrastructure Engineer to join our growing team, where you will have deep ownership over the data layer that underpins everything Peregrine does. You will architect and build the systems that ingest, store, and serve massive volumes of real-time operational data — enabling our customers to make critical decisions with speed and confidence.

This is a senior individual contributor role for someone who thrives on hard technical problems and brings the experience and judgment to shape foundational infrastructure decisions. You will tackle a wide range of complex challenges, including:

  • Designing and operating a high-throughput, real-time data integration platform across diverse customer environments
  • Architecting a scalable open table format layer for reliable data storage at petabyte scale
  • Building and optimizing distributed data processing pipelines with Apache Spark and adjacent streaming technologies
  • Driving performance, reliability, and cost efficiency across the full data infrastructure stack
  • Collaborating with platform and product engineering teams to define data contracts, schemas, and integration patterns
  • Establishing best practices, tooling, and patterns that raise the quality bar for data infrastructure across the organization

Our stack is constantly evolving but is built on AWS GovCloud, Apache Iceberg, Apache Spark, Apache Kafka, Airflow, Kubernetes, and more.

About You
  • Deep passion for data infrastructure — you care about building systems that are correct, fast, and resilient at scale
  • Thrive on ambiguity and are energized by defining the right solution to hard, open-ended problems
  • Strong technical vision with the ability to translate complex data requirements into clean, durable infrastructure designs
  • Desire to own significant portions of the data stack end-to-end, from ingestion to serving
  • Committed to operational excellence — you build things you’re proud to operate
What We Look For
  • 8+ years of experience architecting and operating large-scale data infrastructure systems in production environments
  • Deep expertise with open table formats, particularly Apache Iceberg — including schema evolution, partitioning strategies, compaction, and time travel
  • Extensive hands-on experience with Apache Spark for batch and streaming data processing at scale
  • Strong background in real-time data integration and stream processing, leveraging technologies such as Apache Kafka, Apache Flink, or equivalents
  • Solid experience with data pipeline orchestration using Airflow or similar tools
  • Strong software engineering fundamentals in Python and/or Scala, with a track record of writing production-quality code
  • Extensive experience with AWS or comparable cloud platforms, including S3-based data lake architectures
  • Experience with Kubernetes and containerized deployment of data workloads
  • Degree in Computer Science, Engineering, or a related field, or equivalent practical experience
  • Located in New York and open to working in office
Annual Salary + Benefits + Equity (if applicable) + Bonus (if applicable)
$200,000$275,000 USD

Actual compensation is influenced by a wide array of factors including but not limited to skill set, level of experience, certifications or licenses, and specific work location. Information on the benefits offered is here.

Peregrine Technologies is committed to creating an inclusive environment for all employees. We celebrate diversity and are a proud equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, gender, gender identity or expression, sexual orientation, national origin, genetics, disability, age, or veteran status.

Similar Jobs

4 Days Ago
In-Office
New York, NY, USA
Senior level
Senior level
Angel or VC Firm • Fintech
Build and operate Helios’s acquisition and document-processing infrastructure, including web crawlers, connectors, fault-tolerant pipelines, multilingual document and media processing, OCR, extraction, transcription, and enrichment. Establish data contracts, provenance, versioning, observability, security, and correction propagation across distributed systems. Scale batch and real-time CPU/GPU workloads while supporting international sources and restricted environments.
Top Skills: Air-Gapped EnvironmentsAudio And Video ProcessingCloud InfrastructureCpu ComputingData ConnectorsDistributed ProcessingDocument ConversionGovcloudGpu ComputingLayout AnalysisMultilingual TranscriptionOcrSpeaker DiarizationWeb Crawlers
26 Days Ago
In-Office
New York City, NY, USA
185K-285K Annually
Senior level
185K-285K Annually
Senior level
Artificial Intelligence • Information Technology • Software
Build and operate Helios’s acquisition and document-processing infrastructure, including web crawlers, connectors, fault-tolerant pipelines, multilingual document and media processing, OCR, extraction, enrichment, provenance, replay, and backfills. Scale secure CPU/GPU inference systems across cloud and restricted environments. Develop real-time multilingual audio intelligence pipelines with transcription, speaker attribution, timestamp alignment, and tone analysis while maintaining reliability, observability, cost efficiency, and data quality.
Top Skills: Air-Gapped EnvironmentsCloud InfrastructureCpuDocument ConversionGovcloudGpuLayout AnalysisOcrReal-Time TranscriptionSpeaker DiarizationStructural ExtractionWeb Crawlers
One Month Ago
In-Office or Remote
New York, NY, USA
405K-485K Annually
Expert/Leader
405K-485K Annually
Expert/Leader
Artificial Intelligence • Natural Language Processing • Generative AI
Design, build, and operate scalable, secure data infrastructure including access control, financial reporting pipelines, cloud storage reliability, and data platform tooling. Collaborate with data scientists, analysts, and stakeholders to ensure data integrity, performance, cost-efficiency, compliance, and rapid recovery at petabyte scale.
Top Skills: AirflowAWSBigQueryBigtableCloud ComposerDbtFivetranGCPGcsGoJavaKubernetesPulumiPythonS3SegmentSparkSQLTerraform

What you need to know about the NYC Tech Scene

As the undisputed financial capital of the world, New York City is an epicenter of startup funding activity. The city has a thriving fintech scene and is a major player in verticals ranging from AI to biotech, cybersecurity and digital media. It also has universities like NYU, Columbia and Cornell Tech attracting students and researchers from across the globe, providing the ecosystem with a constant influx of world-class talent. And its East Coast location and three international airports make it a perfect spot for European companies establishing a foothold in the United States.

Key Facts About NYC Tech

  • Number of Tech Workers: 549,200; 6% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Capgemini, Bloomberg, IBM, Spotify
  • Key Industries: Artificial intelligence, Fintech
  • Funding Landscape: $25.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Greycroft, Thrive Capital, Union Square Ventures, FirstMark Capital, Tiger Global Management, Tribeca Venture Partners, Insight Partners, Two Sigma Ventures
  • Research Centers and Universities: Columbia University, New York University, Fordham University, CUNY, AI Now Institute, Flatiron Institute, C.N. Yang Institute for Theoretical Physics, NASA Space Radiation Laboratory

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account