Remote • Paid 6-Week Internship • Expected Start: November 1, 2026
ABOUT GOVSIGNALSGovSignals is the AI system of work for government contracting. It can take the government longer to buy a capability than an adversary takes to field one, and we exist to close that gap.
We're the only startup managing government contract data with AI in both FedRAMP High and DoW Impact Level 5 environments. Our platform monitors 5,000+ live government data sources, 100,000+ federal and state agencies, and 2,000,000+ government contracts in real time. Our customers in government include offices like DIU and SOCOM, and our customers in industry range from small contractors to Fortune 500 primes with billions in annual awards. In the past 18 months, we've earned FedRAMP High and IL5 authorizations, joined GSA MAS and the MDA SHIELD IDIQ, and signed some of the biggest names in government contracting.
ABOUT THE ROLE
Signal Sources Intern
GovSignals is looking for third- and fourth-year undergraduate students and recent graduates to help expand the public data sources that power our platform. This role combines software development, web research, and hands-on ownership of data quality.
WHAT YOU’LL DO
Build and maintain data scrapers and ingestion pipelines for government data sources across the United States
Use internal tooling and AI agents to research sources, develop extraction approaches, and accelerate implementation
Validate data for accuracy, completeness, consistency, and reliability before it reaches the platform
Monitor existing sources, investigate failures or source changes, and improve scraper resilience
Document source behavior, extraction logic, and data-quality checks so systems are maintainable
Support other technical and operational work across the company as needed
SKILLS & EXPERIENCE
Experience with TypeScript; Python is also valuable for data collection and analysis
Familiarity with web scraping or browser automation using tools such as BeautifulSoup, Playwright, or Selenium
Comfortable working with HTML, APIs, structured and unstructured data, and common formats such as JSON and CSV
Strong attention to detail and an ownership mindset for data quality
Familiarity with Git, debugging, and writing clear technical documentation
Interest in using AI agents effectively while validating their output carefully
Curiosity about public-sector data and the persistence to work through inconsistent websites and source formats
You don't need to check every box. We're looking for someone who is technically curious, learns quickly, and enjoys figuring out how things work. Experience with web scraping is helpful, but strong software fundamentals and a willingness to dig into unfamiliar systems are equally valuable.
WHAT YOU’LL LEARN
You’ll see how a high-velocity startup builds trustworthy data products: defining reliable processes, anticipating downstream effects, and making careful tradeoffs that protect data quality.
Similar Jobs
What you need to know about the NYC Tech Scene
Key Facts About NYC Tech
- Number of Tech Workers: 549,200; 6% of overall workforce (2024 CompTIA survey)
- Major Tech Employers: Capgemini, Bloomberg, IBM, Spotify
- Key Industries: Artificial intelligence, Fintech
- Funding Landscape: $25.5 billion in venture capital funding in 2024 (Pitchbook)
- Notable Investors: Greycroft, Thrive Capital, Union Square Ventures, FirstMark Capital, Tiger Global Management, Tribeca Venture Partners, Insight Partners, Two Sigma Ventures
- Research Centers and Universities: Columbia University, New York University, Fordham University, CUNY, AI Now Institute, Flatiron Institute, C.N. Yang Institute for Theoretical Physics, NASA Space Radiation Laboratory


