DEFCON AI Logo

DEFCON AI

ML Engineer, Retrieval & Grounded Generation

Posted 8 Hours Ago
Remote
Hiring Remotely in USA
165K-200K Annually
Senior level
Remote
Hiring Remotely in USA
165K-200K Annually
Senior level
Build and operate production retrieval-augmented generation systems, including embeddings, vector retrieval, grounded language-model generation, citation validation, prompt and schema design, model packaging, serving, rollback, and telemetry. Develop bounded model assistance for narrative extraction with provenance and rule-context tracking. Maintain resilient serving paths in secure government cloud environments, including restricted or air-gapped deployments.
The summary above was generated by AI

ABOUT DEFCON AI

RESILIENCE IN THE FACE OF DISRUPTION. DEFCON AI is an insights company that leverages artificial intelligence, mathematical optimization, data analytics, and software engineering for resilient optimization of complex systems.
In today’s dynamically changing world, DEFCON AI’s technology aligns outcomes with operational goals, better decision making, and empowers customers to anticipate assess, and mitigate the impacts of disruptions.

About the Role 

You'll join the analytics and AI engineering team behind a system that genuinely matters: an AI-assisted platform that pulls together records from dozens of external feeds, resolves them to the right person, surfaces what a human reviewer should look at first, and explains every recommendation in plain, defensible terms — running inside a secure government cloud environment. It's the kind of problem where the details you get right are the ones that count, which is exactly what makes it worth doing well. 

As ML Engineer, Retrieval & Grounded Generation, you'll build embeddings, vector storage, and retrieval at scale, and integrate language models so that every piece of generated text is bound to cited source records and citation failures are tested for rather than assumed away. You'll also own prompt and output-schema design; model packaging, versioning, serving, and rollback; and the telemetry hooks that make later measurement possible without manual reconstruction - real infrastructure for a real production system, not a demo. 

This is a fully remote role, with occasional travel to DEFCON AI HQ, customer sites, and vendor facilities as required. 

Key Responsibilities 

  • Implement embeddings, vector storage, and retrieval across a large, provenance-tracked evidence base 
  • Integrate language models so generated text is bound to cited source records; test for citation failures rather than assuming them away 
  • Design prompts and output schemas 
  • Own model packaging, versioning, serving, and rollback 
  • Instrument telemetry for retrieval and generation quality, recommendation/version attribution, overrides, abstentions, grounding failures, latency, throughput, and measurement events defined with Model Test 
  • Provide bounded model assistance for difficult narrative extraction where deterministic processing is insufficient, with every output tied to its source passage 
  • Supply the recorded rule context to every model-assisted step, so each output carries the exact rule versions and ordered context it received 
  • Maintain a modular in-boundary serving path, self-hosted or managed, alongside the primary managed inference service, so the platform does not depend on one provider’s availability or approval 

Required Qualifications 

  • 5+ years of experience, including a production or near-production retrieval-augmented (RAG) system you built yourself 
  • Ability to speak in detail to your retrieval design, which vector store you used and why, how you tested grounding, what citation failures looked like in practice, and how rollback worked 
  • Strong Python, with hands-on experience in embeddings and vector retrieval at scale 
  • Clarity on what actually shipped in past work — prototype, proposal, or deployed code — since that distinction matters more here than the title on a resume 
  • US Citizenship Required 
  • Active US Secret clearance required to start.  

Preferred Qualifications 

  • Experience deploying models into restricted or air-gapped environments 
  • Self-hosted or open-weight model operation 
  • Fine-tuning, adapters, or custom embeddings 
  • Federal DevSecOps, RMF, ATO, or DoW cloud environment experience 
  • Active Top Secret clearance 

What Success Looks Like 

  • Generated explanations that assert no more than the sources support, with the citation path intact and citation failures tested rather than assumed away 
  • A retrieval system that performs at scale on a large, provenance-tracked evidence base 
  • Model rollback that works when it's needed, with telemetry complete enough that measurement does not require manual reconstruction 

What We Offer 

  • A fully remote, results-based environment 
  • Competitive salary, bonus, and equity package 
  • 100% employer paid, comprehensive health insurance including medical, dental, and vision for you and your family 
  • Unlimited PTO, with your manager's approval 
  • Flexible work environment where you manage your work day 
  • 14 weeks of fully-paid parental leave 

Salary Range: $165,000–$200,000. This represents the typical salary range for this position based on experience, skills, and other factors. 

We’re an Equal Opportunity Employer: You’ll receive consideration for employment without regard to race, sex, color, religion, sexual orientation, gender identity, national origin, protected veteran status, or on the basis of disability. 
Applicant Data Disclosure   
By submitting an application, you acknowledge that Defcon AI uses third-party service providers to facilitate its recruitment and hiring processes. These providers include applicant tracking systems, candidate verification platforms, and fraud detection tools (collectively, "Hiring Platforms"). Your application materials, including your résumé, cover letter, work samples, responses to application questions, and any other information you submit, may be transmitted to and processed by these Hiring Platforms for the following purposes:  
  • Managing and administering your application throughout the hiring process; 
  • Verifying the accuracy and authenticity of application materials, including by cross-referencing information you provide against publicly available sources and proprietary databases; 
  • Identifying indicators of potentially fraudulent, fabricated, or materially misleading application content, including but not limited to discrepancies between submitted materials and publicly available professional profiles, geographic anomalies, and fabricated work histories. 
Applications that are flagged through this process as containing indicators of fraud or material misrepresentation may be declined from further consideration. If you have questions about the status of your application or the evaluation process, please contact [email protected].  
 
Defcon AI requires its Hiring Platform providers to process your information solely for the purposes described above and in accordance with applicable law. Your information will be retained only for as long as necessary to fulfill these purposes and any applicable legal obligations, after which it will be deleted in accordance with Defcon AI's data retention policies.
For more information about how your data is used, please refer to our Privacy Policy and Applicant Privacy Notice.  

 

Similar Jobs

2 Minutes Ago
Remote or Hybrid
45K-85K Annually
Junior
45K-85K Annually
Junior
Artificial Intelligence • Fintech • Insurance • Marketing Tech • Software • Analytics
Handles inbound calls and warm leads, consults customers on insurance needs, recommends appropriate Property and Casualty products, and converts prospects into policyholders. The role includes paid licensing and training, customer-focused sales, schedule flexibility, and remote work from an approved setup. Representatives must meet state licensing requirements, communicate persuasively, maintain strong organizational and computer skills, and work four weekdays plus one weekend day.
Top Skills: PcWired High-Speed Internet
18 Minutes Ago
In-Office or Remote
120K-140K Annually
Senior level
120K-140K Annually
Senior level
Artificial Intelligence • Cloud • Software • Infrastructure as a Service (IaaS)
Own end-to-end projects across DigitalOcean’s go-to-market systems, including implementations, integrations, automations, data flows, Salesforce configuration, and AI-powered workflows. Handle tier 2 escalations, troubleshoot cross-system issues, coordinate with vendors, contribute to solution design, document reusable patterns, and mentor an engineer. The role requires deep Salesforce expertise, integration and API experience, SQL and warehouse fluency, familiarity with code, and the ability to deliver durable production solutions.
Top Skills: ApexAPIsClayDealhubFivetranJavaScriptKaiaLlmsMcp IntegrationsOutreachPlanhatPythonSalesforceSnowflakeSQLSumble
19 Minutes Ago
In-Office or Remote
120K-140K Annually
Senior level
120K-140K Annually
Senior level
Artificial Intelligence • Cloud • Software • Infrastructure as a Service (IaaS)
Owns end-to-end projects across DigitalOcean’s go-to-market systems, including implementations, integrations, automations, data flows, Salesforce configuration, tier 2 troubleshooting, and AI-powered workflows. The role partners with technical program management and engineering stakeholders on solution design, coordinates with vendors, resolves root causes, documents reusable patterns, and mentors an engineer. Systems include Salesforce, Outreach, Planhat, DealHub, Clay, Snowflake, and related integrations.
Top Skills: ApexAPIsClayDealhubFivetranJavaScriptKaiaLlmsMcpOutreachPlanhatPythonSalesforceSnowflakeSQLSumble

What you need to know about the NYC Tech Scene

As the undisputed financial capital of the world, New York City is an epicenter of startup funding activity. The city has a thriving fintech scene and is a major player in verticals ranging from AI to biotech, cybersecurity and digital media. It also has universities like NYU, Columbia and Cornell Tech attracting students and researchers from across the globe, providing the ecosystem with a constant influx of world-class talent. And its East Coast location and three international airports make it a perfect spot for European companies establishing a foothold in the United States.

Key Facts About NYC Tech

  • Number of Tech Workers: 549,200; 6% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Capgemini, Bloomberg, IBM, Spotify
  • Key Industries: Artificial intelligence, Fintech
  • Funding Landscape: $25.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Greycroft, Thrive Capital, Union Square Ventures, FirstMark Capital, Tiger Global Management, Tribeca Venture Partners, Insight Partners, Two Sigma Ventures
  • Research Centers and Universities: Columbia University, New York University, Fordham University, CUNY, AI Now Institute, Flatiron Institute, C.N. Yang Institute for Theoretical Physics, NASA Space Radiation Laboratory

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account