DEFCON AI Logo

DEFCON AI

ML Engineer, Retrieval & Grounded Generation

Posted Yesterday
Remote
Hiring Remotely in USA
165K-200K Annually
Senior level
Remote
Hiring Remotely in USA
165K-200K Annually
Senior level
Build and operate production retrieval-augmented generation systems, including embeddings, vector retrieval, grounded language-model generation, citation validation, prompt and schema design, model packaging, serving, rollback, and telemetry. Develop bounded model assistance for narrative extraction with provenance and rule-context tracking. Maintain resilient serving paths in secure government cloud environments, including restricted or air-gapped deployments.
The summary above was generated by AI

ABOUT DEFCON AI

RESILIENCE IN THE FACE OF DISRUPTION. DEFCON AI is an insights company that leverages artificial intelligence, mathematical optimization, data analytics, and software engineering for resilient optimization of complex systems.
In today’s dynamically changing world, DEFCON AI’s technology aligns outcomes with operational goals, better decision making, and empowers customers to anticipate assess, and mitigate the impacts of disruptions.

About the Role 

You'll join the analytics and AI engineering team behind a system that genuinely matters: an AI-assisted platform that pulls together records from dozens of external feeds, resolves them to the right person, surfaces what a human reviewer should look at first, and explains every recommendation in plain, defensible terms — running inside a secure government cloud environment. It's the kind of problem where the details you get right are the ones that count, which is exactly what makes it worth doing well. 

As ML Engineer, Retrieval & Grounded Generation, you'll build embeddings, vector storage, and retrieval at scale, and integrate language models so that every piece of generated text is bound to cited source records and citation failures are tested for rather than assumed away. You'll also own prompt and output-schema design; model packaging, versioning, serving, and rollback; and the telemetry hooks that make later measurement possible without manual reconstruction - real infrastructure for a real production system, not a demo. 

This is a fully remote role, with occasional travel to DEFCON AI HQ, customer sites, and vendor facilities as required. 

Key Responsibilities 

  • Implement embeddings, vector storage, and retrieval across a large, provenance-tracked evidence base 
  • Integrate language models so generated text is bound to cited source records; test for citation failures rather than assuming them away 
  • Design prompts and output schemas 
  • Own model packaging, versioning, serving, and rollback 
  • Instrument telemetry for retrieval and generation quality, recommendation/version attribution, overrides, abstentions, grounding failures, latency, throughput, and measurement events defined with Model Test 
  • Provide bounded model assistance for difficult narrative extraction where deterministic processing is insufficient, with every output tied to its source passage 
  • Supply the recorded rule context to every model-assisted step, so each output carries the exact rule versions and ordered context it received 
  • Maintain a modular in-boundary serving path, self-hosted or managed, alongside the primary managed inference service, so the platform does not depend on one provider’s availability or approval 

Required Qualifications 

  • 5+ years of experience, including a production or near-production retrieval-augmented (RAG) system you built yourself 
  • Ability to speak in detail to your retrieval design, which vector store you used and why, how you tested grounding, what citation failures looked like in practice, and how rollback worked 
  • Strong Python, with hands-on experience in embeddings and vector retrieval at scale 
  • Clarity on what actually shipped in past work — prototype, proposal, or deployed code — since that distinction matters more here than the title on a resume 
  • US Citizenship Required 
  • Active US Secret clearance required to start.  

Preferred Qualifications 

  • Experience deploying models into restricted or air-gapped environments 
  • Self-hosted or open-weight model operation 
  • Fine-tuning, adapters, or custom embeddings 
  • Federal DevSecOps, RMF, ATO, or DoW cloud environment experience 
  • Active Top Secret clearance 

What Success Looks Like 

  • Generated explanations that assert no more than the sources support, with the citation path intact and citation failures tested rather than assumed away 
  • A retrieval system that performs at scale on a large, provenance-tracked evidence base 
  • Model rollback that works when it's needed, with telemetry complete enough that measurement does not require manual reconstruction 

What We Offer 

  • A fully remote, results-based environment 
  • Competitive salary, bonus, and equity package 
  • 100% employer paid, comprehensive health insurance including medical, dental, and vision for you and your family 
  • Unlimited PTO, with your manager's approval 
  • Flexible work environment where you manage your work day 
  • 14 weeks of fully-paid parental leave 

Salary Range: $165,000–$200,000. This represents the typical salary range for this position based on experience, skills, and other factors. 

We’re an Equal Opportunity Employer: You’ll receive consideration for employment without regard to race, sex, color, religion, sexual orientation, gender identity, national origin, protected veteran status, or on the basis of disability. 
Applicant Data Disclosure   
By submitting an application, you acknowledge that Defcon AI uses third-party service providers to facilitate its recruitment and hiring processes. These providers include applicant tracking systems, candidate verification platforms, and fraud detection tools (collectively, "Hiring Platforms"). Your application materials, including your résumé, cover letter, work samples, responses to application questions, and any other information you submit, may be transmitted to and processed by these Hiring Platforms for the following purposes:  
  • Managing and administering your application throughout the hiring process; 
  • Verifying the accuracy and authenticity of application materials, including by cross-referencing information you provide against publicly available sources and proprietary databases; 
  • Identifying indicators of potentially fraudulent, fabricated, or materially misleading application content, including but not limited to discrepancies between submitted materials and publicly available professional profiles, geographic anomalies, and fabricated work histories. 
Applications that are flagged through this process as containing indicators of fraud or material misrepresentation may be declined from further consideration. If you have questions about the status of your application or the evaluation process, please contact [email protected].  
 
Defcon AI requires its Hiring Platform providers to process your information solely for the purposes described above and in accordance with applicable law. Your information will be retained only for as long as necessary to fulfill these purposes and any applicable legal obligations, after which it will be deleted in accordance with Defcon AI's data retention policies.
For more information about how your data is used, please refer to our Privacy Policy and Applicant Privacy Notice.  

 

Similar Jobs

9 Minutes Ago
In-Office or Remote
United States
100K-233K Annually
Senior level
100K-233K Annually
Senior level
Automotive
Develop and ship native iOS software for Ford’s customer-facing mobile app, serving millions of drivers worldwide. Responsibilities include contributing to product architecture and design reviews, evaluating technologies, writing documentation, collaborating with product and design teams, and delivering production-grade software using agile and CI/CD practices. The role emphasizes Swift, SwiftUI, mobile architecture, testing, observability, debugging, and operational support.
Top Skills: Ci/CdComposable Architecture (Tca)DevOpsFastlaneiOSMvcMvvmRubySite Reliability EngineeringSwiftSwiftui
13 Minutes Ago
Remote
United States
140K-155K Annually
Senior level
140K-155K Annually
Senior level
Computer Vision • Digital Media • Kids + Family • Mobile • Software • Sports
Lead design and delivery of full-stack basketball scoring, statistics, and postgame features. Build scalable backend services and APIs with TypeScript and Node.js, develop React web experiences, support iOS and Android products, and make architectural decisions. Collaborate across product, design, and engineering teams, ensure code quality and system reliability, mentor peers, and contribute to technical standards and tools.
Top Skills: AndroidiOSNode.jsReactTypescript
13 Minutes Ago
Remote
United States
165K-185K Annually
Senior level
165K-185K Annually
Senior level
Computer Vision • Digital Media • Kids + Family • Mobile • Software • Sports
Leads technical strategy and architecture for a scalable advertising platform supporting video, server-side advertising, connected TV, programmatic advertising, and new ad formats. Designs distributed systems, evaluates ad technologies, improves reliability and observability, partners across engineering and product teams, mentors engineers, and guides long-term technical investments and roadmaps.
Top Skills: AWSConnected TvCsaiDistributed SystemsGoogle Ad ManagerHeader BiddingNode.jsPostgresProgrammatic AdvertisingRedisSsaiTypescriptVast

What you need to know about the Austin Tech Scene

Austin has a diverse and thriving tech ecosystem thanks to home-grown companies like Dell and major campuses for IBM, AMD and Apple. The state’s flagship university, the University of Texas at Austin, is known for its engineering school, and the city is known for its annual South by Southwest tech and media conference. Austin’s tech scene spans many verticals, but it’s particularly known for hardware, including semiconductors, as well as AI, biotechnology and cloud computing. And its food and music scene, low taxes and favorable climate has made the city a destination for tech workers from across the country.

Key Facts About Austin Tech

  • Number of Tech Workers: 180,500; 13.7% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Dell, IBM, AMD, Apple, Alphabet
  • Key Industries: Artificial intelligence, hardware, cloud computing, software, healthtech
  • Funding Landscape: $4.5 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Live Oak Ventures, Austin Ventures, Hinge Capital, Gigafund, KdT Ventures, Next Coast Ventures, Silverton Partners
  • Research Centers and Universities: University of Texas, Southwestern University, Texas State University, Center for Complex Quantum Systems, Oden Institute for Computational Engineering and Sciences, Texas Advanced Computing Center

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account