D
Defcon AI
Defcon AI develops AI-driven software for logistics, transportation, and supply chain operations. Its technology combines software modeling and intelligent agents to help organizations plan and respond to disruptions caused by natural disasters, unexpected events, and contested conditions. Working at the intersection of artificial intelligence, mobility, and logistics, the company collaborates with partners to deliver data-driven tools designed to improve operational resilience, decision-making, and response efficiency.

Model Test and Measurement Engineer – United States Remote

Defcon AI is seeking a Model Test and Measurement Engineer to independently validate AI systems. The role covers ground-truth development, model audits, workflow measurement, and evidence for responsible release decisions.

Description

  • Build labeled ground truth for model evaluation
  • Plan and conduct sampled audits
  • Gate releases by model version while maintaining version inventories, evaluation records, and rollback triggers
  • Review model drift and override activity
  • Develop evidence on human oversight, fairness, and disparate effects
  • Remain independent of model-building teams and deliver evaluation evidence without building models, setting thresholds, or approving releases based on that evidence
  • Measure workflow improvements through review time, throughput, backlog movement, override rates, and rework rates
  • Establish evaluation-stage quality controls for sample selection, review units, and fair comparisons across methods or searched sources
  • Specify measurement events required from other teams and verify that the review workspace records workflow events from initial use
  • Collaborate with data scientists, AI engineers, and technical leadership
  • Support release decisions, improvement initiatives, and customer evaluations through independent assessments

Requirements

  • At least 5 years of experience with model validation as a defined responsibility
  • Hands-on experience constructing ground truth and designing sampling approaches
  • Strong Python and SQL skills
  • Ability to work independently from the teams whose models you evaluate
  • US citizenship required
  • Active US Secret clearance
  • Willingness to meet elevated personnel security requirements for portions of the work, as discussed during screening
  • Background in regulated-industry or government model risk
  • Experience testing fairness and disparate effects and documenting human oversight
  • Familiarity with the NIST AI Risk Management Framework or a comparable practice
  • Active Top Secret clearance

Benefits

  • Competitive salary with bonus and equity
  • Fully remote, results-oriented work environment
  • Employer-paid medical, dental, and vision coverage for you and your family
  • Unlimited paid time off with manager approval
  • Flexible scheduling with autonomy over your workday
  • 14 weeks of fully paid parental leave
  • Occasional travel to DEFCON AI headquarters, customer sites, and partner facilities as needed

Related Jobs

Kreato Global | BPO and Language Solutions

Remote English-Spanish OPI/VRI Interpreter

Kreato Global | BPO and Language Solutions
201 – 500 Employees
HealthcareHospitalityLogistics

Interpret remotely between English- and Spanish-speaking people in medical, financial, social service, and customer care settings. Provide language support for Kreato Global across Latin America.

Open
SPERTON - Where Great People Meet

Sales Executive, Elevators and Car Parking Systems

SPERTON - Where Great People Meet
51 – 200 Employees

Drive elevator and car parking system sales across Mumbai’s Western Region. Build client and dealer relationships, and manage deals from initial enquiry through project execution.

Open