7 open positions available
Perform supply chain risk assessments, manage risk dashboards, and develop risk mitigation strategies. | Bachelor's degree plus 5 years or Master's plus 3 years experience in supply chain risk analysis, with knowledge of defense standards and risk frameworks. | Basic Qualifications : Bachelor's degree in a related field or equivalent experience is required plus a minimum of 5 years of relevant experience; or Master's degree and 3 years of relevant experience. CLEARANCE REQUIREMENTS: Applicants selected may be subject to a U.S. Government security investigation and must meet eligibility requirements for access to classified information. Due to the nature of work performed within our facilities, U.S. citizenship is required. Responsibilities for this Position: Job Description: General Dynamics Mission Systems is seeking a proactive and analytical Supply Chain Risk Analyst that is inquisitive, comprehends the complexities of modern supply chains, and possesses the critical thinking skills necessary to develop innovative solutions. The right candidate will bring Supply Chain Risk Management (SCRM) analytical and tactical expertise to the team. You will apply those skills while working closely with other Supply Chain professionals, business units, and cross-functional teams to ensure risk awareness and management are at the forefront of decision-making. You will share risk intelligence, analyze new vulnerabilities, and provide risk strategy recommendations. Knowledge and Experience: • Possesses knowledge of SCRM, software trends, and technologies • Understanding of Supply Chain operations policies, including Supply Chain Risk and business continuity • Understanding in conducting research and using scientific methods to achieve outcomes as demonstrated by research experience, such as previous employment as a research assistant • Ability to work and succeed in a fast-paced, deadline driven environment while demonstrating strong presentation, communication, team building and leadership skills • Knowledge of DoDI01, NDAA 889 (and related DFARS clauses) • Knowledge of National Institute of Standards and Technology (NIST) 800-53, 800-71 • Knowledge of the Defense Industrial Base and the federal procurement process • Experience with risk management frameworks Responsibilities: • Performs risk assessments to support supplier onboarding • Manage supplier risk dashboard • Assist with drafting and publishing monthly risk communication • Assist with enterprise-wide supply chain risk management in collaboration with cross-functional teams to identify and mitigate supply chain risk while complying with industry standards • Support Supply Chain Risk Manager with maintenance and continuing development of tools, metrics, and processes to guide the SCRM team’s work and drive continuous improvements • Conduct research, monitor events, understand market conditions, and report findings related to potential supply chain disruptions along with mitigation recommendations • Develop SCRM strategies and action plans with partners to build resiliency • Assist with training stakeholders on Supply Chain Risk Management, including tools and the identification and mitigation of risks Qualifications: • Experience as a Supply Chain Risk Analyst in ecosystems within global markets • Experience influencing decision making and policy changes across an organization • Ability to pivot quickly and thrive in rapidly changing, high stress environment • Strong candidate is decisive and action oriented • Proven fluency in Microsoft Applications such as Word, PowerPoint, and Excel #LI-Remote Salary Note: This estimate represents the typical salary range for this position based on experience and other factors (geographic location, etc.). Actual pay may vary. This job posting will remain open until the position is filled. Combined Salary Range: USD $95,858.00 - USD $106,342.00 /Yr. Company Overview: General Dynamics Mission Systems (GDMS) engineers a diverse portfolio of high technology solutions, products and services that enable customers to successfully execute missions across all domains of operation. With a global team of 12,000+ top professionals, we partner with the best in industry to expand the bounds of innovation in the defense and scientific arenas. Given the nature of our work and who we are, we value trust, honesty, alignment and transparency. We offer highly competitive benefits and pride ourselves in being a great place to work with a shared sense of purpose. You will also enjoy a flexible work environment where contributions are recognized and rewarded. If who we are and what we do resonates with you, we invite you to join our high-performance team! Equal Opportunity Employer / Individuals with Disabilities / Protected Veterans
Lead systems engineering activities for defense hardware development and program leadership. | Requires advanced systems engineering experience, MBSE proficiency, hardware development skills, and ability to obtain DoD clearance. | Basic Qualifications : Bachelor's degree in Systems Engineering, or a related Science, Engineering or Mathematics field, plus a minimum of 5 years of relevant experience; or Master's degree, plus a minimum of 3 years of relevant experience. CLEARANCE REQUIREMENTS: Must be able to obtain a Department of Defense Secret security clearance within a reasonable amount of time. Applicants selected will be subject to a U.S. Government security investigation and must meet eligibility requirements for access to classified information. Due to the nature of work performed within our facilities, U.S. citizenship is required. Responsibilities for this Position: ROLE AND POSITION OBJECTIVES General Dynamics Mission Systems is seeking an Advanced Systems Engineer 3 to support the D5LE2 program within the SWC Fire Control engineering organization. The successful candidate will serve as a senior technical contributor, applying deep systems engineering expertise across the full system life cycle and leveraging MBSE, digital engineering, and program leadership principles to drive mission-critical outcomes in support of the U.S. Navy's strategic deterrent. This role requires strong independent judgment, the ability to guide major program milestones, and demonstrated leadership in mentoring and developing junior engineers within a complex, long-horizon program environment. This role will be focused on the Hardware Development (Mechanical and Electrical) for the Fire Control System D5LE2 program. Candidate must be able to travel to the Pittsfield, MA office once every 3 weeks. What Sets You Apart • Proven track record leading complex systems engineering efforts across the full development life cycle • Highly proficient with MBSE tools and methodologies, with experience driving model-based deliverables on real programs • Technical leader who mentors junior engineers and builds team capability • Strong communicator and negotiator who influences customer requirements and drives win-win outcomes • Understanding of program management principles (Earned Value, CAM, SPI/CPI) and their intersection with engineering execution • Commitment to shaping engineering processes and tools for continuous improvement • Proficient in hardware development for mission critical computing or communication infrastructure in demanding environments. Key Responsibilities • Lead systems engineering activities including requirements analysis, definition, and management, functional analysis, performance analysis, system design, trade studies, integration and test (verification), validation, and interface definition for D5LE2 subsystems and system elements • Drive MBSE and digital engineering strategy, developing and maintaining system architectures, models, and designs using Cameo/MagicDraw and DOORS in support of upcoming D5LE2 program milestones • Develop hardware and support systems for submarine and shore installations. • Apply expert-level critical thinking to identify, analyze, and resolve system design weaknesses; serve as a technical decision-maker for complex engineering trade-offs • Perform and oversee technical planning, cost and risk analyses, and supportability and effectiveness analyses for subsystems and system elements • Guide the successful completion of major programs and projects; contribute to program milestone reviews and gate decisions • Develop and allocate system requirements to lower-level elements; lead customer requirements analysis and negotiation in alignment with D5LE2 objectives • Conduct detailed technical analyses; evaluate vendor products, COTS, GFE/CFE, specifications to determine feasibility and inform design decisions • Design complete and complex frameworks, systems, or products that set architectural direction • Evaluate systems to ensure designs meet governmental security specifications; lead system accreditation/certification efforts and develop security documentation • Support and lead Modeling and Simulation activities aligned with program objectives • Generate and review technical engineering products; ensure quality and compliance with standards, processes, and tools • Provide leadership, direction, and mentoring to lower-level employees; support talent development within the D5LE2 team • Maintain awareness of technology trends relevant to the D5LE2 mission domain and strategic deterrence • Identify opportunities to apply AI for continuous improvement and innovation Knowledge, Skills & Abilities • Highly proficient understanding and application of systems engineering principles, concepts, and theories in a complex, safety-critical program environment • Highly proficient knowledge of the system development life cycle from concept through disposal, with demonstrated experience leading all phases on a major defense or strategic weapons program • Highly proficient with system modeling tools (Cameo/MagicDraw, DOORS) and recognized architectural patterns (SysML, DoDAF); proven experience driving MBSE methodologies on program deliverables • Highly proficient in hardware development, such as electrical and networking systems, equipment racks, servers, cabling, and user interfaces. • Strong command of digital engineering concepts and their application to program workflows, model management, configuration control, and authoritative sources of truth • Expert-level critical thinking: demonstrated ability to analyze complex, multi-dimensional problems, evaluate alternatives, and independently determine and execute sound solutions • Highly proficient in requirements management tools and Microsoft Office applications • Highly proficient written and verbal communication skills; ability to present to senior leadership, customers, and cross-functional audiences • Understanding and application of project leadership principles including SPI/CPI, Earned Value, Cost Account Management (CAM), and Statistical Process Controls • Ability to build productive internal and external working relationships; effective negotiation skills with customers and stakeholders • Ability to sell concepts and ideas, influence direction, and drive consensus on technical approach • Shows initiative, exercises expert independent judgment, and professionally leads projects to completion • Awareness of safety and surety principles and their critical role in strategic weapons system engineering Workplace Options: Candidate must be able to travel to the Pittsfield, MA office once every 3 weeks. While on-site, you will be a part of the Pittsfield, MA campus Salary Note: This estimate represents the typical salary range for this position based on experience and other factors (geographic location, etc.). Actual pay may vary. This job posting will remain open until the position is filled. Combined Salary Range: USD $124,397.00 - USD $138,003.00 /Yr. Company Overview: General Dynamics Mission Systems (GDMS) engineers a diverse portfolio of high technology solutions, products and services that enable customers to successfully execute missions across all domains of operation. With a global team of 12,000+ top professionals, we partner with the best in industry to expand the bounds of innovation in the defense and scientific arenas. Given the nature of our work and who we are, we value trust, honesty, alignment and transparency. We offer highly competitive benefits and pride ourselves in being a great place to work with a shared sense of purpose. You will also enjoy a flexible work environment where contributions are recognized and rewarded. If who we are and what we do resonates with you, we invite you to join our high-performance team! Equal Opportunity Employer / Individuals with Disabilities / Protected Veterans
Lead systems engineering activities including requirements analysis, design, integration, and testing for defense systems. | Bachelor's degree plus 5 years experience or Master's plus 3 years, ability to obtain DoD Secret clearance, and U.S. citizenship required. | Basic Qualifications : Bachelor's degree in Systems Engineering, or a related Science, Engineering or Mathematics field, plus a minimum of 5 years of relevant experience; or Master's degree, plus a minimum of 3 years of relevant experience. CLEARANCE REQUIREMENTS: Must be able to obtain a Department of Defense Secret security clearance within a reasonable amount of time. Applicants selected will be subject to a U.S. Government security investigation and must meet eligibility requirements for access to classified information. Due to the nature of work performed within our facilities, U.S. citizenship is required. Responsibilities for this Position: ROLE AND POSITION OBJECTIVES General Dynamics Mission Systems is seeking an Advanced Systems Engineer 3 to support the D5LE2 program within the SWC Fire Control engineering organization. The successful candidate will serve as a senior technical contributor, applying deep systems engineering expertise across the full system life cycle and leveraging MBSE, digital engineering, and program leadership principles to drive mission-critical outcomes in support of the U.S. Navy's strategic deterrent. This role requires strong independent judgment, the ability to guide major program milestones, and demonstrated leadership in mentoring and developing junior engineers within a complex, long-horizon program environment. This role will be focused on the Hardware Development (Mechanical and Electrical) for the Fire Control System D5LE2 program. Candidate must be able to travel to the Pittsfield, MA office once every 3 weeks. What Sets You Apart • Proven track record leading complex systems engineering efforts across the full development life cycle • Highly proficient with MBSE tools and methodologies, with experience driving model-based deliverables on real programs • Technical leader who mentors junior engineers and builds team capability • Strong communicator and negotiator who influences customer requirements and drives win-win outcomes • Understanding of program management principles (Earned Value, CAM, SPI/CPI) and their intersection with engineering execution • Commitment to shaping engineering processes and tools for continuous improvement • Proficient in hardware development for mission critical computing or communication infrastructure in demanding environments. Key Responsibilities • Lead systems engineering activities including requirements analysis, definition, and management, functional analysis, performance analysis, system design, trade studies, integration and test (verification), validation, and interface definition for D5LE2 subsystems and system elements • Drive MBSE and digital engineering strategy, developing and maintaining system architectures, models, and designs using Cameo/MagicDraw and DOORS in support of upcoming D5LE2 program milestones • Develop hardware and support systems for submarine and shore installations. • Apply expert-level critical thinking to identify, analyze, and resolve system design weaknesses; serve as a technical decision-maker for complex engineering trade-offs • Perform and oversee technical planning, cost and risk analyses, and supportability and effectiveness analyses for subsystems and system elements • Guide the successful completion of major programs and projects; contribute to program milestone reviews and gate decisions • Develop and allocate system requirements to lower-level elements; lead customer requirements analysis and negotiation in alignment with D5LE2 objectives • Conduct detailed technical analyses; evaluate vendor products, COTS, GFE/CFE, specifications to determine feasibility and inform design decisions • Design complete and complex frameworks, systems, or products that set architectural direction • Evaluate systems to ensure designs meet governmental security specifications; lead system accreditation/certification efforts and develop security documentation • Support and lead Modeling and Simulation activities aligned with program objectives • Generate and review technical engineering products; ensure quality and compliance with standards, processes, and tools • Provide leadership, direction, and mentoring to lower-level employees; support talent development within the D5LE2 team • Maintain awareness of technology trends relevant to the D5LE2 mission domain and strategic deterrence • Identify opportunities to apply AI for continuous improvement and innovation Knowledge, Skills & Abilities • Highly proficient understanding and application of systems engineering principles, concepts, and theories in a complex, safety-critical program environment • Highly proficient knowledge of the system development life cycle from concept through disposal, with demonstrated experience leading all phases on a major defense or strategic weapons program • Highly proficient with system modeling tools (Cameo/MagicDraw, DOORS) and recognized architectural patterns (SysML, DoDAF); proven experience driving MBSE methodologies on program deliverables • Highly proficient in hardware development, such as electrical and networking systems, equipment racks, servers, cabling, and user interfaces. • Strong command of digital engineering concepts and their application to program workflows, model management, configuration control, and authoritative sources of truth • Expert-level critical thinking: demonstrated ability to analyze complex, multi-dimensional problems, evaluate alternatives, and independently determine and execute sound solutions • Highly proficient in requirements management tools and Microsoft Office applications • Highly proficient written and verbal communication skills; ability to present to senior leadership, customers, and cross-functional audiences • Understanding and application of project leadership principles including SPI/CPI, Earned Value, Cost Account Management (CAM), and Statistical Process Controls • Ability to build productive internal and external working relationships; effective negotiation skills with customers and stakeholders • Ability to sell concepts and ideas, influence direction, and drive consensus on technical approach • Shows initiative, exercises expert independent judgment, and professionally leads projects to completion • Awareness of safety and surety principles and their critical role in strategic weapons system engineering Workplace Options: Candidate must be able to travel to the Pittsfield, MA office once every 3 weeks. While on-site, you will be a part of the Pittsfield, MA campus Salary Note: This estimate represents the typical salary range for this position based on experience and other factors (geographic location, etc.). Actual pay may vary. This job posting will remain open until the position is filled. Combined Salary Range: USD $124,397.00 - USD $138,003.00 /Yr. Company Overview: General Dynamics Mission Systems (GDMS) engineers a diverse portfolio of high technology solutions, products and services that enable customers to successfully execute missions across all domains of operation. With a global team of 12,000+ top professionals, we partner with the best in industry to expand the bounds of innovation in the defense and scientific arenas. Given the nature of our work and who we are, we value trust, honesty, alignment and transparency. We offer highly competitive benefits and pride ourselves in being a great place to work with a shared sense of purpose. You will also enjoy a flexible work environment where contributions are recognized and rewarded. If who we are and what we do resonates with you, we invite you to join our high-performance team! Equal Opportunity Employer / Individuals with Disabilities / Protected Veterans
Lead reliability standards, monitoring, incident response, and production readiness for AI services. | Bachelor's degree plus 8 years experience, production SRE/DevOps experience, monitoring tools, scripting, container orchestration, and security clearance required. | Basic Qualifications : Bachelor's degree in Software Engineering, or related Science, Technology, Engineering or Mathematics field, plus a minimum of 8 years of relevant experience; or Master's degree, plus 6 years relevant experience. CLEARANCE REQUIREMENTS: Ability to obtain a Department of Defense Secret security clearance is required at time of hire. Applicants selected will be subject to a U.S. Government security investigation and must meet eligibility requirements for access to classified information. Due to the nature of work performed within our facilities, U.S. citizenship is required. Responsibilities for this Position: What You Will Own • Cross-pod reliability standards. Set the reliability bar and ensure it is met consistently across applications. Collaborate with Functional SREs to connect technical reliability metrics to business-side outcomes. You own the engineering signal; together you tell the full reliability story. • SLOs and reliability metrics. Own definitions of service level objectives for every AI service that goes to production. Establish error budgets and use them to drive engineering decisions — not just measure uptime. • Monitoring and observability. Implement and maintain the full observability stack — logging, metrics, tracing, and dashboards. You will know when something is degrading before users do. • Design and manage alerting infrastructure that tells you what's wrong, not just that something is wrong. Alerts you build catch real problems; they don't cry wolf. • Incident response. Own on-call procedures, escalation paths, and incident management end-to-end. Lead post-incident reviews and maintain the reliability improvement backlog. When something breaks, you coordinate the response and ensure it doesn't break the same way again. • Production Readiness. Define and enforce the criteria that determine whether an AI service is ready for production. You are the gate between "it works in dev" and "it's ready to ship." • Toil elimination. Identify and automate repetitive operational tasks. If a human is doing something a script could do, you fix that. What You Won't Own • Infrastructure provisioning — IT provides the infrastructure; you define what's needed and validate it works • Business process decisions or backlog prioritization • Business-side reliability metrics - you partner with the Functional SRE on those, but they own that domain What Makes This Role Different • AI services have failure modes that traditional applications don't — model drift, token budget exhaustion, prompt injection, upstream data quality degradation. You will build monitoring for problems that most SRE teams have never encountered. • You are applying SRE principles from scratch. There is no existing SRE practice to inherit — you will define it for the platform. • Your production readiness criteria directly determine whether AI services go live. You have real authority to say "not ready." • You operate across projects simultaneously — embedded deeply enough to understand large-scale systems, while maintaining consistent standards across all projects. • Your software engineering background means you can engage directly with development teams at the design level — catching reliability problems before they become operational ones. Required Qualifications • Bachelor’s degree in Computer Science, Software Engineering, or a related field, plus 8 years of experience; or Master’s degree plus 6 years of experience • Production SRE or DevOps experience — you have owned the reliability of systems that real users depended on, not just built CI/CD pipelines • Hands-on experience with monitoring and observability tools — Prometheus, Grafana, Datadog, ELK, CloudWatch, or similar. You have built dashboards and alerts that caught real problems. • Strong scripting and automation skills — Python, Bash, infrastructure-as-code (Terraform, CloudFormation, or similar) • Experience with containerized environments — Docker, Kubernetes, container orchestration at scale • Experience defining and managing SLOs, error budgets, and incident response procedures in production • U.S. citizenship required. Department of Defense Secret security clearance is required at time of hire. Preferred Qualifications • Production SRE or DevOps experience — you have owned the reliability of systems that real users depended on, not just built CI/CD pipelines • Software engineering fundamentals — you can read, write, and meaningfully review production-quality code. You understand how architectural and design decisions made early translate into operational problems later. • Software design experience — you have participated in or led design reviews, defined service interfaces or APIs, and pushed back on design decisions using reliability and operability as criteria • Hands-on experience with monitoring and observability tools — Prometheus, Grafana, Datadog, ELK, CloudWatch, or similar. You have built dashboards and alerts that have caught real problems. • Strong scripting and automation skills — Python, Bash, infrastructure-as-code (Terraform, CloudFormation, or similar) • Experience with containerized environments — Docker, Kubernetes, container orchestration at scale • Experience defining and managing SLOs, error budgets, and incident response procedures in production What Sets You Apart • You build things that work. Your default response to a problem is code, not a document. • You have shipped AI systems that real users depended on in production. • You are comfortable working without detailed specs — you can take a problem statement and figure out the right approach. • You care about reliability as much as capability — you monitor what you deploy. • You move fast without being reckless. You know when to iterate and when to get it right the first time. Details • Remote — 100% telework • 9/80 schedule • Defense industry experience is not required Salary Note: This estimate represents the typical salary range for this position based on experience and other factors (geographic location, etc.). Actual pay may vary. This job posting will remain open until the position is filled. Combined Salary Range: USD $142,696.00 - USD $158,303.00 /Yr. Company Overview: General Dynamics Mission Systems (GDMS) engineers a diverse portfolio of high technology solutions, products and services that enable customers to successfully execute missions across all domains of operation. With a global team of 12,000+ top professionals, we partner with the best in industry to expand the bounds of innovation in the defense and scientific arenas. Given the nature of our work and who we are, we value trust, honesty, alignment and transparency. We offer highly competitive benefits and pride ourselves in being a great place to work with a shared sense of purpose. You will also enjoy a flexible work environment where contributions are recognized and rewarded. If who we are and what we do resonates with you, we invite you to join our high-performance team! Equal Opportunity Employer / Individuals with Disabilities / Protected Veterans
Lead reliability standards, SLOs, monitoring, incident response, and production readiness for AI services. | Requires 8+ years experience in SRE or DevOps with hands-on monitoring, automation, container orchestration, and security clearance eligibility. | Basic Qualifications : Bachelor's degree in Software Engineering, or related Science, Technology, Engineering or Mathematics field, plus a minimum of 8 years of relevant experience; or Master's degree, plus 6 years relevant experience. CLEARANCE REQUIREMENTS: Ability to obtain a Department of Defense Secret security clearance is required at time of hire. Applicants selected will be subject to a U.S. Government security investigation and must meet eligibility requirements for access to classified information. Due to the nature of work performed within our facilities, U.S. citizenship is required. Responsibilities for this Position: What You Will Own • Cross-pod reliability standards. Set the reliability bar and ensure it is met consistently across applications. Collaborate with Functional SREs to connect technical reliability metrics to business-side outcomes. You own the engineering signal; together you tell the full reliability story. • SLOs and reliability metrics. Own definitions of service level objectives for every AI service that goes to production. Establish error budgets and use them to drive engineering decisions — not just measure uptime. • Monitoring and observability. Implement and maintain the full observability stack — logging, metrics, tracing, and dashboards. You will know when something is degrading before users do. • Design and manage alerting infrastructure that tells you what's wrong, not just that something is wrong. Alerts you build catch real problems; they don't cry wolf. • Incident response. Own on-call procedures, escalation paths, and incident management end-to-end. Lead post-incident reviews and maintain the reliability improvement backlog. When something breaks, you coordinate the response and ensure it doesn't break the same way again. • Production Readiness. Define and enforce the criteria that determine whether an AI service is ready for production. You are the gate between "it works in dev" and "it's ready to ship." • Toil elimination. Identify and automate repetitive operational tasks. If a human is doing something a script could do, you fix that. What You Won't Own • Infrastructure provisioning — IT provides the infrastructure; you define what's needed and validate it works • Business process decisions or backlog prioritization • Business-side reliability metrics - you partner with the Functional SRE on those, but they own that domain What Makes This Role Different • AI services have failure modes that traditional applications don't — model drift, token budget exhaustion, prompt injection, upstream data quality degradation. You will build monitoring for problems that most SRE teams have never encountered. • You are applying SRE principles from scratch. There is no existing SRE practice to inherit — you will define it for the platform. • Your production readiness criteria directly determine whether AI services go live. You have real authority to say "not ready." • You operate across projects simultaneously — embedded deeply enough to understand large-scale systems, while maintaining consistent standards across all projects. • Your software engineering background means you can engage directly with development teams at the design level — catching reliability problems before they become operational ones. Required Qualifications • Bachelor’s degree in Computer Science, Software Engineering, or a related field, plus 8 years of experience; or Master’s degree plus 6 years of experience • Production SRE or DevOps experience — you have owned the reliability of systems that real users depended on, not just built CI/CD pipelines • Hands-on experience with monitoring and observability tools — Prometheus, Grafana, Datadog, ELK, CloudWatch, or similar. You have built dashboards and alerts that caught real problems. • Strong scripting and automation skills — Python, Bash, infrastructure-as-code (Terraform, CloudFormation, or similar) • Experience with containerized environments — Docker, Kubernetes, container orchestration at scale • Experience defining and managing SLOs, error budgets, and incident response procedures in production • U.S. citizenship required. Department of Defense Secret security clearance is required at time of hire. Preferred Qualifications • Production SRE or DevOps experience — you have owned the reliability of systems that real users depended on, not just built CI/CD pipelines • Software engineering fundamentals — you can read, write, and meaningfully review production-quality code. You understand how architectural and design decisions made early translate into operational problems later. • Software design experience — you have participated in or led design reviews, defined service interfaces or APIs, and pushed back on design decisions using reliability and operability as criteria • Hands-on experience with monitoring and observability tools — Prometheus, Grafana, Datadog, ELK, CloudWatch, or similar. You have built dashboards and alerts that have caught real problems. • Strong scripting and automation skills — Python, Bash, infrastructure-as-code (Terraform, CloudFormation, or similar) • Experience with containerized environments — Docker, Kubernetes, container orchestration at scale • Experience defining and managing SLOs, error budgets, and incident response procedures in production What Sets You Apart • You build things that work. Your default response to a problem is code, not a document. • You have shipped AI systems that real users depended on in production. • You are comfortable working without detailed specs — you can take a problem statement and figure out the right approach. • You care about reliability as much as capability — you monitor what you deploy. • You move fast without being reckless. You know when to iterate and when to get it right the first time. Details • Remote — 100% telework • 9/80 schedule • Defense industry experience is not required Salary Note: This estimate represents the typical salary range for this position based on experience and other factors (geographic location, etc.). Actual pay may vary. This job posting will remain open until the position is filled. Combined Salary Range: USD $142,696.00 - USD $158,303.00 /Yr. Company Overview: General Dynamics Mission Systems (GDMS) engineers a diverse portfolio of high technology solutions, products and services that enable customers to successfully execute missions across all domains of operation. With a global team of 12,000+ top professionals, we partner with the best in industry to expand the bounds of innovation in the defense and scientific arenas. Given the nature of our work and who we are, we value trust, honesty, alignment and transparency. We offer highly competitive benefits and pride ourselves in being a great place to work with a shared sense of purpose. You will also enjoy a flexible work environment where contributions are recognized and rewarded. If who we are and what we do resonates with you, we invite you to join our high-performance team! Equal Opportunity Employer / Individuals with Disabilities / Protected Veterans
Lead site reliability engineering efforts including SLOs, monitoring, incident response, and production readiness for AI services. | Requires 8+ years experience, production SRE/DevOps expertise, strong scripting and automation skills, container orchestration experience, and U.S. citizenship with security clearance eligibility. | Basic Qualifications : Bachelor's degree in Software Engineering, or related Science, Technology, Engineering or Mathematics field, plus a minimum of 8 years of relevant experience; or Master's degree, plus 6 years relevant experience. CLEARANCE REQUIREMENTS: Ability to obtain a Department of Defense Secret security clearance is required at time of hire. Applicants selected will be subject to a U.S. Government security investigation and must meet eligibility requirements for access to classified information. Due to the nature of work performed within our facilities, U.S. citizenship is required. Responsibilities for this Position: What You Will Own • Cross-pod reliability standards. Set the reliability bar and ensure it is met consistently across applications. Collaborate with Functional SREs to connect technical reliability metrics to business-side outcomes. You own the engineering signal; together you tell the full reliability story. • SLOs and reliability metrics. Own definitions of service level objectives for every AI service that goes to production. Establish error budgets and use them to drive engineering decisions — not just measure uptime. • Monitoring and observability. Implement and maintain the full observability stack — logging, metrics, tracing, and dashboards. You will know when something is degrading before users do. • Design and manage alerting infrastructure that tells you what's wrong, not just that something is wrong. Alerts you build catch real problems; they don't cry wolf. • Incident response. Own on-call procedures, escalation paths, and incident management end-to-end. Lead post-incident reviews and maintain the reliability improvement backlog. When something breaks, you coordinate the response and ensure it doesn't break the same way again. • Production Readiness. Define and enforce the criteria that determine whether an AI service is ready for production. You are the gate between "it works in dev" and "it's ready to ship." • Toil elimination. Identify and automate repetitive operational tasks. If a human is doing something a script could do, you fix that. What You Won't Own • Infrastructure provisioning — IT provides the infrastructure; you define what's needed and validate it works • Business process decisions or backlog prioritization • Business-side reliability metrics - you partner with the Functional SRE on those, but they own that domain What Makes This Role Different • AI services have failure modes that traditional applications don't — model drift, token budget exhaustion, prompt injection, upstream data quality degradation. You will build monitoring for problems that most SRE teams have never encountered. • You are applying SRE principles from scratch. There is no existing SRE practice to inherit — you will define it for the platform. • Your production readiness criteria directly determine whether AI services go live. You have real authority to say "not ready." • You operate across projects simultaneously — embedded deeply enough to understand large-scale systems, while maintaining consistent standards across all projects. • Your software engineering background means you can engage directly with development teams at the design level — catching reliability problems before they become operational ones. Required Qualifications • Bachelor’s degree in Computer Science, Software Engineering, or a related field, plus 8 years of experience; or Master’s degree plus 6 years of experience • Production SRE or DevOps experience — you have owned the reliability of systems that real users depended on, not just built CI/CD pipelines • Hands-on experience with monitoring and observability tools — Prometheus, Grafana, Datadog, ELK, CloudWatch, or similar. You have built dashboards and alerts that caught real problems. • Strong scripting and automation skills — Python, Bash, infrastructure-as-code (Terraform, CloudFormation, or similar) • Experience with containerized environments — Docker, Kubernetes, container orchestration at scale • Experience defining and managing SLOs, error budgets, and incident response procedures in production • U.S. citizenship required. Department of Defense Secret security clearance is required at time of hire. Preferred Qualifications • Production SRE or DevOps experience — you have owned the reliability of systems that real users depended on, not just built CI/CD pipelines • Software engineering fundamentals — you can read, write, and meaningfully review production-quality code. You understand how architectural and design decisions made early translate into operational problems later. • Software design experience — you have participated in or led design reviews, defined service interfaces or APIs, and pushed back on design decisions using reliability and operability as criteria • Hands-on experience with monitoring and observability tools — Prometheus, Grafana, Datadog, ELK, CloudWatch, or similar. You have built dashboards and alerts that have caught real problems. • Strong scripting and automation skills — Python, Bash, infrastructure-as-code (Terraform, CloudFormation, or similar) • Experience with containerized environments — Docker, Kubernetes, container orchestration at scale • Experience defining and managing SLOs, error budgets, and incident response procedures in production What Sets You Apart • You build things that work. Your default response to a problem is code, not a document. • You have shipped AI systems that real users depended on in production. • You are comfortable working without detailed specs — you can take a problem statement and figure out the right approach. • You care about reliability as much as capability — you monitor what you deploy. • You move fast without being reckless. You know when to iterate and when to get it right the first time. Details • Remote — 100% telework • 9/80 schedule • Defense industry experience is not required Salary Note: This estimate represents the typical salary range for this position based on experience and other factors (geographic location, etc.). Actual pay may vary. This job posting will remain open until the position is filled. Combined Salary Range: USD $142,696.00 - USD $158,303.00 /Yr. Company Overview: General Dynamics Mission Systems (GDMS) engineers a diverse portfolio of high technology solutions, products and services that enable customers to successfully execute missions across all domains of operation. With a global team of 12,000+ top professionals, we partner with the best in industry to expand the bounds of innovation in the defense and scientific arenas. Given the nature of our work and who we are, we value trust, honesty, alignment and transparency. We offer highly competitive benefits and pride ourselves in being a great place to work with a shared sense of purpose. You will also enjoy a flexible work environment where contributions are recognized and rewarded. If who we are and what we do resonates with you, we invite you to join our high-performance team! Equal Opportunity Employer / Individuals with Disabilities / Protected Veterans
Lead reliability engineering for AI services including SLOs, monitoring, incident response, and production readiness. | Bachelor's degree plus 8 years experience, production SRE/DevOps experience, monitoring tools expertise, scripting and automation skills, container orchestration experience, and ability to obtain DoD Secret clearance. | Basic Qualifications : Bachelor's degree in Software Engineering, or related Science, Technology, Engineering or Mathematics field, plus a minimum of 8 years of relevant experience; or Master's degree, plus 6 years relevant experience. CLEARANCE REQUIREMENTS: Ability to obtain a Department of Defense Secret security clearance is required at time of hire. Applicants selected will be subject to a U.S. Government security investigation and must meet eligibility requirements for access to classified information. Due to the nature of work performed within our facilities, U.S. citizenship is required. Responsibilities for this Position: What You Will Own • Cross-pod reliability standards. Set the reliability bar and ensure it is met consistently across applications. Collaborate with Functional SREs to connect technical reliability metrics to business-side outcomes. You own the engineering signal; together you tell the full reliability story. • SLOs and reliability metrics. Own definitions of service level objectives for every AI service that goes to production. Establish error budgets and use them to drive engineering decisions — not just measure uptime. • Monitoring and observability. Implement and maintain the full observability stack — logging, metrics, tracing, and dashboards. You will know when something is degrading before users do. • Design and manage alerting infrastructure that tells you what's wrong, not just that something is wrong. Alerts you build catch real problems; they don't cry wolf. • Incident response. Own on-call procedures, escalation paths, and incident management end-to-end. Lead post-incident reviews and maintain the reliability improvement backlog. When something breaks, you coordinate the response and ensure it doesn't break the same way again. • Production Readiness. Define and enforce the criteria that determine whether an AI service is ready for production. You are the gate between "it works in dev" and "it's ready to ship." • Toil elimination. Identify and automate repetitive operational tasks. If a human is doing something a script could do, you fix that. What You Won't Own • Infrastructure provisioning — IT provides the infrastructure; you define what's needed and validate it works • Business process decisions or backlog prioritization • Business-side reliability metrics - you partner with the Functional SRE on those, but they own that domain What Makes This Role Different • AI services have failure modes that traditional applications don't — model drift, token budget exhaustion, prompt injection, upstream data quality degradation. You will build monitoring for problems that most SRE teams have never encountered. • You are applying SRE principles from scratch. There is no existing SRE practice to inherit — you will define it for the platform. • Your production readiness criteria directly determine whether AI services go live. You have real authority to say "not ready." • You operate across projects simultaneously — embedded deeply enough to understand large-scale systems, while maintaining consistent standards across all projects. • Your software engineering background means you can engage directly with development teams at the design level — catching reliability problems before they become operational ones. Required Qualifications • Bachelor’s degree in Computer Science, Software Engineering, or a related field, plus 8 years of experience; or Master’s degree plus 6 years of experience • Production SRE or DevOps experience — you have owned the reliability of systems that real users depended on, not just built CI/CD pipelines • Hands-on experience with monitoring and observability tools — Prometheus, Grafana, Datadog, ELK, CloudWatch, or similar. You have built dashboards and alerts that caught real problems. • Strong scripting and automation skills — Python, Bash, infrastructure-as-code (Terraform, CloudFormation, or similar) • Experience with containerized environments — Docker, Kubernetes, container orchestration at scale • Experience defining and managing SLOs, error budgets, and incident response procedures in production • U.S. citizenship required. Department of Defense Secret security clearance is required at time of hire. Preferred Qualifications • Production SRE or DevOps experience — you have owned the reliability of systems that real users depended on, not just built CI/CD pipelines • Software engineering fundamentals — you can read, write, and meaningfully review production-quality code. You understand how architectural and design decisions made early translate into operational problems later. • Software design experience — you have participated in or led design reviews, defined service interfaces or APIs, and pushed back on design decisions using reliability and operability as criteria • Hands-on experience with monitoring and observability tools — Prometheus, Grafana, Datadog, ELK, CloudWatch, or similar. You have built dashboards and alerts that have caught real problems. • Strong scripting and automation skills — Python, Bash, infrastructure-as-code (Terraform, CloudFormation, or similar) • Experience with containerized environments — Docker, Kubernetes, container orchestration at scale • Experience defining and managing SLOs, error budgets, and incident response procedures in production What Sets You Apart • You build things that work. Your default response to a problem is code, not a document. • You have shipped AI systems that real users depended on in production. • You are comfortable working without detailed specs — you can take a problem statement and figure out the right approach. • You care about reliability as much as capability — you monitor what you deploy. • You move fast without being reckless. You know when to iterate and when to get it right the first time. Details • Remote — 100% telework • 9/80 schedule • Defense industry experience is not required Salary Note: This estimate represents the typical salary range for this position based on experience and other factors (geographic location, etc.). Actual pay may vary. This job posting will remain open until the position is filled. Combined Salary Range: USD $142,696.00 - USD $158,303.00 /Yr. Company Overview: General Dynamics Mission Systems (GDMS) engineers a diverse portfolio of high technology solutions, products and services that enable customers to successfully execute missions across all domains of operation. With a global team of 12,000+ top professionals, we partner with the best in industry to expand the bounds of innovation in the defense and scientific arenas. Given the nature of our work and who we are, we value trust, honesty, alignment and transparency. We offer highly competitive benefits and pride ourselves in being a great place to work with a shared sense of purpose. You will also enjoy a flexible work environment where contributions are recognized and rewarded. If who we are and what we do resonates with you, we invite you to join our high-performance team! Equal Opportunity Employer / Individuals with Disabilities / Protected Veterans
Create tailored applications specifically for General Dynamics Mission Systems, Inc with our AI-powered resume builder
Get Started for Free