MNC InsiderMNC Insider

Staff AI Engineer (Hybrid)

Stryker

Staff AI Engineer (Hybrid)

full-timePosted: Aug 26, 2026Updated: Aug 28, 2026California, Menlo Park

Job Description

Work Flexibility: HybridWe're hiring a Staff AI Engineer to build GenAI and voice agents for medical devices, deployed both on-device and in the cloud. You'll own the technical direction for these systems — connecting clinical use cases to the models behind them (ASR, TTS, SLMs, speech-to-speech) while working within tight on-device limits latency, memory, and reliability. This is a hands-on applied ML role, and the core challenge is making stochastic models behave predictably enough for clinical use: bounding them with deterministic architecture, building evaluation sets and frameworks, and designing safety guardrails that hold up in a regulated environment. At the staff level, you'll set the architecture and evaluation standards the rest of the team builds against, take on the hardest technical bets first, and align device software, data, validation, clinical, and regulatory teams around the safety, effectiveness, and quality of AI-enabled features across the product lifecycle — consistent with FDA guidance and good machine learning practice. What You Will DoOwn the technical direction of GenAI and multimodal agent (voice, text, vision) capabilities: translate product needs into robust, testable AI system designs, drive the architecture across components, and carry the highest-risk pieces from prototype through validation-ready implementation. Architect stateful agentic systems (intent handling, tool/function calling, dialog and device-state management, interruption handling and recovery) that behave deterministically where safety requires it, with well-defined interface contracts between AI components and device software. Manage and mitigate stochastic model behavior: design layered guardrails (deterministic validation, plausibility bounds, model-based checks), define what the model is and is not permitted to decide, and make those boundaries testable. Set the standard for evaluation of multimodal GenAI systems: comprehensive test sets and automated harnesses covering task and intent accuracy, robustness under realistic clinical audio conditions, conversational quality, and responsiveness. Develop and evaluate real-time speech and language components (speech recognition, synthesis, and dialog/turn handling), balancing model quality against the latency, memory, and reliability constraints of medical hardware. Evaluate, select, integrate, and fine-tune off-the-shelf and small-footprint models (SLMs, domain-adapted ASR) for domain-specific terminology; own the buy/adapt/build decisions and their justification. Establish safety, bias, and performance metrics tailored to voice and generative systems, and produce documentation supporting QMS and regulatory submissions. Instrument systems for traceability: structured logging, audit trails of agent actions, and reproducible evaluation runs suitable for a regulated development process. Mentor junior engineers, ensure engineering quality through design and code review, and communicate AI constraints and trade-offs clearly to product, clinical, and regulatory stakeholders. Stay abreast of the rapidly evolving GenAI model, speech, and agent-architecture landscape; identify which advances matter for the roadmap and pragmatically incorporate them. What You Need (Minimum Required Qualifications)Bachelor's Degree in Computer Science, Machine Learning, Electrical Engineering, Biomedical Engineering, Mathematics, or related field.4+ years of AI/ML engineering experience, OR Master's Degree in the above fields and 2+ years. Preferred Qualifications (Strongly Desired)Strong proficiency in Python; optionally, working proficiency in C++ or other relevant languages for performance-critical and embedded integration work Hands-on experience designing custom evaluation metrics, evaluation harnesses, and test-set generation (both synthetic and human) for stochastic AI systems. Experience bringing statistical rigor to AI evaluation and validation (acceptance criteria, confidence intervals, subgroup and robustness analysis), ideally in collaboration with validation or quality functions. Experience building production systems around LLM function calling / tool use, including schema design, context management for small-context models, and prompt/configuration versioning. Demonstrated recent experience in at least one of: NLP/LLM systems, speech processing (ASR/TTS), or the development and evaluation of generative AI applications. Experience fine-tuning ASR or small language models for domain-specific vocabulary. Experience with real-time voice interfaces: streaming audio, turn-taking and interruption handling, and robustness to noisy environments. Track record of technical leadership: setting direction for a significant system or product area and delivering it through/with other engineers. Experience setting technical direction across multiple teams or leading the architecture of a multi-component product area. Experience deploying and optimizing models on edge compute devices under latency/memory constraints. Strong problem-solving, detail orientation, and critical-thinking skills; excellent communication and interpersonal skills, with the ability to effectively communicate complex technical concepts to non-technical stakeholders; general knowledge of the healthcare market. $133,400 - $222,300 USD Annual Travel Percentage: 10%Stryker Corporation is an equal opportunity employer. Qualified applicants will receive consideration for employment without regard to race, ethnicity, color, religion, sex, gender identity, sexual orientation, national origin, disability, or protected veteran status. Stryker is an EO employer – M/F/Veteran/Disability.Stryker Corporation will not discharge or in any other manner discriminate against employees or applicants because they have inquired about, discussed, or disclosed their own pay or the pay of another employee or applicant. However, employees who have access to the compensation information of other employees or applicants as a part of their essential job functions cannot disclose the pay of other employees or applicants to individuals who do not otherwise have access to compensation information, unless the disclosure is (a) in response to a formal complaint or charge, (b) in furtherance of an investigation, proceeding, hearing, or action, including an investigation conducted by the employer, or (c) consistent with the contractor’s legal duty to furnish information.

Locations

  • California, Menlo Park

Salary

133,400 - 222,300 USD / yearly

Skills Required

  • Pythonintermediate
  • C++intermediate
  • production systems around LLM function calling / tool useintermediate
  • at least one of: NLP/LLM systemsintermediate
  • real-time voice interfaces: streaming audiointermediate
  • healthcare marketintermediate

Required Qualifications

  • Bachelor's Degree in Computer Science, Machine Learning, Electrical Engineering, Biomedical Engineering, Mathematics, or related field. (degree in computer science)
  • 4+ years of AI/ML engineering experience, OR Master's Degree in the above fields and 2+ years. (experience, 4 years)
  • Bachelor's Degree in Computer Science, Machine Learning, Electrical Engineering, Biomedical Engineering, Mathematics, or related field. (degree in computer science)
  • 4+ years of AI/ML engineering experience, OR Master's Degree in the above fields and 2+ years. (experience, 4 years)

Preferred Qualifications

  • Strong proficiency in Python; optionally, working proficiency in C++ or other relevant languages for performance-critical and embedded integration work (experience)
  • Strong proficiency in Python; optionally, working proficiency in C++ or other relevant languages for performance-critical and embedded integration work (experience)
  • Hands-on experience designing custom evaluation metrics, evaluation harnesses, and test-set generation (both synthetic and human) for stochastic AI systems. (experience)
  • Hands-on experience designing custom evaluation metrics, evaluation harnesses, and test-set generation (both synthetic and human) for stochastic AI systems. (experience)
  • Experience bringing statistical rigor to AI evaluation and validation (acceptance criteria, confidence intervals, subgroup and robustness analysis), ideally in collaboration with validation or quality functions. (experience)
  • Experience bringing statistical rigor to AI evaluation and validation (acceptance criteria, confidence intervals, subgroup and robustness analysis), ideally in collaboration with validation or quality functions. (experience)
  • Experience building production systems around LLM function calling / tool use, including schema design, context management for small-context models, and prompt/configuration versioning. (experience)
  • Experience building production systems around LLM function calling / tool use, including schema design, context management for small-context models, and prompt/configuration versioning. (experience)
  • Demonstrated recent experience in at least one of: NLP/LLM systems, speech processing (ASR/TTS), or the development and evaluation of generative AI applications. (experience)
  • Demonstrated recent experience in at least one of: NLP/LLM systems, speech processing (ASR/TTS), or the development and evaluation of generative AI applications. (experience)
  • Experience fine-tuning ASR or small language models for domain-specific vocabulary. (experience)
  • Experience fine-tuning ASR or small language models for domain-specific vocabulary. (experience)
  • Experience with real-time voice interfaces: streaming audio, turn-taking and interruption handling, and robustness to noisy environments. (experience)
  • Experience with real-time voice interfaces: streaming audio, turn-taking and interruption handling, and robustness to noisy environments. (experience)
  • Track record of technical leadership: setting direction for a significant system or product area and delivering it through/with other engineers. (experience)
  • Track record of technical leadership: setting direction for a significant system or product area and delivering it through/with other engineers. (experience)
  • Experience setting technical direction across multiple teams or leading the architecture of a multi-component product area. (experience)
  • Experience setting technical direction across multiple teams or leading the architecture of a multi-component product area. (experience)
  • Experience deploying and optimizing models on edge compute devices under latency/memory constraints. (experience)
  • Experience deploying and optimizing models on edge compute devices under latency/memory constraints. (experience)
  • Strong problem-solving, detail orientation, and critical-thinking skills; excellent communication and interpersonal skills, with the ability to effectively communicate complex technical concepts to non-technical stakeholders; general knowledge of the healthcare market. (experience)
  • Strong problem-solving, detail orientation, and critical-thinking skills; excellent communication and interpersonal skills, with the ability to effectively communicate complex technical concepts to non-technical stakeholders; general knowledge of the healthcare market. (experience)
  • $133,400 - $222,300 USD Annual (experience)

Responsibilities

  • Own the technical direction of GenAI and multimodal agent (voice, text, vision) capabilities: translate product needs into robust, testable AI system designs, drive the architecture across components, and carry the highest-risk pieces from prototype through validation-ready implementation.
  • Own the technical direction of GenAI and multimodal agent (voice, text, vision) capabilities: translate product needs into robust, testable AI system designs, drive the architecture across components, and carry the highest-risk pieces from prototype through validation-ready implementation.
  • Architect stateful agentic systems (intent handling, tool/function calling, dialog and device-state management, interruption handling and recovery) that behave deterministically where safety requires it, with well-defined interface contracts between AI components and device software.
  • Architect stateful agentic systems (intent handling, tool/function calling, dialog and device-state management, interruption handling and recovery) that behave deterministically where safety requires it, with well-defined interface contracts between AI components and device software.
  • Manage and mitigate stochastic model behavior: design layered guardrails (deterministic validation, plausibility bounds, model-based checks), define what the model is and is not permitted to decide, and make those boundaries testable.
  • Manage and mitigate stochastic model behavior: design layered guardrails (deterministic validation, plausibility bounds, model-based checks), define what the model is and is not permitted to decide, and make those boundaries testable.
  • Set the standard for evaluation of multimodal GenAI systems: comprehensive test sets and automated harnesses covering task and intent accuracy, robustness under realistic clinical audio conditions, conversational quality, and responsiveness.
  • Set the standard for evaluation of multimodal GenAI systems: comprehensive test sets and automated harnesses covering task and intent accuracy, robustness under realistic clinical audio conditions, conversational quality, and responsiveness.
  • Develop and evaluate real-time speech and language components (speech recognition, synthesis, and dialog/turn handling), balancing model quality against the latency, memory, and reliability constraints of medical hardware.
  • Develop and evaluate real-time speech and language components (speech recognition, synthesis, and dialog/turn handling), balancing model quality against the latency, memory, and reliability constraints of medical hardware.
  • Evaluate, select, integrate, and fine-tune off-the-shelf and small-footprint models (SLMs, domain-adapted ASR) for domain-specific terminology; own the buy/adapt/build decisions and their justification.
  • Evaluate, select, integrate, and fine-tune off-the-shelf and small-footprint models (SLMs, domain-adapted ASR) for domain-specific terminology; own the buy/adapt/build decisions and their justification.
  • Establish safety, bias, and performance metrics tailored to voice and generative systems, and produce documentation supporting QMS and regulatory submissions.
  • Establish safety, bias, and performance metrics tailored to voice and generative systems, and produce documentation supporting QMS and regulatory submissions.
  • Instrument systems for traceability: structured logging, audit trails of agent actions, and reproducible evaluation runs suitable for a regulated development process.
  • Instrument systems for traceability: structured logging, audit trails of agent actions, and reproducible evaluation runs suitable for a regulated development process.
  • Mentor junior engineers, ensure engineering quality through design and code review, and communicate AI constraints and trade-offs clearly to product, clinical, and regulatory stakeholders.
  • Mentor junior engineers, ensure engineering quality through design and code review, and communicate AI constraints and trade-offs clearly to product, clinical, and regulatory stakeholders.
  • Stay abreast of the rapidly evolving GenAI model, speech, and agent-architecture landscape; identify which advances matter for the roadmap and pragmatically incorporate them.
  • Stay abreast of the rapidly evolving GenAI model, speech, and agent-architecture landscape; identify which advances matter for the roadmap and pragmatically incorporate them.

Target Your Resume for "Staff AI Engineer (Hybrid)" , Stryker

Get personalized recommendations to optimize your resume specifically for Staff AI Engineer (Hybrid). Takes only 15 seconds!

AI-powered keyword optimization
Skills matching & gap analysis
Experience alignment suggestions

Check Your ATS Score for "Staff AI Engineer (Hybrid)" , Stryker

Find out how well your resume matches this job's requirements. Get comprehensive analysis including ATS compatibility, keyword matching, skill gaps, and personalized recommendations.

ATS compatibility check
Keyword optimization analysis
Skill matching & gap identification
Format & readability score

Tags & Categories

GeneralGeneral

Answer 10 quick questions to check your fit for Staff AI Engineer (Hybrid) @ Stryker.

Quiz Challenge
10 Questions
~2 Minutes
Instant Score

Related Books and Jobs

No related jobs found at the moment.