Distinguished Machine Learning Engineer, AISWP (Hybrid)

CiscoSan Jose, CA
$342,700 - $433,400Hybrid

About The Position

Cisco’s AI Software & Platform team (AISWP) is building the infrastructure that defines how the internet is managed. As a Distinguished Machine Learning Engineer, you will have a seat at the table where the future of autonomous, agentic networking is designed. Working directly alongside senior executive leadership, you will lead the architecture for Cisco Cloud Control and Canvas—the cornerstone platforms that are shifting Cisco from a collection of products to a unified, AI-driven Agentic Ops powerhouse. This is a rare opportunity to move the needle on a global scale, solving high-stakes challenges that directly impact our customers' ability to run their most critical infrastructure. You will be the bridge between cutting-edge AI research and massive-scale engineering. You will collaborate with our most senior AI researchers, distributed systems architects, and product visionaries to turn abstract technical possibilities into production reality. You will serve as the primary technical advocate, driving alignment across Cisco’s broad engineering organization to ensure our AI services are secure, scalable, and genuinely transformative. As a Distinguished Machine Learning Engineer, you will own the technical trajectory of our platform. Your influence will be felt in every layer of the stack, from model context strategies to the security guardrails that keep customer production environments safe.

Requirements

  • Bachelor’s degree in computer science or a related field with 17+ years of total software engineering experience; Master's with 14+ years; or PhD with 10+ years.
  • 5+ years of experience operating at a "Principal" level or above, with documented proof of influencing technical strategy across at least 3 distinct product organizations.
  • 2+ years of experience in the design, development, and production-level deployment of LLM-based systems serving at least 10,000+ daily active users or processing 500+ requests per second.
  • Successfully architected and shipped at least 1 production-grade multi-agent system using frameworks such as LangGraph, Temporal, or an equivalent complex state-machine stack.
  • 10+ years of deep experience writing production-grade code in Go, Python, or Rust, specifically within hyperscale environments (e.g., managing architectures supporting 100+ microservices or petabyte-scale data pipelines).
  • Experience designing and implementing at least 3 critical safety or observability features in a production environment (e.g., indirect prompt injection defense, human-in-the-loop gates, or automated rollbacks).

Nice To Haves

  • Experience establishing AI governance frameworks (e.g., NIST AI RMF, EU AI Act) and ensuring enterprise compliance across the model lifecycle.
  • Expertise in advanced post-training methodologies, including reinforcement learning from verifiable rewards (RLHF/DPO) and supervised fine-tuning.
  • Deep knowledge of inference serving optimizations, such as KV/prefix caching, continuous batching, and model routing strategies.
  • Contributions to the external technical community, including published research, patents, or active leadership roles in industry open-source standards bodies.
  • Experience in AIOps or network management, specifically in creating correlated context across disparate network and cloud infrastructure layers.

Responsibilities

  • Define and implement multi-year technical strategies for agentic systems, ensuring unified architectural alignment across Cisco’s broad product portfolio and global engineering organizations.
  • Serve as the primary internal and external technical authority, negotiating critical architectural decisions with senior executives and driving the adoption of industry standards, such as the Model Context Protocol (MCP).
  • Architect complex systems that use hybrid and graph-based retrieval, creating unique, scalable solutions for cross-product telemetry, root cause analysis, and automated remediation.
  • Establish the company-wide standards for AI safety, observability, and compliance, ensuring that all agent-based autonomous actions remain measurable, auditable, and resilient to production risks.
  • Cultivate a high-performance engineering culture by mentoring Principal and Senior-level leaders, fostering innovation, and representing Cisco’s technical excellence at industry forums, standards bodies, and global conferences.

Benefits

  • medical, dental and vision insurance
  • a 401(k) plan with a Cisco matching contribution
  • paid parental leave
  • short and long-term disability coverage
  • basic life insurance
  • 10 paid holidays per full calendar year, plus 1 floating holiday for non-exempt employees
  • 1 paid day off for employee’s birthday, paid year-end holiday shutdown, and 4 paid days off for personal wellness determined by Cisco
  • 16 days of paid vacation time per full calendar year, accrued at rate of 4.92 hours per pay period for full-time employees (non-exempt)
  • flexible vacation time off program (exempt)
  • 80 hours of sick time off provided on hire date and each January 1st thereafter, and up to 80 hours of unused sick time carried forward from one calendar year to the next
  • Additional paid time away may be requested to deal with critical or emergency issues for family members
  • Optional 10 paid days per full calendar year to volunteer
  • annual bonuses
  • performance-based incentive pay
  • restricted stock units
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service