Gen-AI Engineer

  • Location:
    RTP, North Carolina, US
  • Alternate Location
    San Jose, CA
  • Area of Interest
    Engineer - Software
  • Compensation Range
    159800 USD - 203200 USD
  • Job Type
    Professional
  • Technology Interest
    AI or Artificial Intelligence
  • Job Id
    1430181

Who We Are 


The Cisco IT team is changing the way we run Cisco's operations by leveraging the power of technology, the best of business processes, and utilizing outstanding data insights. We are redefining how Cisco designs and delivers the employee, partner, and customer experience based on a culture that values customer service. We strive for speed and agility in all that we do. Above all else, we are kind to each other. We aspire to be an industry-leading IT team, with a strong focus on AI and security to differentiate us and foster innovation. To achieve simplicity and the best employee and customer experience, we need excellent talent and the right abilities to succeed.


Who You'll Work With 


You will work with our amazing Data Infrastructure & Platforms team as part of the Infrastructure & Cloud Services. You will build solutions and support them across our portfolio of capabilities, primarily focusing on enabling AI/ML and Generative AI capabilities in a multi-functional team setup. You will collaborate with other engineers, architects, and organization leadership.


Who You Are


We are looking for a highly skilled Senior GenAI Engineer to lead the deployment and management of on-premise Large Language Models (LLMs), with a focus on Retrieval Augmented Generation (RAG). This role requires expertise in developing and supporting large-scale GenAI and ML platforms, with a strong emphasis on document management, security, vector databases, and workflow orchestration. The successful candidate will have extensive experience in responsible AI practices, including toxicity screening and ensuring regulatory compliance across AI solutions. 


What You’ll Do

  • Deploy and manage Large Language Models (LLMs) for on-prem environments, focusing on Retrieval Augmented Generation (RAG) and ensuring high-performance infrastructure.
  • Build and optimize secure, scalable AI/ML platforms that support enterprise-level document management and data security protocols.
  • Design, implement, and manage workflows to ensure seamless document processing, ingestion, classification, and retrieval within AI models.
  • Implement and manage vector databases to support efficient document search and retrieval within AI workflows.
  • Ensure compliance with organizational and regulatory data security standards, including encryption, access control, and auditing of sensitive documents used within AI models.
  • Implement and maintain responsible AI practices, including toxicity screening, bias detection, and regulatory compliance to ensure ethical and safe AI usage.
  • Collaborate with cross-functional teams to ensure data privacy and information security requirements are met when processing documents through AI models.
  • Continuously evaluate and improve infrastructure to support evolving AI/ML needs, particularly focusing on document ingestion, classification, and retrieval.
  • Develop and maintain automated pipelines for LLM deployment and secure document processing with real-time monitoring and alerts.
  • Work closely with compliance, legal, and governance teams to ensure AI models are aligned with security and regulatory frameworks.
  • Stay updated on the latest advancements in AI/ML, document security, and responsible AI practices.

Minimum Qualifications 

  • Bachelor's in Computer Science, Computer Engineering, Electrical Engineering, or a related STEM field.
  • 7+ years of experience in engineering, with at least 2 years specifically in AI/ML engineering
  • Proficiency in programming languages such as Python, Java, C++, or similar.
  • Hands-on experience with ML frameworks such as Kubeflow, AI operators, and/or Langchain, with proficiency in Python for AI operations.
  • Experience with MLOps principles, including model deployment, versioning, and/or monitoring in secure environments.


Preferred Qualifications

  • Master’s Degree
  • Focus on on-prem LLM deployments with an emphasis on RAG
  • Expertise in building large-scale, secure AI/ML platforms with an emphasis on document management and security protocols.
  • Experience with vector databases, such as Pinecone, Weaviate, Milvus, etc.
  • Experience with multi-instance GPUs and containerized AI/ML workflows.
  • Proven ability to collaborate with cross-functional teams, including legal, compliance, and security experts.
  • Understanding of document lifecycle management, particularly in the context of AI model ingestion, classification, and retrieval.

Why Cisco
#WeAre Cisco, where each person is unique, but we bring our talents to work as a team and make a difference powering an inclusive future for all.
We embrace digital and help our customers implement change in their digital businesses. Some may think we're "old" (39 years strong) and only about hardware, but we're also a software company. And a security company. We even invented an intuitive network that adapts, predicts, learns, and protects. No other company can do what we do - you can't put us in a box!

But "Digital Transformation" is an empty buzz phrase without a culture that allows for innovation, creativity, and yes, even failure (if you learn from it).

Day to day, we focus on the give and take. We give our best, give our egos a break, and give of ourselves (because giving back is built into our DNA). We take accountability, bold steps, and take difference to heart. Because without diversity of thought and a dedication to equality for all, there is no moving forward.

So, you have colorful hair? Don't care. Tattoos? Show off your ink. Like polka dots? That's cool. Pop culture geek? Many of us are. Passion for technology and world changing? Be you, with us!

Message to applicants applying to work in the U.S. and/or Canada:

When available, the salary range posted for this position reflects the projected hiring range for new hire, full-time salaries in U.S. and/or Canada locations, not including equity or benefits. For non-sales roles the hiring ranges reflect base salary only; employees are also eligible to receive annual bonuses. Hiring ranges for sales positions include base and incentive compensation target. Individual pay is determined by the candidate's hiring location and additional factors, including but not limited to skillset, experience, and relevant education, certifications, or training. Applicants may not be eligible for the full salary range based on their U.S. or Canada hiring location. The recruiter can share more details about compensation for the role in your location during the hiring process.

U.S. employees have access to quality medical, dental and vision insurance, a 401(k) plan with a Cisco matching contribution, short and long-term disability coverage, basic life insurance and numerous wellbeing offerings.

Employees receive up to twelve paid holidays per calendar year, which includes one floating holiday (for non-exempt employees), plus a day off for their birthday. Non-Exempt new hires accrue up to 16 days of vacation time off each year, at a rate of 4.92 hours per pay period. Exempt new hires participate in Cisco’s flexible Vacation Time Off policy, which does not place a defined limit on how much vacation time eligible employees may use, but is subject to availability and some business limitations. All new hires are eligible for Sick Time Off subject to Cisco’s Sick Time Off Policy and will have eighty (80) hours of sick time off provided on their hire date and on January 1st of each year thereafter.  Up to 80 hours of unused sick time will be carried forward from one calendar year to the next such that the maximum number of sick time hours an employee may have available is 160 hours. Employees in Illinois have a unique time off program designed specifically with local requirements in mind. All employees also have access to paid time away to deal with critical or emergency issues. We offer additional paid time to volunteer and give back to the community.

Employees on sales plans earn performance-based incentive pay on top of their base salary, which is split between quota and non-quota components. For quota-based incentive pay, Cisco typically pays as follows:

.75% of incentive target for each 1% of revenue attainment up to 50% of quota;

1.5% of incentive target for each 1% of attainment between 50% and 75%;

1% of incentive target for each 1% of attainment between 75% and 100%; and once performance exceeds 100% attainment, incentive rates are at or above 1% for each 1% of attainment with no cap on incentive compensation.

For non-quota-based sales performance elements such as strategic sales objectives, Cisco may pay up to 125% of target. Cisco sales plans do not have a minimum threshold of performance for sales incentive compensation to be paid.

Share