Gen-AI Engineer

2 months ago


Raleigh NC United States Cisco Systems, Inc. Full time
Who We Are
The Cisco IT team is changing the way we run Cisco's operations by leveraging the power of technology, the best of business processes, and utilizing outstanding data insights. We are redefining how Cisco designs and delivers the employee, partner, and customer experience based on a culture that values customer service. We strive for speed and agility in all that we do. Above all else, we are kind to each other. We aspire to be an industry-leading IT team, with a strong focus on AI and security to differentiate us and foster innovation. To achieve simplicity and the best employee and customer experience, we need excellent talent and the right abilities to succeed.
Who You'll Work With
You will work with our amazing Data Infrastructure & Platforms team as part of the Infrastructure & Cloud Services. You will build solutions and support them across our portfolio of capabilities, primarily focusing on enabling AI/ML and Generative AI capabilities in a multi-functional team setup. You will collaborate with other engineers, architects, and organization leadership.
Who You Are
We are looking for a highly skilled Senior GenAI Engineer to lead the deployment and management of on-premise Large Language Models (LLMs), with a focus on Retrieval Augmented Generation (RAG). This role requires expertise in developing and supporting large-scale GenAI and ML platforms, with a strong emphasis on document management, security, vector databases, and workflow orchestration. The successful candidate will have extensive experience in responsible AI practices, including toxicity screening and ensuring regulatory compliance across AI solutions.
What You'll Do
  • Deploy and manage Large Language Models (LLMs) for on-prem environments, focusing on Retrieval Augmented Generation (RAG) and ensuring high-performance infrastructure.
  • Build and optimize secure, scalable AI/ML platforms that support enterprise-level document management and data security protocols.
  • Design, implement, and manage workflows to ensure seamless document processing, ingestion, classification, and retrieval within AI models.
  • Implement and manage vector databases to support efficient document search and retrieval within AI workflows.
  • Ensure compliance with organizational and regulatory data security standards, including encryption, access control, and auditing of sensitive documents used within AI models.
  • Implement and maintain responsible AI practices, including toxicity screening, bias detection, and regulatory compliance to ensure ethical and safe AI usage.
  • Collaborate with cross-functional teams to ensure data privacy and information security requirements are met when processing documents through AI models.
  • Continuously evaluate and improve infrastructure to support evolving AI/ML needs, particularly focusing on document ingestion, classification, and retrieval.
  • Develop and maintain automated pipelines for LLM deployment and secure document processing with real-time monitoring and alerts.
  • Work closely with compliance, legal, and governance teams to ensure AI models are aligned with security and regulatory frameworks.
  • Stay updated on the latest advancements in AI/ML, document security, and responsible AI practices.
    Minimum Qualifications
    • Bachelor's in Computer Science, Computer Engineering, Electrical Engineering, or a related STEM field.
    • 7+ years of experience in engineering, with at least 2 years specifically in AI/ML engineering
    • Proficiency in programming languages such as Python, Java, C++, or similar.
    • Hands-on experience with ML frameworks such as Kubeflow, AI operators, and/or Langchain, with proficiency in Python for AI operations.
    • Experience with MLOps principles, including model deployment, versioning, and/or monitoring in secure environments.

      Preferred Qualifications
      • Master's Degree
      • Focus on on-prem LLM deployments with an emphasis on RAG
      • Expertise in building large-scale, secure AI/ML platforms with an emphasis on document management and security protocols.
      • Experience with vector databases, such as Pinecone, Weaviate, Milvus, etc.
      • Experience with multi-instance GPUs and containerized AI/ML workflows.
      • Proven ability to collaborate with cross-functional teams, including legal, compliance, and security experts.
      • Understanding of document lifecycle management, particularly in the context of AI model ingestion, classification, and retrieval.
        Why Cisco
        #WeAre Cisco, where each person is unique, but we bring our talents to work as a team and make a difference powering an inclusive future for all.
        We embrace digital and help our customers implement change in their digital businesses. Some may think we're "old" (39 years strong) and only about hardware, but we're also a software company. And a security company. We even invented an intuitive network that adapts, predicts, learns, and protects. No other company can do what we do - you can't put us in a box

        But "Digital Transformation" is an empty buzz phrase without a culture that allows for innovation, creativity, and yes, even failure (if you learn from it).

        Day to day, we focus on the give and take. We give our best, give our egos a break, and give of ourselves (because giving back is built into our DNA). We take accountability, bold steps, and take difference to heart. Because without diversity of thought and a dedication to equality for all, there is no moving forward.

        So, you have colorful hair? Don't care. Tattoos? Show off your ink. Like polka dots? That's cool. Pop culture geek? Many of us are. Passion for technology and world changing? Be you, with us

  • Gen-AI Engineer

    6 days ago


    Raleigh, United States Cisco Systems, Inc. Full time

    Who We Are The Cisco IT team is changing the way we run Cisco's operations by leveraging the power of technology, the best of business processes, and utilizing outstanding data insights. We are redefining how Cisco designs and delivers the employee, partner, and customer experience based on a culture that values customer service. We strive for speed and...


  • Charlotte, NC, United States Cognizant Full time

    Test Lead/Manager with Gen AI testing exp. at Cognizant summary: The Test Lead/Manager at Cognizant Technology Solutions is responsible for overseeing GEN AI Testing projects while ensuring quality and compliance. This role requires a strong background in Gen AI Testing along with extensive experience in test automation, project management, and...


  • Charlotte, NC, United States Cognizant Full time

    Test Lead/Manager with Gen AI testing exp. at Cognizant summary: The Test Lead/Manager with Gen AI testing experience at Cognizant Technology Solutions leads and oversees projects focused on GEN AI Testing, ensuring quality and compliance in deliverables. With a strong background in test automation and a deep understanding of the Life and Annuities...

  • Gen-AI Engineer

    10 hours ago


    Raleigh, United States CISCO Systems Full time

    Who We Are The Cisco IT team is changing the way we run Cisco's operations by leveraging the power of technology, the best of business processes, and utilizing outstanding data insights. We are redefining how Cisco designs and delivers the employee, partner, and customer experience based on a culture that values customer service. We strive for speed and...


  • Charlotte, NC, United States Cognizant Full time

    Cognizant Technology Solutions is looking for “Test Manager with GEN AI Testing exp.” to join the team of IT professionals in a permanent role. If you meet our background requirements and skills and are looking for an opportunity with these skills and expertise, here is the ideal opportunity for you! About Cognizant’s QEA Practice: We are the...

  • Gen-AI Engineer

    4 weeks ago


    San Jose, CA, United States Cisco Systems, Inc. Full time

    The Cisco IT team is changing the way we run Cisco's operations by leveraging the power of technology, the best of business processes, and utilizing outstanding data insights. We are redefining how Cisco designs and delivers the employee, partner, and customer experience based on a culture that values customer service. We strive for speed and agility in all...

  • Gen-AI Engineer

    6 days ago


    San Jose, CA, United States Cisco Systems, Inc. Full time

    Who We Are The Cisco IT team is changing the way we run Cisco's operations by leveraging the power of technology, the best of business processes, and utilizing outstanding data insights. We are redefining how Cisco designs and delivers the employee, partner, and customer experience based on a culture that values customer service. We strive for speed and...


  • San Francisco, CA, United States Truva AI Full time

    Why Join Truva.ai Truva stands at the forefront of SaaS innovation, specializing in automating tasks, optimizing workflows, and delivering unparalleled operational efficiency with LLMs. Truva is backed by top VCs such as YCombinator and Fintech Collective and led by Gaurav - 2x founder and an alumnus of Stanford, and Anuja - an alumnus of Haas MBA from UC...


  • Raleigh, North Carolina, United States Snorkel AI Full time

    **Company Overview**Snorkel AI is revolutionizing the field of artificial intelligence by developing a data-first AI development platform. The company has seen significant growth since its inception as a research project in the Stanford AI Lab, and is now working with some of the world's largest organizations to empower scientists, engineers, financial...

  • Backend Engineer

    4 weeks ago


    San Francisco, CA, United States Hamming AI Full time

    We are a fast-growing voice AI testing company. We are winning (8Xed our revenue last month) and are hiring a backend and infra engineer to help us win faster. Here's what you'll do: Scale current products and infra to support 100x growth. This includes optimizing and productizing processes that humans currently do. LET’S CHAT IF YOU: Are a power user...

  • Product Engineer

    4 weeks ago


    San Francisco, CA, United States Hamming AI Full time

    We are a fast-growing voice AI testing company. We are winning (8Xed our revenue last month) and are hiring a product engineer to help us win faster. Here's what you'll do: 0 to 1 Build new products extremely quickly that make our customer’s voice agents more reliable. Our customers want new features, and we don’t have enough time to satisfy current...

  • Founding Engineer

    4 weeks ago


    San Francisco, CA, United States Hamming AI Full time

    We are a fast-growing voice AI testing company. We are winning (8Xed our revenue last month) and are hiring a founding engineer to help us win faster. Here's what you'll do: 0 to 1 Build new products extremely quickly that make our customer’s voice agents more reliable. Our customers want new features, and we don’t have enough time to satisfy current...


  • Redwood City, CA, United States C3 AI Full time

    We are looking for a highly skilled and experienced lead software engineer experienced in the field of machine learning and artificial intelligence, and passionate about Generative AI technology and building next-generation software platforms. As a member of C3 AI’s Generative AI team, you will be tasked with developing the infrastructure and tools to...


  • Raleigh, North Carolina, United States Snorkel AI Full time

    About the PositionWe're excited to announce the opening for a high-caliber AI Solution Architect to join our fast-growing GTM team at Snorkel AI. As a key contributor, you will act as a trusted advisor to prospective and existing customers, providing value-based evaluation scenarios and aiding in closing business.Key ResponsibilitiesOwn the technical aspects...


  • Palo Alto, CA, United States Ai Brainer Full time

    The company is committed to leveraging AI to develop innovative features and products that enhance user experience in their matchmaking services. The role involves conducting applied research in Generative AI, developing prototypes, and implementing AI-driven features in production. Collaboration with other engineers and product managers is essential to...

  • AI Solution Director

    4 weeks ago


    Atlanta, TX, United States C3 AI Full time

    The AI Solution Director role is a combination of Business and Product Development with high level of domain expertise. AI Solution Directors will work with our customers to identify and quantify high-impact business opportunities and scope them into problems that can be solved by C3 AI’s platform, then drive the delivery of solutions with compelling...

  • Cloud AI Engineer

    2 weeks ago


    Raleigh, North Carolina, United States Emonics LLC Full time

    About Emonics LLCEmonics LLC is a leading company that provides innovative solutions in data science and artificial intelligence.Job Title: Cloud AI EngineerWe are seeking an experienced Cloud AI Engineer to join our team. The ideal candidate will have a strong background in machine learning and data science, with a focus on Google Cloud Platform (GCP) and...


  • Dallas, TX, United States AMERICAN AIRLINES Full time

    IntroAre you ready to explore a world of possibilities, both at work and duringyour time off? Join our American Airlines family, and you’ll travel the world, grow your expertise and become the best version of you.  As you embark on a new journey, you’ll tackle challenges with flexibility and grace, learning new skills and advancing your career while...


  • Mountain View, CA, United States ZipRecruiter Full time

    Principal Engineer/Researcher – AI Platform Solutions Location : Mountain View, CA 94043Pay Rate: OpenAdditional Benefits : Health, Dental, Vision, 401(k), company-paid holidays and more!Type of hire : Contract, long-term with ongoing potential.Travel: No Position Summary: Seeking a passionate and highly motivated Principal Engineer/Researcher for...


  • New York, NY, United States C3 AI Full time

    Endless inspiration, meaningful work, talented team C3 AI is opening a new software development center in Guadalajara, Mexico. Software engineers in this center will help design, develop and maintain performant and scalable full-stack enterprise AI applications to solve what was previously unsolvable. Senior Software Engineer, Full-Stack C3 AI is looking for...