Senior Generative AI Engineer - USA

CogniifySan Francisco, CA
$150,000 - $170,000Remote

About The Position

We are looking for a hands-on Generative AI and LLM Engineer to design, build and deploy production-grade AI applications. The role will focus on developing LLM-powered products, Retrieval-Augmented Generation (RAG) pipelines and intelligent agent workflows using Python, modern AI frameworks and cloud infrastructure. You will own solutions from requirement understanding and architecture through development, deployment and production monitoring. The ideal candidate combines strong Python backend development experience with practical knowledge of LLMs, vector databases, API integrations, Docker and at least one major cloud platform.

Requirements

  • Bachelor's or Master's degree in Computer Science, Engineering or a related discipline, or equivalent practical software-development experience.
  • 6-9 years of professional software-development experience, including strong hands-on experience with Python.
  • Experience developing backend services and integrating REST APIs.
  • Hands-on experience building LLM or Generative AI applications.
  • Practical experience implementing RAG using embeddings, semantic search and vector databases.
  • Experience with LangChain, LangGraph, LlamaIndex or a comparable LLM application framework.
  • Experience integrating foundation models through APIs such as OpenAI, Anthropic Claude, Gemini or Azure OpenAI.
  • Working knowledge of vector databases such as Pinecone, Weaviate, Milvus, Qdrant, Chroma, FAISS or pgvector.
  • Experience with Docker and deployment on at least one cloud platform: AWS, Azure or GCP.
  • Understanding of prompt engineering, hallucination reduction, output validation and LLM evaluation.
  • Strong understanding of software engineering practices, Git, testing, debugging and clean code.
  • Ability to communicate technical solutions clearly to both technical and non-technical stakeholders.

Nice To Haves

  • Experience building agentic or multi-agent workflows using LangGraph, AutoGen, CrewAI, Semantic Kernel or similar frameworks.
  • Experience with Hugging Face Transformers, PyTorch or fine-tuning techniques such as LoRA or QLoRA.
  • Knowledge of model serving and inference frameworks such as vLLM, TGI or Ollama.
  • Experience with Kubernetes, CI/CD pipelines and infrastructure automation.
  • Familiarity with observability or LLMOps tools such as LangSmith, Langfuse, Arize Phoenix, MLflow or Weights & Biases.
  • Knowledge of reranking, hybrid search, chunking strategies and retrieval evaluation.
  • Experience implementing AI guardrails, PII protection, prompt-injection prevention and responsible AI practices.

Responsibilities

  • Design and develop scalable LLM-powered applications using Python.
  • Build RAG pipelines using document processing, embeddings, vector databases, semantic search and reranking.
  • Develop AI-agent and multi-agent workflows with tool calling, memory, orchestration and human approval steps.
  • Integrate LLMs with internal systems, external APIs, databases and enterprise applications.
  • Evaluate and select suitable foundation models based on accuracy, latency, cost, security and business requirements.
  • Improve prompt quality, retrieval accuracy, response time and token usage.
  • Implement safety guardrails, output validation, access controls and fallback mechanisms.
  • Build automated evaluation frameworks to measure response quality, hallucination, relevance and reliability.
  • Containerize applications using Docker and deploy them on AWS, Azure or GCP.
  • Implement monitoring and LLMOps practices for model performance, cost, latency, errors and production usage.
  • Collaborate with product, engineering and business teams to convert requirements into reliable AI solutions.
  • Document technical architecture, design decisions, APIs and operational processes.

Benefits

  • Unlimited PTO
  • Generous parental leave
  • Entrepreneurial culture
  • Open communication with management and company leadership
  • Small, dynamic teams = massive impact
  • Medical, Dental and Vision coverage for employees
  • Access to Disability & Life insurance
  • Mental health and wellbeing support
  • Annual bonus program
  • Employer Stock Purchase Program (ESPP)
  • Yearly team building experiences
  • Mentorship and sponsorship opportunities
  • Manager resources and support
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service