← Back to results

rag-retrieval-augmented-generation jobs in San Diego

Posted 1 day ago

Lead the design, development, and deployment of enterprise-scale Generative AI and LLM solutions across Azure and AWS environments. Define AI architecture frameworks, standards, and governance practices while architecting RAG systems, agentic AI, multi-agent systems, and vector database solutions. Partner with business leaders and engineering teams to translate business requirements into scalable, secure, and responsible AI platforms. Mentor architects and data science teams while evaluating emerging AI technologies and driving innovation initiatives.

San DiegoLast seen today
$122,600 – $177,900 · Posted 1 day ago

SHEIN is seeking a Senior Data Engineer to build and productionize GenAI/LLM solutions for data engineering workflows, including code assistance, metadata discovery, and knowledge access. The role focuses on developing RAG and agentic workflows, automating incident triage and log analysis, and integrating AI capabilities into internal developer tools and services. You'll own scoped projects from problem definition through production support, requiring 3+ years of hands-on production experience with strong Python/SQL fundamentals, proven GenAI/LLM application development, and practical knowledge of databases, APIs, data pipelines, and distributed systems.

San DiegoLast seen today
$202,500 – $274,000 · Posted 3 days ago

Staff Data Engineer role at Credit Karma building petabyte-scale data infrastructure and streaming platforms on GCP. You will design and operate Kafka-based streaming infrastructure, build cloud-native data pipelines using Dataflow/Beam, Flink, and Spark, and create persistence frameworks across Spanner, MySQL, and BigQuery. The role requires 7+ years backend/data systems experience in JVM languages (Scala/Java), expertise with high-throughput distributed systems, and deep knowledge of streaming platforms, Apache Beam, CDC patterns, encryption, and data governance frameworks. You'll also integrate AI/ML and generative AI technologies including LLMs, RAG, semantic search, and knowledge graphs into the data platform.

San DiegoLast seen 1 day ago
$216,100 – $378,200 · Posted 3 days ago

Principal Engineer leading the technical direction for ServiceNow's Conversational Analytics platform, which enables customers to query data in natural language using AI. The role combines hands-on architecture and prototyping with mentorship of engineering teams, requiring deep expertise in AI/ML systems, vector retrieval, agentic architectures, and knowledge graphs. You'll collaborate with product and leadership to design cloud-based solutions integrating LLMs and AI into enterprise workflows, with accountability for customer success metrics and evaluation frameworks.

San DiegoLast seen 1 day ago
$122,600 – $177,900 · Posted 3 days ago

Senior Data Engineer role focused on building and productionizing GenAI/LLM solutions for data engineering workflows. You will develop RAG and agentic workflows, automate incident triage and root-cause analysis, integrate AI capabilities into internal developer tools, and own projects from problem definition through production support. Required: 3+ years of production systems experience, strong Python/SQL, hands-on GenAI/LLM application building, and software engineering fundamentals including testing, CI/CD, and monitoring.

San DiegoLast seen 1 day ago
$264,000 – $330,000 · Posted 4 days ago

Principal Machine Learning Engineer at AppFolio to architect and lead mission-critical AI systems across the Realm-X property management platform. You'll design advanced AI agentic systems combining reasoning, planning, and execution; establish ML platform primitives for end-to-end production workflows; and drive the transition toward autonomous property management using LLMs, fine-tuning, and reinforcement learning. Requires 10+ years building software systems with deep expertise in traditional ML, deep learning, generative AI, LLM post-training (SFT, RLHF, DPO, RL), and production ML at scale—plus a Master's or Ph.D. in Computer Science or related field.

San DiegoLast seen 2 days ago
$120,000 – $165,000 · Posted 4 days ago

Sr. AI Developer at XiFin will design and build AI-powered applications, agents, and automation tools that enhance software development workflows and improve engineering productivity. The role involves developing intelligent testing solutions integrated with CI/CD pipelines, implementing AI-assisted developer tools using Claude and AWS Bedrock, and evaluating emerging AI technologies. You'll work hands-on across Python, LLMs, generative AI, and cloud-native AWS infrastructure, partnering with Engineering, QA, and DevOps teams to prototype and deploy scalable, production-ready solutions. The position requires 4–7+ years of software engineering experience with strong Python skills, LLM/generative AI application experience, and familiarity with modern DevOps practices.

San DiegoLast seen 2 days ago
$95,500 – $95,500 · Posted 4 days ago

Design and build production AI solutions including copilots, agents, RAG applications, and workflow automations on Azure OpenAI, Claude, and Microsoft platforms. Write custom code in Python, JavaScript/TypeScript, or C# and leverage low-code platforms like Power Automate and Logic Apps where appropriate, following solid engineering practices (version control, testing, CI/CD, AI output evaluation). Integrate solutions with enterprise systems and APIs with security-first design, own projects end-to-end from deployment through adoption and improvement, and work with security, architecture, governance, and compliance teams to ensure production readiness. Communicate clearly to both technical and business audiences, document designs, and coach teammates on AI capabilities, limits, and risks.

San DiegoLast seen 3 days ago
$135,600 – $226,100 · Posted 4 days ago

Lead the design, deployment, and governance of an enterprise AI platform integrating AWS Bedrock, Microsoft 365 Copilot, and structured/unstructured data sources. Act as the primary technical authority for AI infrastructure, RAG architecture, security controls, and AI adoption across business units, with responsibility for platform budget, vendor contracts, and internal capability building through mentoring and community-of-practice leadership. Requires hands-on production experience with Bedrock (Agents, Knowledge Bases, Guardrails), M365 Copilot enterprise deployment, and proven ability to operate as a solo platform owner in a matrixed organization without direct reports.

San DiegoLast seen 2 days ago
$180,000 – $230,000 · Posted 5 days ago

Zensar is seeking an experienced AI Architect to design, develop, and deploy enterprise-scale AI and Generative AI solutions. The role requires deep expertise in Large Language Models (Claude, Gemini, OpenAI), multi-cloud platforms (Azure, AWS), and enterprise architecture, with 10+ years in software/cloud architecture and 5+ years designing AI/ML solutions. Responsibilities include defining AI strategy and architecture standards, designing end-to-end GenAI systems (RAG, agentic AI, multi-agent systems), architecting secure infrastructure across Azure and AWS, establishing governance and responsible AI frameworks, and leading adoption programs. The ideal candidate will mentor teams, evaluate emerging technologies, and translate business requirements into scalable, production-grade AI systems.

San DiegoLast seen 4 days ago
$141,200 – $278,300 · Posted 6 days ago

Lead the design and delivery of enterprise AI platforms and applications on Google Cloud, leveraging Vertex AI, Gemini, and cloud-native technologies. Design, fine-tune, and govern LLM solutions; build RAG and agentic systems; and define end-to-end architectures spanning data pipelines, feature engineering, model lifecycle, APIs, and MLOps/LLMOps. Architect cloud-native applications on GKE, Cloud Run, and managed services while implementing security, governance, and production-grade monitoring for AI/ML systems at scale.

San DiegoLast seen 4 days ago
$155,600 – $306,800 · Posted 6 days ago

Senior Microsoft Forward Deployed Engineer who will design, build, and deploy GenAI/LLM-powered solutions for government clients, translating business problems into production-grade AI systems. The role requires 7+ years of software/data engineering experience, 1+ years hands-on with GenAI/LLM solutions, and deep expertise in Microsoft Azure (AI Foundry, OpenAI, AI Search), Python/TypeScript/C#, and Copilot extensibility. You will lead project workstreams, work directly with client technical teams in fast-paced environments, and mentor others while building reliable, maintainable, well-documented code. 50% travel required.

San DiegoLast seen 6 days ago
$121,625 – $217,711 · Posted 7 days ago

The AI Architect II is a senior technical leadership role responsible for establishing architectural standards and patterns for AI solutions, including RAG pipelines, agentic workflows, and model serving. The role involves designing scalable AWS cloud architectures, defining MLOps practices, implementing responsible AI governance, and translating business requirements into enterprise AI and cloud strategies. The position requires hands-on technical expertise to evaluate and integrate AI/ML services, lead systems design across microservices and event-driven architectures, and mentor engineering teams on cloud and AI best practices. This role drives technology decisions with real business impact while modernizing the company's platform through cloud-native and AI transformation initiatives.

San DiegoLast seen 4 days ago
$140,800 – $211,200 · Posted 7 days ago

Design, develop, and optimize machine learning systems and models for production AI platforms, with focus on inference optimization, scalable ML infrastructure, and deployment. Responsibilities include building ML pipelines, optimizing model inference across hardware environments, integrating LLMs and models into APIs and microservices, designing data pipelines for ingestion and feature engineering, and collaborating cross-functionally on end-to-end ML solutions. Requires strong software engineering fundamentals combined with deep ML expertise, proficiency in Python and at least one systems language (C++, Rust, or Go), and solid understanding of ML frameworks, transformer architectures, and model deployment systems.

San DiegoLast seen 4 days ago
$140,800 – $211,200 · Posted 8 days ago

Design, develop, and optimize machine learning systems for production AI platforms, including model development, inference optimization, and scalable ML infrastructure. Build training-to-deployment pipelines, optimize model serving for latency and cost, and integrate LLMs and generative AI models into microservices and APIs. Develop data pipelines for ingestion, preprocessing, and feature engineering while collaborating cross-functionally with product, platform, and hardware teams to deliver end-to-end ML solutions.

San DiegoLast seen 7 days ago
Posted 8 days ago

The AI Architect will lead the design, development, and deployment of enterprise-scale generative AI solutions across Azure and AWS, partnering with business and engineering leaders to define AI strategy and architecture standards. The role requires deep expertise in modern LLMs (Claude, Gemini, OpenAI), multi-cloud AI platforms, and enterprise architecture, with 10+ years in software/cloud architecture and 5+ years architecting AI/ML solutions. Key responsibilities include designing end-to-end GenAI solutions leveraging RAG, agentic AI, and multi-agent systems; establishing cloud-native deployment patterns; and defining governance, compliance, and responsible AI frameworks. The ideal candidate will serve as a trusted technical advisor to executives, mentor engineering teams, and drive AI adoption through reusable frameworks and reference architectures.

San DiegoLast seen 6 days ago
Posted 9 days ago

Design, build, and ship LLM-powered capabilities end to end—from prototyping and fine-tuning models to deploying production agents and retrieval systems. Own the full stack: prompt and context engineering, multi-step agent design with tool calling, RAG systems (embeddings, chunking, hybrid search, reranking), fine-tuning on multi-GPU with LoRA/QLoRA, evaluation and tracing infrastructure, and clean APIs for other engineers. Work in a secure, distributed environment where the platform runs on customer compute, cloud, or hybrid setups, requiring expertise with both commercial and self-hosted models.

San DiegoLast seen 7 days ago
Posted 9 days ago

S. Navy. a Developer, you will design, develop, and maintain scalable ETL/ELT pipelines and cloud-based data platforms that integrate enterprise systems and support AI/ML initiatives for the U.S. Navy. You will perform data engineering tasks including cleansing, transformation, and optimization of SQL and Apache Spark workloads on AWS, while also developing predictive models and generative AI solutions using Python-based technologies. You will build interactive dashboards and visualizations in Tableau or Qlik, translating technical findings into actionable insights for stakeholders. An active Secret clearance, U.S. citizenship, and a bachelor's degree in a technical field (or 4+ years of relevant professional experience) are required.

San DiegoLast seen 7 days ago
$121,625 – $217,711 · Posted 10 days ago

The Applied AI Engineer develops, deploys, and maintains generative AI solutions for insurance products and internal workflows. The role involves implementing end-to-end AI pipelines using AWS services (SageMaker, Lambda, ECS/EKS, S3), building feature engineering workflows in Snowflake, and collaborating with cloud and MLOps teams. Candidates need 3+ years of AI/ML engineering experience with 1–2 years focused on generative AI or LLMs, proficiency in Python and frameworks like PyTorch or Hugging Face Transformers, and hands-on cloud deployment experience. The position requires monitoring model performance, ensuring security and compliance, and supporting AI architecture reviews.

San DiegoLast seen 8 days ago
$55,300 – $126,000 · Posted 11 days ago

Junior AI Software Developer supporting Navy cybersecurity and mission operations in San Diego. You will participate in problem discovery with mission stakeholders, contribute to technical solution design, and deliver hands-on implementation of scripts, APIs, automation, and AI/LLM-enabled workflows. Work involves integrating systems via REST/HTTP APIs, writing and testing Python code, using version control, and documenting technical solutions for both technical and non-technical audiences while adhering to RMF, security controls, and responsible AI practices in classified or controlled environments.

San DiegoLast seen 9 days ago
$86,900 – $198,000 · Posted 11 days ago

Senior AI Software Developer who will design, implement, and integrate AI/LLM-enabled solutions and automation for Navy cybersecurity and mission workflows. You'll work in small, multidisciplinary teams translating mission requirements into technical approaches, delivering hands-on code (Python, APIs, data integration), and ensuring solutions align with RMF, security controls, and compliance boundaries. The role blends advisory conversations with focused implementation of secure, deployable systems that reduce cyber risk and improve operational efficiency. Success requires 4+ years of software development or integration experience, Python proficiency, API and data store expertise, cloud or enterprise deployment knowledge, and ability to communicate across technical and non-technical stakeholders.

San DiegoLast seen 9 days ago
$69,300 – $158,000 · Posted 11 days ago

An AI Software Developer at Booz Allen Hamilton will design, implement, and integrate AI-enabled automation and analytics solutions for Navy cybersecurity and mission workflows. The role requires hands-on Python development, API integration, and secure system design in classified or controlled environments, with close collaboration between cybersecurity specialists, engineers, and mission stakeholders. Success demands translating mission requirements into technical tasks, delivering credible implementations (not slide-only recommendations), and operating within RMF, compliance, and secure engineering practices. The ideal candidate combines software engineering rigor (version control, code review, testing) with AI/LLM experience and the ability to work in small, structured teams on multidisciplinary problem-solving.

San DiegoLast seen 9 days ago
$140,000 – $180,000 · Posted 12 days ago

Senior Applied AI Engineer at West Health will design, build, and deploy production-quality AI-powered tools and workflows using large language models, multi-agent architectures, and orchestration platforms. The role involves building intelligent automation pipelines with tools like n8n, LangChain/LangGraph, and Microsoft Copilot Studio, integrating AI capabilities with existing data infrastructure (Snowflake, Python), and conducting data analysis to identify high-impact AI opportunities. The ideal candidate combines deep practical expertise in applied AI with the ability to rapidly move from concept to fully working, well-documented applications that create measurable value for healthcare delivery and senior care initiatives.

San DiegoLast seen 10 days ago
$128,100 – $192,100 · Posted 14 days ago

Design, develop, and deploy end-to-end generative AI solutions including GenAI agents, LLMs, and applications to solve complex business challenges. Build data pipelines, design APIs for AI accessibility, and collaborate with cross-functional teams (developers, data scientists, business) to implement enterprise GenAI solutions. Requires 4+ years of full-stack application development (Java, Python, JavaScript) and 3+ years with data structures, algorithms, and data stores. Stay current with AI research and communicate technical concepts to non-technical stakeholders.

San DiegoLast seen 10 days ago
$75,380 – $85,280 · Posted 14 days ago

Staff Software Developer who will design, develop, and maintain AI-enabled knowledge management and decision-support platforms for environmental and engineering applications. Responsibilities include building data ingestion workflows, implementing retrieval-augmented generation (RAG) systems with vector embeddings and semantic search, developing backend services and APIs, creating secure web-based UIs, and working with cloud services (Azure AI Search, Blob Storage, PostgreSQL). The role requires strong software development skills demonstrated through a public code portfolio, familiarity with LLMs and prompt engineering, and the ability to collaborate across technical teams to translate project needs into practical features.

San DiegoLast seen 10 days ago