← Back to results

ai-model-evaluation jobs in San Diego

Posted 4 days ago

Design and deploy generative AI solutions across firmware, embedded OS, and toolchains to automate workflows and optimize code. Partner with software, silicon, and DevOps teams to integrate AI copilots, code analysis, and test generation into existing development pipelines.… Evaluate AI models, develop scalable inference services, and measure their impact on developer productivity. Drive best practices for secure and compliant AI adoption in a fast-paced environment.

San DiegoLast seen 3 days ago
Posted 11 days ago

Lead end-to-end automation testing for web and mobile applications using Java, Selenium, Playwright, and TestNG, with specialized expertise in AI/LLM model evaluation and Generative AI testing. Design and maintain robust automation frameworks for regression, functional, and integration testing while creating comprehensive test plans for both traditional software and AI-based applications.… Evaluate LLM outputs against expected behavior, validate prompt/response quality, and document model-quality findings. Collaborate with developers, product owners, and stakeholders using Shift-Left Testing practices and Agile methodologies, leveraging AI-powered tools like GitHub Copilot and ChatGPT to enhance testing efficiency.

San DiegoLast seen 9 days ago