Get new jobs by email
- ...We are looking for a Python Engineer – Evaluator Library to design and implement reusable evaluation components that ensure the quality, safety, and compliance of enterprise AI agents and LLM-powered workflows. You will build custom evaluation capabilities used across AI...SuggestedContract work
- ...We are looking for an Agent Evaluation Engineer to design and build enterprise-grade evaluation frameworks for agentic AI systems. In this role, you will be responsible for creating automated evaluation pipelines, deployment quality gates, and reliability measurements that...Suggested
- ...We are looking for a Platform Engineer – CI/CD Gate & Online Evaluation to help establish automated quality controls for enterprise AI platforms and agent-based systems. In this role, you will integrate evaluation workflows into deployment pipelines, configure online quality...Suggested
- ...We are looking for a Senior ML / Evaluation Engineer to help define and implement quality standards for enterprise-grade AI agents and LLM-powered applications. In this role, you will design evaluation frameworks, build custom evaluation pipelines, and establish automated...Suggested
- ...Design token validation, claims mapping, and authorization decision workflows. Implement audit logging and traceability for policy evaluation and access decisions. Partner with security, architecture, and platform teams to support enterprise security reviews and...Suggested
- ...Partner with AI engineering teams to define reliable integration points for LLM- and agent-based capabilities , including retrieval, evaluation, guardrails, observability, latency, cost, and data-protection boundaries. Shape healthcare data architecture, including FHIR-...SuggestedContract work
- ...suites using Python and pytest. Validate permit/deny decisions, policy precedence rules, parameter-level access controls, and policy evaluation correctness. Test transitions between LOG_ONLY and ENFORCE policy modes and verify expected enforcement behavior. Perform API...Suggested
- ...Experience building reusable AI skills, agents, or AI-assisted engineering workflows for data engineering. Familiarity with LLM evaluation, guardrails, data privacy, and determining when deterministic/rule-based processing is preferable to an LLM-based solution. What...SuggestedPermanent employment
- ...operating their application, ensuring a set of compliance and best practices. What you will do Design and implement automated evaluation frameworks for LangGraph-based agent workflows and orchestration pipelines. Develop build-time evaluation suites covering agent behavior...Suggested
- ...troubleshoot, and optimize integration performance across distributed environments. Develop performance and latency tests for policy evaluation and agent interaction flows. Collaborate with platform, security, and AI engineering teams to deliver reliable and scalable agent...Suggested
- ..., anomaly detection, or data quality scoring. Practical experience integrating LLMs into production data workflows , including evaluation, reliability, latency, cost, and data-protection considerations. Experience with AI engineering tools such as Claude Code, GitHub...Suggested
- ...maintain secure healthcare APIs using modern authentication, authorization, encryption, and API security patterns. Drive technical evaluation and selection of interoperability platforms, frameworks, tools, and technology solutions. Ensure scalability, reliability,...SuggestedRemote job