Get new jobs by email
- ...We are looking for a Python Engineer – Evaluator Library to design and implement reusable evaluation components that ensure the quality, safety, and compliance of enterprise AI agents and LLM-powered workflows. You will build custom evaluation capabilities used across AI...SuggestedContract work
- ...We are looking for a QA / ML Tester – Evaluation Framework to ensure the quality, reliability, and correctness of evaluation systems used across enterprise AI and agent-based platforms. In this role, you will design and execute validation strategies for evaluation frameworks...Suggested
- ...We are looking for a Platform Engineer – CI/CD Gate & Online Evaluation to help establish automated quality controls for enterprise AI platforms and agent-based systems. In this role, you will integrate evaluation workflows into deployment pipelines, configure online quality...Suggested
- ...We are looking for a Senior ML / Evaluation Engineer to help define and implement quality standards for enterprise-grade AI agents and LLM-powered applications. In this role, you will design evaluation frameworks, build custom evaluation pipelines, and establish automated...Suggested
- ...strategies using VPCs, PrivateLink, and cloud-native security controls. Perform security reviews, threat assessments, and architecture evaluations for enterprise services. Conduct security testing activities, including penetration testing, OWASP-based assessments, and...Suggested
- ...suites using Python and pytest. Validate permit/deny decisions, policy precedence rules, parameter-level access controls, and policy evaluation correctness. Test transitions between LOG_ONLY and ENFORCE policy modes and verify expected enforcement behavior. Perform API...Suggested
- ...equivalent) • Stateful distributed system patterns (checkpointing, idempotency) Will be a plus: • AWS AgentCore Runtime maturity evaluation experience • AWS Bedrock Memory API • Enterprise HITL approval workflow integration What it’s like to work at Intellias At...Suggested
- ...Partner with AI engineering teams to define reliable integration points for LLM- and agent-based capabilities , including retrieval, evaluation, guardrails, observability, latency, cost, and data-protection boundaries. Shape healthcare data architecture, including FHIR-...SuggestedContract work
- ...troubleshoot, and optimize integration performance across distributed environments. Develop performance and latency tests for policy evaluation and agent interaction flows. Collaborate with platform, security, and AI engineering teams to deliver reliable and scalable agent...Suggested
- ..., anomaly detection, or data quality scoring. Practical experience integrating LLMs into production data workflows , including evaluation, reliability, latency, cost, and data-protection considerations. Experience with AI engineering tools such as Claude Code, GitHub...Suggested
- ...Experience building reusable AI skills, agents, or AI-assisted engineering workflows for data engineering. Familiarity with LLM evaluation, guardrails, data privacy, and determining when deterministic/rule-based processing is preferable to an LLM-based solution. What...SuggestedPermanent employment
- ...intelligent multi-agent systems running on AWS Bedrock. The role focuses on agent orchestration, delegation patterns, framework integration, evaluation workflows, and engineering best practices for production-grade AI applications. The ideal candidate has strong Python skills, hands-...Suggested