The AI Engineer, Quality & Evals role at Cortea AI focuses on building evaluation and observability systems for production-grade AI agents used in audits. This position entails developing both online and offline evaluation systems, creating automated quality gates for testing, and analyzing large datasets to improve agent performance. Ideal candidates will have strong Python and backend engineering experience, alongside a solid understanding of LLM evaluation systems.