Submitting more applications increases your chances of landing a job.

Here’s how busy the average job seeker was last month:

Opportunities viewed

Applications submitted

Keep exploring and applying to maximize your chances!

Looking for employers with a proven track record of hiring women?

Click here to explore opportunities now!

We Value Your Feedback

You are invited to participate in a survey designed to help researchers understand how best to match workers to the types of jobs they are searching for

Would You Be Likely to Participate?

If selected, we will contact you via email with further instructions and details about your participation.

You will receive a $7 payout for answering the survey.

https://bayt.page.link/2XfCxVb5V5aYDGpV8

Back to the job results

AI Evaluation Engineer

- FNZ Group
- India

3 days ago 2026/10/01

Complete Questionnaire

Apply on company site

Other Business Support Services

Create a job alert for similar positions

Job alert turned off. You won’t receive updates for this search anymore.

Undo

Job description

AI Evaluation Engineer

Location: Pune, India

Seniority: Mid-level (3-6 years)

Purpose: Execute comprehensive evaluations of FNZ's AI agents across the six-pillar framework, working as a generalist while developing specialist expertise in 1-2 pillars.

Key Responsibilities:

Design and conduct evaluations covering Task Performance, Safety, Efficiency, Groundedness, Robustness, and Suitability
Create "golden sets" of test examples representing expert judgment on desired agent behaviour
Develop evaluation rubrics and scoring criteria aligned to FNZ Evaluation Framework principles
Build comprehensive test suites covering happy paths, edge cases, and adversarial inputs
Evaluate multi-step agentic workflows: planning, tool selection, execution, error handling
Assess agent groundedness: verify outputs against knowledge bases, detect hallucinations
Document findings with clear evidence; collaborate with development teams on remediation
Contribute to building automated evaluation platform and CI/CD integration

Skills and Experience:

3-6 years in software testing, QA engineering, AI/ML development, or data science
Hands-on test automation skills, experience with ML frameworks highly valuable
Practical experience evaluating LLM applications, RAG systems, or AI agents
Understanding of prompt engineering, retrieval-augmented generation, and agent architectures
Analytical mindset to decompose complex agent behaviours and identify failure modes
Strong documentation and presentation skill

About FNZ

FNZ is committed to opening up wealth so that everyone, everywhere can invest in their future on their terms. We know the foundation to do that already exists in the wealth management industry, but complexity holds firms back.

We created wealth’s growth platform to help. We provide a global, end-to-end wealth management platform that integrates modern technology with business and investment operations. All in a regulated financial institution.

We partner with the world’s leading financial institutions, with over US$2.4 trillion in assets on platform (AoP).
Together with our clients, we empower nearly 30 million people across all wealth segments to invest in their future.

This job post has been translated by AI and may contain minor differences or errors.

Apply on company site Email to Friend Complete Questionnaire

Compare your profile with other applicants

Cancel

You’ve reached the maximum limit of 15 job alerts. To create a new alert, please delete an existing one first.

MANAGE

Job alert created for this search. You’ll receive updates when new jobs match.

Manage alerts

Are you sure you want to unapply?

You'll no longer be considered for this role and your application will be removed from the employer's inbox.