Ubique Systems

TRCI-26-05480

Quality Engineer Lead Expert (AI/ML & LLM QA)

Hybrid contract role in Barcelona (1–2 days onsite) for a senior QA lead to own test strategy and automation for AI systems including ML models, LLM/RAG applications and autonomous agents. The engineer will validate functional and non-functional quality (accuracy, bias/fairness, robustness, explainability), perform security/adversarial testing (prompt injection/red teaming), benchmark performance (latency/throughput/concurrency), and build automated regression frameworks (including synthetic data) with audit-ready evidence in an insurance context, working closely with AI engineers, data scientists and DevOps.

Position summary

Location
Spain
Workplace
On-site
Employment
Contract
Experience
Minimum 3 years and Maximum 10 years
Apply now

Role overview

Why This Role Matters.

Hybrid contract role in Barcelona (1–2 days onsite) for a senior QA lead to own test strategy and automation for AI systems including ML models, LLM/RAG applications and autonomous agents. The engineer will validate functional and non-functional quality (accuracy, bias/fairness, robustness, explainability), perform security/adversarial testing (prompt injection/red teaming), benchmark performance (latency/throughput/concurrency), and build automated regression frameworks (including synthetic data) with audit-ready evidence in an insurance context, working closely with AI engineers, data scientists and DevOps.

Your Impact

Deliver Enterprise Value

Help organisations solve complex business problems through modern technology, consulting expertise and measurable outcomes.

Collaboration

Work Across Teams

Collaborate with consultants, architects, engineers and client stakeholders throughout the project lifecycle.

Growth

Learn Continuously

Gain exposure to enterprise technologies, certifications, mentoring and real-world project experience.

Career Path

Grow With Ubique

Build a long-term consulting career with opportunities to take on greater responsibility and leadership over time.

Responsibilities

What You'll Be Doing.

Every role at Ubique contributes directly to solving meaningful business challenges for our clients.

01

Design and execute test strategies for ML models, LLMs and agents (accuracy, fairness, bias, explainability, robustness)

02

Validate LLM and RAG outputs: grounding, hallucination, retrieval accuracy, vector stores, chunking logic

03

Benchmark latency, throughput, concurrency and multi-agent performance; validate fallback and retry logic

04

Build automated test frameworks for AI/ML pipelines, including synthetic data and regression testing of retrained models

05

Define model quality KPIs, produce audit-ready evidence and mentor junior testers

Technology stack

Tools & Technologies.

The platforms and technologies you'll use to build modern, enterprise-grade solutions.

01

QA

02

Quality Engineering

03

AI testing

04

ML testing

05

LLM testing

06

RAG

07

retrieval augmented generation

08

autonomous agents

09

multi-agent

10

Python

11

test automation

12

API testing

13

performance testing

14

latency

15

throughput

16

concurrency

17

security testing

18

prompt injection

19

red teaming

20

adversarial testing

Requirements

Skills & Experience.

We value curiosity, collaboration and continuous learning. If you don't meet every requirement but believe you can make an impact, we'd still love to hear from you.

Essential

Required Qualifications

Quality assurance/testing for AI systems (ML models, LLM applications, RAG pipelines, agents)

Test strategy design and execution for AI/ML quality dimensions (accuracy, robustness, explainability, fairness/bias)

LLM/RAG validation (grounding, hallucination checks, retrieval accuracy, vector stores, chunking logic)

Adversarial testing for LLMs including prompt-injection and red teaming

Python for test automation, data validation, and AI testing scripts

API testing of AI services/model endpoints

Performance testing (latency, throughput, concurrency) and validation of fallback/retry logic

Security testing of AI services/model endpoints

AI ethics, data privacy, and regulatory compliance testing

Testing AI on cloud platforms and big data ecosystems at scale

Define model quality KPIs and produce audit-ready evidence

Mentoring junior testers

Preferred

Nice to Have

Experience in insurance domain QA/compliance

Experience with multi-agent performance benchmarking

Experience building synthetic data generation for testing

Experience with regression testing of retrained models (model drift/quality gates)

What you'll gain

More Than Just A Job.

We're committed to helping every team member grow professionally, personally and technically while working on meaningful projects.

Global Exposure

Collaborate with international clients and multicultural teams on enterprise programmes.

Continuous Learning

Expand your expertise through mentoring, certifications and hands-on project experience.

Career Growth

Take ownership, develop leadership skills and grow your consulting career over time.

Flexible Working

Hybrid and remote collaboration designed around trust and delivering exceptional outcomes.

People First

Join a supportive culture where collaboration, respect and long-term relationships come first.

Enterprise Projects

Work on meaningful technology initiatives for leading organisations across industries.

Apply

Apply for Quality Engineer Lead Expert (AI/ML & LLM QA)

One page, about two minutes. We only ask for what we actually need to have a first conversation.

Your CV

Optional. A line or two is plenty.