TRCI-26-05480
Quality Engineer Lead Expert (AI/ML & LLM QA)
Hybrid contract role in Barcelona (1–2 days onsite) for a senior QA lead to own test strategy and automation for AI systems including ML models, LLM/RAG applications and autonomous agents. The engineer will validate functional and non-functional quality (accuracy, bias/fairness, robustness, explainability), perform security/adversarial testing (prompt injection/red teaming), benchmark performance (latency/throughput/concurrency), and build automated regression frameworks (including synthetic data) with audit-ready evidence in an insurance context, working closely with AI engineers, data scientists and DevOps.
Position summary
- Location
- Spain
- Workplace
- On-site
- Employment
- Contract
- Experience
- Minimum 3 years and Maximum 10 years
Role overview
Why This Role Matters.
Hybrid contract role in Barcelona (1–2 days onsite) for a senior QA lead to own test strategy and automation for AI systems including ML models, LLM/RAG applications and autonomous agents. The engineer will validate functional and non-functional quality (accuracy, bias/fairness, robustness, explainability), perform security/adversarial testing (prompt injection/red teaming), benchmark performance (latency/throughput/concurrency), and build automated regression frameworks (including synthetic data) with audit-ready evidence in an insurance context, working closely with AI engineers, data scientists and DevOps.
Your Impact
Deliver Enterprise Value
Help organisations solve complex business problems through modern technology, consulting expertise and measurable outcomes.
Collaboration
Work Across Teams
Collaborate with consultants, architects, engineers and client stakeholders throughout the project lifecycle.
Growth
Learn Continuously
Gain exposure to enterprise technologies, certifications, mentoring and real-world project experience.
Career Path
Grow With Ubique
Build a long-term consulting career with opportunities to take on greater responsibility and leadership over time.
Responsibilities
What You'll Be Doing.
Every role at Ubique contributes directly to solving meaningful business challenges for our clients.
Validate LLM and RAG outputs: grounding, hallucination, retrieval accuracy, vector stores, chunking logic
Benchmark latency, throughput, concurrency and multi-agent performance; validate fallback and retry logic
Build automated test frameworks for AI/ML pipelines, including synthetic data and regression testing of retrained models
Define model quality KPIs, produce audit-ready evidence and mentor junior testers
Technology stack
Tools & Technologies.
The platforms and technologies you'll use to build modern, enterprise-grade solutions.
QA
Quality Engineering
AI testing
ML testing
LLM testing
RAG
retrieval augmented generation
autonomous agents
multi-agent
Python
test automation
API testing
performance testing
latency
throughput
concurrency
security testing
prompt injection
red teaming
adversarial testing
Requirements
Skills & Experience.
We value curiosity, collaboration and continuous learning. If you don't meet every requirement but believe you can make an impact, we'd still love to hear from you.
Essential
Required Qualifications
Quality assurance/testing for AI systems (ML models, LLM applications, RAG pipelines, agents)
Test strategy design and execution for AI/ML quality dimensions (accuracy, robustness, explainability, fairness/bias)
LLM/RAG validation (grounding, hallucination checks, retrieval accuracy, vector stores, chunking logic)
Adversarial testing for LLMs including prompt-injection and red teaming
Python for test automation, data validation, and AI testing scripts
API testing of AI services/model endpoints
Performance testing (latency, throughput, concurrency) and validation of fallback/retry logic
Security testing of AI services/model endpoints
AI ethics, data privacy, and regulatory compliance testing
Testing AI on cloud platforms and big data ecosystems at scale
Define model quality KPIs and produce audit-ready evidence
Mentoring junior testers
Preferred
Nice to Have
Experience in insurance domain QA/compliance
Experience with multi-agent performance benchmarking
Experience building synthetic data generation for testing
Experience with regression testing of retrained models (model drift/quality gates)
What you'll gain
More Than Just A Job.
We're committed to helping every team member grow professionally, personally and technically while working on meaningful projects.
Global Exposure
Collaborate with international clients and multicultural teams on enterprise programmes.
Continuous Learning
Expand your expertise through mentoring, certifications and hands-on project experience.
Career Growth
Take ownership, develop leadership skills and grow your consulting career over time.
Flexible Working
Hybrid and remote collaboration designed around trust and delivering exceptional outcomes.
People First
Join a supportive culture where collaboration, respect and long-term relationships come first.
Enterprise Projects
Work on meaningful technology initiatives for leading organisations across industries.
Apply
Apply for Quality Engineer Lead Expert (AI/ML & LLM QA)
One page, about two minutes. We only ask for what we actually need to have a first conversation.