Evals / model evaluation jobs in Spain
As of Oct 8, 2026, 41 job listings in Spain mention Evals / model evaluation. 22 of them (54%) say the job is remote. 33 have AI in the job title.
Most recent listings
Technical Product Lead - AI Investing
Bjak · Spain · Remote · Posted Oct 6, 2026
- AI in title
- AI literacy / AI fluency
- Evals / model evaluation
- Human-in-the-loop
- Large Language Models (LLMs)
Technical Product Manager - AI Stockbroking
Bjak · Spain · Remote · Posted Oct 6, 2026
- AI in title
- AI literacy / AI fluency
- Evals / model evaluation
- Human-in-the-loop
- Large Language Models (LLMs)
Technical Product Lead - AI Neobank
Bjak · Spain · Remote · Posted Oct 6, 2026
- AI in title
- AI literacy / AI fluency
- Evals / model evaluation
- Human-in-the-loop
- Large Language Models (LLMs)
Technical Product Manager - AI Neobank
Bjak · Spain · Remote · Posted Oct 6, 2026
- AI in title
- AI literacy / AI fluency
- Evals / model evaluation
- Human-in-the-loop
- Large Language Models (LLMs)
Technical Product Lead - AI Finance
Bjak · Spain · 2 openings · Remote · Posted Oct 6, 2026
- AI in title
- AI literacy / AI fluency
- Evals / model evaluation
- Human-in-the-loop
- Large Language Models (LLMs)
Technical Product Manager - AI Investing
Bjak · Spain · Remote · Posted Oct 6, 2026
- AI in title
- AI literacy / AI fluency
- Evals / model evaluation
- Human-in-the-loop
- Large Language Models (LLMs)
Technical Product Lead - AI Stockbroking
Bjak · Spain · Remote · Posted Oct 6, 2026
- AI in title
- AI literacy / AI fluency
- Evals / model evaluation
- Human-in-the-loop
- Large Language Models (LLMs)
Software Engineer Specialist - Data and AI
Sanofi · Barcelona · Posted Oct 6, 2026
- AI in title
- AI literacy / AI fluency
- Amazon Bedrock
- Evals / model evaluation
- Generative AI
- Large Language Models (LLMs)
- Machine learning
- Retrieval-augmented generation (RAG)
- Snowflake
Evinova Product Director – Study Operations
AstraZeneca · Spain - Barcelona · Posted Oct 5, 2026
- AI agents / agentic AI
- AI automation / workflow automation
- AI literacy / AI fluency
- Evals / model evaluation
- Fine-tuning
- Generative AI
- Human-in-the-loop
- Machine learning
- Prompt engineering
- Retrieval-augmented generation (RAG)
- Vector databases / embeddings
Manager, AI Governance
Dow Jones · Spain - Barcelona · Posted Oct 2, 2026
- AI in title
- AI agents / agentic AI
- AI governance / responsible AI
- AI risk management
- Evals / model evaluation
- Generative AI
- Human-in-the-loop
- Large Language Models (LLMs)
- Machine learning
- Model Context Protocol (MCP)
- Retrieval-augmented generation (RAG)
Lead Test Framework Architect (M/F/D)
Agilent · Spain-Barcelona · Posted Oct 1, 2026
- AI agents / agentic AI
- AI literacy / AI fluency
- Claude / Anthropic
- Claude Code
- Context engineering
- Cursor
- Evals / model evaluation
- Generative AI
- GitHub Copilot
- Human-in-the-loop
- Large Language Models (LLMs)
- Microsoft Copilot
- Model Context Protocol (MCP)
- Retrieval-augmented generation (RAG)
- Vector databases / embeddings
Lead AI & Data Engineer – APD Platform
AstraZeneca · Spain - Barcelona · Posted Oct 1, 2026
- AI in title
- AI agents / agentic AI
- AI governance / responsible AI
- Evals / model evaluation
- Generative AI
- Knowledge graphs
- LangChain / LangGraph
- Large Language Models (LLMs)
- MLOps / LLMOps
- Machine learning
- Prompt engineering
- Retrieval-augmented generation (RAG)
- Snowflake
Director of AI Research, AI for Oncology Clinical Development
AstraZeneca · Spain - Barcelona · Posted Sep 29, 2026
- AI in title
- AI agents / agentic AI
- Evals / model evaluation
- Machine learning
- Multimodal AI
Software Engineer III, Google Cloud, Google Threat Intelligence
Google · Málaga, Spain · Posted Sep 29, 2026
- AI agents / agentic AI
- Evals / model evaluation
- Large Language Models (LLMs)
- NLP
AI Research Lead, AI for Oncology Clinical Development
AstraZeneca · Spain - Barcelona · Posted Sep 28, 2026
- AI in title
- AI agents / agentic AI
- Evals / model evaluation
- Machine learning
- Multimodal AI
Applied AI Engineer
Bjak · Spain · 2 openings · Remote · Posted Sep 28, 2026
- AI in title
- AI literacy / AI fluency
- ChatGPT / OpenAI
- Deep learning
- Evals / model evaluation
- Fine-tuning
- Large Language Models (LLMs)
- Machine learning
LLM Application Engineer
Bjak · Spain · Remote · Posted Sep 28, 2026
- AI in title
- AI agents / agentic AI
- AI literacy / AI fluency
- ChatGPT / OpenAI
- Context engineering
- Evals / model evaluation
- Generative AI
- Large Language Models (LLMs)
- Vector databases / embeddings
Agentic Architect
Kyndryl · Madrid HQ (KES51610) | Bilbao, Vizcaya, Spain | Barcelona, Spain · Posted Sep 24, 2026
- AI in title
- AI agents / agentic AI
- Azure OpenAI / Azure AI
- ChatGPT / OpenAI
- Claude / Anthropic
- Claude Code
- Evals / model evaluation
- LangChain / LangGraph
- Large Language Models (LLMs)
- Model Context Protocol (MCP)
- Vector databases / embeddings
Head of GTM Systems & AI
360Learning · Spain, Remote · Posted Sep 16, 2026
- AI in title
- AI agents / agentic AI
- AI literacy / AI fluency
- Evals / model evaluation
- Large Language Models (LLMs)
Engineer - AI Platform
Ebury · Madrid · Posted Sep 16, 2026
- AI in title
- AI agents / agentic AI
- Evals / model evaluation
- Large Language Models (LLMs)
- Vertex AI
Principal Data Scientist - Deep Learning
Jampp · Spain - Remote · Posted Sep 16, 2026
- AI in title
- Deep learning
- Evals / model evaluation
- Machine learning
- Vector databases / embeddings
Product Manager, Procure-to-Pay
Zone & Co · Spain · Posted Sep 11, 2026
- AI agents / agentic AI
- AI literacy / AI fluency
- Evals / model evaluation
- Human-in-the-loop
Developer Experience - Engineering Manager
Ebury · Madrid · Posted Sep 10, 2026
- AI agents / agentic AI
- Claude / Anthropic
- Evals / model evaluation
- Gemini
- Human-in-the-loop
- Large Language Models (LLMs)
- Model Context Protocol (MCP)
Technical Product Lead - AI Investing App
Bjak · Spain · Remote · Posted Sep 1, 2026
- AI in title
- AI literacy / AI fluency
- Evals / model evaluation
- Human-in-the-loop
- Large Language Models (LLMs)
Technical Product Lead - AI Stockbroking App
Bjak · Spain · Remote · Posted Sep 1, 2026
- AI in title
- AI literacy / AI fluency
- Evals / model evaluation
- Human-in-the-loop
- Large Language Models (LLMs)
Evals / model evaluation jobs in other countries
- United States2,106
- India328
- United Kingdom180
- Canada147
- Singapore84
- Germany53
- Ireland48
- Poland44
- China38
- Philippines37
- South Korea35
- Brazil32
- Taiwan28
- Japan28
- France28
- Israel27
- Portugal26
- Australia26
- Netherlands25
- Hong Kong24
- Vietnam23
- Indonesia23
- Sweden22
- Mexico21
Browse more topics
- Jobs in North America
- Jobs in Latin America
- Jobs in Europe
- Jobs in the Middle East & Africa
- Jobs in Asia Pacific
- Remote jobs
- AI agents / agentic AI jobs
- Large Language Models (LLMs) jobs
- Machine learning jobs
- AI literacy / AI fluency jobs
- Generative AI jobs
- Claude / Anthropic jobs
- AI automation / workflow automation jobs
- Vector databases / embeddings jobs
- Microsoft Copilot jobs
- ChatGPT / OpenAI jobs
- Retrieval-augmented generation (RAG) jobs
- Snowflake jobs
How this list works: we read the public job boards that employers publish on Greenhouse, Lever, Ashby, SmartRecruiters and Workday, plus Google's and Microsoft's own career sites. Locations, regions and remote status come from what each posting says. AI skill labels are matched from each posting's text. Velza is not the employer and does not take applications. Always confirm details on the employer's site.