Evals / model evaluation jobs in Poland
As of Oct 8, 2026, 44 job listings in Poland mention Evals / model evaluation. 20 of them (45%) say the job is remote. 33 have AI in the job title.
Most recent listings
Technical Product Manager - AI Neobank
Bjak · Poland · Remote · Posted Oct 6, 2026
- AI in title
- AI literacy / AI fluency
- Evals / model evaluation
- Human-in-the-loop
- Large Language Models (LLMs)
Technical Product Lead - AI Stockbroking
Bjak · Poland · Remote · Posted Oct 6, 2026
- AI in title
- AI literacy / AI fluency
- Evals / model evaluation
- Human-in-the-loop
- Large Language Models (LLMs)
Technical Product Lead - AI Neobank
Bjak · Poland · Remote · Posted Oct 6, 2026
- AI in title
- AI literacy / AI fluency
- Evals / model evaluation
- Human-in-the-loop
- Large Language Models (LLMs)
Technical Product Manager - AI Investing
Bjak · Poland · Remote · Posted Oct 6, 2026
- AI in title
- AI literacy / AI fluency
- Evals / model evaluation
- Human-in-the-loop
- Large Language Models (LLMs)
Technical Product Lead - AI Finance
Bjak · Poland · 2 openings · Remote · Posted Oct 6, 2026
- AI in title
- AI literacy / AI fluency
- Evals / model evaluation
- Human-in-the-loop
- Large Language Models (LLMs)
Technical Product Manager - AI Stockbroking
Bjak · Poland · 2 openings · Remote · Posted Oct 6, 2026
- AI in title
- AI literacy / AI fluency
- Evals / model evaluation
- Human-in-the-loop
- Large Language Models (LLMs)
Technical Product Lead - AI Investing
Bjak · Poland · Remote · Posted Oct 6, 2026
- AI in title
- AI literacy / AI fluency
- Evals / model evaluation
- Human-in-the-loop
- Large Language Models (LLMs)
Senior Data Scientist (Supply Chain Planning)
Procter & Gamble · WARSAW - EU SUPPLY NETWORK HUB · Posted Sep 30, 2026
- AI in title
- Databricks
- Evals / model evaluation
Senior Software Engineer - Integrations 3P - Platform
Elastic · Poland · Posted Sep 29, 2026
- AI agents / agentic AI
- Evals / model evaluation
- Large Language Models (LLMs)
- Model Context Protocol (MCP)
Director, Product Engineering
Fortrea · Warsaw · Posted Sep 29, 2026
- AI agents / agentic AI
- AI automation / workflow automation
- AI governance / responsible AI
- Azure OpenAI / Azure AI
- ChatGPT / OpenAI
- Claude / Anthropic
- Claude Code
- Cursor
- Evals / model evaluation
- GitHub Copilot
- Large Language Models (LLMs)
- Machine learning
- Microsoft Copilot
- Prompt engineering
- Retrieval-augmented generation (RAG)
- Vertex AI
Manager, Engineering Excellence & Quality Engineering
Fortrea · Warsaw · Posted Sep 29, 2026
- AI agents / agentic AI
- AI automation / workflow automation
- AI governance / responsible AI
- ChatGPT / OpenAI
- Claude / Anthropic
- Claude Code
- Cursor
- Evals / model evaluation
- GitHub Copilot
- Large Language Models (LLMs)
- Machine learning
- Microsoft Copilot
- Prompt engineering
- Retrieval-augmented generation (RAG)
- Vector databases / embeddings
Senior Technical Architect
Fortrea · Warsaw · Posted Sep 29, 2026
- AI agents / agentic AI
- AI automation / workflow automation
- AI governance / responsible AI
- Azure OpenAI / Azure AI
- ChatGPT / OpenAI
- Claude / Anthropic
- Claude Code
- Cursor
- Evals / model evaluation
- GitHub Copilot
- Large Language Models (LLMs)
- Machine learning
- Microsoft Copilot
- Model Context Protocol (MCP)
- Prompt engineering
- Retrieval-augmented generation (RAG)
- Vector databases / embeddings
- Vertex AI
Customer Support Analyst, AI ML Application Support
Workday · Poland, Warsaw · Posted Sep 29, 2026
- AI in title
- AI agents / agentic AI
- Evals / model evaluation
- Large Language Models (LLMs)
Middle Full Stack Developer (Node.js/React)
airSlate · Poland · Remote · Posted Sep 29, 2026
- Evals / model evaluation
- LangChain / LangGraph
- Large Language Models (LLMs)
- Model Context Protocol (MCP)
- Retrieval-augmented generation (RAG)
- Vector databases / embeddings
LLM Application Engineer
Bjak · Poland · Remote · Posted Sep 28, 2026
- AI in title
- AI agents / agentic AI
- AI literacy / AI fluency
- ChatGPT / OpenAI
- Context engineering
- Evals / model evaluation
- Generative AI
- Large Language Models (LLMs)
- Vector databases / embeddings
Applied AI Engineer
Bjak · Poland · 2 openings · Remote · Posted Sep 28, 2026
- AI in title
- AI literacy / AI fluency
- ChatGPT / OpenAI
- Deep learning
- Evals / model evaluation
- Fine-tuning
- Large Language Models (LLMs)
- Machine learning
Lead Data Scientist
Danaher · Krakow, Poland · Posted Sep 25, 2026
- AI in title
- Evals / model evaluation
- Machine learning
Lead AI Scientist
Danaher · Krakow, Poland · Posted Sep 25, 2026
- AI in title
- AI agents / agentic AI
- Evals / model evaluation
- Generative AI
- Machine learning
- Multimodal AI
Senior Risk Data Scientist
GSK · Poznan Pastelowa · Posted Sep 23, 2026
- AI in title
- Databricks
- Evals / model evaluation
- Machine learning
Senior Java Developer - AI/NLP Solutions
State Street · Gdansk, Poland | BIG - Zielinskiego Krakow · Posted Sep 18, 2026
- AI in title
- AI governance / responsible AI
- Evals / model evaluation
- Human-in-the-loop
- Large Language Models (LLMs)
- Machine learning
- NLP
- Retrieval-augmented generation (RAG)
- Snowflake
- Vector databases / embeddings
Software Engineer III, Context Intelligence, Knowledge Catalog
Google · Warsaw, Poland · Posted Sep 17, 2026
- AI agents / agentic AI
- Evals / model evaluation
- NLP
AI Software Engineer
Cisco · Krakow, Poland | Warsaw, Poland · Posted Sep 16, 2026
- AI in title
- AI agents / agentic AI
- Evals / model evaluation
- Fine-tuning
- LangChain / LangGraph
- Large Language Models (LLMs)
- Machine learning
- Prompt engineering
- Retrieval-augmented generation (RAG)
- Vector databases / embeddings
Senior AI Engineer (Remote, Contract)
INFUSE · Poland · Posted Sep 16, 2026
- AI in title
- AI agents / agentic AI
- AI literacy / AI fluency
- Amazon Bedrock
- ChatGPT / OpenAI
- Claude / Anthropic
- Claude Code
- Evals / model evaluation
- LangChain / LangGraph
- Large Language Models (LLMs)
- Model Context Protocol (MCP)
- Retrieval-augmented generation (RAG)
Senior Staff Software Engineer, AI Agent Platform, Vertex
Google · Warsaw, Poland · Posted Sep 15, 2026
- AI in title
- AI agents / agentic AI
- Evals / model evaluation
- Fine-tuning
- Generative AI
AI Builder
Everfield · Poland- Warsaw | Remote · Posted Sep 9, 2026
- AI in title
- AI agents / agentic AI
- AI literacy / AI fluency
- Claude / Anthropic
- Claude Code
- Evals / model evaluation
- Large Language Models (LLMs)
Evals / model evaluation jobs in other countries
- United States2,110
- India341
- United Kingdom180
- Canada147
- Singapore84
- Germany53
- Ireland48
- Spain41
- China39
- Philippines37
- South Korea35
- Brazil34
- Taiwan28
- Japan28
- France28
- Israel27
- Portugal26
- Australia26
- Netherlands25
- Hong Kong24
- Vietnam23
- Indonesia23
- Sweden22
- Mexico21
Browse more topics
- Jobs in North America
- Jobs in Latin America
- Jobs in Europe
- Jobs in the Middle East & Africa
- Jobs in Asia Pacific
- Remote jobs
- AI agents / agentic AI jobs
- Large Language Models (LLMs) jobs
- Machine learning jobs
- AI literacy / AI fluency jobs
- Generative AI jobs
- Claude / Anthropic jobs
- AI automation / workflow automation jobs
- Vector databases / embeddings jobs
- Microsoft Copilot jobs
- ChatGPT / OpenAI jobs
- Retrieval-augmented generation (RAG) jobs
- Snowflake jobs
How this list works: we read the public job boards that employers publish on Greenhouse, Lever, Ashby, SmartRecruiters and Workday, plus Google's and Microsoft's own career sites. Locations, regions and remote status come from what each posting says. AI skill labels are matched from each posting's text. Velza is not the employer and does not take applications. Always confirm details on the employer's site.