Evals / model evaluation jobs in Canada
As of Oct 8, 2026, 147 job listings in Canada mention Evals / model evaluation. 47 of them (32%) say the job is remote. 113 have AI in the job title.
Most recent listings
Senior Machine Learning Engineer
Autodesk · Toronto, ON, CAN · Posted Oct 7, 2026
- AI in title
- AI agents / agentic AI
- Evals / model evaluation
- MLOps / LLMOps
- Machine learning
- Snowflake
Senior Director, Analyst – Software Engineering for AI and Agentic Applications (Remote - Canada) / Directeur principal, analyste – Ingénierie logicielle pour l’IA et les applications agentiques (Télétravail – Canada)
Gartner · Remote - Canada · Posted Oct 7, 2026
- AI in title
- AI agents / agentic AI
- AI literacy / AI fluency
- Evals / model evaluation
- Large Language Models (LLMs)
AVP, AI
Manulife / John Hancock · Toronto, Ontario · Posted Oct 7, 2026
- AI in title
- AI agents / agentic AI
- AI governance / responsible AI
- AI risk management
- Evals / model evaluation
- Generative AI
- Large Language Models (LLMs)
- MLOps / LLMOps
- Machine learning
- Vector databases / embeddings
Principal Product Manager - Workday AI - Agent Factory
Workday · Canada, BC, Vancouver · Posted Oct 6, 2026
- AI in title
- AI agents / agentic AI
- Evals / model evaluation
- Human-in-the-loop
Developer Lead (AI-Enabled Systems)
BMO · Toronto, ON, CAN · Posted Oct 5, 2026
- AI in title
- AI agents / agentic AI
- Amazon Bedrock
- Azure OpenAI / Azure AI
- ChatGPT / OpenAI
- Claude / Anthropic
- Evals / model evaluation
- Human-in-the-loop
- LangChain / LangGraph
- Large Language Models (LLMs)
- Machine learning
- Microsoft Copilot
- Prompt engineering
- Retrieval-augmented generation (RAG)
- Vector databases / embeddings
GenAI Software Developer
BMO · Calgary, AB, CAN · Posted Oct 5, 2026
- AI in title
- AI agents / agentic AI
- AI automation / workflow automation
- AI governance / responsible AI
- Azure OpenAI / Azure AI
- ChatGPT / OpenAI
- Evals / model evaluation
- Generative AI
- GitHub Copilot
- Large Language Models (LLMs)
- Microsoft Copilot
- Model Context Protocol (MCP)
- Prompt engineering
- Retrieval-augmented generation (RAG)
- Vector databases / embeddings
Full Stack Engineer (AI-Enabled)
BMO · Calgary, AB, CAN · Posted Oct 5, 2026
- AI in title
- AI agents / agentic AI
- Amazon Bedrock
- Azure OpenAI / Azure AI
- ChatGPT / OpenAI
- Claude / Anthropic
- Evals / model evaluation
- Human-in-the-loop
- LangChain / LangGraph
- Large Language Models (LLMs)
- Machine learning
- Microsoft Copilot
- Prompt engineering
- Retrieval-augmented generation (RAG)
- Vector databases / embeddings
Senior Data Scientist
Momentum Financial Services Group · Toronto, Canada · Posted Oct 5, 2026
- AI in title
- Databricks
- Evals / model evaluation
- MLOps / LLMOps
- Machine learning
- Snowflake
Lead Data Scientist, Customer & Growth Analytics
Thomson Reuters · Canada, Toronto, Ontario · Posted Oct 5, 2026
- AI in title
- Evals / model evaluation
- Machine learning
Sr / Principal Product Manager - AI Agent Products
Workday · Canada, BC, Vancouver · Posted Oct 5, 2026
- AI in title
- AI agents / agentic AI
- AI automation / workflow automation
- AI literacy / AI fluency
- Claude / Anthropic
- Cursor
- Evals / model evaluation
- Human-in-the-loop
- Large Language Models (LLMs)
Stagiaire en science des données -- Data Science Intern
TobogganLabs · Montréal, Quebec, Canada · Posted Oct 4, 2026
- AI in title
- Evals / model evaluation
MCP/AI Developer
Autodesk · Toronto, ON, CAN · Posted Oct 2, 2026
- AI in title
- AI agents / agentic AI
- Evals / model evaluation
- Human-in-the-loop
- Large Language Models (LLMs)
- Model Context Protocol (MCP)
- Retrieval-augmented generation (RAG)
- Vector databases / embeddings
Senior Software Engineer - Security Integrations - Prompt Engineering
Elastic · Canada · Posted Oct 2, 2026
- AI in title
- Evals / model evaluation
- Generative AI
- Prompt engineering
Senior Software Developer, ML Platform & Infrastructure
Wealthsimple Technologies · Toronto Headquarters · Posted Oct 2, 2026
- AI in title
- Amazon Bedrock
- Evals / model evaluation
- Generative AI
- Large Language Models (LLMs)
- MLOps / LLMOps
- Machine learning
- Prompt engineering
- Snowflake
- Vector databases / embeddings
Senior Cloud Platform Engineer (AWS), AI Infrastructure - Evinova
AstraZeneca · Canada - Mississauga · Posted Oct 1, 2026
- AI in title
- AI agents / agentic AI
- Amazon Bedrock
- ChatGPT / OpenAI
- Claude / Anthropic
- Evals / model evaluation
- Gemini
- Generative AI
- Large Language Models (LLMs)
- Machine learning
- Model Context Protocol (MCP)
- Retrieval-augmented generation (RAG)
- Vertex AI
Intern, AI Data Developer (Winter)
Autodesk · Toronto, ON, CAN · Posted Oct 1, 2026
- AI in title
- Claude / Anthropic
- Claude Code
- Cursor
- Evals / model evaluation
- Generative AI
- GitHub Copilot
- Large Language Models (LLMs)
- Machine learning
- Microsoft Copilot
- Retrieval-augmented generation (RAG)
Intern, AI/ML Platform (Winter)
Autodesk · Toronto, ON, CAN · Posted Oct 1, 2026
- AI in title
- Evals / model evaluation
- Generative AI
- Large Language Models (LLMs)
- MLOps / LLMOps
- Vector databases / embeddings
Staff Product Manager
Klue · Vancouver, British Columbia | Toronto, Ontario · Remote · Posted Oct 1, 2026
- AI literacy / AI fluency
- Evals / model evaluation
Applied AI Engineer – GenAI Systems
Manulife / John Hancock · Toronto, Ontario · Posted Oct 1, 2026
- AI in title
- AI agents / agentic AI
- Databricks
- Evals / model evaluation
- Generative AI
- LangChain / LangGraph
- Large Language Models (LLMs)
- Machine learning
- NLP
- Retrieval-augmented generation (RAG)
- Vector databases / embeddings
AgentCore Platform Engineer
Appnovation Technologies · Toronto, Montreal · Posted Sep 30, 2026
- Amazon Bedrock
- ChatGPT / OpenAI
- Evals / model evaluation
- LangChain / LangGraph
- Large Language Models (LLMs)
- LlamaIndex
- Model Context Protocol (MCP)
- Retrieval-augmented generation (RAG)
Principal Platform Engineer
SkyWatch · Hybrid/Kitchener · Remote · Posted Sep 30, 2026
- AI agents / agentic AI
- AI literacy / AI fluency
- Evals / model evaluation
- Large Language Models (LLMs)
Group Product Manager - Home Financing
BMO · Toronto, ON, CAN · Posted Sep 29, 2026
- Evals / model evaluation
Staff Data Scientist, Applied ML
Jobber · Remote | Vancouver | Kitchener-Waterloo | Edmonton | Toronto · Posted Sep 29, 2026
- AI in title
- Deep learning
- Evals / model evaluation
- Knowledge graphs
- Large Language Models (LLMs)
- MLOps / LLMOps
- Machine learning
- Retrieval-augmented generation (RAG)
- Snowflake
Lead Data Engineer (AI & Data Strategy)
Mastercard · Vancouver, Canada · Posted Sep 29, 2026
- AI in title
- AI agents / agentic AI
- Databricks
- Evals / model evaluation
- Machine learning
Senior Software Engineer II - Agentic Intelligence
Honeycomb.io · Remote - Canada · Posted Sep 25, 2026
- AI in title
- AI agents / agentic AI
- Amazon Bedrock
- Evals / model evaluation
- Fine-tuning
- Large Language Models (LLMs)
- Model Context Protocol (MCP)
- Prompt engineering
- Retrieval-augmented generation (RAG)
Evals / model evaluation jobs in other countries
- United States2,110
- India341
- United Kingdom180
- Singapore84
- Germany53
- Ireland48
- Poland44
- Spain41
- China39
- Philippines37
- South Korea35
- Brazil34
- Taiwan28
- Japan28
- France28
- Israel27
- Portugal26
- Australia26
- Netherlands25
- Hong Kong24
- Vietnam23
- Indonesia23
- Sweden22
- Mexico21
Browse more topics
- Jobs in North America
- Jobs in Latin America
- Jobs in Europe
- Jobs in the Middle East & Africa
- Jobs in Asia Pacific
- Remote jobs
- AI agents / agentic AI jobs
- Large Language Models (LLMs) jobs
- Machine learning jobs
- AI literacy / AI fluency jobs
- Generative AI jobs
- Claude / Anthropic jobs
- AI automation / workflow automation jobs
- Vector databases / embeddings jobs
- Microsoft Copilot jobs
- ChatGPT / OpenAI jobs
- Retrieval-augmented generation (RAG) jobs
- Snowflake jobs
How this list works: we read the public job boards that employers publish on Greenhouse, Lever, Ashby, SmartRecruiters and Workday, plus Google's and Microsoft's own career sites. Locations, regions and remote status come from what each posting says. AI skill labels are matched from each posting's text. Velza is not the employer and does not take applications. Always confirm details on the employer's site.