eBay Logo

eBay

Senior Applied Researcher

Posted 8 Days Ago
Be an Early Applicant
In-Office
Toronto, ON, CAN
Senior level
In-Office
Toronto, ON, CAN
Senior level
Lead applied research on evaluation, monitoring, and safety for generative AI (LLMs, VLMs, agentic systems). Develop scientifically grounded metrics, AI sandbox environments, benchmark methodologies, and safety guardrails. Partner with engineering and product teams to deploy evaluation, monitoring, and governance practices into production, and translate Responsible AI requirements into measurable, scalable solutions.
The summary above was generated by AI

At eBay, we're more than a global ecommerce leader — we’re changing the way the world shops and sells. Our platform empowers millions of buyers and sellers in more than 190 markets around the world. We’re committed to pushing boundaries and leaving our mark as we reinvent the future of ecommerce for enthusiasts.

Our customers are our compass, authenticity thrives, bold ideas are welcome, and everyone can bring their unique selves to work — every day. We're in this together, sustaining the future of our customers, our company, and our planet.

Join a team of passionate thinkers, innovators, and dreamers — and help us connect people and build communities to create economic opportunity for all.

About the team and the role:  

The AI Systems Performance & Governance team at eBay AI, Research and Innovation is looking for a highly qualified Senior Applied Researcher to join our team in Toronto.

In this role, you will work at the intersection of applied research and AI governance for Generative AI systems (LLMs, VLMs, and agentic systems). You will focus on the science of evaluation, monitoring, and safety for GenAI applications, while contributing to the development of AI sandbox environments, evaluation frameworks, and governance best practices.

You will help ensure that GenAI and agentic systems are scientifically grounded, properly evaluated, safe, and production-ready across their full lifecycle. 

What you will accomplish:

  • Drive the fundamental science behind evaluation, creating robust methodologies and metrics to accurately assess the performance, safety, and reliability of LLMs, VLMs, and agentic systems, including offline metrics, human evaluation, and online experimentation.

  • Define scientifically grounded metrics for complex GenAI systems, including RAG pipelines and multi-agent workflows (e.g., hallucination, grounding, robustness, agent reliability, safety).

  • Provide scientific leadership across complex and ambiguous GenAI initiatives by setting research direction, reviewing evaluation and safety methodologies, and driving alignment on technical standards across teams.

  • Build and evolve AI Sandbox environments for safe experimentation, benchmarking, and validation of GenAI and agentic systems.

  • Advance the science behind building content moderation and safety guardrails, developing novel approaches for agentic safety steering to ensure autonomous and generative systems operate safely and ethically.

  • Establish scientific best practices and anti-patterns for GenAI and agentic application development, covering evaluation, safety, and system design.

  • Partner with engineering and product teams to ensure evaluation, safety, and monitoring approaches are applied consistently in production systems.

  • Translate Responsible AI requirements into measurable, testable, and scalable evaluation and safety solutions.

What you will bring: 

  • Master’s degree or PhD in Computer Science, Engineering, Mathematics, or a related field.

  • Proven experience in machine learning, with strong hands-on experience building at least one of the following: LLMs, VLMs, Conversational Search systems, Agentic Systems, or Content Moderation solutions.

  • Proficiency in Python and frameworks such as PyTorch, TensorFlow, Langfuse, LangGraph, or similar.

  • Solid understanding of machine learning algorithms, model architectures, training techniques, and building performant inference pipelines. Experience with model inference optimization techniques and libraries is a plus.

  • Experience with data preprocessing, feature engineering, model evaluation metrics, and large-scale data processing frameworks such as Spark.

  • Solid understanding of the ML lifecycle, including experimentation, validation, and post-deployment monitoring.

  • Excellent analytical and problem-solving skills, and the ability to work in a fast-paced, dynamic environment.

  • Demonstrated experience independently leading complex applied research initiatives from problem formulation through production adoption, with evidence of influencing technical direction, mentoring others, and establishing methodologies or standards used beyond an individual project.

  • Strong communication and collaboration skills, with the ability to explain complex technical concepts to non-technical collaborators, propose creative solutions, and support tracking and delivery within release plans.

  • Publication record in top AI conferences or journals is a strong plus.

  • Experience with AI safety, LLM/VLM/agent guards, content moderation, and policy-driven GenAI evaluation is a strong plus.

What We Offer

  • An opportunity to work on cutting-edge research in GenAI evaluation and monitoring, making significant contributions to both the field and real-world applications.

  • A collaborative and supportive work environment where innovation and creativity are encouraged.

  • Access to state-of-the-art resources and tools to support your research and development work.

  • A culture that values diversity, inclusion, and the professional growth of its members.

  • Competitive compensation and benefits package, tailored to attract the best talent in the field.

 

Additional Details

This job posting relates to an existing vacancy within eBay.

eBay is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, national origin, sex, sexual orientation, gender identity, and disability, or other legally protected status. If you have a need that requires accommodation, please contact us at [email protected]. We will make every effort to respond to your request for accommodation as soon as possible. View our accessibility statement to learn more about eBay's commitment to ensuring digital accessibility.


We use cookies to enhance your experience and may use AI tools for administrative tasks in the hiring process. To learn how we handle your personal data and use AI responsibly, please visit our Talent Privacy Notice, Privacy Center, and AI Hiring Guidelines.

Similar Jobs

3 Days Ago
Remote or Hybrid
CA
Senior level
Senior level
Angel or VC Firm • Artificial Intelligence • Information Technology • Software
Join a talent network connecting senior/staff applied AI scientists to VC-backed startups. Responsibilities include researching and developing ML methods, fine-tuning foundation models, designing experiments and evaluation frameworks, building prototypes, collaborating with engineering/product to productionize models, improving model robustness and efficiency, curating datasets, investigating failures, mentoring peers, and communicating results to stakeholders.
Top Skills: AnthropicDatabricksDistributed TrainingGoogleGpusHugging FaceJaxKubernetesMetaOpenaiPythonPyTorchScikit-LearnSnowflakeSparkTensorFlowVector Databases
An Hour Ago
Remote or Hybrid
Canada
Mid level
Mid level
HR Tech • Information Technology • Professional Services • Sales • Software
Own the full sales cycle for mid-market SaaS customers in Canada, from prospecting and relationship development through product demonstrations, forecasting, negotiation, and closing. Build pipeline through outbound activity, engage decision-makers, assess buying readiness, manage sales data in Salesforce, and network with consultants, partners, influencers, and industry events. Collaborate with Sales Engineers and Business Development representatives to drive new business.
Top Skills: HrisSaaSSalesforce
3 Hours Ago
Remote or Hybrid
Ontario, ON, CAN
Senior level
Senior level
Digital Media • Gaming • Information Technology • Software • Sports • Esports • Big Data Analytics
Develop and maintain relationships with VIP gaming customers, drive sales metrics and net revenue, promote loyalty, and deliver exceptional player experiences. The role executes responsible gaming policies, analyzes player data and trends, recommends customer solutions, and resolves escalations with internal teams. Candidates need a related bachelor's degree or five years of relevant gaming, hospitality, sales, or marketing experience, plus the ability to obtain and maintain required gaming licenses.
Top Skills: AI

What you need to know about the Toronto Tech Scene

Although home to some of the biggest names in tech, including Google, Microsoft and Amazon, Toronto has established itself as one of the largest startup ecosystems in the world. And with over 2,000 startups — more than 30 percent of the country's total startups — Toronto continues to attract new businesses. Be it helping entrepreneurs manage their finances, simplifying business operations by automating payroll or assisting pharmaceutical companies in launching new drugs, the city's tech scene is just getting started.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account