Nexxa.ai Logo

Nexxa.ai

Backend AI Engineer

Posted 27 Days Ago
Be an Early Applicant
In-Office or Remote
Hiring Remotely in Toronto, ON, CAN
Senior level
In-Office or Remote
Hiring Remotely in Toronto, ON, CAN
Senior level
Design and operate production backend services and AI infrastructure, including model serving, inference, data pipelines, vector stores, RAG systems, APIs, and microservices. Architect scalable real-time and batch systems, integrate and deploy LLM, computer vision, and machine learning models, and ensure reliability, observability, testing, and CI/CD. Collaborate with ML, product, and customer-facing teams while documenting systems and mentoring engineers.
The summary above was generated by AI

Nexxa is building the best AI systems for heavy industries — enabling machines, systems and operations to think, decide and act autonomously across manufacturing, large-scale infrastructure, logistics and legacy environments.
Our mission is to translate deep technical breakthroughs into operational reality, solving some of the hardest systems-level problems in industry.

Role Overview

We're looking for Backend AI Engineers to design, build, and own the core AI infrastructure and services that power Nexxa's products in production. Where our Forward Deployed Engineers embed with customers to deliver solutions on the ground, this role builds the systems that make those solutions possible at scale: model-serving pipelines, inference and orchestration layers, data pipelines, and the APIs and microservices that connect Generative AI, Computer Vision, and Machine Learning models to real enterprise environments.

This role is a blend of backend software engineering, ML infrastructure, and systems architecture. You'll design distributed systems, integrate and serve models in production, and build the reusable platform capabilities that our customer-facing and product teams depend on.

Key Responsibilities
  • Design, build, and maintain backend services and APIs that power GenAI, LLM, and Computer Vision model integrations across Nexxa's products.

  • Build and own core AI/ML infrastructure: model-serving pipelines, inference services, data pipelines, and embedding/vector stores.

  • Architect scalable, production-grade systems for real-time and batch AI workloads across manufacturing, infrastructure, and logistics domains.

  • Implement and optimize RAG systems, prompt/context pipelines, and orchestration layers connecting models to enterprise and operational data sources.

  • Build robust APIs, microservices, and integration layers connecting AI systems to customer data, legacy systems, and existing infrastructure.

  • Own the reliability, performance, and observability of backend AI systems — logging, monitoring, testing, and CI/CD for ML services.

  • Collaborate closely with Forward Deployed Engineers, ML engineers, and product teams to translate customer and field requirements into reusable, hardened backend capabilities.

  • Evaluate and integrate ML/CV/LLM models into production backend systems; manage model versioning, rollout, and deployment pipelines.

  • Produce clear technical documentation: architecture diagrams, API specs, and runbooks for internal and customer-facing teams.

  • Mentor engineers and contribute to internal backend engineering best practices.

Qualifications
  • 4–8+ years of experience in backend software engineering, ML/platform engineering, or similar roles.

  • Strong proficiency in TypeScript/Node.js (our primary backend language), with strong API and microservice design skills. Working proficiency in Python is a plus for ML/model integration work.

  • Hands-on experience building and operating production backend systems at scale — distributed systems, databases, message queues.

  • Experience integrating ML or Generative AI models (LLMs, multimodal models) into backend services — inference, orchestration, and evaluation.

  • Solid understanding of cloud infrastructure (AWS, GCP, or Azure) and containerization (Docker, Kubernetes).

  • Experience designing and operating data pipelines (batch and/or streaming) across structured and unstructured data.

  • Hands-on experience building retrieval-augmented generation (RAG) systems and AI memory architectures — retrieval pipelines, vector stores, context management, and long-term/session memory for LLM applications.

  • Strong grasp of system design fundamentals: scalability, reliability, security, and observability.

  • Comfortable working cross-functionally with ML engineers, product, and customer-facing teams.

  • Bachelor's degree (or higher) in Computer Science or a related field.

Preferred
  • Familiarity with ML frameworks (PyTorch, TensorFlow, OpenCV) sufficient to integrate, serve, or evaluate models, even without training them yourself.

  • Experience with MLOps tooling: model registries, feature stores, CI/CD for ML, and monitoring/observability for ML systems.

  • Background in event-driven or real-time systems (Kafka, gRPC, WebSockets).

  • Experience in industrial, IoT, or operational technology (OT) environments.

  • Experience in startup or high-growth environments.

What We're Looking For
  • A backend engineer who wants to build the infrastructure powering real-world autonomous AI systems.

  • Someone who can architect for scale and reliability while still moving fast.

  • A systems thinker who enjoys turning ambiguous AI capabilities into dependable, production-grade backend services.

  • A strong collaborator who partners well with ML engineers, Forward Deployed teams, and product.

Why Join Nexxa.AI?

Innovative Environment: Build the foundational systems behind groundbreaking AI and automation technologies transforming heavy industries.

Collaborative Culture: Be part of a team that values innovation, discipline, and continuous improvement.

Professional Growth: Benefit from significant opportunities for career development and advancement.

Competitive Compensation: Enjoy a comprehensive salary and equity package reflective of your expertise and contributions.

If you're passionate about backend engineering and eager to build the infrastructure powering advanced AI solutions, we'd love to connect.

Similar Jobs

One Month Ago
Easy Apply
Remote
Canada
Easy Apply
Entry level
Entry level
Cloud • Security • Software • Cybersecurity • Automation
Build and maintain the Python-based backend engine for GitLab Duo Chat, including agentic LangGraph flows, the Flow Registry, and Duo Workflow Service. Integrate chat capabilities with GitLab’s Rails monolith and GraphQL API, develop secure and tested high-scale features, optimize performance, review code, address technical debt, and collaborate cross-functionally. Participate in weekday, weekend, and occasional nighttime Tier 2 or Tier 3 on-call rotations.
Top Skills: Duo Workflow ServiceGitlab RailsGraphQLLangchainLanggraphLarge Language ModelsPythonRspecRuby On Rails
6 Days Ago
Remote
Ontario, ON, CAN
Senior level
Senior level
Artificial Intelligence • Legal Tech • Software
Build and scale backend systems and AI workflows for Spellbook’s legal technology platform. Responsibilities include designing reliable search, inference, orchestration, and RAG retrieval systems; managing rate limits, retries, fallbacks, permissions, and data isolation; optimizing MongoDB performance; partnering with Product and Design; and participating in on-call response. The role requires strong distributed-systems expertise, production backend experience, and the ability to turn ambiguous problems into reliable shipped solutions.
Top Skills: AnthropicAWSAws CdkDockerExpressLexical SearchLlmsMongoDBNode.jsOpenaiRagTrpcTypescriptVector Search
One Month Ago
Easy Apply
Remote
Canada
Easy Apply
Senior level
Senior level
Cloud • Security • Software • Cybersecurity • Automation
Lead technical direction and architecture for DAP Repository Flows, design and ship autonomous and scheduled agents that operate on customer repositories, improve DAP onboarding, collaborate across teams on platform interfaces, mentor engineers, and solve high-scale technical problems to raise engineering quality.
Top Skills: Autonomous AgentsDuo Agent Platform (Dap)GitlabLarge Language Models (Llms)RubyRuby On RailsSlack

What you need to know about the Toronto Tech Scene

Although home to some of the biggest names in tech, including Google, Microsoft and Amazon, Toronto has established itself as one of the largest startup ecosystems in the world. And with over 2,000 startups — more than 30 percent of the country's total startups — Toronto continues to attract new businesses. Be it helping entrepreneurs manage their finances, simplifying business operations by automating payroll or assisting pharmaceutical companies in launching new drugs, the city's tech scene is just getting started.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account