Instrumentl Logo

Instrumentl

Senior Backend Engineer

Reposted 10 Days Ago
Remote
Hiring Remotely in Canada
Senior level
Remote
Hiring Remotely in Canada
Senior level
Build and ship production AI features (LLM agents, tool/function calling, RAG pipelines) and reliable backend services. Implement observability, tests, rollback/fallbacks, and evaluation to keep models grounded, safe, and cost-effective. Collaborate with product, design, and GTM; run experiments and raise engineering standards.
The summary above was generated by AI
Hello, we’re Instrumentl. 👋

Nonprofits do some of the most important work in the world, and most of them are still managing grants in spreadsheets. We're fixing that.

Instrumentl is a profitable, hypergrowth, YC-backed SaaS platform building the operating system for grant-funded organizations. More than 5,500 nonprofits use Instrumentl to discover, track, and win grant funding, from local community organizations to the San Diego Zoo and the University of Alaska. Collectively they’ve moved over $1 billion through our platform.

We're growing quickly, customers love us (check out our G2 reviews!), and we're hiring people who want to build something that matters.


About the role

    We're hiring a Senior Backend Engineer to own AI features end to end, from rapid prototype to production and the evaluation that keeps them honest. You'll build the APIs, tool-using agents, and RAG pipelines that turn frontier LLMs into grant discovery, application drafting, and research tools our 5,500+ nonprofits rely on every day. It's a high-ownership seat on a small team, where what you ship reaches customers fast and you help shape how we build AI here.

What you'll do

    Ship AI to production

  • Build tool-using LLM agents (task planning, function and tool calling, multi-step workflows, guardrails) for grant discovery, application drafting, and research assistance.
  • Turn prototypes into resilient, observable services with clear SLAs, rollback and fallback strategies, and cost and latency budgets.
  • Stand up evaluation and observability so our AI stays grounded, safe, and cost-effective.
  • Build trustworthy backends

  • Write high-quality, thoroughly tested code across the backend and the data pipelines that power retrieval and evaluation.
  • Contribute to reliability practices: alerts, dashboards, and incident response.
  • Collaborate and raise the bar

  • Partner with Product, Design, and GTM on scoping, UX, and measurement.
  • Run experiments (A/B, canaries), interpret results, and iterate.
  • Raise engineering standards through clear, maintainable code, tests, docs, and thoughtful review.

What we're looking for

    Required

  • 7+ years building and shipping production backend systems in Python (FastAPI, Celery, or equivalent), taking features from prototype to production with real reliability practices like tests, observability, and rollback.
  • Hands-on experience building LLM features in production: tool and function calling, multi-step agent workflows, and the guardrails and evals that keep them grounded, safe, and cost-effective. This is the core of the role.
  • Strong data fundamentals: SQL, schema design, and building pipelines that power retrieval and evaluation.
  • Thrives in a fast, scrappy startup environment with high ownership and a bias for action, speed, quality, and simplicity.
  • Nice to have

  • TypeScript and Node, plus familiarity with Ruby on Rails (our core platform) or a willingness to learn it.
  • Experience with AWS or GCP, Docker, CI/CD, and observability (logs, metrics, traces).
  • RAG depth: document ingestion, chunking and windowing, embeddings, hybrid search (keyword plus vector), re-ranking, and grounded citations.
  • Experience with re-rankers and cross-encoders, hybrid retrieval tuning, or search and recommendation systems.
  • Evaluation mindset: designing eval suites (RAG/QA, extraction, summarization) using automated and human-in-the-loop methods, with familiarity with frameworks like Ragas, DeepEval, or OpenAI Evals.
  • Orchestration frameworks: LangChain or LangGraph, LlamaIndex, Semantic Kernel, or custom orchestration.

Compensation & Benefits

    For US-based candidates, the target salary range for this role is 175,000 - $220,000 USD, plus equity. Final compensation is determined based on experience, skillset, scope of responsibility, interview performance, and geographic location. We’re committed to paying competitively and equitably.

    For candidates based in Canada, compensation varies by province and will be shared by your recruiter early in the process.

    Benefits

  • 100% covered health, dental, and vision insurance for employees (50% for dependents)
  • Generous PTO, including parental leave
  • 401(k)
  • Company laptop and home-office stipend
  • Bi-annual company retreats
  • Instrumentl is evolving rapidly. You’ll always have new challenges and opportunities to grow here.

Instrumentl is an equal opportunity employer. We are committed to building an inclusive workplace and do not discriminate based on race, color, religion, sex, national origin, age, disability, veteran status, sexual orientation, gender identity or expression, genetic information, or any other legally protected status. We encourage candidates from all backgrounds to apply. If you need a reasonable accommodation during the application or interview process, please let us know.

Similar Jobs

Yesterday
Easy Apply
Remote
Canada
Easy Apply
Senior level
Senior level
Cloud • Security • Software • Cybersecurity • Automation
Design and build tooling, workflows, and guidance to make GitLab easier to deploy, validate, and operate across cloud-native, self-managed, and ephemeral environments. Shape deployment architecture, topology, scaling principles, and platform contracts while collaborating with application, infrastructure, and delivery teams.
Top Skills: Ci/CdCloud NativeGitlabKubernetes
13 Days Ago
Easy Apply
Remote
Canada
Easy Apply
Senior level
Senior level
Cloud • Security • Software • Cybersecurity • Automation
Build and maintain backend AI capabilities for GitLab Duo Chat: design flow components and agentic flows in Python/LangGraph, integrate LLMs and multi-agent orchestration, implement and maintain APIs in the Rails monolith, improve observability and reliability, and participate in on-call rotations and cross-functional collaboration to deliver safe, scalable AI chat features.
Top Skills: Agent FrameworksGraphQLLangchainLanggraphLarge Language ModelsMulti-Agent OrchestrationPostgresPytestPythonRestRspecRuby On RailsSQLTool/Function Calling
2 Days Ago
Easy Apply
Remote
Canada
Easy Apply
Senior level
Senior level
Cloud • Security • Software • Cybersecurity • Automation
Design and build a unified analytics instrumentation platform using Go, Kubernetes, ClickHouse, and NATS. Improve data emission, transit, and quality for usage and billing, provide incident response and root cause analysis, collaborate with product and data teams to drive adoption, make tradeoffs for reliability and performance, and mentor engineers through reviews and documentation.
Top Skills: ClickhouseGoGoJavaScriptKubernetesNatsPython

What you need to know about the Toronto Tech Scene

Although home to some of the biggest names in tech, including Google, Microsoft and Amazon, Toronto has established itself as one of the largest startup ecosystems in the world. And with over 2,000 startups — more than 30 percent of the country's total startups — Toronto continues to attract new businesses. Be it helping entrepreneurs manage their finances, simplifying business operations by automating payroll or assisting pharmaceutical companies in launching new drugs, the city's tech scene is just getting started.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account