Skip to main content
TryApplyNow logo
Innodata Inc. logo

Applied Data Scientist, Finance AI Evaluation & Datasets

Innodata Inc.
Be an Early ApplicantFull TimejuniorRemote
Remote - CanadaRemotePosted Today

Role Overview

Innodata Inc. is hiring a entry-level Applied Data Scientist, Finance AI Evaluation & Datasets. This is a full-time remote role, with the team based in Remote - Canada. Part of Innodata Inc.'s Risk hiring, posted today. applications are still in the early window, before most candidates have applied. Full responsibilities, required qualifications, and the apply link are listed in the description below.

Salary Context

Salary is not disclosed in this posting. Market median for Junior-level Risk roles is $87k-$115k (based on 117 comparable listings). Many employers share specifics during the interview process or after an initial screen.

Resume Keywords to Include

Make sure these keywords appear in your resume to improve ATS scoring

PythonSQLPandasPyTorchscikit-learnGAAPIFRSAuditing

Job description

Innodata (Nasdaq: INOD) is a global data engineering company. We believe that data and Artificial Intelligence (AI) are inextricably linked. Our mission is to enable the responsible advancement of artificial intelligence by providing the data, evaluation frameworks, and human expertise required to build AI systems that can be trusted at scale. We provide a range of transferable solutions, platforms, and services for Generative AI / AI builders and adopters. In every relationship, we honor our 36+ year legacy delivering the highest quality data and outstanding outcomes for our customers.

 

Scope of the Role: 

Financial services is one of the highest-stakes domains for generative AI. Numerical accuracy, regulatory compliance, model risk management, auditability, and customer harm prevention, among other concerns, are the bar for shipping anything real. Innodata partners with foundation model labs, banks, asset managers, fintechs, and other enterprise AI teams building LLMs, multimodal systems, and AI agents for financial workflows. 

As an Applied Data Scientist, Financial AI Evaluation & Datasets, you own the design, measurement quality, and domain validity of the datasets used to train, fine-tune, evaluate, and monitor financial-domain LLMs, vision-language models, multimodal document models, and AI agents. You bring financial-domain fluency and data science rigor: you can read a risk policy, financial statement, or customer transcript, among other financial-services documents; turn it into a measurable dataset and evaluation specification; define what correct, grounded, compliant, and safe mean for the use case; and produce evidence that sophisticated financial-services customers, model-risk teams, and AI governance stakeholders can trust. 

This role has a special emphasis on unstructured and multimodal financial data — PDFs, scanned documents, spreadsheets, charts, call transcripts, and other mixed-document workflows where text, numbers, visuals, and metadata all matter. You will work in a pod with a Technical Solutions Architect (scopes the engagement), an Applied Research Scientist (shapes evaluation methodology), an AI/ML Research Engineer (builds training and evaluation infrastructure), and Language Data Scientists (run annotation at scale), making sure what the team produces is domain-valid, statistically defensible, compliant, auditable, and useful for evaluation and post-training. 

What You’ll Own:

  • Translate customer goals — such as improving financial reasoning, building an eval suite for earnings-call summarization, or evaluating an AML/fraud copilot — into concrete dataset specifications, taxonomies, rubrics, and acceptance criteria.
  • Design training and evaluation datasets across the financial AI surface: financial QA, filings and earnings analysis, credit and underwriting, fraud/AML investigation, and compliance, among other financial workflows.
  • Foreground unstructured and multimodal financial data in dataset design — PDFs, scanned statements, tables, charts, and call transcripts — used by analysts, advisors, compliance reviewers, and operations teams.
  • Design datasets and evaluations for retrieval-augmented and source-grounded systems: evidence citation and faithfulness to source documents, data freshness, conflict resolution across sources, and failure modes caused by incomplete or incorrectly parsed context.
  • Evaluate agentic and workflow-integrated financial AI systems: tool use, retrieval, transaction boundaries, escalation behavior, and controls that prevent unsafe or unauthorized actions.
  • Develop evaluation methodology that goes beyond surface accuracy — numerical consistency, hallucination rates on high-risk claims, refusal and escalation appropriateness, robustness under ambiguity, and fairness across protected or sensitive customer segments.
  • Define sampling strategies, label schemas, and adjudication workflows with Language Data Scientists and finance SMEs; write annotation guidelines that make subjective finance-domain judgments explicit, calibratable, and auditable.
  • Build the statistical and ML tooling that makes large financial datasets trustworthy: stratified sampling across products, markets, and modalities; bias analysis; leakage detection; and distribution shift checks, among other reliability checks.
  • Build evaluation and dataset-quality evidence to support financial-services model risk management: assumptions, limitations, validation results, and residual risks, packaged as reproducible evidence.
  • Partner with the AI/ML Research Engineer to instrument datasets into training, evaluation, and monitoring pipelines — rubric-grounded LLM-as-judge prompts, regression suites, and continuous monitoring.
  • Own data quality end-to-end, from intake through delivery: PII handling, provenance tracking, versioning, and modality-specific QA checks.
  • Reason about financial workflow context: where AI outputs enter analyst, advisor, compliance, risk, or customer-facing workflows; what evidence a reviewer needs to trust them; and when uncertainty must be surfaced.
  • Support the Technical Solutions Architect during customer discovery and proposals: scoping dataset programs, sizing annotation effort, and explaining methodology to client stakeholders.
  • Stay current on the financial AI landscape: regulatory developments, benchmark releases, and emerging evaluation methodology for finance-domain models.
  • Contribute to Innodata internal IP: reusable taxonomies, evaluation rubrics, golden datasets, and methodology templates. 

You’ll Thrive in This Role If You Have:

  • 5+ years of data science experience, with at least 2+ years in financial services, fintech, banking, or a comparable regulated data environment.
  • Real working knowledge of financial data and workflows: financial statements, SEC filings, transaction data, and other common financial-services document types.
  • Hands-on experience with unstructured and multimodal financial data — some combination of PDFs, scanned documents, spreadsheets, charts, or call transcripts.
  • Familiarity with financial standards or protocols such as XBRL, ISO 20022, or GAAP/IFRS reporting concepts, etc. is strongly preferred.
  • Hands-on experience designing datasets for ML — not just consuming them. You have written annotation guidelines, sized cohorts, set quality thresholds, and shipped data that downstream teams could actually train, evaluate, or monitor on.
  • Familiarity with LLM-based and multimodal financial AI workflows: prompt design, rubric-based evaluation, RAG, LLM-as-judge methods, and the limitations of automated evaluation in high-stakes contexts.
  • Strong Python and SQL; comfort with pandas, scikit-learn, or equivalent; working familiarity with Hugging Face, PyTorch, or model APIs.
  • Statistical literacy: sampling design, inter-annotator agreement metrics (e.g., Cohen's kappa), confidence intervals, and the ability to push back when a number is being over-interpreted.
  • Solid grasp of financial services privacy, compliance, and governance: PII handling, GLBA or equivalent privacy regimes, MNPI sensitivity, and documentation fit for regulated AI programs.
  • Excellent collaboration skills — upstream with a Technical Solutions Architect, sideways with research scientists and engineers, and downstream with SME annotators and quality teams.
  • A bias toward financial workflow realism. You would rather build a smaller dataset that reflects what analysts, advisors, or customers actually see than a larger one that looks impressive on paper but fails in practice.
  • Degree in a relevant field — statistics, data science, economics, finance, or a related quantitative field, or equivalent demonstrated experience. Formal finance credentials aren't required, but CFA, FRM, or MBA backgrounds, etc. are especially encouraged.
  • Experience designing evaluations for LLMs, VLMs, or multimodal models in financial reasoning, filings analysis, or fraud/AML contexts.
  • Experience with document AI, OCR/post-OCR quality, or table and chart extraction for complex financial documents.
  • Familiarity with agentic evaluation, AI observability, experiment tracking, or tools such as Weights & Biases or LangFuse.
  • Familiarity with model risk management frameworks, validation documentation, fairness/bias auditing, or consumer protection analysis.
  • Experience with multilingual or cross-border financial data, or published/open-source work in financial AI or model governance. 

The expected salary range for this position is $210,000 – $245,000 CAD per year, based on experience, skills, and qualifications.

 

 

Please be aware of recruitment scams involving individuals or organizations falsely claiming to represent employers. Innodata will never ask for payment, banking details, or sensitive personal information during the application process. To learn more on how to recognize job scams, please visit the Federal Trade Commission’s guide at https://consumer.ftc.gov/articles/job-scams. 

If you believe you’ve been targeted by a recruitment scam, please report it to Innodata at verifyjoboffer@innodata.com and consider reporting it to the FTC at ReportFraud.ftc.gov.

About Innodata Inc.

Innodata Inc. logo

Innodata Inc.

innodata.com

RiskHires remote

93 other open roles at Innodata Inc. on TryApplyNow.

Frequently Asked Questions

How do I apply for the Applied Data Scientist, Finance AI Evaluation & Datasets position at Innodata Inc.?

Use the Apply button above to submit your application directly to Innodata Inc.. Most applications take less than 5 minutes if your resume and contact details are ready, and you'll be routed to the employer's official application system to finish.

Is the Applied Data Scientist, Finance AI Evaluation & Datasets role at Innodata Inc. remote?

Yes. This is a remote role. The team is based in Remote - Canada, but the position itself does not require relocating to that office.

What does a Applied Data Scientist, Finance AI Evaluation & Datasets at Innodata Inc. earn?

Innodata Inc. has not disclosed a salary range in this posting. Many employers share specifics later in the interview process; you can also ask during a recruiter screen if compensation transparency is important to you.

When was the Applied Data Scientist, Finance AI Evaluation & Datasets role at Innodata Inc. posted?

This role was posted on July 23, 2026 (today). It's still listed as actively hiring; we re-confirm openings against the source system multiple times per day and remove closed roles.

Is the Applied Data Scientist, Finance AI Evaluation & Datasets role at Innodata Inc. entry-level?

Yes. This is an entry-level position. Strong candidates typically have 0-2 years of relevant work experience, internships, or significant project work. Read the full description for any specific qualification requirements Innodata Inc. has listed.

AI-powered job search

Get every job scored to your resume

Upload your resume and get jobs ranked, your resume tailored, and employee contacts found automatically.

Get started free

No credit card to start