SonicJobs Logo
Left arrow iconBack to search

Lead AI ML Engineer

UFS LLC
Posted 2 months ago, valid for 12 days
Salary

Competitive

Contract type

Full Time

By applying, a Sonicjobs account will be created for you. Sonicjobs's Privacy Policy and Terms & Conditions will apply.

SonicJobs' Terms & Conditions and Privacy Policy also apply.

Sonic Summary

info
  • The Lead AI/ML Engineer at Navanta is responsible for the core functionalities of the Navanta AI platform, including retrieval, model serving, and evaluation harnesses.
  • This role requires 6–10+ years of software development experience, with at least 2–3 years of experience in shipping production LLM, RAG, or NLP systems.
  • Candidates should possess strong Python skills and a solid understanding of software engineering principles, along with a focus on accuracy and evaluation.
  • The position offers a salary range of $130,000 to $180,000, depending on experience and qualifications.
  • Successful candidates will collaborate closely with data and platform engineering teams to ensure a governed and auditable AI system.

The Lead AI/ML Engineer owns the brain of the Navanta AI platform — retrieval, text-to-metrics, model serving, tool orchestration, and the evaluation harness that keeps answers honest. Working under the SVP of Technology and Commercial AI and in close collaboration with the data, platform, and product teams, this role makes ā€œcorrect and verifiableā€ the product’s default — the foundation of trust in a regulated banking environment where a confident wrong number loses the account.

Key Responsibilities

Ā·Ā  Ā Ā Build Navanta’s retrieval and verifications over data systems, with shown queries and citations for every answer

Ā·Ā  Ā Ā Stand up self-hosted open-weight models serving and embeddings inside each bank’s environment or shared environments for Navanta; evolve RAG to a dedicated standard

Ā·Ā  Ā Ā Design the MCP tool layer that exposes a small, audited set of read-only tools (metrics, documents, customer 360), eventually growing into read/write tools with heavy amounts of regulated, highly sensitive data

Ā·Ā  Ā Ā Build and maintain the evaluation harness — golden-question regression, groundedness and retrieval metrics, explicit ā€œI don’t knowā€ behavior — and make it a release gate

Ā·Ā  Ā Ā Implement LLM guardrails: PII redaction in prompts and context, prompt-injection defenses, and cost and row limits aligned to regulatory security expectations

Ā·Ā  Ā Ā Partner with data teams so the model selects governed metrics from the semantic layer rather than improvising SQL

Ā·Ā  Ā Ā Document model architecture, evaluation methodology, and guardrail controls to support customer security reviews and audit readiness

Ā·Ā  Ā Ā Track latency, cost, and quality trade-offs across model versions and deployment configurations

Core Competencies

Ā·Ā  Ā Ā Accuracy and evaluation orientation — a demonstrated focus on verifiability and groundedness, not just compelling demos

Ā·Ā  Ā Ā Production LLM/RAG engineering: retrieval pipelines, tool orchestration, prompt engineering, and guardrail implementation

Ā·Ā  Ā Ā Security and compliance mindset: PII handling, prompt-injection defense, and least-privilege tool access aligned to NIST CSF 2.0 principles

Ā·Ā  Ā Ā Cross-functional collaboration with data and platform engineering to deliver a governed, auditable AI system

Key Performance Indicators (KPIs)

Ā·Ā  Ā Ā Golden-question accuracy — maintained or improved release over release against the verified question set

Ā·Ā  Ā Ā Groundedness rate: percentage of assistant answers fully supported by retrieved context

Ā·Ā  Ā Ā PII redaction coverage and zero prompt-injection incidents in production

Ā·Ā  Ā Ā Model serving latency and cost per query within defined targets

Ā·Ā  Ā Ā Evaluation harness adoption as a release gate — zero releases without passing the regression suite

Qualifications

To perform this job successfully, an individual must be able to perform each essential duty satisfactorily. The requirements listed below are representative of the knowledge, skill, and/or ability required.

Ā·Ā  Ā Ā 6–10+ years building software, with 2–3+ years shipping production LLM, RAG, or NLP systems used by real people — not prototypes

Ā·Ā  Ā Ā A demonstrated focus on accuracy and evaluation, not just demos

Ā·Ā  Ā Ā Strong Python and solid software-engineering fundamentals

Ā·Ā  Ā Ā Comfort operating self-hosted open-weight models and reasoning about latency, cost, and quality trade-offs

Core Technologies

Ā·Ā  Ā Ā Languages: Python

Ā·Ā  Ā Ā Serving & inference: vLLM, Ollama; GPU / CUDA familiarity, NVIDIA Enterprise (NVAIE)

Ā·Ā  Ā Ā RAG & retrieval: LlamaIndex or Haystack; Qdrant, pgvector; embeddings

Ā·Ā  Ā Ā Orchestration: MCP, tool / function calling

Ā·Ā  Ā Ā Structured querying: text-to-SQL; semantic layers (Cube / dbt MetricFlow)

Ā·Ā  Ā Ā Evaluation & guardrails: groundedness and eval frameworks, PII redaction, prompt-injection defense

Nice to Have

Ā·Ā  Ā Ā Experience in regulated or high-stakes domains where a wrong answer is costly

Ā·Ā  Ā Ā Fine-tuning, adapters, and retrieval-quality optimization

Ā·Ā  Ā Ā Familiarity with banking and finance terminology

Education and/or Experience

Ā·Ā  Ā Ā Bachelor’s degree in computer science, mathematics, or a related technical field, or equivalent hands-on experience

Ā·Ā  Ā Ā Experience in the financial services industry or a regulated, high-accuracy AI application environment strongly preferred

Work Structure & Expectations

Ā·Ā  Ā Ā Full-time role combining ongoing model operations and evaluation with initiative-based build-out of the data retrieval, guardrail, and serving infrastructure

Ā·Ā  Ā Ā Close collaboration with data engineering, platform engineering, and product teams; on-call rotation covering reliability in production

Physical Demands

The physical demands described here are representative of those that must be met by an employee to successfully perform the essential functions of this job. Reasonable accommodations may be made to enable individuals with disabilities to perform the essential functions.

While performing the duties of this job, the employee is regularly required to sit and use hands to finger, handle, or touch objects, tools, or controls. The employee frequently is required to talk or hear. The employee is occasionally required to stand; walk; and stoop, kneel, crouch, or crawl. The employee must occasionally lift and/or move up to 10 pounds, usually waist high, up to 50 feet away. Specific vision abilities required by this job include close vision and the ability to adjust focus.

Work Environment

The work environment characteristics described here are representative of those an employee encounters while performing the essential functions of this job. Reasonable accommodations may be made to enable individuals with disabilities to perform the essential functions.

•  Ā Ā Typical office environment

•  Ā Ā Up to 20% travel time may be required

Who is Navanta?

Navanta is the trusted technology and services partner for community financial institutions, unifying critical systems, security, cloud infrastructure, and support into one seamless, purpose built experience. With more than 35 years of banking expertise — from Managed IT to Core Banking, CRM, and Advisory Services — Navanta helps institutions simplify complexity, reduce risk, and strengthen daily operations. Navanta empowers community bankers and their people to thrive together. Go Bankers, Go.ā„¢





Learn more about this Employer on their Career Site

Apply now in a few quick clicks

By applying, a Sonicjobs account will be created for you. Sonicjobs's Privacy Policy and Terms & Conditions will apply.

SonicJobs' Terms & Conditions and Privacy Policy also apply.