{bc}
linkedin

Senior AI Engineer

DV Trading LLC
Chicago, USA
Full-time
Mid-Senior
Onsite
USD 200000-300000 yearly
Discovered 1 weeks ago
PythonLLM fine-tuningModel distillationOpen-weight modelsLLM serving and inferencevLLM or TGI
Free

Job Fit Check

Base Career helps you apply smarter for this job.

?%
Ready to Scan

Key skills for this role

PythonLLM fine-tuningModel distillation
Smart Apply

Full Job Posting

About the Company

DV Trading is an independent proprietary trading firm using its own capital, trading strategies, and risk management methodologies to provide liquidity in global financial markets.

The DV Group has more than 600 employees across North America, Europe, and Asia.

Role Overview

DV Trading is building a centralized AI function focused on developing and operating its own model capability.

The role covers open-model fine-tuning and distillation, on-premises inference infrastructure, and a model gateway routing across open and closed providers.

Responsibilities

  • Build and operate a model gateway with cost, latency, and quality tracking.
  • Design distillation pipelines that generate task-specific training data from frontier model outputs.
  • Fine-tune and evaluate open-weight models including Llama, Qwen, and Mistral.
  • Deploy and maintain on-premises inference infrastructure using vLLM, TGI, Triton, or equivalent technologies on Kubernetes.
  • Build evaluation frameworks for quality, cost, latency, and regression.
  • Define tooling and criteria for open-model versus closed-API selection.
  • Partner with agent engineering to support agent workloads.

Required Qualifications

  • Five or more years of software engineering experience and strong Python skills required.
  • Production fine-tuning or distillation of open-weight models required.
  • Experience serving LLMs on premises required.
  • Production GPU infrastructure management experience, including provisioning, scheduling, and utilization monitoring, required.
  • Production model evaluation and regression testing experience required.
  • Kubernetes and GPU workload management experience required.
  • Understanding of tradeoffs between open and closed models across cost, quality, latency, and data sensitivity required.

Preferred Qualifications

  • Quantization, PEFT/LoRA, or other efficient training techniques.
  • Model gateway or inference proxy design involving routing, fallback, or rate limiting.
  • Financial services or other regulated or sensitive-data experience.
  • Familiarity with Hugging Face, model cards, and open-model licensing.

Compensation

  • The expected base salary range is $200,000 to $300,000 USD yearly.
  • The role is also eligible for a discretionary bonus and the company's benefits package.

Benefits

  • Benefits include medical, dental, and vision insurance; HSA, FSA, and dependent care options; employer-paid group life and AD&D insurance; voluntary LTD, life, and AD&D insurance; flexible vacation; and a retirement plan with employer match.

Apply for this job in 1 click

Skip the repetitive application forms

Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.

Sarah M.James T.Maya R.

Trusted by over 500,000 job seekers on Base Career

Start Free Today