worktrucksolutions logo

AI Data Engineer

worktrucksolutions
Remote
Remote$184k–$215k· about 1 hour ago

Straight from worktrucksolutions’s careers page. Apply on the company site — no recruiter, no middleman.

Explore more remote Data Engineer jobs — salaries, top companies, and the latest openings.See all →

AI Data Engineer

Location: Remote

Staff AI Data Engineer

Location: Remote (U.S.-based) Preference given to CA, TX, and FL
Department: Product
Reports to: Chief Product Officer
Starting Pay Range: $184k - $215k/yr

About Work Truck Solutions

Work Truck Solutions is the operating system for the commercial vehicle industry. While the retail auto market is saturated with software, the $130B+ commercial truck market was operating in the dark—until we turned on the lights.

We are the only platform that connects the entire ecosystem: OEMs, upfitters, dealerships, and fleet buyers. By digitizing complex inventory data (chassis + upfits) and streamlining the supply chain, we dont just help dealers sell trucks; we ensure American businesses get the mission-critical vehicles they need to work. We are a team of innovators, disruptors, and problem-solvers dedicated to one mission: removing the friction from the commercial vehicle industry. We are profitable, growing, and aggressively scaling our technology to remain the undisputed authority in this space.

The Opportunity

We have spent years assembling something no one else in this industry has: a connected view of whats being built, whats sitting on lots, whats moving, and what buyers are looking for and failing to find. Right now, a lot of that value is latent. It lives in our pipelines instead of in our customers hands.

We are seeking a Staff AI Data Engineer to change that. This role exists to turn our data asset into products dealers, upfitters, and OEMs will pay for and rely on—market intelligence, demand and pricing signal, inventory recommendations, automated enrichment, search and matching that actually finds the right truck.

Your primary job is delivering customer-facing results from our data. That said, youll need to be able to build the plumbing when its in your way. The best people for this role dont wait on a ticket queue for a feature pipeline—they architect the ingestion and transformation they need, ship it, and hand it off cleanly. Were looking for someone who is fluent in both directions and clear about which one the moment calls for.

We also want someone genuinely eager about what AI makes possible here. The hardest parts of our data problem—unstructured spec sheets, inconsistent upfit descriptions, entity resolution across feeds that agree on nothing—are exactly the problems where modern AI and LLMs are a step change over what was possible three years ago. You should be excited to reach for those tools, rigorous about proving they worked, and honest when a simpler method wins.

How Youll Spend Your Time

Rough shape of the role, so theres no ambiguity about the emphasis:
 
  • ~60% productizing data. Building models, analyses, and data products that reach customers—forecasting, pricing signal, recommendations, matching, enrichment, market intelligence—and iterating on them based on how they actually get used.
  • ~25% AI-driven capability. Applying LLMs and ML to extract, structure, and enrich the data that makes those products possible, with the evaluation rigor to know its working.
  • ~15% data engineering. Building the ingestion, transformation, and feature pipelines your work depends on, and setting standards others can follow.

Key Responsibilities

Deliver results from our data
 
  • Own customer-facing data products end to end: define the opportunity, build the model or analysis, ship it, measure whether it actually helped, and iterate.
  • Build the intelligence layer of our platform—demand forecasting, pricing and market signal, inventory and configuration recommendations, matching and ranking—on problems where being right has direct commercial consequence for our customers.
  • Partner with product and design on how model output surfaces to a dealer or upfitter, what happens when its wrong, and how much confidence to express.
  • Work directly with customers and the commercial team to understand what decisions theyre actually trying to make, and let that shape what you build.
  • Define and instrument success metrics for everything you ship; be the person who knows whether it worked.

Use AI to unlock the data
 
  • Apply LLMs to the unstructured layer of our business: extraction from spec sheets and vehicle descriptions, classification, taxonomy mapping, enrichment, semantic search and matching.
  • Build the evaluation infrastructure that makes AI output trustworthy—golden datasets, offline and online metrics, monitoring for degradation, and guardrails with sensible fallback behavior.
  • Bring AI into your own workflow aggressively and critically, and raise the practice of the people around you.
  • Make honest calls about where generative approaches beat classical ML or plain deterministic logic, and where they dont.

Build what you need
 
  • Design and build the ingestion, transformation, and feature pipelines your models depend on, rather than waiting for them.
  • Contribute to entity resolution and normalization systems that turn inconsistent supplier data into a trustworthy canonical record.
  • Establish data quality, lineage, and contract standards for the data your products rest on.
  • Partner with data engineering on the platform decisions that outlast any single project, and hand off what you build in a state others can own.

Raise the bar
 
  • Set the standard for analytical and modeling rigor through peer review, and mentor the analysts and engineers around you.
  • Write clearly enough that your findings change decisions and your systems can be maintained by someone else.

Qualifications

Data science and productization
 
  • 8+ years applying data science to real problems, with a track record of models and analyses that shipped to users and changed outcomes—not internal reports that circulated and stalled.
  • Demonstrated ownership of customer-facing data products, where your models output was the product and its quality was visible to people paying for it.
  • Strong statistical foundation: experimental design, regression, uncertainty quantification, and the judgment to know what your assumptions are and what happens when they break.
  • Substantial applied modeling experience across the families that matter here—forecasting, recommendation and ranking, gradient boosting, segmentation, entity resolution and fuzzy matching, anomaly detection.
  • Genuine rigor about evaluation: offline metrics that predict online behavior, correct validation for temporal and grouped data, leakage awareness, and calibration—not just accuracy.
  • Product instinct. You care whether the customers decision got better, not whether the model was interesting.

AI fluency and appetite
 
  • Hands-on production experience applying LLMs to data problems: structured extraction, classification and enrichment, embeddings for similarity and clustering, semantic search and matching over messy real-world text.
  • Experience building evaluation and guardrail systems for probabilistic output; you dont ship a prompt without a way to know when it degrades.
  • Working command of the production tradeoffs—model selection, structured output enforcement, context and token cost, latency, caching, human-in-the-loop review—and able to build a business case that accounts for them.
  • Actively curious about the frontier of these tools and eager to apply them, paired with the discipline to verify rather than assume.
  • Familiarity with agentic patterns and tool use, with a realistic view of where theyre production-ready and where they arent.
  • The judgment to argue against AI when a well-chosen heuristic or a clear dashboard solves the problem more cheaply and more legibly.

Data engineering capability
 
  • Expert SQL and strong Python; you write production code that others maintain comfortably.
  • Able to design and build ingestion and ETL/ELT independently—orchestration and transformation tooling, batch and streaming patterns, and sensible data modeling.
  • Experience in a modern warehouse or lakehouse environment (BigQuery, Snowflake, Databricks, or equivalent), including awareness of cost and performance.
  • Comfortable integrating messy, semi-structured, unreliable third-party sources and building the reconciliation that makes them usable.
  • Version control, code review, CI/CD, and reproducible work. Your output is not a folder of untracked notebooks.

Working style
 
  • Exceptional written communication. At this level, the writing is part of the deliverable.
  • Demonstrated influence without authority across product, engineering, and commercial teams.
  • Comfortable scoping a vague business question into tractable work without being handed the framing.
  • Bias toward shipping and learning, with the discipline to follow through past launch.
  • Bachelors degree in a quantitative field, or equivalent depth demonstrated in practice. Advanced degrees welcome but not required.
  • Currently resides in one of the following states: CA, TX, FL, MN

Nice to Have

  • Experience in automotive, dealership software, logistics, supply chain, fleet, or industrial B2B.
  • Pricing, demand forecasting, or inventory optimization background.
  • Background with catalog, taxonomy, or configuration data at scale.
  • Marketplace experience: liquidity, matching efficiency, supply and demand balance.
  • Experience building a companys first customer-facing data product rather than inheriting a mature one.

Why Join Us?

  • Remote Flexibility: Work from anywhere in the U.S. while staying connected to a collaborative team.
  • Impactful Work: Contribute to products that are reshaping the commercial vehicle industry.
  • Growth Opportunities: Be part of a rapidly growing company with ample opportunities for professional development.
  • Inclusive Culture: Join a team that values diversity, creativity, and innovation.

Ready to Drive Innovation?

If youre tired of building models that never reach a user—and you want to turn a genuinely unique dataset into products an industry runs on—wed love to hear from you. Apply now or reach out to us at jobs@worktrucksolutions.com with your resume.

Similar remote jobs

All Data Engineer jobs →
Tenet3 logo

Tenet3

DevOps Engineer (Hybrid)

Remote
Dayton, OH
✓ From careers page· 9 minutes ago
skylo logo

skylo

Senior Engineer, Cloud Infrastructure and Networking (Remote)

Remote
$125k–$135k
✓ From careers page· 10 minutes ago
Cloudera logo

Cloudera

Sovereign Cloud Solution Architect (Remote)

Remote
✓ From careers page· about 1 hour ago
Grantek Systems Integration logo

Grantek Systems Integration

Technical Solutions Consultant (Remote)

Remote
Burlington, ON$95k–$120k
✓ From careers page· about 1 hour ago

Discover More than 100,000 Hidden Remote Jobs Before Everyone Else

Unlock All Remote Jobs Today

Simple pricing. Big savings on Quarterly and Yearly.

Monthly Access

$19/month
  • Instant access to fresh remote jobs from 500+ companies
  • New opportunities added hourly, often 3-7 days before anywhere else
  • Advanced filtering by role type, stack, pay, and location
  • Priority customer support
Start 7-day trial — $2.95
Most Popular

Yearly Access

$59/year
  • Everything in Monthly
  • Save $169 (~74%) vs paying monthly
  • Average job search takes ~6 months - get covered for the whole journey
  • Less than the cost of one lunch per month for competitive advantage
  • Equivalent to just ~$4.92/month
Start 7-day trial — $2.95