Frontier Supply Co.

Data and environments for AI training

Training data and RL environments, built to spec.

Frontier Supply builds reinforcement learning environments, expert-written data and evaluations for AI labs and the data companies that serve them. Every project starts with a small pilot, so you judge the quality before you commit.

01What we supply

Five things, made to your spec.

We work in your format and your tooling. If you need something that isn't listed, ask.

  • FS-01

    RL environments

    Sandboxed copies of real software and real workflows, with tasks, tools and automatic graders. Tested so a model can't pass by gaming the grader. Containers · task sets · graders · reward functions

  • FS-02

    Expert data

    Demonstrations, step-by-step solutions, preference rankings and critiques, written by people who do the work for a living. SFT pairs · preference data · reasoning traces · critiques

  • FS-03

    Rubrics and evaluations

    Grading rubrics, held-out test sets and benchmarks that measure real skill on hard, multi-step work. Rubrics · eval sets · benchmark design · grading

  • FS-04

    Licensed datasets

    Proprietary data sourced from the organizations that own it, cleared for training, with the chain of rights documented. Sourcing · rights clearance · de-identification · delivery

  • FS-05

    Specialist capacity

    Practitioners and project teams for data and labeling companies that need a specific expertise, or more hands, on a live project. Overflow projects · specialist experts · white-label delivery

02Who we work with

Teams that train and test models.

Each project is staffed with practitioners from the field it covers: finance, healthcare, law, software, operations and science. Contributors are vetted on credentials and a paid test task before they touch your work.

AI labs

Post-training, RL and evaluation teams that need environments and data in a domain they can't staff quickly in-house.

Data and labeling companies

Vendors that need specialist practitioners or extra capacity for a customer project, delivered under your name if you prefer.

AI application companies

Teams fine-tuning or evaluating models for one industry, who need data and tests that reflect how the work is really done.

03How we work

Small pilot first. Scale once you're satisfied.

  1. Scope

    A short call to set the domain, format, volume and quality bar. NDA first if you need one.

  2. Pilot

    We deliver a small sample set. Your team reviews it and we adjust before anything scales.

  3. Produce

    Every item is checked by a second expert. Environments are tested for broken states and reward hacking.

  4. Deliver

    In your schema and tooling, versioned, with a provenance log and a quality report.

04Standards

What you can count on.

Labs ask where every piece of data came from. These rules apply to every project.

Original work

Contributors create new material. No scraped or paywalled content, and nothing from anyone's current or former employer.

Clear provenance

Every item can be traced to who made it and when. Licensed data ships with its chain of rights.

No real personal data

We use synthetic people, patients and companies. Licensed data is de-identified before it reaches you.

Confidential by default

NDAs before specs are shared. Your tasks, data and results are never reused or resold.

Yours to keep

Work made for you comes with full IP assignment.

Written by people

Contributors don't use AI tools to write deliverables. Human work is the product.

05Contact

Tell us what you're training for.

A few lines on the domain, the format and the volume you need is enough to start. We reply within one business day.

contact@frontierdatasupplies.com
  • 01Request a sample set in your domain
  • 02Scope a pilot for a new environment or dataset
  • 03Add specialist capacity to a live project