Skip to content
Livex402Every live per-call product is also available to AI agentsPay list in USDCBaseSolanaNo accountNo API keyStructured JSON inlinePricing starts at $0.02Livex402Every live per-call product is also available to AI agentsPay list in USDCBaseSolanaNo accountNo API keyStructured JSON inlinePricing starts at $0.02
Data-on-Demand · for AI fine-tuning

The dataset is the other half of fine-tuning.

We build a clean, answer-verified dataset and deliver it HuggingFace-ready. Drop the repo straight into Gradients, TRL, Axolotl, or Unsloth. Every checkable answer is verified in code, not trusted from a model.

Order a dataset
See a live sample
row 0428 · verified-math-reasoning-3k

train split

instruction
An item costs $161. It is on sale with a 25% discount. What is the final price in dollars?
output
25% of $161 = $40.25 · $161 − $40.25 = $120.75 · The answer is 120.75

ground truth 120.75computed in Python, not the model

answer matches ground truth, row kept

How it works

Python owns the truth. The model only writes the prose.

Most synthetic datasets trust the language model to be right. This one does not. The correct answer is computed independently before the model writes a word, then the model’s answer is checked against it.

  1. step 01

    Construct

    Each problem is built in code with a known-correct answer as ground truth.

  2. step 02

    Generate

    The model writes step-by-step reasoning and a final answer for the problem.

  3. step 03

    Verify

    The model's answer is checked against ground truth in code. Mismatches are discarded.

  4. step 04

    Deliver

    Deduplicated, split train/val/test, documented, pushed to a HuggingFace repo.

100%

of checkable answers verified against code-computed ground truth

Alpaca

instruction / input / output: Gradients-ready, drop-in for TRL · Axolotl · Unsloth

24h

standard turnaround on a verifiable build · Apache-2.0, commercial use
Agents

Building this into an autonomous agent?

Agents discover and buy this dataset programmatically over MCP. Describe the task, get a quote, pay per-call in USDC via x402. No card, no form.

Agents & developer docs

Order

Tell us the task. The price follows.

Four published tiers. Change the task, domain, or rows and the live price updates. Specialized work is quoted in 24 hours.

  1. 01

    Choose the job

    Task, domain, rows, format.

  2. 02

    Price locks

    List price, or a 24-hour quote.

  3. 03

    We deliver

    Answer-verified HuggingFace repo.

Task type

Domain

Row target

Output format

Describe your task

Email

Buy with card
Agents: pay via x402

Questions before you order? Speak to an expert

Four published builds, priced as a job. Change the order form and the live price follows the tier, unless the work is specialized, in which case it is quoted.

Tier SInstant-buy

Narrow-task LoRA
for first-time fine-tuners

$75 /build

Select plan
  • 1,000: 2,000 answer-verified rows
  • Train / val / test split
  • Dataset card + Apache-2.0 license
  • HuggingFace repo, Gradients-ready
  • Alpaca / ShareGPT / OpenAI messages
  • Scope review in 24 hours
Tier MMost fine-tunes

Domain LoRA, staging-ready
for startups and ML engineers

$150 /build

Select plan
  • 2,000: 5,000 answer-verified rows
  • Everything in S
  • Larger, richer coverage
  • Schema tuned to your trainer
  • Ready to drop into TRL, Axolotl, Unsloth
  • Scope review in 24 hours
Tier LPriority queue

Production domain model
for teams that need to trust it

$300 /build

Select plan
  • 5,000: 10,000 answer-verified rows
  • Everything in M
  • Priority build queue (2-3 days)
  • Custom schema and task review
  • Adversarial + edge-case test set
  • Scope review in 24 hours
Tier XLLabs

Deep domain specialist
for labs that need depth

$600 /build

Select plan
  • 15,000: 20,000 answer-verified rows
  • Everything in L
  • Multi-task schema + task routing
  • Instruction diversity at scale
  • Async delivery (5-7 days)
  • Scope review in 24 hours
CustomQuoted in 24h

25,000+ rows, or a shape that does not fit a tier.
Multi-language, red-team, multi-modal, large corpora.

Price on request /build

Select plan
Scope and guarantee

Every order is reviewed. If we cannot build it, you do not pay.

24-hour scope review

After payment, every order gets a scope review within 24 hours. If the request needs domain expertise, sources, or verification depth beyond the quoted price, we propose an adjusted scope first.

Full-refund fallback

If no acceptable adjustment exists, you get a full refund within 2 business days. Revisions for errors in the delivered dataset are always included. Changes to the task definition itself are a new order.

What we do not build

HIPAA-regulated patient data, classified or restricted-source material, real-time data streams, or expert credentials we do not hold. These are flagged at scope review and automatically refunded.

Next

A data foundry. Not a catalog. Not sales software.

Live

HSH Intelligence

Any data you can describe, built to order, and callable by a human or an agent.

HSH Intelligence