Synthetic data · for data & ML teams

Realistic data from a sentence — or your own CSV.

Describe what you need, or upload real data to learn from. Get a realistic, privacy-safe dataset in minutes.

Four products under one sealed contract and one evidence chain — so every dataset ships with a receipt you (or your auditor) can replay and verify offline. No real records are ever copied.

sub-minutedeterministicquality-gatedsealed evidence
app.radmah.ai / chatlive

What would you like to generate?

Upload fileSelect datasetPrivate
AutoMockSynthesizeOT/SCADAICS Security
Describe the data you need...
History

Teams arrive with one of three problems

A partner integration stuck behind a six-week DPA negotiation.

A training set the fraud team can't get because real data is too sensitive.

A model-validation run that needs realistic data a regulator will accept.

Pick the product that matches the problem, hand us the contract, and walk out with a dataset your CISO signs off on.

◆ FlagshipAutonomous Data Scientist

Plan → run across engines → self-heal → seal.

A 48-module autonomous agent with 43 typed planner tools. It plans before it acts, executes across every engine, reads the evidence bundle after each step, self-heals when a quality gate fails, and chains every decision into a verifiable audit — with a human-approval gate before any credit-spending step.

  • Plans first, executes second — no blind tool calls
  • Cost / credit estimate before each step
  • Same evidence chain as Mock and Synthesize
Open the flagship page
app.radmah.ai / agentlive
Agentic Data Scientist
Autonomous multi-step pipelines with human approval gates
+ New Project
Search by title or goal...
All51Active0Awaiting you1Complete21Failed29
blocked15 June 2026
A6w-ns
[Action panel] Direct execution of mock_data
Progress0/1 steps
Cancel Delete
Open
blocked15 June 2026
A6-ns
[Action panel] Direct execution of mock_data
Progress0/1 steps
Cancel Delete
Open
blocked15 June 2026
A5-ns
[Action panel] Direct execution of mock_data
Progress0/1 steps
Cancel Delete
Open
blocked14 June 2026
A4-noseal
[Action panel] Direct execution of mock_data
Progress0/1 steps
Cancel Delete
Open
complete14 June 2026
A2-seal
[Action panel] Direct execution of mock_data
Progress1/1 steps
Delete
Open
blocked14 June 2026
A2-noseal
[Action panel] Direct execution of mock_data
Progress0/1 steps
Cancel Delete
Open
◆ Evidence pipeline

Six stages from prompt to sealed bundle.

Every dataset from any product here passes the same six stages. The bundle in your bucket can be verified offline by anyone with the verifier CLI.

app.radmah.ai / evidencelive
Concept coverage
PASS

Whether the generated schema actually covers the concepts you asked for. PASS means every extracted concept is represented.

Coverage
100%
Requested
9
Covered
9
Missing
0
Covered concepts
customersinvoicesinvoice line itemsquantityunit priceline total+3 more
Generated data quality
PASS

How readable, deterministic, and policy-clean the synthetic rows are after every post-render repair has run.

Text fallbacks used
0
Placeholder hits
0
Forbidden token hits
Business rule checks
380 passed0 repaired0 failed

Cross-field invariants the engine evaluates against every generated row (arithmetic totals, derived numerics, enum coherence).

Currency value matches declared enum
40/40 pass
Customer Id
20/20 pass
Grand Total
40/40 pass
Invoice Id
40/40 pass
Line Item Id
40/40 pass
Line Total
40/40 pass
Quantity
40/40 pass
01Sealed contract

Schema, constraints, intent and seed sealed into one artefact before any data is generated.

02Engine run

Mock, Synthesize or an ADS-driven pipeline executes against the contract; per-step I/O recorded.

03Quality gates

K-S, Pearson, χ², constraint satisfaction and per-column drift checked. Fail-closed — a regression aborts.

04Cryptographic chain

Each step's inputs and outputs hashed and chained. Tamper-evident, verifiable offline.

05Evidence bundle

A signed .tar.zst with the contract, run-log, quality report, artefact manifest and engine SBOM.

06Tenant isolation

Per-tenant keys, per-tenant artefact prefixes, per-tenant evidence keys — end to end.

◆ Numbers we'll defend

Concrete, and reproducible.

95.69%
tabular fidelity

vs. real data, under an independent QA harness

43
planner tools

the same engines run from chat or the SDK

14
connectors

warehouses, databases, object stores — secrets vaulted

0
plain-text secrets

connector creds auto-vaulted, never written to logs

Benchmark numbers reproduce under the same independent QA harness with matched train / test splits. Full per-release certificates live in the authenticated console.

Bring a CSV. Keep the evidence bundle.

In a 30-minute working session we run a representative dataset through Synthesize and the Autonomous Data Scientist, end to end. You keep the signed evidence bundle and the quality report.