Superpose

Methods

Generation and validation

ATLAS definitions are generated through a reproducible framework, inspected at stored-output level, and published with limitations intact.

01

Define

Declare class catalogs, condition ranges, scene composition, labels, and intended use before production scale.

02

Generate

Use shipped configuration directories and the stock rfgen CLI. Record seeds, realized values, and provenance roles.

03

Measure

Test causal axes, physics, scene statistics, separability, robustness, and—where evidence exists—real-capture behavior.

04

Falsify

Pin defects, rerun experiments, preserve negative results, and narrow claims when assumptions fail.

05

Release

Publish the definition, evidence, source pin, corpus status, and explicit non-claims together.

Provenance roles

Provenance field roles

Label Class or scene truth.

Drawn axis A sampled value proven to affect rendering.

Realized read-back A value measured after rendering.

Not applicable An explicit absence—not a plausible but inert number.

Corpus scale / 10 August 2026

Basis for the largest-dataset claim

“Largest” is measured by paired RF-text records in a released multimodal dataset. One record is a bounded RF example with one or more grounded text annotations. Individual complex-valued signal samples are excluded from the count. Signal ATLAS contains 200 million paired records in each released domain, or 400 million across the two current domains.

RF multimodal dataset comparison, snapshot 10 August 2026
DatasetReported scaleNormalization decision
Signal ATLAS400M paired RF-text recordsIncluded. Two released domains at 200M records each.
RF-Lang ↗288,000 I/Q-language examplesIncluded. Direct RF and structured-language pairing.
RF-Behavior ↗44 participants across gesture, activity, and emotion tasksExcluded from normalized count. RF is paired with sensor modalities, not a reported RF-text record corpus.
RVTALL ↗20 participants; RF, visual, text, audio, laser, and landmarksExcluded from normalized count. No released paired-record total comparable to the declared unit.
OPERAnet ↗Approximately 8 hours of annotated measurementsExcluded from normalized count. RF and vision activity data, not RF-language records.
NIST semantic RF proposal ↗Proposed collection; no released record countExcluded. No released corpus at the snapshot date.

The comparison is versioned by snapshot date because corpus sizes and access conditions change. Different task definitions are not treated as evidence of equivalent scientific scope.