BriefComposer SFT Multi image analytical brief rows composed from completed FireWatch, OceanScout, LandShift, and FloodPulse dataset folders ( metadata/ + images/ ). Each sample stitches 1–4 images and metadata derived headlines into one executive style assistant reply. Record counts (this build) Split JSONL lines train 6307 validation 851 test 842 total 8000 Inputs Source roots: one or more source root directories (each must contain images/ and metadata/ from the profile builders). CLI: build lfm vl brief sft.py samples N controls JSONL line count. Run after the four temporal profile datasets are built so metadata and PNGs exist. Dataset layout data/train.jsonl , data/validation.jsonl , data/test.jsonl — VLM SFT samples ( messages list with system/user/assistant content; images referenced by relative paths under this folder). images/ — PNG chips (typically mixed profile images per row) consumed by the JSONL. metadata/ — one JSON sidecar per tile/sample (scene ids, bbox, optional regions , profile tag). Splits Train / validation / test use a stable hash of the synthetic sample id ( brief ) assigned at compose time. Regenerating locally Run from the nutonic repository root (paths re…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy