A data sourcing pipeline
Takes a list of targets and returns structured, de-duplicated records. It renders each page, uses a language model to pull fields out of what a browser actually sees, resolves duplicates, and batches the result for whatever runs next. Staged, so a run picks up where it stopped.