Skip to content

Future: synthetic address seeding pipeline (E2E) #9

Description

@BleedingDev

Goal

Build our own address dataset over time by generating/scraping realistic addresses via E2E flows, searching for them, and persisting results.

Who Uses This

  • Product/engineering: build a high-coverage address DB without manual imports.
  • Operators: reduce long-term dependency and cost.

What “Done” Means (Business)

  • We can run a repeatable job that:
    • generates or scrapes address candidates
    • calls our /suggest
    • stores results into our DB for later reuse/analysis
  • We can measure coverage improvements over time.

Constraints

  • Must be safe/legal/ethical (no uncontrolled scraping).

Out of Scope

  • Full-scale data pipeline and ML ranking.

Activity

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

No one assigned

    Labels

    area:cache-dbCaching / address databaseepicEpic (parent issue)priority:p3Someday / backlog

    Type

    No type

    Projects

    No projects

      Milestone

      No milestone

      Relationships

      None yet

      Development

      No branches or pull requests

      Issue actions