3 ms·
How to create PostgreSQL test data at scale
- ttymck 6y agoAnyone able to weigh in on what "at scale" means here? Just "N number of records" where N is larger than is tenable by manual curation? What about generating realistic amounts of data for performance testing (millions of records), and pushing that to shared database servers, or making it available for developers to populate their local databases? Are there best practices for this?
- openquery 6y agoThanks for the question. > Anyone able to weigh in on what "at scale" means here? Just "N number of records" where N is larger than is tenable by manual curation? Scale here refers to the complexity of the schema. (I admit perhaps the wording is misleading). If your schema is sufficiently complex, generating coherent data can be a massive PITA - Synth alleviates that pain :). > What about generating realistic amounts of data for performance testing (millions of records), and pushing that to shared database servers, or making it available for developers to populate their local databases? Are there best practices for this? Good question. As Synth currently stands, theoretically it can do millions of rows but it would take a while as it's making calls to a Python runtime to piggy-back off the Faker lib. I think I'll make a follow up post about this as it is quite an interesting question.