Paul and Shipra would like to use Datafaker to generate synthetic Camino data. Their workflow involves using a Parquet file as the input source and exporting the generated synthetic data to a CSV file.
To support this use case, the duckdb.rst documentation should be updated with a step-by-step tutorial that demonstrates the complete workflow from input to output.
Expected outcome: Users should be able to follow the documentation and successfully generate synthetic data from a Parquet file and save the results as a CSV file without requiring additional guidance.
This issue is linked to the User Instructions issue (https://github.com/alan-turing-institute/DataMatryoshka/issues/14).
Reactions are currently unavailable
Paul and Shipra would like to use Datafaker to generate synthetic Camino data. Their workflow involves using a Parquet file as the input source and exporting the generated synthetic data to a CSV file.
To support this use case, the duckdb.rst documentation should be updated with a step-by-step tutorial that demonstrates the complete workflow from input to output.
Expected outcome: Users should be able to follow the documentation and successfully generate synthetic data from a Parquet file and save the results as a CSV file without requiring additional guidance.
This issue is linked to the User Instructions issue (https://github.com/alan-turing-institute/DataMatryoshka/issues/14).