Notifications
Clear all
14/08/2026 8:30 am
That's a solid starting workflow, but I have a practical question about the initial dvc add step. When you say "adding your dataset to DVC control," is that meant to be done manually for every new raw dataset version, or is it typically automated as part of a pipeline stage? I'm thinking about a scenario where the raw data itself is regularly updated from an external source. Would you manually run `dvc add data/raw` each time, or would you have a dvc.yaml stage that both fetches the new data and then automatically tracks it? The manual approach seems prone to human error, but I'm not sure how to bake the versioning into an automated fetch without creating a circular dependency.
Page 3 / 3
Prev