Live data from Hacker News

Data Science Workflow: Overview and Challenges

cacm.acm.org

1–10 of 39 posts

Re: Data Science Workflow: Overview and Challenges

#9
> As an extreme example, my former officemate Cristian Cadar used to archive his experiments by removing the hard drive from his computer after submitting an important paper to ensure that he can re-insert the hard drive months later and reproduce his original results.

Perhaps it would be an idea to use a virtual machine such as VirtualBox instead?

Re: Data Science Workflow: Overview and Challenges

#10
post #9

> As an extreme example, my former officemate Cristian Cadar used to archive his experiments by removing the hard drive from his computer after submitting an important paper to ensure that he can re-insert the hard drive months later and reproduce his original results. Perhaps it would be an idea to use a virtual machine such as VirtualBox instead?

FTA bio link: Cadar's a Stanford CS PhD who runs a software reliability research group, of all things. I doubt he's overlooked VHD backups.

FTA: It's tedious to keep an environment consistent over a long period. You omitted that context.

Post reply on HN