Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

What course are you taking? Imagenet is only 150 GB, and Common Crawl is only 320 TB.

Big data is a moving target, but I’m comfortable defining it as data too large to fit in memory. Obviously, you can always get a bigger node, my rule is thumb is that if you need generators, you are working with big data.



Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: