Powering big data at Pinterest

We currently log 20 terabytes of new data each day, and have around 10 petabytes of data in S3. We use Hadoop to process this data, which enables us to put the most relevant and recent content in front of Pinners through features such as Related Pins, Guided Search, and image processing. It also powers thousands of daily metrics and allows us to put every user-facing change through rigorous experimentation and analysis.

In order to build big data applications quickly, we’ve evolved our single cluster Hadoop infrastructure into a ubiquitous self-serving platform.
making Pinterest

Advertisements

Author: kOoLiNuS

♂, Italian, male, husband, dad of a wonder, “cazzaro”, friendly, blogger, motorcyclist, geek, avid reader, sysadmin, ICT consultant, curious. I come in peace… I'm an active social networker since 1999. I've been using WordPress sice 2004 and WordPress.com since 2006, and I'm currently involved in WordPress and WooCommerce communities in Bari, Apulia.

Leave a Reply

Fill in your details below or click an icon to log in:

WordPress.com Logo

You are commenting using your WordPress.com account. Log Out /  Change )

Google+ photo

You are commenting using your Google+ account. Log Out /  Change )

Twitter picture

You are commenting using your Twitter account. Log Out /  Change )

Facebook photo

You are commenting using your Facebook account. Log Out /  Change )

w

Connecting to %s

This site uses Akismet to reduce spam. Learn how your comment data is processed.