<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Presto on Ryan Orban</title><link>https://ryanorban.com/categories/presto/</link><description>Recent content in Presto on Ryan Orban</description><generator>Hugo</generator><language>en-us</language><managingEditor>me@ryanorban.com (Ryan Orban)</managingEditor><webMaster>me@ryanorban.com (Ryan Orban)</webMaster><copyright>Ryan Orban</copyright><lastBuildDate>Mon, 11 Nov 2013 00:00:00 +0000</lastBuildDate><atom:link href="https://ryanorban.com/categories/presto/index.xml" rel="self" type="application/rss+xml"/><item><title>Presto: Interacting with Petabytes of Data at Facebook</title><link>https://ryanorban.com/notes/presto-facebook-sql-petabytes/</link><pubDate>Mon, 11 Nov 2013 00:00:00 +0000</pubDate><author>me@ryanorban.com (Ryan Orban)</author><guid>https://ryanorban.com/notes/presto-facebook-sql-petabytes/</guid><description>&lt;h3 id="summary" class="scroll-mt-8 group"&gt;
 Summary
 
 &lt;a href="#summary"
 class="no-underline hidden opacity-50 hover:opacity-100 !text-inherit group-hover:inline-block"
 aria-hidden="true" title="Link to this heading" tabindex="-1"&gt;
 &lt;svg
 xmlns="http://www.w3.org/2000/svg"
 width="16"
 height="16"
 fill="none"
 stroke="currentColor"
 stroke-linecap="round"
 stroke-linejoin="round"
 stroke-width="2"
 class="lucide lucide-link w-4 h-4 block"
 viewBox="0 0 24 24"
&gt;
 &lt;path d="M10 13a5 5 0 0 0 7.54.54l3-3a5 5 0 0 0-7.07-7.07l-1.72 1.71" /&gt;
 &lt;path d="M14 11a5 5 0 0 0-7.54-.54l-3 3a5 5 0 0 0 7.07 7.07l1.71-1.71" /&gt;
&lt;/svg&gt;

 &lt;/a&gt;
 
&lt;/h3&gt;
&lt;p&gt;Presto is a distributed SQL query engine built by Facebook to solve a specific problem: Hive queries over HDFS data took hours, which made ad hoc data exploration impractical. Analysts needed answers in seconds or minutes, not hours. Presto was built from scratch as a massively parallel processing (MPP) query engine that could return results from petabyte-scale datasets interactively.&lt;/p&gt;</description></item><item><title>Facebook Unveils Presto for 250 PB Data Warehouse</title><link>https://ryanorban.com/notes/facebook-presto-250pb-warehouse/</link><pubDate>Fri, 07 Jun 2013 00:00:00 +0000</pubDate><author>me@ryanorban.com (Ryan Orban)</author><guid>https://ryanorban.com/notes/facebook-presto-250pb-warehouse/</guid><description>&lt;h3 id="summary" class="scroll-mt-8 group"&gt;
 Summary
 
 &lt;a href="#summary"
 class="no-underline hidden opacity-50 hover:opacity-100 !text-inherit group-hover:inline-block"
 aria-hidden="true" title="Link to this heading" tabindex="-1"&gt;
 &lt;svg
 xmlns="http://www.w3.org/2000/svg"
 width="16"
 height="16"
 fill="none"
 stroke="currentColor"
 stroke-linecap="round"
 stroke-linejoin="round"
 stroke-width="2"
 class="lucide lucide-link w-4 h-4 block"
 viewBox="0 0 24 24"
&gt;
 &lt;path d="M10 13a5 5 0 0 0 7.54.54l3-3a5 5 0 0 0-7.07-7.07l-1.72 1.71" /&gt;
 &lt;path d="M14 11a5 5 0 0 0-7.54-.54l-3 3a5 5 0 0 0 7.07 7.07l1.71-1.71" /&gt;
&lt;/svg&gt;

 &lt;/a&gt;
 
&lt;/h3&gt;
&lt;p&gt;Facebook announced Presto in 2013 — a distributed SQL query engine built to run interactive queries against their 250 petabyte Hadoop-based data warehouse. This was a significant milestone: at the time, Hive was the standard way to query HDFS data, but Hive translated SQL to MapReduce jobs, which required multiple disk write cycles and took minutes to hours for common analytical queries. Presto addressed this directly.&lt;/p&gt;</description></item></channel></rss>