SC

Snowflake

Two Oracle architects rebuilt the warehouse for the cloud with storage apart from compute, so each workload ran and billed on its own; sellers landed one workload and migrated the rest, three clouds and live data sharing gave AWS's Redshift nothing to match, and the migrated pipelines made leaving costly.

How Snowflake won

  1. 1 · 2012–16Innovation: the warehouse rebuilt for the cloud, storage apart from computeTwo Oracle database architects wrote a SQL warehouse from scratch to run only as a service: many compute clusters over one copy of the data, nothing to tune.Novel Architecture
  2. 2 · 2015–21Each workload gets its own compute, billed by useSeparate compute clusters per workload, shut off when idle and paid for only when running; Tide left Redshift for compute that scales in seconds.Usage-Based Pricing
  3. 3 · 2017–22Land one workload, then migrate the restCapital One and a European retailer spread from one workload across the business; net revenue retention above 150% for years, 177% in early 2022.Land and Expand
  4. 4 · 2017–20Neutral across clouds, with live data shared between accountsAzure in 2018 and Google Cloud in 2020; accounts share live data without copies, and hundreds of customers queried one COVID-19 data set.Counter-positioning
  5. 5 · 2018–24Switching costsAnalytics, logs and pipelines consolidated on one platform; leaving means rebuilding pipelines, permissions and reports.Switching costs

Versus Amazon Redshift: Redshift put the older shared-nothing warehouse on AWS and sold it by the node-hour; Snowflake gave each workload its own compute billed by use, ran on every major cloud and let accounts share live data SRX-1 SN2 SN1. By the time Redshift separated storage (2019), shared data (2020) and billed by the second (2022), the workloads had moved SRX-2 SRX-4 SRX-7.

Arena: Market Conditions Before Snowflake

Companies' data and analytics teams · 2012 · United States

Changing technical requirementsEnabling technology shiftHigh setup and upkeep costs

In 2012 a company that wanted to analyse its data had two main choices. An analytic warehouse such as Teradata or Netezza ran on a fixed pool of machines sized in advance, needed administrators to plan capacity, tune it and apply upgrades, and handled poorly the web and application data, often in JSON, that companies were now collecting SN11 SN1 SN2. Hadoop stored anything cheaply but was a toolkit: teams wrote code and hired specialists to get answers from it SN13 SN2. Demand for analysis was also becoming less predictable, with large jobs that needed a lot of computing for a few hours and then none. Public clouds had begun to rent computing and storage on demand, but the warehouse software itself still assumed it owned its machines SN2.

How each step happened

Step 1 of 5 · 2012–16

Innovation: the warehouse rebuilt for the cloud, storage apart from compute

Snowflake's founding bet concerned how a warehouse should use the cloud. Warehouses of the day, Amazon Redshift included, used a shared-nothing design: each node kept a slice of the data on its own disks, so compute and storage grew together and every resize reshuffled data SN2 SN13. Snowflake kept one copy of the data in cloud storage and put as many separate compute clusters against it as customers needed, each able to start, grow or stop without touching the data SN2 SN13. Customers got compute sized to each job, which Redshift could not offer without being rebuilt.

Benoit Dageville and Thierry Cruanes, database architects at Oracle, left in 2012 rather than pitch it there, judging that a company with a mature database could not build it SN1 SN8. With most people expecting Hadoop to win, their 2016 paper calls a classic SQL warehouse written from scratch a contrarian and risky choice SN7 SN2.

They ran it only as a service, with no tuning, table statistics or vacuuming for users and one production version the company could fix quickly SN2. Mike Speiser of Sutter Hill, chief executive until June 2014, kept it in stealth for an 18 to 20 month technology head start, the first salesperson, Chris Degnan, recalls SN1 SN22. Under Bob Muglia the service became generally available on Amazon's cloud in June 2015 SN1 SN2. The separation made possible, first, a new way to buy compute.

Rivals Amazon Redshift launched in November 2012 on technology licensed from ParAccel, with nodes holding 2 or 16 TB, so more compute meant more nodes and redistributed data SRX-1 SN2. It separated storage from compute only with RA3 nodes in December 2019 SRX-2. Hadoop, Muglia said, was a toolkit that needed code and specialists SN13.

Novel ArchitectureFounder Domain ExpertiseManaged Service

Step 2 of 5 · 2015–21

Each workload gets its own compute, billed by use

Because compute no longer held the data, each workload could get its own compute cluster, which Snowflake calls a virtual warehouse. Data loading, analysts' reports and a data science job ran on separate machines, so one did not slow another; each shut off when idle, and customers paid only for the storage and compute they used SN2. Compute is now billed by the second SN25.

It was also easier to run. In 2021 the fintech Tide left Redshift because its cluster kept running out of space and one bad query could force a restart and hours of reloading; on Snowflake it scaled compute in seconds, and it says maintenance, not cost, drove the move SRX-5.

By 2020 most customers signed annual capacity commitments, could use more than they had bought, and rolled unused capacity forward when they bought more SN1. Revenue followed use, not a fixed subscription, so a release that made queries cheaper cut near-term revenue SN1. But every new workload was new revenue, and finding workloads became the sales force's job.

Rivals Redshift was priced by the node-hour: $0.85 an hour on demand, or under $1,000 per terabyte a year reserved SRX-1. It added Concurrency Scaling for queued queries in March 2019, and per-second billing only with Redshift Serverless in July 2022 SRX-3 SRX-7.

Usage-Based PricingEase of Use

Step 3 of 5 · 2017–22

Land one workload, then migrate the rest

Sellers landed one workload, then worked to migrate the rest of the customer's SN1. Capital One signed in June 2017 and moved its analytics workloads; use spread to other lines of business, and Capital One was 17% of revenue in the fiscal year to January 2019 SN1. A European retailer started in 2018 with one brand's analytics and kept moving workloads SN1. Degnan says the team used its own product to spot customers with new uses SN22.

Net revenue retention compares what customers pay now with what the same customers paid a year earlier. Snowflake's exceeded 150% at January 2019 and 2020 and reached 177% in the fourth quarter of fiscal 2022 SN1 SRX-8. That year 44% of migrations came from cloud platforms SRX-8.

Selling to the largest companies

Degnan had called an early enterprise focus a waste of time, and Muglia had never run a direct sales force SN21 SN22. Frank Slootman, formerly of ServiceNow, took over in April 2019; he says modernizing old analytics workloads worked as an entry strategy but was not where Snowflake could stay SN1 SN20. By the 2020 listing a sales force segmented by customer size was aiming at large enterprises, and customers included 146 of the Fortune 500 SN1 SN26.

Rivals AWS set up a dedicated sales team in 2018 to answer Snowflake's growth and Redshift's losses, and in February 2020 Slootman said Snowflake had taken business from Redshift SRX-9. Teradata and other installed warehouses held many of the large accounts Snowflake migrated SN1.

Land and ExpandTop-Down Selling

Step 4 of 5 · 2017–20

Neutral across clouds, with live data shared between accounts

This step ran alongside the sales push from 2017 and gave sellers an answer no cloud's own warehouse had. Redshift ran only on AWS, Azure's warehouse only on Azure and BigQuery only on Google SN1 SN2. Owning no cloud, Snowflake could run on all three. Muglia called Amazon a partner Snowflake competed with every day SN12. Snowflake reached Azure in September 2018 and Google Cloud in February 2020, with databases replicated across clouds SN16 SN19. A layer hiding each cloud's specifics let customers' applications move with it SN17.

Data sharing grew from the same design. Any compute could reach any data with the owner's consent, so from June 2017 one account could share live data with another without copying it; the provider paid nothing and the consumer paid for its compute, so every share was also consumption SN7 SN14. In March 2020 hundreds of customers queried Starschema's COVID-19 data from their own accounts, and in May FactSet put 24 data sets on the platform for shared clients SN1. The prospectus describes the network effect: the more customers adopt, the more data they can exchange SN1. Slootman made this access the center of his Data Cloud strategy SN20.

Rivals Redshift stayed on AWS only, and its data sharing arrived in preview in December 2020, between Redshift clusters on RA3 nodes SRX-4. Databricks now shares data with recipients outside its own platform SN6.

Counter-positioningCollaboration & exchange network effects

Step 5 of 5 · 2018–24

Switching costs

Each workload migrated in step 3 brought pipelines, permissions and reports built around it. Capital One consolidated analytics, marketing and petabytes of logs on one platform used across its lines of business SN1, and DoorDash engineers describe the ETL jobs and dataset tuning they maintain in Snowflake SN5. Leaving means moving the data and rebuilding that work. The prospectus named customers' investment in older warehouses as a brake on adopting Snowflake; once the workloads had moved, the same brake worked for Snowflake SN1.

The barrier has narrowed. Degnan says Snowflake gave Databricks room by building tools for data science teams too slowly SN22, and since June 2024 Iceberg tables let customers keep data in their own storage, readable by other engines SN3. What holds customers now is mostly the work built on the data, which is why step 3's one-workload-at-a-time migration mattered.

Rivals Redshift matched the architecture late: separated storage in December 2019, data sharing in December 2020 and per-second serverless billing in July 2022 SRX-2 SRX-4 SRX-7. Instead, by 2020 AWS was paying its own field staff on Snowflake consumption, a first, Slootman said SRX-9.

Switching costsData Gravity

Key dates

  1. 2012-07Two Oracle architects start a cloud-only warehouse
  2. 2012-11Redshift launches on AWS at $0.85 an hour, on ParAccel technology SRX-1
  3. 2014-10Out of stealth under Muglia
  4. 2015-06Generally available as a service on AWS
  5. 2017-06Live data sharing without copies
  6. 2017-06Capital One signs, then spreads the use
  7. 2018AWS forms a sales team to answer Snowflake SRX-9
  8. 2018-09Azure, the second cloud
  9. 2019-01$96.7M revenue Fiscal year to January 2019 SN1
  10. 2019-04Slootman becomes chief executive
  11. 2019-12Redshift separates storage with RA3 nodes SRX-2
  12. 2020-01$264.7M revenue Fiscal year to January 2020, +174% year over year SN1
  13. 2020-02All three major clouds
  14. 2020-03Hundreds of customers query Starschema's COVID-19 data
  15. 2020-07158% net revenue retention At July 31, 2020; 3,117 customers SN1
  16. 2020-09Snowflake goes public
  17. 2020-12Redshift data sharing in preview, RA3 nodes only SRX-4
  18. 2022-01177% net revenue retention Fourth quarter of fiscal 2022 (FY22) SRX-8
  19. 2022-07Redshift Serverless bills by the second SRX-7
  20. 2024-06Iceberg tables keep data in customers' own storage

Sources

Oldest first.

  1. SRX-1 Amazon Web Services Announces Amazon Redshift. Amazon press center · 2012-11-28 Company press release
  2. SN23 Snowflake Raises $26M in Funding to Reinvent the Data Warehouse. Snowflake (GlobeNewswire) · 2014-10-21 Press release
  3. SN2 The Snowflake Elastic Data Warehouse, SIGMOD 2016. Dageville, Cruanes, Zukowski et al. (Snowflake Computing) · 2016-06-26 Technical paper
  4. SN14 Snowflake Unveils Enterprise Data Warehouse Sharing. Solutions Review · 2017-06-22 Launch report
  5. SN11 Interview: Bob Muglia, Microsoft veteran and Snowflake Computing CEO, on databases and a changing Seattle. GeekWire · 2017-12-06 CEO interview
  6. SN12 When an Innovation Ecosystem Makes Partners out of Competitors. Mack Institute, Wharton · 2018-03-15 CEO talk
  7. SN15 Snowflake Announces Availability on Microsoft Azure. Snowflake (PR Newswire) · 2018-07-12 Press release
  8. SN16 Snowflake Announces General Availability on Microsoft Azure. Snowflake (PR Newswire) · 2018-09-26 Press release
  9. SRX-3 AWS Announces General Availability of Concurrency Scaling for Amazon Redshift. Amazon press center · 2019-03-27 Company press release
  10. SN13 Snowflake CEO Bob Muglia talks cloud data warehouse evolution. TechTarget · 2019-04-30 CEO interview
  11. SN18 Snowflake brings its popular data warehouse service to Google Cloud. SiliconANGLE · 2019-06-04 Launch report
  12. SRX-2 Amazon Redshift introduces RA3 nodes with managed storage enabling independent compute and storage scaling. AWS What's New · 2019-12-03 Company product announcement
  13. SN19 Snowflake Announces General Availability on Google Cloud. Snowflake (Business Wire) · 2020-02-18 Press release
  14. SN17 Delivering a Single Data Experience Across Multiple Clouds and Regions. Snowflake blog (Benoit Dageville) · 2020-06-02 Founder essay
  15. SN1 Snowflake 2020 IPO S-1. Snowflake / SEC · 2020-08-24 IPO filing
  16. SN26 Meet Snowflake, one of the buzziest tech IPOs ever. Fortune · 2020-09-15 Contemporaneous report
  17. SRX-9 Snowflake depends on Amazon, competes with Amazon Redshift. CNBC (Jordan Novet) · 2020-09-16 Trade report with executive quotes
  18. SN10 The Mike Speiser Incubation Playbook. Kevin Kwok (kwokchain) · 2020-09-22 Investor-model analysis
  19. SRX-4 Amazon Redshift introduces data sharing (preview). AWS What's New · 2020-12-09 Company product announcement
  20. SRX-6 Snowflake vs. Redshift: just choose Snowflake. Fivetran blog · 2021-01-11 Vendor-partner practitioner comparison
  21. SRX-5 Tide.co: Migrating from Redshift to Snowflake. Snowflake Builders Blog (Medium), Plamen Ivanov · 2021-01-19 Practitioner migration account
  22. SN9 Scaling a business: the road to the largest software IPO (Marcin Zukowski). Inovo (Tomasz Swieboda), Medium · 2021-04-29 Founder interview
  23. SN5 Overcoming rapid growth challenges for datasets in Snowflake. DoorDash · 2021-06-22 Customer engineering account
  24. SN24 Industry Benchmarks and Competing with Integrity. Snowflake blog (Benoit Dageville and Thierry Cruanes) · 2021-11-12 Founder essay
  25. SN20 LEADERS interview with Frank Slootman. LEADERS Magazine · 2022-01 CEO interview
  26. SRX-8 Snowflake Investor Day 2022. Snowflake Inc. · 2022-06 Investor presentation
  27. SRX-7 Amazon Redshift Serverless is now generally available. AWS What's New · 2022-07-12 Company product announcement
  28. SN28 Amping it up with Snowflake CEO Frank Slootman. SiliconANGLE (Dave Vellante) · 2022-07-17 CEO interview
  29. SN27 Data Visionary Bob Muglia on Data, AI, and New Book, The Datapreneurs (transcript). Madrona (Founded and Funded) · 2023-08-08 Podcast transcript
  30. SN3 June 10, 2024: Apache Iceberg tables, General Availability. Snowflake · 2024-06-10 Release notes
  31. SN4 Snowflake FY2025 annual report. Snowflake / SEC · 2025 Annual report
  32. SN7 Snowflake founders reveal cuckoo cloud vision that disrupted big data. Computer Weekly (Stephen Withers) · 2025-03-20 Founder interview
  33. SN8 Why we had to build Snowflake: an interview with co-founder Benoit Dageville. iTWire (David M Williams) · 2025-06-02 Founder interview
  34. R10 Order.co Takes the Pain Out of Procurement with up to 7x Faster Analytics. Snowflake · 2026 Challenger-hosted named customer selection and migration account
  35. SN6 What is OpenSharing? Databricks on AWS. Databricks · 2026-09-02 Rival documentation
  36. SN21 Going from 0 to 100 with Snowflake's Founding CRO Chris Degnan. Stage 2 Capital · Undated; observed 2026-09-25 Sales-leader interview
  37. SN22 How Chris Degnan Built Snowflake's Sales Org From Scratch (transcript). The Logan Bartlett Show (Redpoint) · Undated; observed 2026-09-25 Podcast transcript
  38. SN25 Understanding compute cost. Snowflake Documentation · Undated; observed 2026-09-25 Product documentation