{"id":23972,"date":"2026-09-03T14:15:47","date_gmt":"2026-09-03T21:15:47","guid":{"rendered":"https:\/\/www.pingcap.com\/?post_type=article&#038;p=23972"},"modified":"2026-09-03T14:15:48","modified_gmt":"2026-09-03T21:15:48","slug":"the-evolution-and-impact-of-distributed-databases","status":"publish","type":"article","link":"https:\/\/www.pingcap.com\/ko\/article\/the-evolution-and-impact-of-distributed-databases\/","title":{"rendered":"The Evolution and Impact of Distributed Databases"},"content":{"rendered":"<p class=\"wp-block-paragraph\">Every generation of distributed database was built to fix what the last one gave up. That is the useful way to read the evolution of distributed databases: not as a march toward a better product, but as a sequence of trades. Scale for consistency, then consistency back at the cost of proprietary hardware, then transactions and analytics in one system. Each step made the previous compromise unnecessary, and the current step is being forced by a workload that did not exist five years ago.<\/p>\n\n\n\n<h2 class=\"wp-block-heading\"><span class=\"ez-toc-section\" id=\"Historical_Development_and_Milestones\"><\/span><strong>Historical Development and Milestones<\/strong><span class=\"ez-toc-section-end\"><\/span><\/h2>\n\n\n\n<p class=\"wp-block-paragraph\">The sequence matters more than the individual systems, because each entry is a response to the limitation of the one above it.<\/p>\n\n\n\n<figure class=\"wp-block-table\"><table class=\"has-fixed-layout\"><thead><tr><th><strong>Era<\/strong><\/th><th><strong>What arrived<\/strong><\/th><th><strong>What it traded away<\/strong><\/th><\/tr><\/thead><tbody><tr><td>Late 1970s to 1980s<\/td><td>Research prototypes, SDD-1 and IBM\u2019s System R* among them<\/td><td>Proved distributed transactions were possible, but the hardware of the era could not make them fast<\/td><\/tr><tr><td>2006 to 2007<\/td><td>Google\u2019s Bigtable and Amazon\u2019s Dynamo papers<\/td><td>Bought scale and availability by giving up strong consistency<\/td><\/tr><tr><td>2008 to 2010<\/td><td>The open-source NoSQL wave: Cassandra, HBase, MongoDB<\/td><td>Gave developers horizontal scale, and moved the consistency problem into application code<\/td><\/tr><tr><td>2012<\/td><td>Google\u2019s Spanner paper and TrueTime<\/td><td>Restored strong consistency at global scale, on infrastructure only Google had<\/td><\/tr><tr><td>2015 to 2017<\/td><td>CockroachDB and TiDB<\/td><td>Brought Spanner\u2019s guarantees to commodity hardware and open source<\/td><\/tr><tr><td>2020 onward<\/td><td>HTAP, in TiDB\u2019s case through the TiFlash columnar engine<\/td><td>Removed the separate analytics warehouse and the pipeline feeding it<\/td><\/tr><tr><td>2024 to 2026<\/td><td>Vector search inside the transactional engine<\/td><td>Removes the standalone vector database and its sync gap<\/td><\/tr><\/tbody><\/table><\/figure>\n\n\n\n<p class=\"wp-block-paragraph\">Two corrections to the version of this history that circulates most often. Spanner was not an early system; it arrived in 2012, after the NoSQL wave and partly in answer to it. And NoSQL did not follow <a href=\"https:\/\/www.pingcap.com\/ko\/compare\/\">CockroachDB<\/a>; it preceded both CockroachDB and TiDB by roughly a decade, which is precisely why both were built.<\/p>\n\n\n\n<h2 class=\"wp-block-heading\"><span class=\"ez-toc-section\" id=\"What_Each_Generation_Left_Unsolved\"><\/span><strong>What Each Generation Left Unsolved<\/strong><span class=\"ez-toc-section-end\"><\/span><\/h2>\n\n\n\n<p class=\"wp-block-paragraph\">Three problems recur at every stage, and how a system handles them is what places it on the timeline above.<\/p>\n\n\n\n<ul class=\"wp-block-list\">\n<li><strong>Scale.<\/strong> Spreading a workload across nodes is the easy part. Doing it without the application knowing which node holds what is the part that took thirty years.<\/li>\n\n\n\n<li><strong>Availability.<\/strong> Replication keeps a system serving through node failure, but only if failover happens without an operator and without losing a committed write.<\/li>\n\n\n\n<li><strong>Consistency.<\/strong> This is the one the NoSQL generation traded away and the Raft and Paxos generation bought back. A <a href=\"https:\/\/www.pingcap.com\/ko\/article\/decentralized-cloud-computing-benefits-and-challenges\/\">consensus protocol<\/a> commits a write only once a majority of replicas acknowledge it, so no node serves a version of the data the others have not agreed to.<\/li>\n<\/ul>\n\n\n\n<p class=\"wp-block-paragraph\">The general trade-offs behind these three, and when distribution is worth its operational cost at all, are covered in our guide to <a href=\"https:\/\/www.pingcap.com\/ko\/article\/distributed-database-use-case\/\">distributed database use cases<\/a>.<\/p>\n\n\n\n<h2 class=\"wp-block-heading\"><span class=\"ez-toc-section\" id=\"What_TiDB_Added_HTAP_in_One_Cluster\"><\/span><strong>What TiDB Added: HTAP in One Cluster<\/strong><span class=\"ez-toc-section-end\"><\/span><\/h2>\n\n\n\n<p class=\"wp-block-paragraph\"><a href=\"https:\/\/www.pingcap.com\/ko\/what-is-tidb\/\">\ud2f0DB<\/a> entered the same lineage as CockroachDB in 2015 and made a different bet. Rather than choosing between transactional and analytical processing, it built both into one system: hybrid transactional\/analytical processing, or HTAP.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">The architecture is three parts. The Placement Driver (PD) manages metadata, schedules where replicas live, and balances load. TiKV is the row-based storage engine that serves transactions, replicating each range of data across nodes as its own Raft group. TiFlash keeps a columnar replica of that same data and serves read-heavy analytical queries, so analytics never competes with transactions for the same engine. The <a href=\"https:\/\/www.pingcap.com\/ko\/article\/exploring-distributed-databases-tidbs-architecture-benefits\/\">component-level detail<\/a> is worth reading if you are evaluating the design rather than the history.<\/p>\n\n\n\n<h2 class=\"wp-block-heading\"><span class=\"ez-toc-section\" id=\"Where_Distributed_Databases_Are_Headed_The_AI_and_Agentic_Era\"><\/span><strong>Where Distributed Databases Are Headed: The AI and Agentic Era<\/strong><span class=\"ez-toc-section-end\"><\/span><\/h2>\n\n\n\n<p class=\"wp-block-paragraph\">The next stage is being driven by what <a href=\"https:\/\/www.pingcap.com\/ko\/ai\/agentic-ai\/\">AI workloads<\/a> ask of the database rather than by transaction volume. An agent needs its transactional state, its tool-call history, and its semantic retrieval index to agree with each other at the same instant, not to be synced on separate schedules across separate systems. That requirement did not exist when Spanner or the first NoSQL systems were designed, which is why HTAP is shifting from a differentiator to a baseline expectation.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">TiDB extends the same architecture one step further: <a href=\"https:\/\/www.pingcap.com\/ko\/ai\/vector-search\/\">\ub124\uc774\ud2f0\ube0c \ubca1\ud130 \uac80\uc0c9<\/a> runs inside the same Raft-consistent SQL engine as TiKV and TiFlash, rather than in a separate vector database that has to be kept in sync. Because the vector index sits on the columnar replica and validates its log index against TiKV before serving, a retrieval query reads data consistent with what the agent just committed. That closes the same class of synchronization gap that made earlier distributed systems fragile whenever a new workload type arrived.<\/p>\n\n\n\n<h2 class=\"wp-block-heading\"><span class=\"ez-toc-section\" id=\"TiDB_in_Production\"><\/span><strong>TiDB in Production<\/strong><span class=\"ez-toc-section-end\"><\/span><\/h2>\n\n\n\n<p class=\"wp-block-paragraph\">Three deployments across different pressures:<\/p>\n\n\n\n<ul class=\"wp-block-list\">\n<li><a href=\"https:\/\/www.pingcap.com\/ko\/case-study\/mnc-bank-supercharges-performance-with-tidb\/\"><strong>MNC Bank<\/strong><\/a> uses TiDB\u2019s transactional integrity and real-time analytics for financial processing, and reported 10x higher throughput, 50% lower latency, and 85% faster backups after adopting it.<\/li>\n\n\n\n<li><a href=\"https:\/\/www.pingcap.com\/ko\/case-study\/flipkart-transforming-database-management-and-reducing-complexity-with-tidb\/\"><strong>Flipkart<\/strong><\/a> handles high-volume transactions and customer traffic through peak sales without manual resharding.<\/li>\n\n\n\n<li><a href=\"https:\/\/www.pingcap.com\/ko\/case-study\/plaid-adopts-tidb-to-reduce-maintenance-effort-96-with-zero-downtime-upgrades\/\"><strong>Plaid<\/strong><\/a> modernized its legacy MySQL estate on TiDB, cutting database-maintenance effort by 96% and gaining zero-downtime upgrades.<\/li>\n<\/ul>\n\n\n\n<h2 class=\"wp-block-heading\"><span class=\"ez-toc-section\" id=\"The_Takeaway\"><\/span><strong>The Takeaway<\/strong><span class=\"ez-toc-section-end\"><\/span><\/h2>\n\n\n\n<p class=\"wp-block-paragraph\">The evolution of distributed databases has been a sequence of reclaimed compromises: scale without consistency, then consistency without commodity hardware, then transactions and analytics without two systems, and now retrieval without a separate store to synchronize. Each generation existed because the last one gave something up.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">If you want to see where the current step landed at the architectural level, the <a href=\"https:\/\/docs.pingcap.com\/tidb\/stable\/overview\/\">TiDB architecture documentation<\/a> covers how the components actually fit together.<\/p>\n\n\n\n<h2 class=\"wp-block-heading\"><span class=\"ez-toc-section\" id=\"FAQs\"><\/span><strong>\uc790\uc8fc \ubb3b\ub294 \uc9c8\ubb38<\/strong><span class=\"ez-toc-section-end\"><\/span><\/h2>\n\n\n\n<h3 class=\"wp-block-heading\"><strong>When were distributed databases invented?<\/strong><\/h3>\n\n\n\n<p class=\"wp-block-paragraph\">The first distributed database prototypes are usually credited to research projects of the late 1970s and early 1980s, SDD-1 and IBM\u2019s System R* among them. They demonstrated that transactions could span machines, but the hardware and networks of the period made them too slow for production. The commercially significant wave came much later, starting with Google\u2019s Bigtable paper in 2006.<\/p>\n\n\n\n<h3 class=\"wp-block-heading\"><strong>Is NoSQL a distributed database?<\/strong><\/h3>\n\n\n\n<p class=\"wp-block-paragraph\">Most NoSQL systems are distributed, but the categories are not the same thing. NoSQL describes the data model and the consistency guarantee, typically non-relational and eventually consistent. Distributed describes how the system spreads data across nodes. A distributed SQL database like TiDB is distributed without being NoSQL: it keeps the relational model and strong consistency.<\/p>\n\n\n\n<h3 class=\"wp-block-heading\"><strong>What problem does Spanner solve that NoSQL did not?<\/strong><\/h3>\n\n\n\n<p class=\"wp-block-paragraph\">Strong consistency at global scale. The NoSQL generation achieved scale by accepting eventual consistency, which pushed conflict resolution into application code. Spanner, published in 2012, showed that a globally distributed database could offer strict transactional guarantees, using TrueTime to order transactions across regions. Its practical limitation was that it depended on Google\u2019s own infrastructure.<\/p>\n\n\n\n<h3 class=\"wp-block-heading\"><strong>What comes after HTAP?<\/strong><\/h3>\n\n\n\n<p class=\"wp-block-paragraph\">Retrieval as a first-class part of the same engine. Agentic applications need vector similarity search, transactional state, and memory to reflect the same moment, which means the retrieval index has to live in the same consistency domain as the transactions rather than in a separate vector store synced on its own schedule.<\/p>","protected":false},"excerpt":{"rendered":"<p>Discover how distributed databases enhance scalability, availability, and fault tolerance, with insights into TiDB&#8217;s innovations.<\/p>","protected":false},"author":218,"featured_media":0,"template":"","class_list":["post-23972","article","type-article","status-publish","hentry"],"acf":[],"yoast_head":"<!-- This site is optimized with the Yoast SEO plugin v28.2 - https:\/\/yoast.com\/product\/yoast-seo-wordpress\/ -->\n<title>The Evolution and Impact of Distributed Databases | TiDB<\/title>\n<meta name=\"description\" content=\"The evolution of distributed databases, from 1970s prototypes through Bigtable and Spanner to TiDB\u2019s HTAP capabilities.\" \/>\n<meta name=\"robots\" content=\"index, follow, max-snippet:-1, max-image-preview:large, max-video-preview:-1\" \/>\n<link rel=\"canonical\" href=\"https:\/\/www.pingcap.com\/ko\/article\/the-evolution-and-impact-of-distributed-databases\/\" \/>\n<meta property=\"og:locale\" content=\"ko_KR\" \/>\n<meta property=\"og:type\" content=\"article\" \/>\n<meta property=\"og:title\" content=\"The Evolution and Impact of Distributed Databases | TiDB\" \/>\n<meta property=\"og:description\" content=\"The evolution of distributed databases, from 1970s prototypes through Bigtable and Spanner to TiDB\u2019s HTAP capabilities.\" \/>\n<meta property=\"og:url\" content=\"https:\/\/www.pingcap.com\/ko\/article\/the-evolution-and-impact-of-distributed-databases\/\" \/>\n<meta property=\"og:site_name\" content=\"TiDB\" \/>\n<meta property=\"article:publisher\" content=\"https:\/\/facebook.com\/pingcap2015\" \/>\n<meta property=\"article:modified_time\" content=\"2026-09-03T21:15:48+00:00\" \/>\n<meta property=\"og:image\" content=\"https:\/\/static.pingcap.com\/files\/2024\/09\/11005522\/Homepage-Ad.png\" \/>\n\t<meta property=\"og:image:width\" content=\"1440\" \/>\n\t<meta property=\"og:image:height\" content=\"714\" \/>\n\t<meta property=\"og:image:type\" content=\"image\/png\" \/>\n<meta name=\"twitter:card\" content=\"summary_large_image\" \/>\n<meta name=\"twitter:site\" content=\"@PingCAP\" \/>\n<meta name=\"twitter:label1\" content=\"\uc608\uc0c1 \ub418\ub294 \ud310\ub3c5 \uc2dc\uac04\" \/>\n\t<meta name=\"twitter:data1\" content=\"6\ubd84\" \/>\n<script type=\"application\/ld+json\" class=\"yoast-schema-graph\">{\"@context\":\"https:\\\/\\\/schema.org\",\"@graph\":[{\"@type\":\"WebPage\",\"@id\":\"https:\\\/\\\/www.pingcap.com\\\/article\\\/the-evolution-and-impact-of-distributed-databases\\\/\",\"url\":\"https:\\\/\\\/www.pingcap.com\\\/article\\\/the-evolution-and-impact-of-distributed-databases\\\/\",\"name\":\"The Evolution and Impact of Distributed Databases | TiDB\",\"isPartOf\":{\"@id\":\"https:\\\/\\\/www.pingcap.com\\\/#website\"},\"datePublished\":\"2026-09-03T21:15:47+00:00\",\"dateModified\":\"2026-09-03T21:15:48+00:00\",\"description\":\"The evolution of distributed databases, from 1970s prototypes through Bigtable and Spanner to TiDB\u2019s HTAP capabilities.\",\"breadcrumb\":{\"@id\":\"https:\\\/\\\/www.pingcap.com\\\/article\\\/the-evolution-and-impact-of-distributed-databases\\\/#breadcrumb\"},\"inLanguage\":\"ko-KR\",\"potentialAction\":[{\"@type\":\"ReadAction\",\"target\":[\"https:\\\/\\\/www.pingcap.com\\\/article\\\/the-evolution-and-impact-of-distributed-databases\\\/\"]}]},{\"@type\":\"BreadcrumbList\",\"@id\":\"https:\\\/\\\/www.pingcap.com\\\/article\\\/the-evolution-and-impact-of-distributed-databases\\\/#breadcrumb\",\"itemListElement\":[{\"@type\":\"ListItem\",\"position\":1,\"name\":\"Home\",\"item\":\"https:\\\/\\\/www.pingcap.com\\\/\"},{\"@type\":\"ListItem\",\"position\":2,\"name\":\"Articles\",\"item\":\"https:\\\/\\\/www.pingcap.com\\\/article\\\/\"},{\"@type\":\"ListItem\",\"position\":3,\"name\":\"The Evolution and Impact of Distributed Databases\"}]},{\"@type\":\"WebSite\",\"@id\":\"https:\\\/\\\/www.pingcap.com\\\/#website\",\"url\":\"https:\\\/\\\/www.pingcap.com\\\/\",\"name\":\"TiDB\",\"description\":\"TiDB | SQL at Scale\",\"publisher\":{\"@id\":\"https:\\\/\\\/www.pingcap.com\\\/#organization\"},\"potentialAction\":[{\"@type\":\"SearchAction\",\"target\":{\"@type\":\"EntryPoint\",\"urlTemplate\":\"https:\\\/\\\/www.pingcap.com\\\/?s={search_term_string}\"},\"query-input\":{\"@type\":\"PropertyValueSpecification\",\"valueRequired\":true,\"valueName\":\"search_term_string\"}}],\"inLanguage\":\"ko-KR\"},{\"@type\":\"Organization\",\"@id\":\"https:\\\/\\\/www.pingcap.com\\\/#organization\",\"name\":\"PingCAP\",\"url\":\"https:\\\/\\\/www.pingcap.com\\\/\",\"logo\":{\"@type\":\"ImageObject\",\"inLanguage\":\"ko-KR\",\"@id\":\"https:\\\/\\\/www.pingcap.com\\\/#\\\/schema\\\/logo\\\/image\\\/\",\"url\":\"https:\\\/\\\/static.pingcap.com\\\/files\\\/2021\\\/11\\\/pingcap-logo.png\",\"contentUrl\":\"https:\\\/\\\/static.pingcap.com\\\/files\\\/2021\\\/11\\\/pingcap-logo.png\",\"width\":811,\"height\":232,\"caption\":\"PingCAP\"},\"image\":{\"@id\":\"https:\\\/\\\/www.pingcap.com\\\/#\\\/schema\\\/logo\\\/image\\\/\"},\"sameAs\":[\"https:\\\/\\\/facebook.com\\\/pingcap2015\",\"https:\\\/\\\/x.com\\\/PingCAP\",\"https:\\\/\\\/linkedin.com\\\/company\\\/pingcap\",\"https:\\\/\\\/youtube.com\\\/channel\\\/UCuq4puT32DzHKT5rU1IZpIA\"]}]}<\/script>\n<!-- \/ Yoast SEO plugin. -->","yoast_head_json":{"title":"The Evolution and Impact of Distributed Databases | TiDB","description":"The evolution of distributed databases, from 1970s prototypes through Bigtable and Spanner to TiDB\u2019s HTAP capabilities.","robots":{"index":"index","follow":"follow","max-snippet":"max-snippet:-1","max-image-preview":"max-image-preview:large","max-video-preview":"max-video-preview:-1"},"canonical":"https:\/\/www.pingcap.com\/ko\/article\/the-evolution-and-impact-of-distributed-databases\/","og_locale":"ko_KR","og_type":"article","og_title":"The Evolution and Impact of Distributed Databases | TiDB","og_description":"The evolution of distributed databases, from 1970s prototypes through Bigtable and Spanner to TiDB\u2019s HTAP capabilities.","og_url":"https:\/\/www.pingcap.com\/ko\/article\/the-evolution-and-impact-of-distributed-databases\/","og_site_name":"TiDB","article_publisher":"https:\/\/facebook.com\/pingcap2015","article_modified_time":"2026-09-03T21:15:48+00:00","og_image":[{"width":1440,"height":714,"url":"https:\/\/static.pingcap.com\/files\/2024\/09\/11005522\/Homepage-Ad.png","type":"image\/png"}],"twitter_card":"summary_large_image","twitter_site":"@PingCAP","twitter_misc":{"\uc608\uc0c1 \ub418\ub294 \ud310\ub3c5 \uc2dc\uac04":"6\ubd84"},"schema":{"@context":"https:\/\/schema.org","@graph":[{"@type":"WebPage","@id":"https:\/\/www.pingcap.com\/article\/the-evolution-and-impact-of-distributed-databases\/","url":"https:\/\/www.pingcap.com\/article\/the-evolution-and-impact-of-distributed-databases\/","name":"The Evolution and Impact of Distributed Databases | TiDB","isPartOf":{"@id":"https:\/\/www.pingcap.com\/#website"},"datePublished":"2026-09-03T21:15:47+00:00","dateModified":"2026-09-03T21:15:48+00:00","description":"The evolution of distributed databases, from 1970s prototypes through Bigtable and Spanner to TiDB\u2019s HTAP capabilities.","breadcrumb":{"@id":"https:\/\/www.pingcap.com\/article\/the-evolution-and-impact-of-distributed-databases\/#breadcrumb"},"inLanguage":"ko-KR","potentialAction":[{"@type":"ReadAction","target":["https:\/\/www.pingcap.com\/article\/the-evolution-and-impact-of-distributed-databases\/"]}]},{"@type":"BreadcrumbList","@id":"https:\/\/www.pingcap.com\/article\/the-evolution-and-impact-of-distributed-databases\/#breadcrumb","itemListElement":[{"@type":"ListItem","position":1,"name":"Home","item":"https:\/\/www.pingcap.com\/"},{"@type":"ListItem","position":2,"name":"Articles","item":"https:\/\/www.pingcap.com\/article\/"},{"@type":"ListItem","position":3,"name":"The Evolution and Impact of Distributed Databases"}]},{"@type":"WebSite","@id":"https:\/\/www.pingcap.com\/#website","url":"https:\/\/www.pingcap.com\/","name":"\ud2f0DB","description":"TiDB | SQL at Scale","publisher":{"@id":"https:\/\/www.pingcap.com\/#organization"},"potentialAction":[{"@type":"SearchAction","target":{"@type":"EntryPoint","urlTemplate":"https:\/\/www.pingcap.com\/?s={search_term_string}"},"query-input":{"@type":"PropertyValueSpecification","valueRequired":true,"valueName":"search_term_string"}}],"inLanguage":"ko-KR"},{"@type":"Organization","@id":"https:\/\/www.pingcap.com\/#organization","name":"PingCAP","url":"https:\/\/www.pingcap.com\/","logo":{"@type":"ImageObject","inLanguage":"ko-KR","@id":"https:\/\/www.pingcap.com\/#\/schema\/logo\/image\/","url":"https:\/\/static.pingcap.com\/files\/2021\/11\/pingcap-logo.png","contentUrl":"https:\/\/static.pingcap.com\/files\/2021\/11\/pingcap-logo.png","width":811,"height":232,"caption":"PingCAP"},"image":{"@id":"https:\/\/www.pingcap.com\/#\/schema\/logo\/image\/"},"sameAs":["https:\/\/facebook.com\/pingcap2015","https:\/\/x.com\/PingCAP","https:\/\/linkedin.com\/company\/pingcap","https:\/\/youtube.com\/channel\/UCuq4puT32DzHKT5rU1IZpIA"]}]}},"card_markup":"        <a class=\"card-article\" href=\"https:\/\/www.pingcap.com\/ko\/article\/the-evolution-and-impact-of-distributed-databases\/\">            <h3>The Evolution and Impact of Distributed Databases<\/h3>            <p>Discover how distributed databases enhance scalability, availability, and fault tolerance, with insights into TiDB's innovations.<\/p>        <\/a>","_links":{"self":[{"href":"https:\/\/www.pingcap.com\/ko\/wp-json\/wp\/v2\/article\/23972","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.pingcap.com\/ko\/wp-json\/wp\/v2\/article"}],"about":[{"href":"https:\/\/www.pingcap.com\/ko\/wp-json\/wp\/v2\/types\/article"}],"author":[{"embeddable":true,"href":"https:\/\/www.pingcap.com\/ko\/wp-json\/wp\/v2\/users\/218"}],"wp:attachment":[{"href":"https:\/\/www.pingcap.com\/ko\/wp-json\/wp\/v2\/media?parent=23972"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}