dados e armazenamento
como se guarda e se lê o que não cabe numa máquina: relacional, lsm, colunar, vetorial.
muda como você pensa
- A Relational Model of Data for Large Shared Data Banks
- The Log-Structured Merge-Tree (LSM-Tree)
- Dynamo: Amazon's Highly Available Key-value Store
- Dremel: Interactive Analysis of Web-Scale Datasets
- Gorilla: A Fast, Scalable, In-Memory Time Series Database
- Amazon Aurora: Design Considerations for High Throughput Cloud-Native Relational Databases
- Amazon DynamoDB: A Scalable, Predictably Performant, and Fully Managed NoSQL Database Service
- What Goes Around Comes Around... And Around...
vale o tempo
- ARIES: A Transaction Recovery Method Supporting Fine-Granularity Locking and Partial Rollbacks Using Write-Ahead Logging
- C-Store: A Column-oriented DBMS
- What Goes Around Comes Around
- The Bw-Tree: A B-tree for New Hardware Platforms
- The Snowflake Elastic Data Warehouse
- SILK: Preventing Latency Spikes in Log-Structured Merge Key-Value Stores
- Umbra: A Disk-Based System with In-Memory Performance
- Evolution of Development Priorities in Key-value Stores Serving Large-scale Applications: The RocksDB Experience
- Amazon Redshift Re-invented
- Vector Database Management Techniques and Systems
- Distributed Transactions at Scale in Amazon DynamoDB
para aprofundar
- On-the-fly Sharing for Streamed Aggregation
- Cassandra: A Decentralized Structured Storage System
- Storing and Querying Tree-Structured Records in Dremel
- In-Memory Performance for Big Data
- Amazon Redshift and the Case for Simpler Data Warehouses
- TiDB: A Raft-based HTAP Database
- Lakehouse: A New Generation of Open Platforms that Unify Data Warehousing and Advanced Analytics
- Manu: A Cloud Native Vector Database Management System
- Intelligent Scaling in Amazon Redshift
- Amazon MemoryDB: A Fast and Durable Memory-First Cloud Database
- Predicate Caching: Query-Driven Secondary Indexing for Cloud Data Warehouses
- Stage: Query Execution Time Prediction in Amazon Redshift