Skip to main content
Ryan Orban

Ryan Orban

Subject
14 entries

Database

Bookmarks

  1. ReadySet: Transparent Database Caching Layer

    ReadySet is a Rust-built wire-compatible caching proxy for MySQL and PostgreSQL that sits between your app and database, incrementally maintaining cached query results via the replication stream. Drop-in deployment with no application code changes.

  2. Benchmarking Postgres vector search: pgvector vs Lantern

    Tembo's benchmark comparing pgvector and Lantern for vector similarity search in PostgreSQL — tests query speed, indexing time, and recall across different dataset sizes. Practical data for choosing a Postgres vector extension.

  3. SuperDuperDB: Bring AI to Your Database

    SuperDuperDB integrates AI models and APIs directly with existing databases — train, manage, and query models where your data already lives rather than moving data to a separate vector database. A database-native alternative to building a separate AI data pipeline.

  4. PostgreSQL Encryption: The Available Options

    Matt Palmer's comprehensive overview of PostgreSQL encryption options — covers full-disk encryption, TLS in transit, column-level encryption with pgcrypto, and the tradeoffs between each. A clear-headed reference for teams navigating database encryption requirements.

  5. ayb: Multi-Tenant Database for Data Ownership

    ayb is a multi-tenant database server built on SQLite that lets individuals own and control their own data — each user gets their own database instance. A philosophical statement about data ownership as much as a technical product.

  6. Prisma Engines: parser-database Source

    The parser-database module inside Prisma's Rust engine — where Prisma Schema Language (PSL) is parsed and validated into a semantic model. Useful reference for understanding how Prisma's schema compiler works internally.

  7. Neon: Serverless Branchable Postgres

    Neon is serverless, fault-tolerant, branchable Postgres — separating storage from compute so you can scale to zero, branch databases like git branches, and get instant provisioning. Positions itself as the Postgres for the serverless/edge computing era.

  8. Grist: Open-Source Spreadsheet-Database

    Grist is an open-source spreadsheet-database hybrid — like Airtable or Notion Database but self-hostable, with Python formulas and a relational data model. Fills the gap between rigid spreadsheets and heavyweight databases for structured data work.

  9. CMU 15-721: Advanced Database Systems

    CMU 15-721 is Andy Pavlo's advanced database systems course covering in-memory databases, query compilation, concurrency control, and storage engines — the internals that most engineers never see. Free lectures, reading list of seminal papers, and a reputation as one of the best systems courses available.

  10. Cleaning Up Your Postgres Database

    Crunchy Data's guide to PostgreSQL database maintenance — identifying bloat, reclaiming space with VACUUM, finding unused indexes, and cleaning up dead connections. Practical operations reference for keeping a Postgres database healthy.

  11. django-seal: Queryset Sealing for Django

    django-seal lets you mark a QuerySet as 'sealed' so that any lazy evaluation attempt (iterating after the context is closed, triggering N+1 queries) raises an exception. It enforces eager loading discipline at the queryset level, catching ORM performance mistakes in development.

  12. Data Cleaning: Problems and Current Approaches (Berkeley/UNECE)

    Joe Hellerstein's Berkeley paper on data cleaning for UNECE — a systematic treatment of the data quality problem from a database research perspective. The academic foundation for what practitioners know as the most time-consuming part of data science work.

  13. CAP Confusion: Problems with 'Partition Tolerance'

    Cloudera's clarification of the most common CAP theorem misreading: partition tolerance isn't a feature you choose — it's a property you must accept because network partitions happen. The real CAP choice is between consistency and availability when partitions occur.

  14. Firebase: Real-Time Backend as a Service

    Firebase (founded 2011, acquired by Google in 2014) was a real-time backend-as-a-service that synced data across clients via WebSockets — eliminating the need to write server-side code for data storage and live updates. It became the template for the 'serverless' and BaaS movement.

All bookmarks