PostgreSQL Performance & Query Optimization · Lezione

Strategie di replica (streaming, logica)

Comprenda e configuri diversi metodi di replica, come la replica streaming e logica, per garantire alta disponibilità e scalabilità delle letture.

Lezione 2 di 412 passaggi

Strategie di replica (streaming, logica) è una lezione PostgreSQL Performance & Query Optimization gratuita su CoddyKit. Questa è la lezione 2 di 4. Puoi leggere la lezione completa qui gratuitamente — poi esercitati direttamente nel browser con un editor di codice integrato e un tutor IA disponibile 24/7. Fa parte del percorso di apprendimento PostgreSQL Performance & Query Optimization, e i tuoi progressi si sincronizzano tra il web e l'app CoddyKit. Il corso PostgreSQL Performance & Query Optimization include 4 lezioni in totale.

Parti di questa lezione non sono ancora state tradotte e vengono mostrate in inglese.

Why Data Replication?

Imagine your database goes down! Replication is key to preventing data loss and ensuring your application stays online. It creates copies of your database.

It's vital for:

  • High Availability (HA): If the main server fails, a copy can take over.
  • Read Scaling: Distribute read queries across multiple servers, reducing load on the primary.

The Primary-Replica Setup

The most common replication setup is a Primary-Replica (or Master-Slave) model. One database server acts as the primary, handling all writes.

  • Primary: Processes all write operations (INSERT, UPDATE, DELETE).
  • Replica(s): Receive copies of data changes from the primary and can serve read-only queries.

This ensures data consistency and offloads read traffic.

Streaming Replication: The Basics

Streaming Replication is PostgreSQL's built-in, physical replication method. It works by continuously shipping the Write-Ahead Log (WAL) records from the primary to one or more replicas.

WAL records are low-level descriptions of every change made to the database. Replicas apply these changes to stay in sync.

WAL Shipping in Action

Here's how streaming replication works:

  • The primary server writes all database changes to its WAL files.
  • A dedicated process (WAL sender) on the primary sends these WAL records to replica servers.
  • A process (WAL receiver) on each replica receives and applies these records.

This keeps the replica's data identical to the primary's, block by block.

Setting Up Streaming Replication (Primary)

To enable streaming replication on the primary, you need to configure postgresql.conf:

  • wal_level = replica (or higher)
  • max_wal_senders = 10 (number of replicas/connections)
  • listen_addresses = '*' (allow connections from replicas)

You'll also need to set up pg_hba.conf to allow replica connections.

ALTER SYSTEM SET wal_level = 'replica';
ALTER SYSTEM SET max_wal_senders = 10;
ALTER SYSTEM SET listen_addresses = '*';

-- After these, restart PostgreSQL

Initializing a Replica with pg_basebackup

Before a replica can stream WAL, it needs a full copy of the primary's data. pg_basebackup is the standard tool for this.

It creates a consistent snapshot of the primary's data directory, which then becomes the replica's initial data.

pg_basebackup -h primary_host -p 5432 -U replication_user -D /var/lib/postgresql/data_replica -F p -Xs stream -P

Configuring the Replica

After pg_basebackup, the replica needs a few files:

  • A standby.signal file in its data directory (for PostgreSQL 12+).
  • A postgresql.conf with hot_standby = on (if you want to run queries).
  • A primary_conninfo setting in postgresql.conf to tell it where to connect to the primary.

This allows it to start up and connect as a replica.

ALTER SYSTEM SET hot_standby = 'on';
ALTER SYSTEM SET primary_conninfo = 'host=primary_host port=5432 user=replication_user password=your_password';

-- Create an empty standby.signal file:
touch /var/lib/postgresql/data_replica/standby.signal

-- Then start the replica PostgreSQL instance

Logical Replication: A Different Approach

Logical Replication is a newer, more flexible replication method introduced in PostgreSQL 10. Unlike streaming replication, it replicates data based on its logical representation (row changes, not physical blocks).

This allows for selective replication of tables or even databases, and cross-version/cross-platform replication.

Publications and Subscriptions

Logical replication uses a publish-subscribe model:

  • A Publication is created on the primary, defining which tables or changes to publish.
  • A Subscription is created on the replica, defining which publications to subscribe to.

The primary then sends logical decoding output (changes) to the subscribers, which apply them.

Setting Up Logical Replication (Example)

Here's a simplified example of setting up logical replication:

On Primary:

CREATE PUBLICATION my_pub FOR TABLE my_table;

On Replica:

CREATE SUBSCRIPTION my_sub CONNECTION 'host=primary_host dbname=mydb user=repl_user password=xyz' PUBLICATION my_pub;

Remember to configure wal_level = logical and max_replication_slots on the primary.

ALTER SYSTEM SET wal_level = 'logical';
ALTER SYSTEM SET max_replication_slots = 5;

-- On Primary:
CREATE PUBLICATION my_pub FOR TABLE my_table;

-- On Replica:
CREATE SUBSCRIPTION my_sub CONNECTION 'host=primary_host dbname=mydb user=repl_user password=xyz' PUBLICATION my_pub;

Replication Strategy Check

You need to set up replication for a PostgreSQL database. You want an exact, block-for-block copy of your entire database on a standby server for high availability. Which replication method is best suited for this goal?

Recap: Streaming vs. Logical

We've explored two key PostgreSQL replication strategies:

  • Streaming Replication: A physical, continuous copy of the entire database using WAL files. Great for high availability and read scaling with identical replicas.
  • Logical Replication: A flexible, publish-subscribe model replicating logical changes (rows). Ideal for selective replication, cross-version upgrades, or integrating with other systems.

Choosing the right strategy depends on your specific needs for consistency, flexibility, and performance.

Gratis per iniziare

Impara SQL con un tutor IA — gratis

Scrivi ed esegui vero codice nel tuo browser, ricevi aiuto istantaneo da un tutor IA disponibile 24/7, e riprendi da dove hai lasciato sul web o nell'app.

Corsi
22
Lezioni
88

Domande Frequenti

La lezione «Strategie di replica (streaming, logica)» è gratuita?

Sì — il testo completo di «Strategie di replica (streaming, logica)» è gratuito qui sul web. Per esercitarvi in modo interattivo (un editor di codice integrato e un tutor IA 24/7) e sbloccare il resto del corso PostgreSQL Performance & Query Optimization, passa a CoddyKit PRO. Il corso PostgreSQL Performance & Query Optimization include 4 lezioni in totale.

Cosa imparerò in «Strategie di replica (streaming, logica)»?

Comprenda e configuri diversi metodi di replica, come la replica streaming e logica, per garantire alta disponibilità e scalabilità delle letture. Eserciti PostgreSQL Performance & Query Optimization con codice pratico che esegui direttamente nel browser, e un tutor IA 24/7 risponde alle tue domande mentre lavori sulla lezione.

Ho bisogno di esperienza per iniziare PostgreSQL Performance & Query Optimization?

Non è richiesta alcuna esperienza precedente. PostgreSQL Performance & Query Optimization su CoddyKit è strutturato per principianti e studenti avanzati, quindi puoi iniziare da qui o dall'inizio e procedere al tuo ritmo. Questa è la lezione 2 di 4.

Quanto tempo richiede la lezione «Strategie di replica (streaming, logica)»?

La maggior parte delle lezioni CoddyKit richiede circa 5–10 minuti. Ogni lezione è breve e interattiva, quindi fai progressi costanti e riprendi esattamente da dove hai lasciato su web e app.

Posso scrivere ed eseguire codice in questa lezione PostgreSQL Performance & Query Optimization?

Sì. Ogni lezione PostgreSQL Performance & Query Optimization include un editor di codice integrato, quindi scrivi ed esegui codice reale direttamente nel tuo browser e ricevi feedback istantaneo dall'IA — nessuna configurazione locale necessaria.

Tutte le lezioni di questo corso

  1. Connection pooling con PgBouncer
  2. Strategie di replica (streaming, logica)
  3. Sharding e PostgreSQL distribuito
  4. Scalare le letture con Hot Standby e bilanciamento del carico
← Torna a PostgreSQL Performance & Query Optimization