Indeksy pokrywające i skanowanie wyłącznie indeksu
Dowiedz się, jak indeksy pokrywające mogą wyeliminować dostęp do tabeli i umożliwić bardzo wydajne skanowanie wyłącznie indeksu.
Indeksy pokrywające i skanowanie wyłącznie indeksu to bezpłatna lekcja PostgreSQL Performance & Query Optimization na CoddyKit. To lekcja 3 z 4. Możesz przeczytać całą lekcję poniżej za darmo — a potem ćwiczyć ją interaktywnie w przeglądarce z wbudowanym edytorem kodu i tutorem AI dostępnym 24/7. To część ścieżki edukacyjnej PostgreSQL Performance & Query Optimization, a Twój postęp synchronizuje się między webem a aplikacją CoddyKit. Kurs PostgreSQL Performance & Query Optimization zawiera 4 lekcji w sumie.
Części tej lekcji nie zostały jeszcze przetłumaczone i są wyświetlane po angielsku.
What are Covering Indexes?
Welcome to our lesson on Covering Indexes and Index-Only Scans! These advanced indexing techniques can dramatically boost your PostgreSQL query performance.
A regular index helps PostgreSQL quickly find rows based on certain column values. Think of it like a book's index: it tells you which page to go to for a topic.
The Cost of Table Access
When you use a standard index, PostgreSQL first finds the row's location (its tuple ID or CTID) in the index. Then, it has to go to the actual table to fetch the rest of the data for that row.
This extra step, called a table lookup or heap fetch, can be slow, especially if many rows are involved or if the table data isn't in memory.
Indexing All You Need
A covering index is an index that contains all the columns needed to fulfill a query, not just the columns in the WHERE clause. This means PostgreSQL can answer the query by reading only the index, without ever touching the main table data.
It 'covers' the query's data requirements entirely.
Building a Covering Index
To create a covering index, you include all the columns your query might select or filter on. PostgreSQL supports including non-key columns using the INCLUDE clause, which stores them without making them part of the primary sort key.
Let's create a table and then a covering index.
CREATE TABLE products (
product_id SERIAL PRIMARY KEY,
name VARCHAR(100) NOT NULL,
price DECIMAL(10, 2) NOT NULL,
category VARCHAR(50)
);
INSERT INTO products (name, price, category) VALUES
('Laptop', 1200.00, 'Electronics'),
('Keyboard', 75.00, 'Accessories'),
('Mouse', 25.00, 'Accessories'),
('Monitor', 300.00, 'Electronics');
-- Covering index for queries on category that also select name and price
CREATE INDEX idx_products_category_cover_name_price
ON products (category) INCLUDE (name, price);Introducing Index-Only Scans
When a query can be fully satisfied by a covering index, PostgreSQL performs an Index-Only Scan. This is a highly efficient operation because the database engine doesn't need to read any data blocks from the main table (the 'heap').
It's like finding all the information you need directly in the book's index, without ever turning to the main chapters.
MVCC & Index-Only Scans
PostgreSQL uses MVCC (Multi-Version Concurrency Control). For an Index-Only Scan to work, PostgreSQL must ensure that the versions of the rows it finds in the index are visible to the current transaction.
It does this by checking the table's visibility map, a special data structure that tracks which table pages contain only 'all-visible' tuples. If a page is marked all-visible, no heap fetch is needed to check row visibility.
Confirming with EXPLAIN
You can verify if PostgreSQL is using an Index-Only Scan by checking the query plan with EXPLAIN ANALYZE. Look for Index Only Scan in the output.
Let's try a query that our covering index should handle.
EXPLAIN ANALYZE
SELECT name, price
FROM products
WHERE category = 'Electronics';Benefits: Speed & Efficiency
Index-Only Scans offer several key advantages:
- Reduced Disk I/O: Fewer disk reads are needed because the main table is skipped.
- Faster Query Execution: Less I/O directly translates to quicker query responses.
- Better Cache Utilization: Indexes are often smaller and more frequently accessed, leading to better use of memory caches.
- Concurrency: Less contention on table data blocks.
Trade-offs & Considerations
While powerful, covering indexes aren't a silver bullet:
- Larger Indexes: Including more columns makes the index physically larger, consuming more disk space.
- Slower Writes: Updates, inserts, and deletes to the main table also require updating the index, which can be slower for larger indexes.
- Maintenance: Larger indexes might take longer to
VACUUM.
Use them strategically for read-heavy queries that frequently access a specific subset of columns.
Practical Covering Index Example
Consider a `users` table where you often query user names and emails based on their registration date. An index on `registration_date` including `name` and `email` would be a perfect candidate for a covering index.
CREATE TABLE users (
user_id SERIAL PRIMARY KEY,
username VARCHAR(50) NOT NULL,
email VARCHAR(100) NOT NULL,
registration_date DATE NOT NULL
);
INSERT INTO users (username, email, registration_date) VALUES
('alice', 'alice@example.com', '2023-01-15'),
('bob', 'bob@example.com', '2023-02-20'),
('charlie', 'charlie@example.com', '2023-01-25');
CREATE INDEX idx_users_regdate_cover_name_email
ON users (registration_date) INCLUDE (username, email);
EXPLAIN ANALYZE
SELECT username, email
FROM users
WHERE registration_date BETWEEN '2023-01-01' AND '2023-01-31';Quick Check: Index Scans
Based on what you've learned, identify the correct statements about Index-Only Scans in PostgreSQL.
Recap: Covering Indexes
You've learned that covering indexes include all columns needed by a query, enabling highly efficient Index-Only Scans. These scans bypass the main table, significantly reducing disk I/O and speeding up read-heavy queries.
Remember to use EXPLAIN ANALYZE to confirm their usage and consider the trade-offs of increased index size and potential write overhead. Happy optimizing!
Często zadawane pytania
Czy lekcja „Indeksy pokrywające i skanowanie wyłącznie indeksu” jest bezpłatna?
Tak — pełny tekst „Indeksy pokrywające i skanowanie wyłącznie indeksu” jest dostępny za darmo tutaj w sieci. Aby ćwiczyć ją interaktywnie (wbudowany edytor kodu i tutor AI dostępny 24/7) i odblokować resztę kursu PostgreSQL Performance & Query Optimization, przejdź na CoddyKit PRO. Kurs PostgreSQL Performance & Query Optimization zawiera 4 lekcji w sumie.
Co nauczysz się w „Indeksy pokrywające i skanowanie wyłącznie indeksu”?
Dowiedz się, jak indeksy pokrywające mogą wyeliminować dostęp do tabeli i umożliwić bardzo wydajne skanowanie wyłącznie indeksu. Ćwiczysz PostgreSQL Performance & Query Optimization z praktycznym kodem, który uruchamiasz bezpośrednio w przeglądarce, a tutor AI dostępny 24/7 odpowiada na Twoje pytania podczas pracy nad lekcją.
Czy potrzebuję doświadczenia, aby zacząć PostgreSQL Performance & Query Optimization?
Nie wymagamy żadnego doświadczenia. PostgreSQL Performance & Query Optimization w CoddyKit jest strukturyzowany dla początkujących i zaawansowanych użytkowników, więc możesz zacząć tutaj lub od początku i uczyć się w swoim tempie. To lekcja 3 z 4.
Ile czasu zajmuje lekcja „Indeksy pokrywające i skanowanie wyłącznie indeksu”?
Większość lekcji CoddyKit trwa około 5–10 minut. Każda lekcja to mały, interaktywny krok, dzięki czemu robisz systematyczne postępy i zawsze wracasz dokładnie do tego samego miejsca — na webie i w aplikacji.
Czy mogę pisać i uruchamiać kod w tej lekcji PostgreSQL Performance & Query Optimization?
Tak. Każda lekcja PostgreSQL Performance & Query Optimization zawiera wbudowany edytor kodu, więc piszesz i uruchamiasz prawdziwy kod bezpośrednio w przeglądarce i od razu otrzymujesz sprzężenie zwrotne od AI — bez konfiguracji na komputerze.
Wszystkie lekcje w tym kursie
- Indeksy Hash, GIN i GiST
- Indeksy częściowe i indeksy wyrażeń
- Indeksy pokrywające i skanowanie wyłącznie indeksu
- Indeksy BRIN dla dużych danych sekwencyjnych