0Pricing
Web Scraping & Bots · Leçon

Introduction à Selenium

Commencez avec Selenium, un outil puissant d’automatisation des navigateurs et d’extraction de contenu généré par JavaScript.

Introduction à Selenium est une leçon Web Scraping & Bots gratuite sur CoddyKit. Ceci est la leçon 1 sur 4. Tu peux lire la leçon complète ci-dessous gratuitement — puis la pratiquer en direct dans le navigateur avec un éditeur de code intégré et un tuteur IA 24/7. Elle fait partie du parcours d'apprentissage Web Scraping & Bots, et ta progression se synchronise sur le web et l'application CoddyKit. Le cours Web Scraping & Bots comprend 4 leçons au total.

Certaines parties de cette leçon n'ont pas encore été traduites et s'affichent en anglais.

Dynamic Content Challenge

Most websites today aren't static pages. They use JavaScript to load content after the initial page renders.

Think of social media feeds, search results that update as you scroll, or interactive forms. Traditional scraping tools (like Requests) often miss this content because they only fetch the initial HTML.

Meet Selenium WebDriver

Selenium WebDriver is a powerful tool designed to automate web browsers. Unlike libraries that just fetch HTML, Selenium actually launches a real browser (like Chrome or Firefox).

This means it can execute JavaScript, interact with elements, and see the web page exactly as a human user would, making it perfect for dynamic content.

Selenium's Core Idea

Selenium works by sending commands to a specific browser driver (e.g., ChromeDriver for Chrome, GeckoDriver for Firefox).

  • Your Python script tells the driver what actions to perform.
  • The driver then controls the actual browser.
  • The browser executes these actions (navigating, clicking, typing) and returns the updated page state or data.

Install Selenium Library

First, let's install the Selenium library for Python. You can do this using pip, Python's package installer.

Open your terminal or command prompt and run:

pip install selenium

This command downloads and installs the necessary Python components to interact with browsers.

Get Your Browser Driver

Selenium needs a specific browser driver to control your browser. For Chrome, you'll need ChromeDriver. For Firefox, it's GeckoDriver.

1. Check your browser version (e.g., Chrome -> Help -> About Google Chrome).

2. Download the matching driver from its official site:

Place the downloaded driver executable (e.g., chromedriver.exe) in a location accessible by your system's PATH, or note its full path.

Launch Your First Browser

Let's write our first script to open a Chrome browser window using Selenium WebDriver.

Make sure your chromedriver is accessible (either in your PATH or specify its path directly in the Service object).

from selenium import webdriver
from selenium.webdriver.chrome.service import Service
import time

def main():
    driver = None
    try:
        # IMPORTANT: Ensure ChromeDriver is installed and in your system PATH.
        # If not, specify its path directly:
        # service = Service("/path/to/your/chromedriver") 
        # driver = webdriver.Chrome(service=service)
        
        driver = webdriver.Chrome() # Assumes chromedriver is in PATH
        print("Chrome browser launched successfully!")
        
        time.sleep(5) # Keep browser open for 5 seconds to observe
        
    except Exception as e:
        print(f"An error occurred: {e}")
    finally:
        if driver:
            driver.quit() # Always close the browser
            print("Browser closed.")

if __name__ == "__main__":
    main()

Go to a Web Page

Once the browser is open, you can tell it to navigate to any URL using the .get() method.

This is like typing a URL into the address bar and pressing Enter. The browser will load the page, including any JavaScript content.

from selenium import webdriver
import time

def main():
    driver = None
    try:
        driver = webdriver.Chrome() 
        print("Browser launched.")
        
        print("Navigating to example.com...")
        driver.get("https://www.example.com") 
        
        print(f"Current page title: {driver.title}")
        print(f"Current URL: {driver.current_url}")
        
        time.sleep(5) # Keep browser open to see the page
        
    except Exception as e:
        print(f"An error occurred: {e}")
    finally:
        if driver:
            driver.quit()
            print("Browser closed.")

if __name__ == "__main__":
    main()

Waiting for Content

Web pages often take time to load completely, especially with dynamic content. Selenium provides ways to wait for elements to appear.

For now, we'll use a simple time.sleep() to pause execution. Later lessons will cover more robust explicit and implicit waits, which are more efficient.

  • time.sleep(seconds): Pauses execution for a fixed duration.

Always Close Your Browser

It's crucial to close the browser window when your script is done. This frees up system resources and prevents lingering browser processes.

Use the .quit() method on your WebDriver instance. This closes the browser and ends the WebDriver session cleanly.

from selenium import webdriver
import time

def main():
    driver = None
    try:
        driver = webdriver.Chrome() 
        print("Browser launched.")
        
        driver.get("https://www.example.com") 
        print(f"Navigated to: {driver.current_url}")
        
        time.sleep(3) # Wait a bit to see the page
        
    except Exception as e:
        print(f"An error occurred: {e}")
    finally:
        if driver:
            driver.quit() # This line closes the browser window
            print("Browser closed successfully!")

if __name__ == "__main__":
    main()

Quick Check

You've launched your first browser with Selenium! What is the primary purpose of using Selenium WebDriver compared to libraries like Requests for web scraping?

Recap & Next Steps

In this lesson, you learned:

  • Why Selenium is essential for handling dynamic web content.
  • How to install the Selenium library and obtain a browser driver.
  • To launch a browser, navigate to a URL, and close the session cleanly.

Next, we'll dive deeper into automating browser interactions like clicks, form submissions, and finding specific elements on a page!

Questions Fréquemment Posées

La leçon « Introduction à Selenium » est-elle gratuite ?

Oui — le texte complet de « Introduction à Selenium » est gratuit à lire ici sur le web. Pour la pratiquer de manière interactive (un éditeur de code intégré et un tuteur IA 24/7) et déverrouiller le reste du cours Web Scraping & Bots, passe à CoddyKit PRO. Le cours Web Scraping & Bots comprend 4 leçons au total.

Qu'est-ce que j'apprendrai dans « Introduction à Selenium » ?

Commencez avec Selenium, un outil puissant d’automatisation des navigateurs et d’extraction de contenu généré par JavaScript. Tu pratiques Web Scraping & Bots avec du code pratique que tu exécutes directement dans le navigateur, et un tuteur IA 24/7 répond à tes questions au fur et à mesure que tu avances dans la leçon.

Dois-je avoir de l'expérience pour commencer Web Scraping & Bots ?

Aucune expérience préalable n'est requise. Web Scraping & Bots sur CoddyKit est structuré pour les débutants jusqu'aux apprenants avancés, donc tu peux commencer ici ou depuis le début et avancer à ton rythme. Ceci est la leçon 1 sur 4.

Combien de temps prend la leçon « Introduction à Selenium » ?

La plupart des leçons CoddyKit prennent environ 5–10 minutes. Chacune est courte et interactive, tu progresses régulièrement et tu repiques exactement où tu t'es arrêté sur le web et l'app.

Peux-tu écrire et exécuter du code dans cette leçon Web Scraping & Bots ?

Oui. Chaque leçon Web Scraping & Bots inclut un éditeur de code intégré, tu écris et exécutes du vrai code directement dans ton navigateur et tu reçois des retours IA instantanés — aucune configuration locale requise.

Toutes les leçons de ce cours

  1. Introduction à Selenium
  2. Automatiser les interactions avec le navigateur
  3. Extraire des données de JavaScript
  4. Stratégies d’attente pour les pages dynamiques
← Retour à Web Scraping & Bots