0Pricing
Web Scraping & Bots · Lección

Introducción a Selenium

Comience a utilizar Selenium, una herramienta potente para automatizar navegadores y extraer contenido generado mediante JavaScript.

Introducción a Selenium es una lección gratuita de Web Scraping & Bots en CoddyKit. Esta es la lección 1 de 4. Puedes leer la lección completa abajo gratuitamente — luego la practicas en el navegador con un editor de código integrado y un tutor de IA 24/7. Forma parte de la ruta de aprendizaje de Web Scraping & Bots, y tu progreso se sincroniza en la web y la app de CoddyKit. El curso de Web Scraping & Bots incluye 4 lecciones en total.

Partes de esta lección aún no han sido traducidas y se muestran en inglés.

Dynamic Content Challenge

Most websites today aren't static pages. They use JavaScript to load content after the initial page renders.

Think of social media feeds, search results that update as you scroll, or interactive forms. Traditional scraping tools (like Requests) often miss this content because they only fetch the initial HTML.

Meet Selenium WebDriver

Selenium WebDriver is a powerful tool designed to automate web browsers. Unlike libraries that just fetch HTML, Selenium actually launches a real browser (like Chrome or Firefox).

This means it can execute JavaScript, interact with elements, and see the web page exactly as a human user would, making it perfect for dynamic content.

Selenium's Core Idea

Selenium works by sending commands to a specific browser driver (e.g., ChromeDriver for Chrome, GeckoDriver for Firefox).

  • Your Python script tells the driver what actions to perform.
  • The driver then controls the actual browser.
  • The browser executes these actions (navigating, clicking, typing) and returns the updated page state or data.

Install Selenium Library

First, let's install the Selenium library for Python. You can do this using pip, Python's package installer.

Open your terminal or command prompt and run:

pip install selenium

This command downloads and installs the necessary Python components to interact with browsers.

Get Your Browser Driver

Selenium needs a specific browser driver to control your browser. For Chrome, you'll need ChromeDriver. For Firefox, it's GeckoDriver.

1. Check your browser version (e.g., Chrome -> Help -> About Google Chrome).

2. Download the matching driver from its official site:

Place the downloaded driver executable (e.g., chromedriver.exe) in a location accessible by your system's PATH, or note its full path.

Launch Your First Browser

Let's write our first script to open a Chrome browser window using Selenium WebDriver.

Make sure your chromedriver is accessible (either in your PATH or specify its path directly in the Service object).

from selenium import webdriver
from selenium.webdriver.chrome.service import Service
import time

def main():
    driver = None
    try:
        # IMPORTANT: Ensure ChromeDriver is installed and in your system PATH.
        # If not, specify its path directly:
        # service = Service("/path/to/your/chromedriver") 
        # driver = webdriver.Chrome(service=service)
        
        driver = webdriver.Chrome() # Assumes chromedriver is in PATH
        print("Chrome browser launched successfully!")
        
        time.sleep(5) # Keep browser open for 5 seconds to observe
        
    except Exception as e:
        print(f"An error occurred: {e}")
    finally:
        if driver:
            driver.quit() # Always close the browser
            print("Browser closed.")

if __name__ == "__main__":
    main()

Go to a Web Page

Once the browser is open, you can tell it to navigate to any URL using the .get() method.

This is like typing a URL into the address bar and pressing Enter. The browser will load the page, including any JavaScript content.

from selenium import webdriver
import time

def main():
    driver = None
    try:
        driver = webdriver.Chrome() 
        print("Browser launched.")
        
        print("Navigating to example.com...")
        driver.get("https://www.example.com") 
        
        print(f"Current page title: {driver.title}")
        print(f"Current URL: {driver.current_url}")
        
        time.sleep(5) # Keep browser open to see the page
        
    except Exception as e:
        print(f"An error occurred: {e}")
    finally:
        if driver:
            driver.quit()
            print("Browser closed.")

if __name__ == "__main__":
    main()

Waiting for Content

Web pages often take time to load completely, especially with dynamic content. Selenium provides ways to wait for elements to appear.

For now, we'll use a simple time.sleep() to pause execution. Later lessons will cover more robust explicit and implicit waits, which are more efficient.

  • time.sleep(seconds): Pauses execution for a fixed duration.

Always Close Your Browser

It's crucial to close the browser window when your script is done. This frees up system resources and prevents lingering browser processes.

Use the .quit() method on your WebDriver instance. This closes the browser and ends the WebDriver session cleanly.

from selenium import webdriver
import time

def main():
    driver = None
    try:
        driver = webdriver.Chrome() 
        print("Browser launched.")
        
        driver.get("https://www.example.com") 
        print(f"Navigated to: {driver.current_url}")
        
        time.sleep(3) # Wait a bit to see the page
        
    except Exception as e:
        print(f"An error occurred: {e}")
    finally:
        if driver:
            driver.quit() # This line closes the browser window
            print("Browser closed successfully!")

if __name__ == "__main__":
    main()

Quick Check

You've launched your first browser with Selenium! What is the primary purpose of using Selenium WebDriver compared to libraries like Requests for web scraping?

Recap & Next Steps

In this lesson, you learned:

  • Why Selenium is essential for handling dynamic web content.
  • How to install the Selenium library and obtain a browser driver.
  • To launch a browser, navigate to a URL, and close the session cleanly.

Next, we'll dive deeper into automating browser interactions like clicks, form submissions, and finding specific elements on a page!

Preguntas frecuentes

¿La lección «Introducción a Selenium» es gratis?

Sí — el texto completo de «Introducción a Selenium» es gratis para leer aquí en la web. Para practicarla de forma interactiva (editor de código integrado y tutor de IA 24/7) y desbloquear el resto del curso de Web Scraping & Bots, actualiza a CoddyKit PRO. El curso de Web Scraping & Bots incluye 4 lecciones en total.

¿Qué aprenderé en «Introducción a Selenium»?

Comience a utilizar Selenium, una herramienta potente para automatizar navegadores y extraer contenido generado mediante JavaScript. Practicas Web Scraping & Bots con código real que ejecutas directamente en el navegador, y un tutor de IA 24/7 responde tus preguntas mientras trabajas en la lección.

¿Necesito experiencia previa para empezar Web Scraping & Bots?

No se requiere experiencia previa. Web Scraping & Bots en CoddyKit está estructurado para principiantes hasta estudiantes avanzados, así que puedes empezar aquí o desde el inicio y avanzar a tu ritmo. Esta es la lección 1 de 4.

¿Cuánto tiempo toma la lección «Introducción a Selenium»?

La mayoría de las lecciones de CoddyKit toman alrededor de 5–10 minutos. Cada una es compacta e interactiva, así que avanzas constantemente y retomas exactamente por donde dejaste en la web y la app.

¿Puedo escribir y ejecutar código en esta lección de Web Scraping & Bots?

Sí. Cada lección de Web Scraping & Bots incluye un editor de código integrado, así que escribes y ejecutas código real directamente en tu navegador y obtienes retroalimentación instantánea de IA — sin configuración local necesaria.

Todas las lecciones de este curso

  1. Introducción a Selenium
  2. Automatización de interacciones del navegador
  3. Extracción de datos de JavaScript
  4. Estrategias de espera para páginas dinámicas
← Volver a Web Scraping & Bots