You must be enrolled in a school or university able to provide a *convention de stage* for the full 6 months.
**Must-haves:**
* Working knowledge of Python: `requests`, `playwright` (or Selenium), plus the basics (loops, functions, files, JSON, virtual environments)
* Basic SQL: you can write a `SELECT` with a `WHERE`, a `GROUP BY`, and a simple `JOIN` without needing to look everything up
* Comfort reading HTML and using browser dev tools (inspecting the DOM, reading the network tab)
* Git basics for version control
* Autonomy: you're comfortable investigating a problem on your own before asking
* Fluency in English (written & spoken); French is a bonus
**Nice-to-haves:**
* Prior experience with web scraping, in any context (personal projects count)
* Familiarity with anti-bot mechanisms, proxies, or headless browsers
* Experience with pandas for quick data checks
**What you get:**
* Real, messy, large-scale data from day one — the kind you can't get from a course project
* Genuine technical depth on scraping: reverse-engineering APIs, dealing with sites that don't want to be scraped, keeping collection reliable at scale
* Autonomy and ownership over the sources you handle
* A flat structure, flexible hours, casual dress, no bureaucracy
Technologies clés
PythonHTMLSQLGitPandas
Conditions en un coup d’œil
Contrat
Stage(déduit)
Durée
6 mois(déduite)
Début
Non précisé
Télétravail
Hybride
Source
Welcome to the Jungle
Suivi de candidature
Connectez-vous pour suivre vos candidatures et garder des notes privées.