CEASaclayNon précisé
Context
The interpretation of human interactions in images or videos has significantly improved with the emergence of Large Language Models (LLMs) and Vision-Language Models (VLMs). However, these large models, whether used directly or distilled into specialized models, still have significant limitations, particularly in accurately attributing interactions to the correct person in dense scenes and discriminating actions in the presence of objects.
Evaluation protocols and databases for this task do not always accurately reflect the true capabilities of the methods due to issues such as annotation imprecision or overly rigid semantic metrics.
What do we expect from you ?
To address these problems, the internship will focus on the following objectives:
# Cea List
Moyens / Méthodes / Logiciels
AI, Deep Neural Network, Computer Vision, Human behavior analysis
Profil du candidat
Profile
Localisation du poste Site
Saclay
Localisation du poste
France, Ile-de-France, Essonne (91)
Ville
Saclay
Critères candidat Diplôme préparé
Bac+5
Formation recommandée
AI, Deep Learning, Computer Vision
Possibilité de poursuite en thèse
Oui
Demandeur Disponibilité du poste
01/02/2027