Perfilado de sección

  • LESSON 13 - ARTIFICIAL INTELLIGENCE FOUNDATIONS FOR BROADENING CULTURAL HORIZONS

    What is the purpose of data collection?

    In this lesson, we analyze the importance of data collection in evaluating the behavior of language models. These evaluations are conducted using benchmarks, test sets designed to measure performance and, in particular, detect bias. Bias benchmarks function as controlled experiments containing stereotyped sentences and scenarios to observe how the model responds. However, much of the current research focuses on English, which means metrics, datasets, and mitigation strategies are rooted in specific cultural contexts and do not always represent realities like those in Latin America.

    Within this framework, we present the HESEIA 2024 Project, an educational initiative involving teachers and students in Argentina that combined critical AI literacy with collaborative data production. Using the EDIA tool, participants explored biases related to social class, gender, and nationality, building a dataset of nearly 50,000 sentences grounded in local and intersectional contexts. Results showed that many regional stereotypes go undetected by current models. HESEIA thus serves as both a pedagogical and technological intervention: fostering critical thinking while contributing to the creation of fairer, more diverse, and more representative benchmarks.

    At the end of the lesson, you are invited to answer a few questions


    👨‍🏫 Instructor: Guido Ivetta

    • Prefer to watch in a language other than Spanish? Just turn on CC, go to Settings (⚙️) > Subtitles > Auto-translate, and select your language!

    • Below you will find the lesson slides, designed as a reference to revisit the ideas, questions, and tensions raised in the video. (Please note: The slides are in Spanish).

      📚 Key Concepts

      Benchmark: A set of standardized, controlled tests used to measure and compare the performance or biases of different AI models.

      Dataset: An organized collection of information (texts, images, questions) used to train or evaluate models.

      A creation by Fundación Vía Libre in collaboration with FAMAF – UNC.
    • Assignment: Idiom Validator

      Below you will find a video to help you carry out this activity.

      If you don't remember how to access the tool, you can rewatch the video from the "Typical Idioms" activity.

      Access the EDIA tool to start the activity: https://edia.ngrok.app/

      Cultural Questions Activity: Let's get started!