back to search

Making Language Models Aligned with Human Intentions and Values

SOT86059Specialisations6 ECTSEnglishwinter semesterDepartment Governance
AI-edited module sheet. Based on the TUMonline module description, edited for readability.Original in TUMonline

What it is about

You learn the foundations and current concepts of AI alignment research for large language models. You engage with model architectures, existing methods, metrics and societal challenges, and you gain practical experience in project design and Python implementation to design and implement strategies to align language models with human intentions and values.

What you will be able to do

  • Understanding the development processes and structures of classical and modern language models
  • Knowledge of central theories and relevant works on AI alignment of large language models
  • Ability to design and implement strategies to constrain or evaluate language models taking human intentions and values into account

What the module consists of

  • SeminarIntroductory session, participant presentations on specific alignment topics and subsequent discussions
  • ProjektarbeitGroup work to develop and implement a project to mitigate misalignments of language models

Teaching method

  • Teilnehmerpräsentationen mit DiskussionDeepening individual alignment topics and promoting critical discussion
  • Gruppenprojekte mit Python-ImplementierungPractical application, project design and implementation of alignment strategies
No dates in the current semester
There are no course dates for this module this semester, or they haven't been matched yet.

Module ratings

No ratings for this module yet.

Rate this module

Only fill in the categories you can judge – for each one, either stars and text together or nothing at all.

Lecture
Tutorial
Exam

Reviews are automatically checked before they are published.

Official page in TUMonline · Details are not binding.