8 October 2026
Teslova ulica 30
Europe/Ljubljana timezone
Prijave so obvezne! / Registrations Obligatory!

Izvajalec / Course provider: Jožef Stefan Institute (JSI)

Predavatelji / Instructors: Matej Martinc (JSI), Boshko Koloski (JSI)

Learning objectives: Gain practical knowledge of Large Language Model (LLM) alignment using reinforcement learning methods to ensure safety, reduce bias, and follow complex user instructions.

Course content:

For the successful implementation of LLMs in business and scientific environments, models require more than just knowledge; they need alignment with ethical standards and specific user preferences. This practical workshop will cover:

  • Fundamentals of Alignment: Why simple next-token prediction is not enough and how to guide models toward helpfulness, honesty, and harmlessness.
  • Reinforcement Learning from Human Feedback (RLHF): An overview of algorithms (e.g., PPO, DPO, GRPO) for fine-tuning models (such as Gemma and GaMS).
  • Debiasing: Techniques for identifying and reducing stereotypes and biases in generated responses.
  • Adaptation to user needs: How to use the alignment process to ensure a model adheres to the specific tone, style, or professional constraints of a particular organization.

 

Learning outcomes:

  • Understanding the process of model alignment using reinforcement learning.
  • The ability to implement techniques to reduce bias in model responses.
  • Practical knowledge of adapting model behavior to the specific requirements of a client or research project.

Conference information

Date/Time

Starts

Ends

All times are in Europe/Ljubljana

Location

Teslova ulica 30
predavalnica Elektronike na Teslovi 30
Go to map

Chairpersons

Extra information

Language: Slovenian; English

Venue: Online / in-person

Prerequisites: /

Target audience: Advanced users of LLMs, data scientists.

Registration
Registration for this event is currently open.