FB18 Electrical Engineering and Information Technology · Offered in SoSe 2026
No grade data reported for this course yet.
Grade distributions arrive when a student who took the course opens their Notenspiegel with the TUPlan extension installed.
No ratings yet — be the first.
Your vote is stored against a random id in this browser. No account, nothing that identifies you.
Überblick über Wahrscheinlichkeitstheorie Markov-Eigenschaft und Markov-Entscheidungsprozesse Das Problem des Mehrarmigen Banditen (MAB) und das vollständige Reinforcement Learning (RL) Problem Taxonomie von MAB-Problemen (z.B. stochastische Rewards vs. adversarial Rewards, kontext-abhängige MAB) Algorithmen für MAB-Probleme (z.B. Upper Confidence Interval (UCB), Epsilon-Greedy, SoftMax, LinUCB) und ihre Anwendung in cyber-physischen Systemen Grundlagen der Dynamischen Programmierung und Bellman-Gleichungen Taxonomie der Lösungsansätze für das vollständige RL-Problem (z.B. Temporal-Difference Learning, Policy Gradient und Actor-Critic) Algorithmen für das vollständige RL-Problem (z.B. Q-Learning, SARSA, Policy Gradient, Actor-Critic) und ihre Anwendung in cyber-physischen Systemen Lineare Funktionsapproximation Nicht-Lineare Funktionsapproximation
Times, rooms and details come from TUCaN and may be out of date. Report something wrong