Stefan, P: Combined Use of Reinforcement Learning and Simula
-
- Englisch ausgewählt
52,99 €
UVP
59,00 €
inkl. gesetzl. MwSt.,
Beschreibung
Produktdetails
(RL) and
simulated annealing (SA) concepts, problems, proposed
solutions,
algorithms and application examples are shown.
RL models a decision maker as a goal-driven agent
aiming to reach
goal states in the problem representation state
space. The agent
takes different choices among the numerous
possibilities, but each
choice can make different impact in the environment.
Each decision
has some effect being expressed in the form of
numeric honor or
dishonor, in a reward value. The agent utilizes the
feedback to
recognize which actions are honored and which are
not. The agent
then tries to govern its decision sequence into the
direction that
maximizes the environment s satisfaction .
The concept of SA is based on the analogy of how
liquids freeze.
There an initially high temperature and disordered
melt is slowly
cooled down and reaches thermal equilibrium.
While in annealing the temperature parameter bounds are
straightforward, in SA they might be dependent on the
problem and
its numeric representation.
This dissertation gives a method which can be used
for defining
temperature bounds in RL environment.
Noch keine Bewertungen vorhanden
Verfassen Sie die erste Bewertung zu diesem Artikel
Helfen Sie anderen Kundinnen und Kunden durch Ihre Meinung.
Kurze Frage zu unserer Seite
Vielen Dank für dein Feedback
Wir nutzen dein Feedback, um unsere Produktseiten zu verbessern. Bitte habe Verständnis, dass wir dir keine Rückmeldung geben können. Falls du Kontakt mit uns aufnehmen möchtest, kannst du dich aber gerne an unseren Kund*innenservice wenden.
zum Kundenservice