During the initial training phase, both algorithms demonstrated a decreasing trend in their average reward.
← all excerpts
Path Planning of a Mobile Robot for a Dynamic Indoor Environment Based on an SAC-LSTM Algorithm.
1
—
—
During the initial training phase, both algorithms demonstrated a decreasing trend in their average reward.