•1 min read•from Machine Learning
How is RLCD (jev) RL? [D]
Just saw the YouTube presentation and I was left wondering this question.
If jev only outputs Choice, Score, or Noul … well those are all perfectly differentiable. (Cross entropy or mse)
I don’t know if I’m missing something or if adding RL is just for marketing.
Like what would an RL environment even look like?
[link] [comments]
Want to read more?
Check out the full article on the original site
Tagged with
#RLCD
#jev
#RL
#Reinforcement Learning
#Choice
#Score
#Noul
#Differentiable
#Cross Entropy
#MSE
#Environment
#Machine Learning
#Marketing
#YouTube Presentation
#Model Output
#Optimization
#Loss Function
#Training
#Algorithm