“Q* is the ideal scenario - it represents the perfect knowledge of the best action to take in every possible state to achieve the highest reward.
Q-learning, on the other hand, is the method used to approximate Q*. It’s a learning process where the agent interacts with its environment, receives feedback in terms of rewards or penalties, and updates its Q-values based on that feedback.
In Q-learning, the agent initially has no knowledge of the environment and starts with arbitrary Q-values. It then updates these values based on the rewards it receives for its actions, using a formula that takes into account both immediate rewards and estimated future rewards.
Over time, through exploration and learning, the Q-values converge to Q*, giving the agent a map of the best actions to take in each state.
So, in summary, Q* is the ideal set of Q-values that define the best possible strategy, while Q-learning is the process through which an agent learns and approximates these Q-values based on experience and feedback
Or put another way, Q* would be the optimal set of choices to make in a specific game of chess and Q learning would’ve been the method by which those choices were discovered.”
~ GPT-4
Things I commonly hear from wealthy friends:
- I wish I worked less
- I wish I had more freedom
- I wish I could be honest about stress
- I would trade some of my revenue for less work
This is why "rich" is such a subjective term.
Build a rich life, not a rich business.
Today marked 30 days of cold showers. I hate to tell you this but it actually has made a big difference for me. It definitely makes the hard things easier.
I’m excited to be attending the 2020 Clio Cloud Conference, taking place online October 13-16, 2020! Check it out at https://t.co/bCoUZ0MXX9 @goclio#ClioCloud9
Regular note-taking didn't work for me.
Notes stay separate—connections are not made.
I fixed this by building a Zettelkasten using @RoamResearch.
(Zettelkasten originates from highly prolific sociologist Niklas Luhmann. He wrote 70 books & 500 scholarly articles.)
Thread 👇
First time in history...
MONDAY: United States Supreme Court -- U.S. Patent and Trademark Office v. https://t.co/bQ3pNBa5BC B.V. Oral Argument - LIVE at 10am ET on C-SPAN, @cspanradio & online here: https://t.co/QSEIezkfT3
#SCOTUS