We hypothesize that maximizing the chance of producing a correct answer using outcome-based RL may incentivize blind guessing. Also, some behaviors like simulating a code tool may improve accuracy on some training tasks, even though they confuse the model on other tasks. (18/)
@minimaxir Haha, I think the main reason may be that people’s expectations for the quality of ordinary blog posts are different from the (higher) quality of the blog posts you publish.
Obsessed with the new “make it more” trend on ChatGPT.
You generate an image of something, and then keep asking for it to be MORE.
For example - spicy ramen getting progressively spicier 🔥 (from u/dulipat)
Each month, we highlight community members who are doing unique and interesting things with #KNIME, or who are sharing useful #datascience tips and tricks. https://t.co/kYrqGLgZzs
3/3
Probably the best thing you'll see today.
In 2017, a group of developers hilariously competed for who could create worst volume control interface in the world.
The results 🧵
1/22
Article @StanfordMed about our lab’s discovery (Clinical Trial) of 5 min Cyclic (Physiological) Sighing as a potent way to reduce anxiety 24/7. For reducing stress/elevating mood Cyclic Sighing outperforms meditation or other practices. https://t.co/qvOjSkBkYx
As they say, if you don't schedule maintenance for your car, it will schedule it for you. Applies to people too. Discipline your life or your life will discipline you.