Register and share your invite link to earn from video plays and referrals.

Search results for reinforcement
reinforcement community
One keyword maps to one global community path.
Create community
People
Not Found
Tweets including reinforcement
@LightOrigins_ latest demo: pure reinforcement learning in simulation, with zero-shot transfer to the real robot.
A new @SciRobotics study highlights a reinforcement learning–based framework that integrates perceptual uncertainty into policy learning to equip humanoid robots with reactive soccer skills and dynamic behaviors.
Show more
Depleted Liberty expected to get reinforcements soon with injury returns, overseas arrivals looming
OPENAI SAYS THAT ON AUGUST 28 IT RESTARTED A LARGE FRONTIER REINFORCEMENT LEARNING RUN THAT WAS PREVIOUSLY PAUSED, AND THAT FOR GPT-5.6 IT IMPROVED THE ROBUSTNESS OF ITS SYSTEM-LEVEL STACK BY ADDING ACTIVATION CLASSIFIERS AND IMPROVING COVERAGE OVER UNIVERSAL JAILBREAKS.
Show more
are any of the big labs training models with RLCF (reinforcement learning from comedians’ feedback)?
Brian Cashman made a risky Yankees bet with lack of reinforcements at trade deadline
Congratulations to the @GoogleDeepMind authors of "Asynchronous Methods for Deep Reinforcement Learning", recipient of the #ICML2026# Test of Time Award. This work shows that asynchronous actor-critic succeeds on a wide variety of continuous motor control problems as well as on a new task of navigating random 3D mazes using a visual input.
Show more
REK is looking to hire a sales/ops and a RL (reinforcement learning) person that is/can be based in SF. Come build real steel with us, will be the coolest thing you ever do.
Congratulations to the authors of "Continuous Control with Deep Reinforcement Learning", recipient of the #ICLR2026# Test of Time Award. This work remains a foundational contribution to robotics. Read the paper: @GoogleDeepMind
Show more
I’m often struck by how human self-development mirrors Reinforcement Learning. With strong reward signals from parents, friends, or bosses, we learn fast. With punishment, we avoid mistakes. Sometimes we get stuck in local optima—chasing short-term wins while missing rewards that take years to reveal. What matters most is: 1) the environment we choose, does it gives good signal to us; 2) can we see through the noise to spot true rewards; and are we willing to explore the uncertain—even when the payoff isn’t clear yet? We only live once. Make every exploration count.
Show more