Sign in

Khurram Javed

@khurramjaved.com
77 followers 33 following 12 posts

Working on scalable and decentralized algorithms for real-time reinforcement learning. Research scientist @ Keen AGI Prev - PhD with Richard S. Sutton

PostsRepliesMedia
Khurram Javed @khurramjaved.com · 11/04/2026
A bunch of us are organizing a workshop at RLC. If your goal is to develop algorithms that allow agents to learn from complex data streams without relying on human data and human designers, then this workshop would be a good fit.
051
Khurram Javed @khurramjaved.com · 05/03/2025
It's about time! The key principles of reinforcement learning (e.g., learning by interaction, TD learning) are fundamental to intelligence.
awards.acm.org
Andrew Barto and Richard Sutton are the recipients of the 2024 ACM A.M. Turing Award for developing the conceptual and algorithmic foundations of reinforcement learning.
Andrew Barto and Richard Sutton as the recipients of the 2024 ACM A.M. Turing Award for developing the conceptual and algorithmic foundations of reinforcement learning. In a series of papers beginning...
030
Reposted by Khurram Javed
Tom Schaul @schaul.bsky.social · 24/02/2025
Some extra motivation for those of you in RLC deadline mode: our line-up of keynote speakers -- as all accepted papers get a talk, they may attend yours! @rl-conference.bsky.social
RLC Keynote speakers: Leslie Kaelbling, Peter Dayan, Rich Sutton, Dale Schuurmans, Joelle Pineau, Michael Littman
03710
Khurram Javed @khurramjaved.com · 12/02/2025
Olympiad problems are designed to have elegant solutions and new problems are often designed around old patterns. As language models get better we should expect them to conquer IMO/IOI problems. Simple problems not designed to have elegant solutions would prove harder for language models.
010
Khurram Javed @khurramjaved.com · 03/02/2025
How to keep the AI hype cycle going: 1. Propose new knowledge based benchmarks and show existing LLMs do poorly on them. 2. Train LLMs on the knowledge required to do well on the new benchmarks. 3. Use improvements on the benchmarks as signs of rapid progress.
050
Khurram Javed @khurramjaved.com · 01/01/2025
Almost all robotics startups are betting on learning from large supervised learning datasets collected by teleoperation. The odds of success for this strategy are small. Like the Sim2Real bubble, this bubble might not burst for years. At-least it's keeping the roboticists employed.
160
Khurram Javed @khurramjaved.com · 04/12/2024
Despite significant growth of the AI community some promising research directions are untouched because we rely on a homogeneous set of tools (e.g., autograd). Ideas that are easy to implement with existing tools win the software lottery and are more thoroughly tested.
030
Reposted by Khurram Javed
Tom Schaul @schaul.bsky.social · 02/12/2024
This year's (first-ever) RL conference was a breath of fresh air! And now that it's established, the next edition is likely to be even better: Consider sending your best and most original RL work there, and then join us in Edmonton next summer!
0193
Reposted by Khurram Javed
Glen Berseth @glenberseth.bsky.social · 22/11/2024
RLC will be held at the Univ. of Alberta, Edmonton, in 2025. I'm happy to say that we now have the conference's website out: rl-conference.cc/index.html Looking forward to seeing you all there! @rl-conference.bsky.social #reinforcementlearning
26019