Sign in

Roy Fox

@royf.org
1.6K followers 113 following 59 posts

Assistant Professor of Computer Science, UC Irvine Website: royf.org

PostsRepliesMedia
Roy Fox @royf.org · 05/05/2026
Happy AISTATS, to those who celebrate! We're celebrating a long-coming paper in gradient-based optimization that we call “Moonwalk🕺: Inverse-Forward Differentiation”. indylab.org/pub/Krylov20... 🧵/5
indylab.org
Moonwalk: Inverse-Forward Differentiation
Backpropagation’s main limitation is its need to store intermediate activations, or residuals, during the forward pass, which restricts the depth of trainable networks. This raises a fundamental quest...
100
Reposted by Roy Fox
ACM Special Interest Group on AI @acmsigai.bsky.social · 11/03/2025
This year's ACM/SIGAI Autonomous Agents Research Award goes to Prof. Shlomo Zilberstein. His work on decentralized Markov Decision Processes laid the foundation for decision-theoretic planning in multi-agent systems and multi-agent reinforcement learning. sigai.acm.org/main/2025/03... #SIGAIAward
sigai.acm.org
Shlomo Zilberstein (2025 Autonomous Agents Research Award) - ACM SIGAI
The selection committee for the ACM/SIGAI Autonomous Agents Research Award is pleased to announce that Professor Shlomo Zilberstein is the recipient of the 2025 award. Shlomo Zilberstein is Professor…
0133
Roy Fox @royf.org · 10/03/2025
I hear that the other site has been undergoing a Distributed Disinterest in Service attack.
000
Reposted by Roy Fox
RLDM @rldmparis2027.bsky.social · 11/02/2025
Exciting news - early bird registration is now open for #RLDM2025! 🔗 Register now: forms.gle/QZS1GkZhYGRF... Register now to save €100 on your ticket. Early bird prices are only available until 1st April.
21715
Roy Fox @royf.org · 03/02/2025
2025 is looking to be the year that information-theoretic principles in sequential decision making, finally make a comeback! (at least for me, I know others never stopped.) already 4 very exciting projects, and counting!
030
Roy Fox @royf.org · 28/01/2025
I received an email from the Department of Energy stating that “DOE is moving aggressively to implement this Executive Order by directing the suspension of [...] DEI policies [...] Community Benefits Plans [... and] Justice40 requirements”. This probably explains the NSF panel suspensions as well.
021
Reposted by Roy Fox
Grace Lindsay @neurograce.bsky.social · 07/01/2025
Want a job in robotics in New York? faunarobotics.com
Screenshot of open roles at Fauna Robotics
12810
Roy Fox @royf.org · 31/12/2024
Our 2024 research review isn't complete without mentioning 2 workshop papers that preview upcoming publications; I'll leave other things happening as surprises for 2025.
110
Roy Fox @royf.org · 31/12/2024
Last in our 2024 research review: control with efficient safety guarantees. Formal verification methods are very slow, but here's a cool trick to use them for safe control, with minimal slowdown and provable safety guarantees.
indylab.org
Verification-Guided Shielding for Deep Reinforcement Learning
In recent years, Deep Reinforcement Learning (DRL) has emerged as an effective approach to solving real-world tasks. However, despite their successes, DRL-based policies suffer from poor reliability, ...
110
Roy Fox @royf.org · 30/12/2024
Next up in our 2024 research overview: reinforcement learning under delays. The usual control loop assumes immediate observation and action in each time step, but that's not always possible, as processing observations and decisions can take time. How can we learn to control delayed systems?
indylab.org
Reinforcement Learning from Delayed Observations via World Models
In standard reinforcement learning settings, agents typically assume immediate feedback about the effects of their actions after taking them. However, in practice, this assumption may not hold true du...
110
Roy Fox @royf.org · 16/12/2024
Way back in 2023, before multimodal foundation models were a thing, we wanted to apply language agents to visual domains. One idea was to use vision models to extract perceptual features and put them into text templates. But “a picture is worth 1000 words” — a big context! Can be slow, distracting.
indylab.org
Selective Perception: Learning Concise State Descriptions for Language Model Actors
It is increasingly common for large language models (LLMs) to be applied as actors in sequential decision making problems in embodied domains such as robotics and games, due to their general world kno...
111
Roy Fox @royf.org · 10/12/2024
Many are posting end-of-year research summaries, good idea! Let's start: You want an agent's behavior that can't be exploited by an adversary (zero-sum Nash equilibrium = NE). The world is big, so you restrict the agent to stochastic mixing of a small population. How should you grow the population?
indylab.org
Toward Optimal Policy Population Growth in Two-Player Zero-Sum Games
In competitive two-agent environments, deep reinforcement learning (RL) methods like Policy Space Response Oracles (PSRO) often increase exploitability between iterations, which is problematic when tr...
260
Roy Fox @royf.org · 18/11/2024
Bluesky is really nice! I'm moving all my social inactivity here.
181