Assistant Professor @Princeton. Developing robots that plan and learn to help people.

Princeton, NJ
Tom Silver retweeted
We put a lot of effort into making the experimental setup clean and relatively strict, and hope the sandbox can be useful for future agent evaluation. Also don’t miss the qualitative examples, where the agents came up with some pretty interesting strategies 👀
Hungry for more Astra robot videos? 🤖 How about a large-scale sandboxed evaluation to go with it? We ran 98,000 evaluations across 28 simulated environments and found some surprisingly clever physical reasoning along the way! New preprint 🧵👇 (1/9)
1
1
10
1,079
Tom Silver retweeted
A number of these problems were unsolvable (or at least very very hard to write down a solution to) ~6 months ago. Wild times!
I’ve seen robots do plenty of unexpected things. That usually means it’s time to fix a bug. This new project is one of only a few times in my career where I’ve seen robots do things that are unexpectedly clever—that none of us on the project anticipated. A few examples...
1
2
16
2,350
Tom Silver retweeted
Agents like Astra has been popular for a while, but “rigorously” what can they actually “solve” at all? Excited to share our latest work on strict, red-teamed sandbox 🔐evaluation on coding agents! (They are indeed surprisingly smart 🤯🤯)
Replying to @merler_m
This is my first paper with @tomssilver's group, and I'm super happy with the results. Thank you @Bw_Li1024 @thosehippos @yichao_liang @WangRobin58334 @YixuanHuang13 for all the help! 🙏 Excited to keep working together! 🚀 🔗 Paper + videos: agenticgentamp.github.io/ (9/9)
1
6
573
Tom Silver retweeted
This is my first paper with @tomssilver's group, and I'm super happy with the results. Thank you @Bw_Li1024 @thosehippos @yichao_liang @WangRobin58334 @YixuanHuang13 for all the help! 🙏 Excited to keep working together! 🚀 🔗 Paper + videos: agenticgentamp.github.io/ (9/9)
1
5
15
1,409
I’ve seen robots do plenty of unexpected things. That usually means it’s time to fix a bug. This new project is one of only a few times in my career where I’ve seen robots do things that are unexpectedly clever—that none of us on the project anticipated. A few examples...
Hungry for more Astra robot videos? 🤖 How about a large-scale sandboxed evaluation to go with it? We ran 98,000 evaluations across 28 simulated environments and found some surprisingly clever physical reasoning along the way! New preprint 🧵👇 (1/9)
4
11
86
10,653
🏀 🗑️
2
2
11
582
The early deadline has passed, but there's still time to submit to LEAP @corl_conf! We're also excited to announce that Dieter Fox has joined our program 🙂 See you in Austin!
9
32
2,481
This week's #PaperILike is "Principles of Animal Cognition for LLM Evaluations: A Case Study on Transitive Inference" (Rane et al., 2025). Still hunting for ways to understand LLMs/agents. This week: animal cognition! PDF: amandaroyka.github.io/ICMLPO… Also: arxiv.org/abs/2503.02882
1
23
1,571
Tom Silver retweeted
Very excited to ramp up our work working with Tom! We're looking for an exceptional postdoctoral researcher (details below) to work jointly with Tom and Basis on MARA project--our effort to build robotic agents that actively learn models of the physical world.
Replying to @BasisOrg
@basisorg and my group at Princeton are recruiting a postdoc to work at the intersection of robotics and code-based world models. We’re looking for someone excited about abstractions & planning, and who knows their way around a real robot. Link 👇 Thanks for boosting!
1
4
24
1,904
Replying to @BasisOrg
@basisorg and my group at Princeton are recruiting a postdoc to work at the intersection of robotics and code-based world models. We’re looking for someone excited about abstractions & planning, and who knows their way around a real robot. Link 👇 Thanks for boosting!
2
12
41
4,332
This week's #PaperILike is "Latent Programming Horizons in Coding Agents" (Silva et al., 2026). Now seems like a good time to better understand what coding agents are doing. Here's a nice example of the kind of analysis one can do. PDF: arxiv.org/abs/2607.05188
4
38
3,311
This week's #PaperILike is "Safe Model-based Reinforcement Learning with Stability Guarantees" (Berkenkamp et al., NeurIPS 2017). Some safe RL guarantees the learned policy is safe; this guarantees safety *during* learning. Important for RL in real! PDF: arxiv.org/abs/1705.08551
1
2
44
3,161
Tom Silver retweeted
Excited to announce the 4th LEAP Workshop at CoRL #CoRL2026! 🚀 We’re looking forward to bringing together the community to discuss learning, reasoning, and planning for robots. Hope to see you there! Learn more: leap-workshop.github.io/
We're excited to announce the 4th Workshop on Learning Effective Abstractions for Planning (LEAP) at #CoRL2026! Previous LEAP papers have gone on to win awards at main conferences (SymSkill, Universal Visual Decomposer). Yes, we'll take all the credit! Workshop link 👇
1
6
722
Tom Silver retweeted
Proud of being part of this organizing team! The topic is more timely than ever! Make sure you submit your contributions and join us!
We're excited to announce the 4th Workshop on Learning Effective Abstractions for Planning (LEAP) at #CoRL2026! Previous LEAP papers have gone on to win awards at main conferences (SymSkill, Universal Visual Decomposer). Yes, we'll take all the credit! Workshop link 👇
1
14
1,592
We're excited to announce the 4th Workshop on Learning Effective Abstractions for Planning (LEAP) at #CoRL2026! Previous LEAP papers have gone on to win awards at main conferences (SymSkill, Universal Visual Decomposer). Yes, we'll take all the credit! Workshop link 👇
1
8
38
4,407
leap-workshop.github.io/ Deadlines: Sep 18 (early), Oct 7 (late) Organized w/ @shah__naman , @GregoryJStein , @GeorgiaChal , David Paulius, @YixuanHuang13 , @utkarshm0410 , @YiqingXu6 , @ambermli , and Franziska Herbert
2
337