Reinforcement Learning
Search documents
Chelsea Finn: This is the State of the Art in Robotics
Y Combinator· 2026-08-12 15:41
Everyone, today I'm going to be talking about the state-of-the-art of physical intelligence. And in particular, two years ago, I founded a company called physical intelligence. And uh we're really interested in how we can basically uh develop any robot allow any robot to do any task in the real world.Uh and I was actually spoke at this event a year ago uh last year and at the event last year I shared some of our progress in uh at the company at physical intelligence where we could do things really complicat ...
Scaling Compute on Context — Jack Morris, Engram
AI Engineer· 2026-08-12 15:30
[music] >> All right. Hi, everybody. Uh my name's Jack.I'm here to talk about scaling compute on context and also our startup N gram, which launched last week. Um more This isn't going to be like a super detail-oriented talk where I go through a lot of experiments we've been running or talk too much about what our models do. I just want to frame like the high-level problem of what we call scaling compute on context.People have many names for this. It's maybe like a a sub problem of continual learning or may ...
Scaling up Continual Learning — Ronak Malde, Trajectory
AI Engineer· 2026-08-12 14:30
[music] All right. Hey everyone. Uh, thanks for coming by over here.Hope you enjoyed the chat by part on uh, how to measure continual learning. Now I'm here to talk about how we scale it up. So a bit of background about me.I uh went to this company called WinSurf where I was growing the research team over there. Uh we trained this model called sui 1 that ended up leading to the two billion acquisition at deep mind and then I ended up giving up all the acquisition money to start trajectory where we're buildi ...
RL Environments Explained: How AI Agents Learn Real-World Work | Brendan Foody, Mercor
Sequoia Capital· 2026-08-12 13:00
Um, Record, I think you guys grew from a 1 to a 2 billion dollar revenue run rate in the last 4 months or so. Um, so this company's off to the races and I think you were just so front and center to how companies are thinking about uh post training their own models, uh building their own intelligence. So, thank you for joining us for this conversation.Um, format-wise what we're going to do is we've 15 minutes or so of content from Brendan. He's going to talk about uh RL environments in particular, which I th ...
Post-Training Is How You Keep Your Taste | Fireworks CEO Lin Qiao
Sequoia Capital· 2026-08-12 12:00
Okay, up next we have Lynn, uh close close friend and collaborator of Brendan's. Um Lynn, uh actually show of hands, who here who here does post training. Okay, good amount of the room.And who here uses Firework. Okay, good amount of the room, too. So, Lynn, you have you have some friendlies in the audience.Um and so for this next talk, what we're going to do is we're going to focus on post training. Um Lynn, si- similar setup to to Brendan's talk. We're going to do 15 minutes on how to approach the problem ...