Reinforcement Learning: Machine Learning Meets Control Theory

Sparse Identification of Nonlinear Dynamics (SINDy): Sparse Machine Learning Models 5 Years Later!

Why Does Diffusion Work Better than Auto-Regression?

Which team will win? Incredibox Sprunki Phase 1 vs Phase 3?!

ชาวบ้านผวา! หนุ่มบุกยิงประธาน อบต. สาหัส | ข่าวเย็นช่องวัน | สำนักข่าววันนิวส์

🔴LIVE เชียร์สด : แมนเชสเตอร์ ยูไนเต็ด พบ เอฟเวอร์ตัน | ผีแดงปะทะท็อฟฟี่สีน้ำเงิน MW13

SINDy-RL: Interpretable and Efficient Model-Based Reinforcement Learning

Steve Brunton

มุมมอง 18 465

เพิ่มลงใน
- เพลย์ลิสต์ของฉัน
- ดูภายหลัง
แชร์

แชร์

ฝัง

ขนาดวิดีโอ:

แสดงแผงควบคุมโปรแกรมเล่น

เล่นอัตโนมัติ

เล่นใหม่

เผยแพร่เมื่อ 2 ธ.ค. 2024

ความคิดเห็น • 22

@deltax7159 6 หลายเดือนก่อน ⁺⁷
you guys are so brilliant. such a great idea, would love to hear a podcast with you guys talking about how you came up with these ideas/ the life cycle of SINDY rl.
@musicarroll 6 หลายเดือนก่อน ⁺³
Nick: Excellent work! This is genuine progress in AI to integrate state estimation SOTA with decision making (RL). Would love to see this further refined using POMDPs ( Partially Oberservable Markov Decision Processes).
@sai4007 6 หลายเดือนก่อน
Checkout PlaNet and dreamer models
@Idonai 6 หลายเดือนก่อน ⁺³
Thanks for the presentation. Do I understand correctly that this whole process could be automated making highly efficient agents or do some aspects of this process require manual work? Also, how well does it scale to significantly harder RL problems? Does this technique stay computationally efficient (e.g. compared to PPO) in these harder ernvironments? Could this be combined with Reinforcement learning from human feedback (RLHF) in a practical manner?
@jimlbeaver 6 หลายเดือนก่อน ⁺³
Great presentation, very interesting approach. I’m curious about the intuition behind the ensemble…eager to read more. Thanks!
@Eigensteve 6 หลายเดือนก่อน ⁺¹
Thanks Jim! The ensembling gives us way more robustness to noisy data and also to very few data samples, so it can let us train models much more quickly than NN models.
@JoshuaSheppard-pp5iz 6 หลายเดือนก่อน
Bold steps ... thrilling work! I look forward to working through the implementation.
@brianarbuckle2284 6 หลายเดือนก่อน ⁺¹
Great work. This is fantastic!
@TheRubencho176 6 หลายเดือนก่อน
Impressive! Thank you very much for sharing and for the inspiration.
@drj92 6 หลายเดือนก่อน ⁺²
Has your lab considered experimenting with Kolmogorov-Arnold Networks in combination with SINDy? It feels like a potentially excellent match.
Their approach to network sparsification, in particular, seems like it could be automated in a very interesting way via SINDy. In the recent paper they fix and prune activation functions by hand, but it seems that you could instead use SINDy to automatically fix a particular activation function once it fit a dictionary term beyond some threshold.
Love the presentation!
@Eigensteve 6 หลายเดือนก่อน ⁺¹
Neat idea -- definitely thinking about ways of connecting these topics. Thanks!
@xueqiu6384 5 หลายเดือนก่อน
curious about how fitting can accelerate the training process. Any assumptions for action space/ state space / environment? Thanks for your attention.
@srikanthtupurani6316 6 หลายเดือนก่อน ⁺²
This is so amazing I don't have words. Deepmind made computers play go game, chess game. It uses reinforcement learning. It is simply superb.
@ClicheKHFan 6 หลายเดือนก่อน
Amazing. I've been looking for something like this.
@awsomeguy563 6 หลายเดือนก่อน
Absolutely brilliant
@Student-ve5ug 6 หลายเดือนก่อน
Dear Sir,
If we want to use reinforcement learning (RL) in a specific environment, I am concerned that the trial-and-error method will result in many errors, some of which may have negative consequences. Furthermore, I am unsure how many attempts the RL model will need to reach the optimal and correct decision. How can this challenge be addressed?
@nvjt101 6 หลายเดือนก่อน ⁺³
Real AI is RL
@Pedritox0953 6 หลายเดือนก่อน
Great video!
@kevinarancibiacalderon9039 6 หลายเดือนก่อน
Gracias por el video!
@alexxxcanz 6 หลายเดือนก่อน
Great!
@SrZonne 6 หลายเดือนก่อน
Amazing
@juleswombat5309 6 หลายเดือนก่อน
Interesting

ต่อไป

เล่นอัตโนมัติ

Reinforcement Learning: Machine Learning Meets Control Theory

Reinforcement Learning: Machine Learning Meets Control Theory

Sparse Identification of Nonlinear Dynamics (SINDy): Sparse Machine Learning Models 5 Years Later!

Sparse Identification of Nonlinear Dynamics (SINDy): Sparse Machine Learning Models 5 Years Later!

Why Does Diffusion Work Better than Auto-Regression?

Why Does Diffusion Work Better than Auto-Regression?

Which team will win? Incredibox Sprunki Phase 1 vs Phase 3?!

Which team will win? Incredibox Sprunki Phase 1 vs Phase 3?!

ชาวบ้านผวา! หนุ่มบุกยิงประธาน อบต. สาหัส | ข่าวเย็นช่องวัน | สำนักข่าววันนิวส์

ชาวบ้านผวา! หนุ่มบุกยิงประธาน อบต. สาหัส | ข่าวเย็นช่องวัน | สำนักข่าววันนิวส์

🔴LIVE เชียร์สด : แมนเชสเตอร์ ยูไนเต็ด พบ เอฟเวอร์ตัน | ผีแดงปะทะท็อฟฟี่สีน้ำเงิน MW13

🔴LIVE เชียร์สด : แมนเชสเตอร์ ยูไนเต็ด พบ เอฟเวอร์ตัน | ผีแดงปะทะท็อฟฟี่สีน้ำเงิน MW13

ทหารไทยอันดับ25โลก! พม่าถล่มเรือประมงไทย | HOTSHOT เดลินิวส์ 01/12/67

ทหารไทยอันดับ25โลก! พม่าถล่มเรือประมงไทย | HOTSHOT เดลินิวส์ 01/12/67

Visualizing transformers and attention | Talk for TNG Big Tech Day '24

Visualizing transformers and attention | Talk for TNG Big Tech Day '24

Differentiable Convex Modeling for Robotic Planning and Control | PhD Defense

Differentiable Convex Modeling for Robotic Planning and Control | PhD Defense

Why Choose Model-Based Reinforcement Learning?

Why Choose Model-Based Reinforcement Learning?

Miles Cranmer - The Next Great Scientific Theory is Hiding Inside a Neural Network (April 3, 2024)

Miles Cranmer - The Next Great Scientific Theory is Hiding Inside a Neural Network (April 3, 2024)

Deep Reinforcement Learning for Fluid Dynamics and Control

Deep Reinforcement Learning for Fluid Dynamics and Control

AI/ML+Physics Part 1: Choosing what to model [Physics Informed Machine Learning]

AI/ML+Physics Part 1: Choosing what to model [Physics Informed Machine Learning]

You're Probably Wrong About Rainbows

You're Probably Wrong About Rainbows

Large Language Models explained briefly

Large Language Models explained briefly

AI/ML+Physics: Preview of Upcoming Modules and Bootcamps [Physics Informed Machine Learning]

AI/ML+Physics: Preview of Upcoming Modules and Bootcamps [Physics Informed Machine Learning]

路飞做的坏事被拆穿了 #路飞#海贼王

路飞做的坏事被拆穿了 #路飞#海贼王

ไฮไลท์ฟุตบอล พรีเมียร์ลีก 2024/25 สัปดาห์ที่ 13 : แมนเชสเตอร์ ยูไนเต็ด พบ เอฟเวอร์ตัน

ไฮไลท์ฟุตบอล พรีเมียร์ลีก 2024/25 สัปดาห์ที่ 13 : แมนเชสเตอร์ ยูไนเต็ด พบ เอฟเวอร์ตัน

He Surrendered 😐 ft @KingArtyom

He Surrendered 😐 ft @KingArtyom

บุกร้านแตงโมแซ่บเวอร์ สาขากรุงเทพ..สลวนสุดๆๆๆ !!!!| Nisamanee.Nutt

บุกร้านแตงโมแซ่บเวอร์ สาขากรุงเทพ..สลวนสุดๆๆๆ !!!!| Nisamanee.Nutt

Tuna 🍣 ⁠@patrickzeinali ⁠@ChefRush

Tuna 🍣 ⁠@patrickzeinali ⁠@ChefRush

Pineapple pizza 🍕@Lionfield @albert_cancook #pizza #italian #food #funny #cooking #viralvideo

Pineapple pizza 🍕@Lionfield @albert_cancook #pizza #italian #food #funny #cooking #viralvideo

[ ออกกอง ] วัยหนุ่ม 2544 โลกหลังลูกกรง..ทัวร์กองถ่ายหนังใน "คุกของจริง" | JUSTดูIT.

[ ออกกอง ] วัยหนุ่ม 2544 โลกหลังลูกกรง..ทัวร์กองถ่ายหนังใน "คุกของจริง" | JUSTดูIT.

🔴LIVE เชียร์สด : ลิเวอร์พูล พบ แมนเชสเตอร์ ซิตี้ | บิ๊กแมตช์ หงส์แดงดวลเรือใบสีฟ้า MW13

🔴LIVE เชียร์สด : ลิเวอร์พูล พบ แมนเชสเตอร์ ซิตี้ | บิ๊กแมตช์ หงส์แดงดวลเรือใบสีฟ้า MW13