Diagnosing Non-Intermittent Anomalies in Reinforcement Learning Policy Executions (Short Paper)

TM
Trevor McFedries
@trevvyboi

Due to the safety risks and training sample inefficiency, it is often preferred to develop controllers in simulation. However, minor differences between the simulation and the real world can cause a significant sim-to-real gap. This gap can reduce the effec...

Uploaded
Uploaded Jul 12, 2026
Queried
Queried 0 times

No preview text is available for this document yet.

Want to learn more?

Ask a question