Diagnosing Non-Intermittent Anomalies in Reinforcement Learning Policy Executions (Short Paper)
TM
Trevor McFedries
@trevvyboi
Due to the safety risks and training sample inefficiency, it is often preferred to develop controllers in simulation. However, minor differences between the simulation and the real world can cause a significant sim-to-real gap. This gap can reduce the effec...
- Uploaded
- Uploaded Jul 12, 2026
- Queried
- Queried 0 times
No preview text is available for this document yet.
Want to learn more?
Ask a question