Submitted by Chanyoung Kim 75 LIBERO-Para: A Diagnostic Benchmark and Metrics for Paraphrase Robustness in VLA Models Human-centered AI Laboratory 28 5