MARS policy: Multimodality only when it matters
We present MARS, a generative policy that adaptively invokes stochasticity only when behavioral branching truly matters and reverts to deterministic learning otherwise, yielding 16.67% higher success and 83.20% lower inference latency in real-world tests.