SF-Push: Safe and Fair Pose-to-Pose Collaborative Transport with Multiple Quadruped Robots
Safety and fairness restrict how a robot team moves without rewarding task completion, so a single scalar reward has to fix their trade-off in advance. SF-Push keeps one policy gradient per robot and per objective, takes its direction from a scale-free minimum-norm combination and its step length from the task gradient alone — no trade-off weights to tune.