Multi-Agent Deep Deterministic Policy Gradient Algorithm Based on Classification Experience Replay | AMiner