A Two-timescale Resource Allocation Method Based on Deep Reinforcement Learning for 6G Networks | AMiner