Enhancing LM’s Task Adaptability: Powerful Post-training Framework with Reinforcement Learning from Model Feedback | AMiner