HyperAIHyperAI

Command Palette

Search for a command to run...

Paper - On the Generalization of SFT: A Reinforcement Learning Perspective with Reward Rectification | Papers | HyperAI