Man-Lu
Lu-Man
AI & ML interests
None yet
Recent Activity
upvoted a paper about 1 month ago
CoRT: Counterfactual Replay for Token-Level Rubric-Guided Policy Optimization upvoted a paper 10 months ago
A Theoretical Study on Bridging Internal Probability and
Self-Consistency for LLM Reasoning liked a Space 11 months ago
WNJXYK/RPCOrganizations
None yet