Zhi Zheng
zz1358m
AI & ML interests
LLM reasoning, Trustworthy LLM, LLM application, Neural combinatorial optimization.
Recent Activity
upvoted a paper about 15 hours ago
Best Practice Critic Optimization liked a model 7 days ago
zz1358m/Qwen3.5-4B-MATH-ReAct-Agentic-ESOpt upvoted a paper 7 days ago
Agentic ESOpt: Fine-Tuning Long-Horizon LLM Agents with Minimal GPU Requirements