Scholar
Junkang Wu
Google Scholar ID: deBwV5oAAAAJ
University of Science and Technology of China
llm
contrastive learning
Follow
Homepage
↗
Google Scholar
↗
Citations & Impact
All-time
Citations
466
H-index
11
i10-index
11
Publications
20
Co-authors
9
list available
Contact
No contact links provided.
Publications
23 items
Know When to Stop, Where to Restart: Accelerating Multi-Turn Agentic On-Policy Distillation
2026
Cited
0
PEA-DPO: Perception-Enhanced Alignment Direct Preference Optimization for MLLMs Alignment
2026
Cited
0
ARMOR: Stabilizing On-Policy LLM RL with Off-Policy Anchor Samples
2026
Cited
0
Experience Augmented Policy Optimization for LLM Reasoning
2026
Cited
0
R^2-Mem: Reflective Experience for Memory Search
2026
Cited
0
Beyond Where to Look: Trajectory-Guided Reinforcement Learning for Multimodal RLVR
2026
Cited
0
Bridging Perception and Reasoning: Token Reweighting for RLVR in Multimodal LLMs
2026
Cited
0
On the Direction of RLVR Updates for LLM Reasoning: Identification and Exploitation
2026
Cited
0
Load more
Resume (English only)
Co-authors
9 total
Xiangnan He
University of Science and Technology of China
Jiancan Wu
University of Science and Technology of China
Xiang Wang
University of Science and Technology of China
Wentao Shi
University of Science and Technology of China
Zhengyi Yang
PhD Student, University of Science and Technology of China
Xue Wang
DAMO Academy, Alibaba
Yuexiang Xie
Alibaba Group
Kexin Huang
University of Science and Technology of China