Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Hao Jiang's picture

Hao Jiang

Lutalica
8 7
bowiehsu's profile picture 21world's profile picture
·
https://rewindl.github.io/
  • RewindL

AI & ML interests

Multimodal LLMs, LLM Reasoning, Reinforcement Learning, Efficient Inference

Organizations

AGI Lab's profile picture Sun Yat-Sen University's profile picture

upvoted 3 papers 3 months ago

Pyramid Texture Filtering

Paper • 2305.06525 • Published May 11, 2023 • 1

Recovering Policy-Induced Errors: Benchmarking and Trajectory Synthesis for Robust GUI Agents

Paper • 2605.29447 • Published May 28 • 21

Long Live The Balance: Information Bottleneck Driven Tree-based Policy Optimization

Paper • 2605.28109 • Published May 27 • 23
upvoted a paper 6 months ago

MASQuant: Modality-Aware Smoothing Quantization for Multimodal Large Language Models

Paper • 2603.04800 • Published Mar 5 • 25
upvoted a paper 7 months ago

D-CORE: Incentivizing Task Decomposition in Large Reasoning Models for Complex Tool Use

Paper • 2602.02160 • Published Feb 2 • 14
upvoted a paper about 1 year ago

Reasoning or Memorization? Unreliable Results of Reinforcement Learning Due to Data Contamination

Paper • 2507.10532 • Published Jul 14, 2025 • 90
upvoted a paper over 1 year ago

Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?

Paper • 2504.13837 • Published Apr 18, 2025 • 141
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs