Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Ting Zou's picture

Ting Zou

tingzou
10 8
·

AI & ML interests

Reinforcement learning, reward modeling, RLHF, policy optimization, offline RL

Recent Activity

liked a model about 23 hours ago
lobonexequiel/reinforcement-learning
liked a model about 23 hours ago
Chaew00n/test-policy-optimization-query-rewrite-0522
upvoted a paper about 23 hours ago
Recursive self-improvement of AI research agents
View all activity

Organizations

None yet

liked 2 models about 23 hours ago

lobonexequiel/reinforcement-learning

Reinforcement Learning • Updated Feb 20, 2023 • 20 • 5

Chaew00n/test-policy-optimization-query-rewrite-0522

Updated May 22, 2025 • 1
liked a model 5 days ago

Dhanraj1503/deep_reinforcement_learning

Reinforcement Learning • Updated Jan 15, 2024 • 11 • 6
liked 3 datasets 6 days ago

youinwww/reinforcement_learning

Viewer • Updated Jul 7, 2025 • 20 • 91 • 4

Srishti280992/repro-exact-unlearning-in-reinforcement-learning-traces

Traces • Updated Aug 3 • 97 • 4

D21W12/Diffusion-Deep-Reinforcement-Learning

Updated Jun 27 • 7 • 1
liked 2 models 6 days ago

JonusNattapong/Reinforcement-Learning-for-Gold-Trading-Model

Reinforcement Learning • Updated Dec 23, 2025 • 109 • 16

mineworker/reward_modeling

Updated Jul 31, 2024 • 5
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs