Jincenzi commited on
Commit
df8ed64
·
verified ·
1 Parent(s): 7628efa

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +2 -2
README.md CHANGED
@@ -17,9 +17,9 @@ pipeline_tag: text-generation
17
 
18
  # SocialR1-4B
19
 
20
- **SocialR1-4B** is a social reasoning model built on [Qwen3-4B](https://huggingface.co/Qwen/Qwen3-4B), trained with trajectory-level reinforcement learning (GRPO) using the **Social-R1** framework. It enhances Theory-of-Mind (ToM) and social inference capabilities by aligning reasoning processes with the Social Information Processing (SIP) theory.
21
 
22
- 📄 **Paper**: [Social-R1: Enhancing Social Reasoning in LLMs through Trajectory-Level Reinforcement Learning](https://arxiv.org/abs/2603.09249) (NeurIPS 2026)
23
 
24
  ## Highlights
25
 
 
17
 
18
  # SocialR1-4B
19
 
20
+ **SocialR1-4B** is a social reasoning model built on [Qwen3-4B](https://huggingface.co/Qwen/Qwen3-4B), trained with trajectory-level reinforcement learning (GRPO) using the **Social-R1** framework. It enhances social reasoning capabilities by aligning reasoning processes with the Social Information Processing (SIP) theory.
21
 
22
+ 📄 **Paper**: [Social-R1: Enhancing Social Reasoning in LLMs through Trajectory-Level Reinforcement Learning](https://arxiv.org/abs/2603.09249)
23
 
24
  ## Highlights
25