Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Seongryong Jung
SeongryongJung
1
3
4
Follow
sangminan's profile picture
suwanive's profile picture
2 followers
·
1 following
https://jungseongryong.github.io/
jungseongryong
seongryongjung
AI & ML interests
Post-training, Knowledge Distillation, Self-Evolving AI
Recent Activity
updated
a dataset
about 1 month ago
SeongryongJung/chemistry-jsdhist3-summary
published
a dataset
about 1 month ago
SeongryongJung/chemistry-jsdhist3-summary
updated
a collection
about 1 month ago
Qwen3-4B Chemistry TR Methods
View all activity
Organizations
None yet
SeongryongJung
's models
112
Sort: Recently updated
SeongryongJung/Qwen3-4B-Chemistry-RLRT-TR
Reinforcement Learning
•
4B
•
Updated
Jul 7
•
25
SeongryongJung/Qwen3-8B-ToolUse-RLRT-TR
Reinforcement Learning
•
8B
•
Updated
Jul 7
•
6
SeongryongJung/Qwen3-8B-Materials-RLRT-TR
Reinforcement Learning
•
8B
•
Updated
Jul 7
•
6
SeongryongJung/Qwen3-8B-Biology-RLRT-TR
Reinforcement Learning
•
8B
•
Updated
Jul 7
•
7
SeongryongJung/Qwen3-8B-Physics-RLRT-TR
Reinforcement Learning
•
8B
•
Updated
Jul 7
•
6
SeongryongJung/Qwen3-8B-Chemistry-RLRT-TR
Reinforcement Learning
•
8B
•
Updated
Jul 7
•
20
SeongryongJung/Qwen3-4B-ToolUse-RLRT-TR
Reinforcement Learning
•
4B
•
Updated
Jul 7
•
5
SeongryongJung/Qwen3-4B-Materials-RLRT-TR
Reinforcement Learning
•
4B
•
Updated
Jul 7
•
6
SeongryongJung/Qwen3-4B-Biology-RLRT-TR
Reinforcement Learning
•
4B
•
Updated
Jul 7
•
7
SeongryongJung/Qwen3-4B-Physics-RLRT-TR
Reinforcement Learning
•
4B
•
Updated
Jul 7
•
6
SeongryongJung/Qwen3-8B-Tooluse-GRPO-TR
Text Generation
•
8B
•
Updated
Jul 7
•
8
SeongryongJung/Qwen3-8B-Tooluse-RLSD-TR
Text Generation
•
8B
•
Updated
Jul 3
•
34
SeongryongJung/Qwen3-8B-Material-RLSD-TR
Text Generation
•
8B
•
Updated
Jul 3
•
31
•
1
SeongryongJung/Qwen3-8B-Material-GRPO-TR
Text Generation
•
8B
•
Updated
Jul 3
•
17
SeongryongJung/Qwen3-8B-Biology-RLSD-TR
Text Generation
•
8B
•
Updated
Jul 3
•
74
•
1
SeongryongJung/Qwen3-8B-Biology-GRPO-TR
Text Generation
•
8B
•
Updated
Jul 3
•
74
SeongryongJung/Qwen3-8B-ToolUse-SRPO-TR
Reinforcement Learning
•
8B
•
Updated
Jul 3
•
5
•
1
SeongryongJung/Qwen3-8B-ToolUse-SDPO-TR
Reinforcement Learning
•
8B
•
Updated
Jul 3
•
5
SeongryongJung/Qwen3-8B-Materials-SRPO-TR
Reinforcement Learning
•
8B
•
Updated
Jul 3
•
5
SeongryongJung/Qwen3-8B-Physics-RLSD-TR
Text Generation
•
8B
•
Updated
Jul 3
•
32
SeongryongJung/Qwen3-8B-Physics-GRPO-TR
Text Generation
•
8B
•
Updated
Jul 3
•
78
SeongryongJung/Qwen3-8B-Materials-SDPO-TR
Reinforcement Learning
•
8B
•
Updated
Jul 3
•
5
SeongryongJung/Qwen3-8B-Physics-SRPO-TR
Reinforcement Learning
•
8B
•
Updated
Jul 3
•
6
SeongryongJung/Qwen3-8B-Physics-SDPO-TR
Reinforcement Learning
•
8B
•
Updated
Jul 3
•
4
SeongryongJung/Qwen3-8B-Chemistry-SRPO-TR
Reinforcement Learning
•
8B
•
Updated
Jul 3
•
21
SeongryongJung/Qwen3-8B-Chemistry-SDPO-TR
Reinforcement Learning
•
8B
•
Updated
Jul 3
•
25
SeongryongJung/Qwen3-8B-Biology-SRPO-TR
Reinforcement Learning
•
8B
•
Updated
Jul 3
•
4
SeongryongJung/Qwen3-8B-Chemistry-RLSD-TR
Text Generation
•
8B
•
Updated
Jul 3
•
101
SeongryongJung/Qwen3-8B-Biology-SDPO-TR
Reinforcement Learning
•
8B
•
Updated
Jul 3
•
5
SeongryongJung/Qwen3-8B-Chemistry-GRPO-TR
Text Generation
•
8B
•
Updated
Jul 3
•
110
Previous
1
2
3
4
Next