Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Nguyễn Minh Phúc
DatPySci
6
Follow
dark-pen's profile picture
Oztobuzz's profile picture
2 followers
·
2 following
AI & ML interests
Reinforcement learning, NLP
Recent Activity
updated
a dataset
5 days ago
DatPySci/Self-Distillation-Qwen3-4B-Thinking-2507-medreason-tokenized
published
a dataset
5 days ago
DatPySci/Self-Distillation-Qwen3-4B-Thinking-2507-medreason-tokenized
updated
a dataset
5 days ago
DatPySci/Self-Distillation-Qwen3-4B-Instruct-2507-medreason-tokenized
View all activity
Organizations
DatPySci
's datasets
68
Sort: Recently updated
DatPySci/tldr_gpt2_w2s_feedback
Viewer
•
Updated
Jan 4, 2025
•
46.4k
•
18
DatPySci/gpt2-medium_dpo_tldr_temp_1_2
Viewer
•
Updated
Jan 2, 2025
•
8k
•
13
DatPySci/gpt2_dpo_tldr_temp_1_0
Viewer
•
Updated
Jan 2, 2025
•
3.88k
•
10
DatPySci/gpt2-large_dpo_tldr_temp_1_0
Viewer
•
Updated
Jan 2, 2025
•
3.88k
•
12
DatPySci/gpt2-medium_dpo_tldr_temp_1_0
Viewer
•
Updated
Jan 2, 2025
•
3.88k
•
17
DatPySci/weak_to_strong_reward_tldr
Viewer
•
Updated
Dec 30, 2024
•
94.8k
•
14
•
1
DatPySci/weak_to_strong_reward_hh
Viewer
•
Updated
Dec 30, 2024
•
110k
•
53
•
1
DatPySci/gpt2_dpo_anthropic_hh_pref
Viewer
•
Updated
Dec 28, 2024
•
128k
•
23
DatPySci/base_llama3-1b_anthropic_hh
Viewer
•
Updated
Dec 26, 2024
•
8k
•
25
DatPySci/gpt2-large_dpo_anthropic_hh
Viewer
•
Updated
Dec 25, 2024
•
8k
•
23
DatPySci/gpt2-medium_dpo_anthropic_hh
Viewer
•
Updated
Dec 25, 2024
•
8k
•
28
DatPySci/gpt2_dpo_anthropic_hh
Viewer
•
Updated
Dec 25, 2024
•
8k
•
24
DatPySci/base_anthropic_hh
Viewer
•
Updated
Dec 25, 2024
•
219k
•
33
DatPySci/llama3-1b_dpo_tldr
Viewer
•
Updated
Dec 21, 2024
•
8k
•
18
DatPySci/gpt2-large_dpo_tldr
Viewer
•
Updated
Dec 21, 2024
•
8k
•
13
DatPySci/gpt2-medium_dpo_tldr
Viewer
•
Updated
Dec 21, 2024
•
8k
•
15
DatPySci/gpt2_dpo_tldr
Viewer
•
Updated
Dec 21, 2024
•
8k
•
16
DatPySci/tldr-preference-llama3-8b
Viewer
•
Updated
Dec 21, 2024
•
103k
•
33
DatPySci/gsm8k_llama3-1b-instruct
Viewer
•
Updated
Dec 17, 2024
•
7.47k
•
20
DatPySci/weak_Llama-3.2-3B_hh
Viewer
•
Updated
Dec 10, 2024
•
8k
•
9
DatPySci/weak_Llama-3.2-1B_hh
Viewer
•
Updated
Dec 10, 2024
•
8k
•
9
DatPySci/weak_gpt2-large_hh
Viewer
•
Updated
Dec 10, 2024
•
8k
•
10
DatPySci/weak_gpt2-medium_hh
Viewer
•
Updated
Dec 10, 2024
•
8k
•
11
DatPySci/weak_gpt2_hh
Viewer
•
Updated
Dec 10, 2024
•
8k
•
12
DatPySci/weak_gpt2-large_tldr_synthetic
Viewer
•
Updated
Nov 24, 2024
•
115k
•
11
DatPySci/weak_gpt2-medium_tldr_synthetic
Viewer
•
Updated
Nov 24, 2024
•
115k
•
14
DatPySci/weak_gpt2_tldr_synthetic
Viewer
•
Updated
Nov 24, 2024
•
115k
•
26
DatPySci/synthetic_tldr_sft
Viewer
•
Updated
Nov 18, 2024
•
50k
•
61
DatPySci/synthetic_tldr_step_72000
Viewer
•
Updated
Nov 18, 2024
•
50k
•
78
DatPySci/synthetic_tldr_step_32400
Viewer
•
Updated
Nov 18, 2024
•
50k
•
14
Previous
1
2
3
Next