Aleksei Dorkin PRO
adorkin
AI & ML interests
Computational Linguistics
Recent Activity
liked
a model
about 17 hours ago
tokyotech-llm/Qwen3-Swallow-8B-SFT-v0.2
liked
a model
about 17 hours ago
tokyotech-llm/Qwen3-Swallow-8B-CPT-v0.2
upvoted
a
paper
about 19 hours ago
Project Alexandria: Towards Freeing Scientific Knowledge from Copyright
Burdens via LLMs
Organizations
Math Datasets
Code SFT Datasets
Llama 3(.1) 8B Finetunes
-
nvidia/Llama-3.1-Nemotron-Nano-8B-v1
Text Generation • Updated • 89.5k • • 218 -
nvidia/llama-3.1-nemoguard-8b-topic-control
Text Classification • Updated • 501 • 16 -
nvidia/llama-3.1-nemoguard-8b-content-safety
Text Classification • Updated • 1.25k • 32 -
nvidia/Llama-3.1-Nemotron-Nano-VL-8B-V1
Image-Text-to-Text • Updated • 896k • 175
Multilingual Text Embedding Models
Code RL Datasets
Reward Models
-
nvidia/Llama-3.3-Nemotron-70B-Reward-Multilingual
Text Generation • 71B • Updated • 49 • 10 -
nvidia/Llama-3.3-Nemotron-70B-Reward-Principle
Text Generation • 71B • Updated • 218 • 6 -
nvidia/Qwen-3-Nemotron-32B-Reward
Text Classification • 32B • Updated • 972 • 19 -
Skywork/Skywork-Reward-V2-Llama-3.1-8B
Text Classification • 8B • Updated • 19k • 38
My Shared Task Papers
-
TartuNLP @ SIGTYP 2024 Shared Task: Adapting XLM-RoBERTa for Ancient and Historical Languages
Paper • 2404.12845 • Published -
TartuNLP at EvaLatin 2024: Emotion Polarity Detection
Paper • 2405.01159 • Published -
TartuNLP @ AXOLOTL-24: Leveraging Classifier Output for New Sense Detection in Lexical Semantics
Paper • 2407.03861 • Published -
TartuNLP at SemEval-2025 Task 5: Subject Tagging as Two-Stage Information Retrieval
Paper • 2504.21547 • Published
Multilingual Text Encoders
Multilingual Text Embedding Models
Math Datasets
Code RL Datasets
Code SFT Datasets
Reward Models
-
nvidia/Llama-3.3-Nemotron-70B-Reward-Multilingual
Text Generation • 71B • Updated • 49 • 10 -
nvidia/Llama-3.3-Nemotron-70B-Reward-Principle
Text Generation • 71B • Updated • 218 • 6 -
nvidia/Qwen-3-Nemotron-32B-Reward
Text Classification • 32B • Updated • 972 • 19 -
Skywork/Skywork-Reward-V2-Llama-3.1-8B
Text Classification • 8B • Updated • 19k • 38
Llama 3(.1) 8B Finetunes
-
nvidia/Llama-3.1-Nemotron-Nano-8B-v1
Text Generation • Updated • 89.5k • • 218 -
nvidia/llama-3.1-nemoguard-8b-topic-control
Text Classification • Updated • 501 • 16 -
nvidia/llama-3.1-nemoguard-8b-content-safety
Text Classification • Updated • 1.25k • 32 -
nvidia/Llama-3.1-Nemotron-Nano-VL-8B-V1
Image-Text-to-Text • Updated • 896k • 175
My Shared Task Papers
-
TartuNLP @ SIGTYP 2024 Shared Task: Adapting XLM-RoBERTa for Ancient and Historical Languages
Paper • 2404.12845 • Published -
TartuNLP at EvaLatin 2024: Emotion Polarity Detection
Paper • 2405.01159 • Published -
TartuNLP @ AXOLOTL-24: Leveraging Classifier Output for New Sense Detection in Lexical Semantics
Paper • 2407.03861 • Published -
TartuNLP at SemEval-2025 Task 5: Subject Tagging as Two-Stage Information Retrieval
Paper • 2504.21547 • Published