-
BytedTsinghua-SIA/Sequential-Qwen3-1.7B
Text Generation • 2B • Updated • 271 -
BytedTsinghua-SIA/QuestA-R1-7B
Text Generation • 8B • Updated • 787 -
BytedTsinghua-SIA/JustRL-R1-7B
Text Generation • 8B • Updated • 691 -
BytedTsinghua-SIA/JustRL-Qwen3-4B
Text Generation • 4B • Updated • 686
AI & ML interests
None defined yet.
Recent Activity
View all activity
Resources for the Enigmata Project: https://seed-enigmata.github.io.
-
BytedTsinghua-SIA/Enigmata-Qwen2.5-32B
33B • Updated • 11 • 3 -
Enigmata: Scaling Logical Reasoning in Large Language Models with Synthetic Verifiable Puzzles
Paper • 2505.19914 • Published • 46 -
BytedTsinghua-SIA/Enigmata-Eval
Viewer • Updated • 4.76k • 1.54k • 2 -
BytedTsinghua-SIA/Enigmata-Data
Preview • Updated • 211 • 4
memo
-
BytedTsinghua-SIA/RL-MemoryAgent-7B
8B • Updated • 265 • 8 -
BytedTsinghua-SIA/RL-MemoryAgent-14B
15B • Updated • 3.25k • 31 -
BytedTsinghua-SIA/hotpotqa
Updated • 436 • 14 -
MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent
Paper • 2507.02259 • Published • 6
-
BytedTsinghua-SIA/DAPO-Math-17k
Viewer • Updated • 1.79M • 12.5k • 185 -
BytedTsinghua-SIA/AIME-2024
Viewer • Updated • 960 • 5.34k • 11 -
BytedTsinghua-SIA/DAPO-Qwen-32B
Text Generation • 33B • Updated • 71 • • 12 -
DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Paper • 2503.14476 • Published • 146
-
BytedTsinghua-SIA/Sequential-Qwen3-1.7B
Text Generation • 2B • Updated • 271 -
BytedTsinghua-SIA/QuestA-R1-7B
Text Generation • 8B • Updated • 787 -
BytedTsinghua-SIA/JustRL-R1-7B
Text Generation • 8B • Updated • 691 -
BytedTsinghua-SIA/JustRL-Qwen3-4B
Text Generation • 4B • Updated • 686
memo
-
BytedTsinghua-SIA/RL-MemoryAgent-7B
8B • Updated • 265 • 8 -
BytedTsinghua-SIA/RL-MemoryAgent-14B
15B • Updated • 3.25k • 31 -
BytedTsinghua-SIA/hotpotqa
Updated • 436 • 14 -
MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent
Paper • 2507.02259 • Published • 6
Resources for the Enigmata Project: https://seed-enigmata.github.io.
-
BytedTsinghua-SIA/Enigmata-Qwen2.5-32B
33B • Updated • 11 • 3 -
Enigmata: Scaling Logical Reasoning in Large Language Models with Synthetic Verifiable Puzzles
Paper • 2505.19914 • Published • 46 -
BytedTsinghua-SIA/Enigmata-Eval
Viewer • Updated • 4.76k • 1.54k • 2 -
BytedTsinghua-SIA/Enigmata-Data
Preview • Updated • 211 • 4
-
BytedTsinghua-SIA/DAPO-Math-17k
Viewer • Updated • 1.79M • 12.5k • 185 -
BytedTsinghua-SIA/AIME-2024
Viewer • Updated • 960 • 5.34k • 11 -
BytedTsinghua-SIA/DAPO-Qwen-32B
Text Generation • 33B • Updated • 71 • • 12 -
DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Paper • 2503.14476 • Published • 146