-
BytedTsinghua-SIA/Sequential-Qwen3-1.7B
Text Generation • 2B • Updated • 288 -
BytedTsinghua-SIA/QuestA-R1-7B
Text Generation • 8B • Updated • 801 -
BytedTsinghua-SIA/JustRL-R1-7B
Text Generation • 8B • Updated • 706 -
BytedTsinghua-SIA/JustRL-Qwen3-4B
Text Generation • 4B • Updated • 707
AI & ML interests
None defined yet.
Recent Activity
View all activity
Resources for the Enigmata Project: https://seed-enigmata.github.io.
-
BytedTsinghua-SIA/Enigmata-Qwen2.5-32B
33B • Updated • 12 • 3 -
Enigmata: Scaling Logical Reasoning in Large Language Models with Synthetic Verifiable Puzzles
Paper • 2505.19914 • Published • 46 -
BytedTsinghua-SIA/Enigmata-Eval
Viewer • Updated • 4.76k • 1.32k • 2 -
BytedTsinghua-SIA/Enigmata-Data
Preview • Updated • 205 • 4
memo
-
BytedTsinghua-SIA/RL-MemoryAgent-7B
8B • Updated • 336 • 8 -
BytedTsinghua-SIA/RL-MemoryAgent-14B
15B • Updated • 3.28k • 31 -
BytedTsinghua-SIA/hotpotqa
Updated • 418 • 14 -
MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent
Paper • 2507.02259 • Published • 6
-
BytedTsinghua-SIA/DAPO-Math-17k
Viewer • Updated • 1.79M • 13.5k • 185 -
BytedTsinghua-SIA/AIME-2024
Viewer • Updated • 960 • 5.67k • 11 -
BytedTsinghua-SIA/DAPO-Qwen-32B
Text Generation • 33B • Updated • 68 • • 12 -
DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Paper • 2503.14476 • Published • 146
-
BytedTsinghua-SIA/Sequential-Qwen3-1.7B
Text Generation • 2B • Updated • 288 -
BytedTsinghua-SIA/QuestA-R1-7B
Text Generation • 8B • Updated • 801 -
BytedTsinghua-SIA/JustRL-R1-7B
Text Generation • 8B • Updated • 706 -
BytedTsinghua-SIA/JustRL-Qwen3-4B
Text Generation • 4B • Updated • 707
memo
-
BytedTsinghua-SIA/RL-MemoryAgent-7B
8B • Updated • 336 • 8 -
BytedTsinghua-SIA/RL-MemoryAgent-14B
15B • Updated • 3.28k • 31 -
BytedTsinghua-SIA/hotpotqa
Updated • 418 • 14 -
MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent
Paper • 2507.02259 • Published • 6
Resources for the Enigmata Project: https://seed-enigmata.github.io.
-
BytedTsinghua-SIA/Enigmata-Qwen2.5-32B
33B • Updated • 12 • 3 -
Enigmata: Scaling Logical Reasoning in Large Language Models with Synthetic Verifiable Puzzles
Paper • 2505.19914 • Published • 46 -
BytedTsinghua-SIA/Enigmata-Eval
Viewer • Updated • 4.76k • 1.32k • 2 -
BytedTsinghua-SIA/Enigmata-Data
Preview • Updated • 205 • 4
-
BytedTsinghua-SIA/DAPO-Math-17k
Viewer • Updated • 1.79M • 13.5k • 185 -
BytedTsinghua-SIA/AIME-2024
Viewer • Updated • 960 • 5.67k • 11 -
BytedTsinghua-SIA/DAPO-Qwen-32B
Text Generation • 33B • Updated • 68 • • 12 -
DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Paper • 2503.14476 • Published • 146