-
BytedTsinghua-SIA/Sequential-Qwen3-1.7B
Text Generation • 2B • Updated • 249 -
BytedTsinghua-SIA/QuestA-R1-7B
Text Generation • 8B • Updated • 739 -
BytedTsinghua-SIA/JustRL-R1-7B
Text Generation • 8B • Updated • 641 -
BytedTsinghua-SIA/JustRL-Qwen3-4B
Text Generation • 4B • Updated • 627
AI & ML interests
None defined yet.
Recent Activity
View all activity
Resources for the Enigmata Project: https://seed-enigmata.github.io.
-
BytedTsinghua-SIA/Enigmata-Qwen2.5-32B
33B • Updated • 9 • 3 -
Enigmata: Scaling Logical Reasoning in Large Language Models with Synthetic Verifiable Puzzles
Paper • 2505.19914 • Published • 46 -
BytedTsinghua-SIA/Enigmata-Eval
Viewer • Updated • 4.76k • 1.89k • 2 -
BytedTsinghua-SIA/Enigmata-Data
Preview • Updated • 223 • 4
memo
-
BytedTsinghua-SIA/RL-MemoryAgent-7B
8B • Updated • 348 • 8 -
BytedTsinghua-SIA/RL-MemoryAgent-14B
15B • Updated • 3.24k • 31 -
BytedTsinghua-SIA/hotpotqa
Updated • 451 • 14 -
MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent
Paper • 2507.02259 • Published • 6
-
BytedTsinghua-SIA/DAPO-Math-17k
Viewer • Updated • 1.79M • 10.6k • 185 -
BytedTsinghua-SIA/AIME-2024
Viewer • Updated • 960 • 4.51k • 11 -
BytedTsinghua-SIA/DAPO-Qwen-32B
Text Generation • 33B • Updated • 82 • • 12 -
DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Paper • 2503.14476 • Published • 146
-
BytedTsinghua-SIA/Sequential-Qwen3-1.7B
Text Generation • 2B • Updated • 249 -
BytedTsinghua-SIA/QuestA-R1-7B
Text Generation • 8B • Updated • 739 -
BytedTsinghua-SIA/JustRL-R1-7B
Text Generation • 8B • Updated • 641 -
BytedTsinghua-SIA/JustRL-Qwen3-4B
Text Generation • 4B • Updated • 627
memo
-
BytedTsinghua-SIA/RL-MemoryAgent-7B
8B • Updated • 348 • 8 -
BytedTsinghua-SIA/RL-MemoryAgent-14B
15B • Updated • 3.24k • 31 -
BytedTsinghua-SIA/hotpotqa
Updated • 451 • 14 -
MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent
Paper • 2507.02259 • Published • 6
Resources for the Enigmata Project: https://seed-enigmata.github.io.
-
BytedTsinghua-SIA/Enigmata-Qwen2.5-32B
33B • Updated • 9 • 3 -
Enigmata: Scaling Logical Reasoning in Large Language Models with Synthetic Verifiable Puzzles
Paper • 2505.19914 • Published • 46 -
BytedTsinghua-SIA/Enigmata-Eval
Viewer • Updated • 4.76k • 1.89k • 2 -
BytedTsinghua-SIA/Enigmata-Data
Preview • Updated • 223 • 4
-
BytedTsinghua-SIA/DAPO-Math-17k
Viewer • Updated • 1.79M • 10.6k • 185 -
BytedTsinghua-SIA/AIME-2024
Viewer • Updated • 960 • 4.51k • 11 -
BytedTsinghua-SIA/DAPO-Qwen-32B
Text Generation • 33B • Updated • 82 • • 12 -
DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Paper • 2503.14476 • Published • 146