L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning
L3 Lab
university
AI & ML interests
None defined yet.
Recent Activity
View all activity
models 11
l3lab/L1-Qwen3-8B-Max
8B • Updated • 21
l3lab/L1-Qwen3-8B-Exact
8B • Updated • 39 • 1
l3lab/L1-Qwen-7B-Max
8B • Updated • 24
l3lab/L1-Qwen-7B-Exact
8B • Updated • 15 • 1
l3lab/L1-1.5B-Short
2B • Updated • 8
l3lab/all-distilroberta-v1-lr2e-4-bs256-nneg3-ml-ne2
Updated • 179
l3lab/L1-Qwen-1.5B-Exact
2B • Updated • 167 • 6
l3lab/L1-Qwen-1.5B-Max
2B • Updated • 113 • 15
l3lab/ntp-mathlib-context-deepseek-coder-1.3b
Text Generation • Updated • 96 • 3
l3lab/ntp-mathlib-st-deepseek-coder-1.3b
Text Generation • Updated • 58
datasets 10
l3lab/far-explorer-2026-09-30
Viewer • Updated • 103k • 69
l3lab/miniCTX-v2
Viewer • Updated • 668 • 138 • 3
l3lab/miniCTX-v2-data
Updated • 11
l3lab/Massive-Math-455K-Verified
Viewer • Updated • 455k • 160 • 1
l3lab/lean-premises
Updated • 136 • 3
l3lab/miniCTX
Viewer • Updated • 662 • 1.07k • 3
l3lab/ntp-mathlib-instruct-context-fullproof
Viewer • Updated • 144k • 52 • 1
l3lab/ntp-mathlib-instruct-context
Viewer • Updated • 614k • 109 • 1
l3lab/ntp-mathlib
Viewer • Updated • 213k • 100 • 2
l3lab/ntp-mathlib-instruct-st
Viewer • Updated • 307k • 92