LensVLM: Selective Context Expansion for Compressed Visual Representation of Text Paper • 2605.07019 • Published May 7 • 8
view article Article Fine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps +1 iamleonie, burtenshaw, sergiopaniego • 28 days ago • 141
Terminal-Universe: Turning Agent Trajectories into Scalable Terminal Environments Paper • 2609.04148 • Published 28 days ago • 248
RRSI: Regularized Recursive Self-Improvement of Agent Harnesses Paper • 2609.24972 • Published 10 days ago • 219
It Takes Two to Match: Co-Evolving Generative Retriever with Reinforcement Learning Paper • 2609.00638 • Published about 1 month ago • 66
DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression Paper • 2609.19969 • Published 14 days ago • 191
On the Design of Qwen3.8-Next Architecture: Evaluation, Efficiency, and Training Stability Paper • 2608.30320 • Published Aug 31 • 63
view article Article NeoMME: an efficient Multimodal-native and Multilingual Encoder Hcompany • 28 days ago • 120