SurviAGI
← All updates
Tencent · 2026-09-24 · Research result

Tencent Hunyuan research extends critical-batch-size theory to online LLM RL

Tencent Hunyuan published new research revisiting classical critical-batch-size theory and extending it to online LLM reinforcement learning.

This update bears on 1 kinds of work in 1 markets. Highest level accepted: L1 Assisted.

Highest accepted
L1Assisted
Work
1
Markets
1
Views
14.4K

The work it bears on

Highest level first. Open a kind of work to see everything else that moved it.

Published at