Liang
Andynsn
·
AI & ML interests
None yet
Recent Activity
updated a collection 22 days ago
DLLM updated a collection about 1 month ago
latent updated a collection about 2 months ago
self - evolveOrganizations
RL
RL
-
Revisiting On-Policy Distillation: Empirical Failure Modes and Simple Fixes
Paper • 2603.25562 • Published • 19 -
Self-Distilled Agentic Reinforcement Learning
Paper • 2605.15155 • Published • 118 -
DRAGON: Distributional Rewards Optimize Diffusion Generative Models
Paper • 2504.15217 • Published • 11 -
Diffusion Policy Policy Optimization
Paper • 2409.00588 • Published • 20
DLLM
-
Continuous Latent Diffusion Language Model
Paper • 2605.06548 • Published • 85 -
SANA-WM: Efficient Minute-Scale World Modeling with Hybrid Linear Diffusion Transformer
Paper • 2605.15178 • Published • 91 -
A^2RD: Agentic Autoregressive Diffusion for Long Video Consistency
Paper • 2605.06924 • Published • 16 -
Diffusion Policy Policy Optimization
Paper • 2409.00588 • Published • 20
Agentic
latent
-
Unified Latents (UL): How to train your latents
Paper • 2602.17270 • Published • 62 -
The Latent Space: Foundation, Evolution, Mechanism, Ability, and Outlook
Paper • 2604.02029 • Published • 152 -
Dynamic Large Concept Models: Latent Reasoning in an Adaptive Semantic Space
Paper • 2512.24617 • Published • 67 -
Cross-Tokenizer LLM Distillation through a Byte-Level Interface
Paper • 2604.07466 • Published • 7
self - evolve
memory
rec
latent
-
Unified Latents (UL): How to train your latents
Paper • 2602.17270 • Published • 62 -
The Latent Space: Foundation, Evolution, Mechanism, Ability, and Outlook
Paper • 2604.02029 • Published • 152 -
Dynamic Large Concept Models: Latent Reasoning in an Adaptive Semantic Space
Paper • 2512.24617 • Published • 67 -
Cross-Tokenizer LLM Distillation through a Byte-Level Interface
Paper • 2604.07466 • Published • 7
RL
RL
-
Revisiting On-Policy Distillation: Empirical Failure Modes and Simple Fixes
Paper • 2603.25562 • Published • 19 -
Self-Distilled Agentic Reinforcement Learning
Paper • 2605.15155 • Published • 118 -
DRAGON: Distributional Rewards Optimize Diffusion Generative Models
Paper • 2504.15217 • Published • 11 -
Diffusion Policy Policy Optimization
Paper • 2409.00588 • Published • 20
self - evolve
DLLM
-
Continuous Latent Diffusion Language Model
Paper • 2605.06548 • Published • 85 -
SANA-WM: Efficient Minute-Scale World Modeling with Hybrid Linear Diffusion Transformer
Paper • 2605.15178 • Published • 91 -
A^2RD: Agentic Autoregressive Diffusion for Long Video Consistency
Paper • 2605.06924 • Published • 16 -
Diffusion Policy Policy Optimization
Paper • 2409.00588 • Published • 20
memory
Agentic