Delin Qu's picture

4 7 7

Delin Qu

delinqu

·

https://delinqu.github.io/

AI & ML interests

Embodied AI, 3D Vision

Recent Activity

updated a model 1 day ago

IPEC-COMMUNITY/spatialvla-4b-mix-224-pt

updated a model 1 day ago

IPEC-COMMUNITY/spatialvla-4b-224-pt

upvoted a paper 6 days ago

SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features

View all activity

Organizations

delinqu's activity

upvoted a paper 6 days ago

SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features

Paper • 2502.14786 • Published 8 days ago • 118

upvoted an article 8 days ago

Article

π0 and π0-FAST: Vision-Language-Action Models for General Robot Control

25 days ago

• 109

upvoted a paper 8 days ago

LiveScene: Language Embedding Interactive Radiance Fields for Physical Scene Rendering and Control

Paper • 2406.16038 • Published Jun 23, 2024 • 1

upvoted a paper 11 days ago

Large Language Diffusion Models

Paper • 2502.09992 • Published 14 days ago • 94

upvoted a collection 14 days ago

Foundation Vision-language-action Model

2 items • Updated 4 days ago • 3

upvoted 2 papers 14 days ago

SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model

Paper • 2501.15830 • Published Jan 27 • 14

Exploring the Potential of Encoder-free Architectures in 3D LMMs

Paper • 2502.09620 • Published 15 days ago • 25