-
An Empirical Study of Scaling Instruct-Tuned Large Multimodal Models
Paper • 2309.09958 • Published • 20 -
TextBind: Multi-turn Interleaved Multimodal Instruction-following
Paper • 2309.08637 • Published • 7 -
AnyMAL: An Efficient and Scalable Any-Modality Augmented Language Model
Paper • 2309.16058 • Published • 56 -
Qwen Technical Report
Paper • 2309.16609 • Published • 39
WANG Jiong
wjwow
·
AI & ML interests
None yet
Recent Activity
upvoted a paper 4 days ago
N_0-TWAM: Scaling Tactile-Native World-Action Model for Contact-Rich Manipulation upvoted a paper 4 days ago
N_0-VTLA: Scaling Vision-Tactile-Language-Action Model with Latent Tactile Tokens upvoted a paper 2 months ago
ResearchClawBench: A Benchmark for End-to-End Autonomous Scientific ResearchOrganizations
None yet