fulong ye
Alon77777
AI & ML interests
vision and language, diffusion model, text-to-image generation, image-to-text generation, referring expression generation and comprehension
Recent Activity
upvoted a paper about 2 hours ago
LightNav-0: Eliciting VLM Spatial Intelligence for Generalist Embodied Navigation authored a paper 5 months ago
SPRING: Situated Conversation Agent Pretrained with Multimodal Questions
from Incremental Layout Graph authored a paper 5 months ago
AltCLIP: Altering the Language Encoder in CLIP for Extended Language
Capabilities