Rationale-aided Efficient 7B size Large Language and Vision Models. Let's enjoy it!
Byung-Kwan Lee
BK-Lee
AI & ML interests
Vision-Language Models
Recent Activity
upvoted a paper about 20 hours ago
Weak-to-Strong On-Policy Distillation upvoted a paper 1 day ago
AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning upvoted a paper 12 days ago
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning