arxiv:2606.11176
🤝 Open to Collab
Kevin Lin
KevinQHLin
AI & ML interests
Vision-Language Model, Video Understanding, Agent
Recent Activity
upvoted a paper 7 days ago
HumanCLAW: Can Vision-Language Models Act Through a Body? upvoted a paper 30 days ago
EdgeBench: Unveiling Scaling Laws of Learning from Real-World Environments upvoted a paper 30 days ago
Parallelized Autoregressive Decoding for Omni-Modal Dense Video Captioning