Magma: A Foundation Model for Multimodal AI Agents Paper • 2502.13130 • Published 3 days ago • 42 • 4
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents Paper • 2502.09560 • Published 8 days ago • 32
microsoft/llmlingua-2-xlm-roberta-large-meetingbank Token Classification • Updated Jan 8 • 37.8k • 18
microsoft/llmlingua-2-bert-base-multilingual-cased-meetingbank Token Classification • Updated Jan 8 • 153k • 26
SCBench: A KV Cache-Centric Analysis of Long-Context Methods Paper • 2412.10319 • Published Dec 13, 2024 • 10
SCBench: A KV Cache-Centric Analysis of Long-Context Methods Paper • 2412.10319 • Published Dec 13, 2024 • 10
Wolf: Captioning Everything with a World Summarization Framework Paper • 2407.18908 • Published Jul 26, 2024 • 32
MindSearch: Mimicking Human Minds Elicits Deep AI Searcher Paper • 2407.20183 • Published Jul 29, 2024 • 42
view article Article MInference 1.0: 10x Faster Million Context Inference with a Single GPU By liyucheng • Jul 11, 2024 • 13
view article Article How to Optimize TTFT of 8B LLMs with 1M Tokens to 20s By iofu728 • Jul 21, 2024 • 2
MInference 1.0: Accelerating Pre-filling for Long-Context LLMs via Dynamic Sparse Attention Paper • 2407.02490 • Published Jul 2, 2024 • 23
MInference 1.0: Accelerating Pre-filling for Long-Context LLMs via Dynamic Sparse Attention Paper • 2407.02490 • Published Jul 2, 2024 • 23