Huiqiang Jiang's picture

Huiqiang Jiang

iofu728

·

https://hqjiang.com/

AI & ML interests

None yet

Recent Activity

authored a paper about 2 months ago

Chain-of-Model Learning for Language Model

upvoted a paper about 2 months ago

Chain-of-Model Learning for Language Model

authored a paper about 2 months ago

RetroInfer: A Vector-Storage Approach for Scalable Long-Context LLM Inference

View all activity

Organizations

authored 2 papers about 2 months ago

Chain-of-Model Learning for Language Model

Paper • 2505.11820 • Published May 17 • 119

RetroInfer: A Vector-Storage Approach for Scalable Long-Context LLM Inference

Paper • 2505.02922 • Published May 5 • 27

authored a paper 7 months ago

SCBench: A KV Cache-Centric Analysis of Long-Context Methods

Paper • 2412.10319 • Published Dec 13, 2024 • 10

authored a paper 10 months ago

RetrievalAttention: Accelerating Long-Context LLM Inference via Vector Retrieval

Paper • 2409.10516 • Published Sep 16, 2024 • 44

authored a paper about 1 year ago

MInference 1.0: Accelerating Pre-filling for Long-Context LLMs via Dynamic Sparse Attention

Paper • 2407.02490 • Published Jul 2, 2024 • 28

authored 3 papers over 1 year ago

LLMLingua-2: Data Distillation for Efficient and Faithful Task-Agnostic Prompt Compression

Paper • 2403.12968 • Published Mar 19, 2024 • 26

LongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt Compression

Paper • 2310.06839 • Published Oct 10, 2023 • 3

LLMLingua: Compressing Prompts for Accelerated Inference of Large Language Models

Paper • 2310.05736 • Published Oct 9, 2023 • 4