Loading the SOTA2 catalog…
PruneHal: Reducing Hallucinations in Multi-modal Large Language Models through Adaptive KV Cache Pruning · SOTA2 Research