CAF-分数:用 LALM 校正 CLAP 用于无参考音频字幕评价
CAF-Score: Calibrating CLAP with LALMs for Reference-free Audio Captioning Evaluation

作者
Authors Insung Lee|Taeyoung Jeong|Haejun Yoo|Du-Seong Chang|Myoung-Wan Koo

期刊
Journal arXiv

年份
Year 2026

分类
Category 人工智能
Artificial Intelligence

国家
Country 德国Germany

🔗 访问原文
🔗 Access Paper

📝 摘要
Abstract

While Large Audio-Language Models (LALMs) have advanced audio captioning, robust evaluation remains difficult. Reference-based metrics are expensive and often fail to assess acoustic fidelity, while Contrastive Language-Audio Pretraining (CLAP)-based approaches frequently overlook syntactic errors and fine-grained details. We propose CAF-Score, a reference-free metric that calibrates CLAP's coarse-grained semantic alignment with the fine-grained comprehension and syntactic awareness of LALMs. By combining contrastive audio-text embeddings with LALM reasoning, CAF-Score effectively detects syntactic inconsistencies and subtle hallucinations. Experiments on the BRACE benchmark demonstrate that our approach achieves the highest correlation with human judgments, even outperforming reference-based baselines in challenging scenarios. These results highlight the efficacy of CAF-Score for reference-free audio captioning evaluation. Code and results are available at https://github.com/inseong00/CAF-Score.

📊 文章统计
Article Statistics

基础数据
Basic Stats

430 浏览
Views

0 下载
Downloads

39 引用
Citations

引用趋势
Citation Trend

阅读国家分布
Country Distribution

阅读机构分布
Institution Distribution

月度浏览趋势
Monthly Views

影响因子分析
Impact Analysis

6.90 综合评分
Overall Score

引用影响力
Citation Impact

浏览热度
View Popularity

下载频次
Download Frequency

CAF-分数:用 LALM 校正 CLAP 用于无参考音频字幕评价
CAF-Score: Calibrating CLAP with LALMs for Reference-free Audio Captioning Evaluation

📝 摘要
Abstract

📊 文章统计
Article Statistics

基础数据
Basic Stats

引用趋势
Citation Trend

阅读国家分布
Country Distribution

阅读机构分布
Institution Distribution

月度浏览趋势
Monthly Views

相关关键词
Related Keywords

影响因子分析
Impact Analysis

📄 相关文章
Related Articles

CAF-分数:用 LALM 校正 CLAP 用于无参考音频字幕评价CAF-Score: Calibrating CLAP with LALMs for Reference-free Audio Captioning Evaluation

📝 摘要Abstract

📊 文章统计Article Statistics

基础数据Basic Stats

引用趋势Citation Trend

阅读国家分布Country Distribution

阅读机构分布Institution Distribution

月度浏览趋势Monthly Views

相关关键词Related Keywords

影响因子分析Impact Analysis

📄 相关文章Related Articles

海洋智能分析Ocean AI Analysis

CAF-分数:用 LALM 校正 CLAP 用于无参考音频字幕评价
CAF-Score: Calibrating CLAP with LALMs for Reference-free Audio Captioning Evaluation

📝 摘要
Abstract

📊 文章统计
Article Statistics

基础数据
Basic Stats

引用趋势
Citation Trend

阅读国家分布
Country Distribution

阅读机构分布
Institution Distribution

月度浏览趋势
Monthly Views

相关关键词
Related Keywords

影响因子分析
Impact Analysis

📄 相关文章
Related Articles