9786 条条目 · 106 个活跃源
2026年9月11日
04:00
arXiv cs.CL

Automated Identification of Competing Narratives in Political Discourse on Social Media

04:00
arXiv cs.CL

OmniHallu: Unified Hallucination Detection for Cross-Modal Comprehension and Generation in Multimodal Large Language Models

04:00
arXiv cs.CL

MultiHuSE: A Multimodal Dataset for Humour Styles and Emotions

04:00
arXiv cs.CL

ReGround: Grounding Reviewer Comments in Multimodal Evidence

04:00
arXiv cs.CL

Automatic Lyric Transcription for Greek Songs: Scaling and Task Composition Effects in Whisper Adaptation

04:00
arXiv cs.CL

The Illusion of Balanced Multimodal Sentiment Analysis: Beyond the Limits of Optimization-Based Methods

04:00
arXiv cs.CL

SEAR: Segment-Evidence-Aware Routing for Weak-to-Strong Multilingual Speech MCQ

04:00
arXiv cs.CL

E-CONAN (Entailment, CONtradition And Neutral) Benchmarks: Arabic Textual Entailment and Natural Inference Datasets

04:00
arXiv cs.CL

On the Impact of Anonymization on the Performance of Large Language Models

04:00
arXiv cs.CL

SWRouter: Similarity-Contractive Window Routing for Multi-Turn Large Language Model Conversations

04:00
arXiv cs.CL

TransClean: A Benchmark for Detecting and Extracting Clean Translations from Large Language Model Outputs

04:00
arXiv cs.CL

Cross-Lingual Clinical Annotation Projection as Constrained Text Generation: A Six-Language Study

04:00
arXiv cs.CL

Structural priors for data-efficient language learning

04:00
arXiv cs.CL

Complex-Text Robustness Evaluation and Failure Diagnosis for Low-Resource Multilingual Text-to-Speech

04:00
arXiv cs.CL

A Training-Free, Alignment-Free Approach to Corporate Intelligence: Application to SEC Filings

04:00
arXiv cs.CL

Recognizing Is Not Reversing: A Controlled Inversion Test of Fact-Preserving News Framing

04:00
arXiv cs.CL

Structured Transforms for Low-Overhead Quantization of Language Models

04:00
arXiv cs.CL

Negative Self-Distillation: Learning to Reason by Avoiding Flaws

04:00
arXiv cs.CL

The Eloquence submission for Task 2 of the Interspeech 2026 MLC-SLM challenge

04:00
arXiv cs.CL

LOCUS: Task-Aware Low-Rank Post-Training for Token-Efficient Language Generation

04:00
arXiv cs.CL

RAG-Safety-Bench: Reliable Evaluation of Retrieval-Augmented LLM Safety

04:00
arXiv cs.CL

Component-Aware Differential Privacy for Federated Multilingual Speech-LLMs

04:00
arXiv cs.CL

More than half of recent astronomy papers are written with language-model assistance

04:00
arXiv cs.CL

Beyond Word Error Rate: A Switch Aware Evaluation of ASR and Audio Language Models on English Yoruba Code-Switched Speech

04:00
arXiv cs.CL

Epistemic orientation predicts legislative effectiveness among members of the US Congress

04:00
arXiv cs.CL

Target leakage, not model class, explains reported accuracy in survey-based cardiovascular screening: a leakage-tiered audit of glass-box and tabular foundation models

04:00
arXiv cs.CL

The widening evaluation gap in medical large language model research 2023 to 2026

04:00
arXiv cs.CL

Domain-Specific Hallucination Detection in Large Language Models

04:00
arXiv cs.CL

Nuha-Speech: Building General-Purpose Arabic Speech-LLMs

04:00
arXiv cs.CL

Augustinian BabyLM: What Ostensive Definition Can and Cannot Teach a Small Language Model

04:00
arXiv cs.CL

Distance generalization in transformers: why bother with positional encoding?

04:00
arXiv cs.CL

IndicTriMix: Developing Language Identification Datasets and Models for Tri-Language Code-Mixing

04:00
arXiv cs.CL

Studying Without a Syllabus: Task-Agnostic Environment Preprocessing

04:00
arXiv cs.CL

Empirical Evaluation of Membership Inference Attacks on NLP Text Classifiers: A Baseline Study on SST-2

04:00
arXiv cs.CL

The Semantic Elevation Operator and the Closure of the Undecidable Class under Preservation

04:00
arXiv cs.CL

(Whose defaults?) Is artificial intelligence reorienting archaeological methods?

04:00
arXiv cs.CL

KuaiRP Series Role-playing Models Technical Report

04:00
arXiv cs.CL

Same Day, Same Story; One Day Ahead, a Different Signal: The Dual Validity of Financial Sentiment

04:00
arXiv cs.CL

The Oligarch Barely Steers Model Collapse in Multi-Model Ecosystems

04:00
arXiv cs.CL

INDRA: A New AI Tool for Exploring Tobacco, Fossil Fuel, and Chemical Industry Archives

04:00
arXiv cs.CL

A Voice-Interactive Multi-Agent System for Smart Operating Rooms: Architecture Design and Key Technologies

04:00
arXiv cs.CL

Xiaomi-CocktailASR-1 Technical Report

04:00
arXiv cs.CL

VikingRAG: Accurate and Token-efficient Retrieval-augmented Generation over Structured Documents

04:00
arXiv cs.CL

A Short Survey of Viewing Large Language Models in Legal Aspect

04:00
arXiv cs.CL

SpecGuard: Inference-Time Backdoor Detection For Free

04:00
arXiv cs.CL

RetroThinker: Enabling Retrospective Thinking in Speech LLMs

04:00
arXiv cs.CL

A Unified Per-Token Gating Family for On-Policy Distillation: FKL/RKL Mixing with Multi-Channel and Bias Coefficients

04:00
arXiv cs.CL

Biology-in-the-loop: Amortized Adaptive Hit Discovery in CRISPR Screens

04:00
arXiv cs.CL

MindTopo: Can Foundation Models Reason in Topological Space?

04:00
arXiv cs.CL

SIRF: A Spec-Internalized Risk Foundation Model for Industrial Content Risk Control