arXiv cs.CL
arXiv cs.CL

arXiv cs.CL

arXiv cs.CL publishes articles covering LLM, AI. A trusted source for AI and technology insights.

Profile generated by AI for Anything

Human-in-the-Loop Large Language Model Framework for Identification of Cutaneous Immune-Related Adverse Events

Human-in-the-Loop Large Language Model Framework for Identification of Cutaneous Immune-Related Adverse Events

21h

What is Good? Extracting and Testing Implicit Theories of Literary Quality from LLM Reasoning Traces

What is Good? Extracting and Testing Implicit Theories of Literary Quality from LLM Reasoning Traces

21h

Moir: Let the Model Direct Its Own Story for Robust Cross-Domain Knowledge Editing

Moir: Let the Model Direct Its Own Story for Robust Cross-Domain Knowledge Editing

21h

LLM-INSTRUCT at UZH Shared Task 2026: Constraint-Aware Retrieval and Selective Debate for Paragraph-Level Argument Mining

LLM-INSTRUCT at UZH Shared Task 2026: Constraint-Aware Retrieval and Selective Debate for Paragraph-Level Argument Mining

21h

Knowledge Injection Exists in MoE? Exploring Expert-Aware Contrast Decoding in MoE for Mitigating LLMs'Hallucinations

Knowledge Injection Exists in MoE? Exploring Expert-Aware Contrast Decoding in MoE for Mitigating LLMs'Hallucinations

21h

Is MoE Routing a Huffman Code? Discovering the Frequency-Diversity Law in Chain-of-Thought

Is MoE Routing a Huffman Code? Discovering the Frequency-Diversity Law in Chain-of-Thought

21h

More Is Not More: What Matters for Diversity in LLM Opinions?

More Is Not More: What Matters for Diversity in LLM Opinions?

21h

Position: Natural Language Should Not Fully Replace Formal Languages

Position: Natural Language Should Not Fully Replace Formal Languages

21h

The original title is "null" so I need to create a headline from the summary. The key facts are:

The original title is "null" so I need to create a headline from the summary. The key facts are:

21h

Skill-Contracted Agents for Evidence-Aware Materials Literature Analysis

Skill-Contracted Agents for Evidence-Aware Materials Literature Analysis

21h

Scaling Laws for Hypernetwork-Based Knowledge Injection in Large Language Models

Scaling Laws for Hypernetwork-Based Knowledge Injection in Large Language Models

1d

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play

1d

Multi-Mask Diffusion Language Models for Few-Step Generation

Multi-Mask Diffusion Language Models for Few-Step Generation

1d

Stateful Guardrails for Multi-Turn LLM Systems: A Conversational Risk Accumulation Framework

Stateful Guardrails for Multi-Turn LLM Systems: A Conversational Risk Accumulation Framework

1d

Reference-Free Evaluation of Reasoning in Open-Ended Question Answering

Reference-Free Evaluation of Reasoning in Open-Ended Question Answering

1d

SLPO: Scaling Latent Reasoning via a Surrogate Policy

SLPO: Scaling Latent Reasoning via a Surrogate Policy

1d

On the Computational Complexity of Structural Generalization

On the Computational Complexity of Structural Generalization

1d

Adaptive Capitulation: A Structural Failure Mode of LLM Responses in Vulnerability Contexts

Adaptive Capitulation: A Structural Failure Mode of LLM Responses in Vulnerability Contexts

1d

Lightweight Person-Place Relation Extraction from Historical Newspapers with Dependency Graphs and Proximity Features

Lightweight Person-Place Relation Extraction from Historical Newspapers with Dependency Graphs and Proximity Features

1d

Task Competence Is Not Instruction Following: Evaluating Instruction-Conflicting Behavior in Small Language Models

Task Competence Is Not Instruction Following: Evaluating Instruction-Conflicting Behavior in Small Language Models

1d

A Classifier That Teaches Itself: Self-Improving, Frozen-gate Training (SIFT) for Dynamic Document Classification

A Classifier That Teaches Itself: Self-Improving, Frozen-gate Training (SIFT) for Dynamic Document Classification

2d

Convolution for Large Language Models

Convolution for Large Language Models

2d

Building a European Multilingual Evaluation Dataset: The MMLU Localisation Project within the EMT Network

Building a European Multilingual Evaluation Dataset: The MMLU Localisation Project within the EMT Network

2d

Fine-tuned LLMs classify vulnerability in 3,000 UK police logs

Fine-tuned LLMs classify vulnerability in 3,000 UK police logs

2d