Back to feed
arXiv cs.CL
arXiv cs.CL
7/23/2026
Task Competence Is Not Instruction Following: Evaluating Instruction-Conflicting Behavior in Small Language Models

Task Competence Is Not Instruction Following: Evaluating Instruction-Conflicting Behavior in Small Language Models

Short summary

This study evaluates whether small language models follow instructions that conflict with their usual task behavior across MCQA, sentiment classification, and math QA tasks. Using an Instruction-Following Failure Rate (IFFR) metric, the authors find that small models stay competent but routinely ignore conflicting instructions, while larger models show a clearer gap between standard and non-standard instruction compliance. The key finding is that task competence and instruction following are distinct abilities, and standard accuracy metrics hide instruction-following failures.

  • Small models maintain task competence but routinely ignore conflicting non-standard instructions
  • Larger models show a clearer gap between standard accuracy and instruction-following compliance
  • Task competence and instruction following are distinct abilities; standard accuracy masks IFFR failures

Generated with AI, which can make mistakes.

Is this a good recommendation for you?

Comments

Failed to load comments. Please try again.

Explore more