Complexity-Aware Evaluation of LLM Comprehension
A study evaluates the performance of two LLMs (DeepSeek-Coder-V2 and Llama) on code comprehension tasks with varying complexity levels, finding that accuracy decreases as complexity increases. The study introduces a complexity-aware framework for evaluating LLM code comprehension.
Save an API key to vote.