Research1 min read
PerfReasoning benchmark evaluates LLMs on hardware performance reasoning
PerfReasoning assesses LLMs as performance reasoners and code generators, with models achieving up to 90% accuracy on reasoning tasks. Construction of models remains challenging, especially for smaller configurations.
From arXiv cs.AI
