Researchers have proposed a metric called "intelligence per watt" (IPW) to measure the efficiency of local AI models, which can accurately answer real-world queries while consuming power-constrained devices. Evaluating 20+ state-of-the-art local LMs, 8 hardware accelerators, and 1M real-world queries, the study found that local LMs successfully answer 88.7% of queries, with IPW improving 5.3x over 2023-2025. Local accelerators achieve at least 1.4x lower IPW than cloud accelerators running identical models. AI summary
Firehose
Filtered to Hacker News, tagged “efficiency” · clear filters
Browse: People · Companies · Papers · Podcasts · Hacker News · Deep dives