PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
February 20, 2026Machine Learning and Knowledge Extraction2 citationsOpen Access

Introducing LEAF: LLM Edge Assessment Framework for Generative AI on the Edge

View Full Paper
MAMustafa AbdulkadhimSRSándor R. Répás

Key Points

  • This research aims to create a framework that benchmarks generative AI for edge computing, focusing on sustainability and performance.
  • Introduced the LEAF framework for evaluating edge deployments of generative AI.
  • Assessed performance across five metrics: Circular Economy Score, Energy Efficiency, Performance Speed, Semantic Accuracy, and End-to-End Latency.
  • Conducted experimental analysis using various hardware classes, including embedded IoT devices and professional edge servers.
  • Repurposed consumer hardware outperformed modern edge SoCs in speed and efficiency.
  • The legacy NVIDIA GTX 1050 Ti demonstrated a 20× speedup over the Raspberry Pi 4.
  • Achieved better energy-per-task efficiency compared to low-power ARM architectures.

Abstract

The transition of Large Language Models (LLMs) from centralized clouds to edge environments is critical for addressing privacy concerns, latency bottlenecks, and operational costs. However, existing edge benchmarking frameworks remain tailored to discriminative Deep Learning tasks (e.g., object detection), failing to capture the multidimensional challenges of generative AI, specifically the trade-offs between token generation speed, semantic accuracy, and hardware sustainability. To address this gap, we introduce LEAF (LLM Edge Assessment Framework), a novel evaluation methodology that integrates Circular Economy principles directly into performance metrics. LEAF assesses edge deployments across five synergistic pillars: Circular Economy Score, Energy Efficiency (Joules/Token), Performance Speed (Tokens/Second), semantic accuracy (BERTScore), and End-to-End Latency. We validate LEAF through an extensive experimental analysis of five distinct hardware classes, ranging from embedded IoT devices (Raspberry Pi 4 and 5, NVIDIA Jetson Nano) to professional edge servers (NVIDIA T400) and repurposed legacy workstations (NVIDIA GTX 1050 Ti). Utilizing 4-bit quantized models via the Ollama runtime, our results reveal a counterintuitive insight: repurposed consumer hardware significantly outperforms modern purpose-built edge SoCs. The legacy GTX 1050 Ti achieved a 20× speedup over the Raspberry Pi 4 and maintained superior energy-per-task efficiency compared to low-power ARM architectures by minimizing active runtime. These findings challenge the prevailing narrative that newer silicon is essential for Edge AI, demonstrating that sustainable, high-performance inference can be achieved by extending the lifecycle of existing hardware. LEAF thus provides a blueprint for a “Green Edge” ecosystem that balances computational capability with environmental responsibility.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Abdulkadhim et al. (2026) studied this question.

synapsesocial.com/papers/6997f9b8ad1d9b11b3452648https://doi.org/10.3390/make8020048
Ask AI
Helpful
Bookmark
Share
View Full Paper