A Comparative Analysis of ARM and x86-64 Laptop-Class Processors: Architecture, Assembly-Level Performance, and Energy Efficiency

📅 2026-04-20
📈 Citations: 0
✨ Influential: 0
📄 PDF
🤖 AI Summary
This study investigates the performance and energy efficiency differences between ARM and x86-64 laptop processors, demonstrating that these disparities stem not only from instruction set architecture (ISA) but also significantly from system-level design choices. For the first time, the authors conduct a comprehensive evaluation on real-world laptop platforms—Apple M3 and AMD Ryzen 7 3750H—combining fine-grained power measurements with microarchitectural analysis. Using assembly-level benchmarks (recursive Fibonacci, integer matrix multiplication), cross-platform performance counters, and portable C-based probes, they systematically assess the impact of architectural and integration factors on energy efficiency. Results reveal that while the Ryzen platform excels in branch-intensive workloads, the Apple platform achieves substantially superior energy efficiency, reducing energy-per-computation by 5.82× and 6.38× respectively, thereby highlighting the critical role of non-ISA design elements.

Technology Category

Machine Learning: Hardware-aware MLCognitive Modeling & Cognitive Systems: (Computational) Cognitive ArchitecturesHumans and AI: Other Foundations of Human Computation & AI

Application Category

Systems and Infrastructure for Web, Mobile and WoT: Web performance, measurement, and characterizationSearch and Retrieval-Augmented AI: Efficiency and scalability of Web search enginesSecurity and Privacy: Large-scale security measurements
📝 Abstract
ARM-based and x86-64 laptop processors differ not only in instruction-set design, but also in memory hierarchy, core organization, system integration, and power-management mechanisms. This study presents a combined architectural and experimental comparison of an Apple M3 system and an AMD Ryzen 7 3750H system. The architectural analysis contrasts AArch64's fixed-width load-store design with the variable-length, memory-operand-rich x86-64 instruction model, and discusses how register organization, calling conventions, heterogeneous core organization, memory behavior, and low-power mechanisms shape observed performance and energy characteristics. The experimental part uses two native assembly benchmarks: a recursive Fibonacci workload and an integer matrix-multiplication workload. The analysis combines repeated timing measurements, processor-energy measurements, and cross-platform microarchitectural counter measurements from matched portable-C profiling runs. The Ryzen platform is decisively faster on the branch-heavy Fibonacci benchmark, while matrix multiplication shows no meaningful timing advantage for either platform in the present measurements. In contrast, the Apple platform is markedly more energy-efficient, reducing energy-to-solution by approximately 5.82$\times$ on Fibonacci and 6.38$\times$ on matrix multiplication. These results are interpreted as platform-level findings rather than as pure ISA-only effects, reflecting differences in implementation, system integration, and measurement methodology in addition to instruction-set structure.
Problem

Research questions and friction points this paper is trying to address.

ARM
x86-64
performance
energy efficiency
processor architecture
Innovation

Methods, ideas, or system contributions that make the work stand out.

instruction set architecture (ISA)
energy efficiency
microarchitectural analysis
heterogeneous cores
cross-platform benchmarking
🔎 Similar Papers
No similar papers found.
💼 Related Jobs
No related jobs found.
M
Mustafa Mert Özyılmaz
Sorbonne Université, Master of Computer Science, Paris Île-de-France, France, 75005