Not with the same model: according to the Intelligence per Watt study, which was carried out on a Mac Studio, the chip of a desktop computer needs 1.6 to 2.3 times more energy per request than a data centre card. A local AI uses less when the installed model is smaller: 0.03 to 0.13 Wh per answer on a laptop for models with 2 to 8 billion parameters, against about 2.3 Wh for a model with 70 billion parameters on a server, according to another study.
Read the source article: Eco-friendly AI chatbot: how to choose one in 2026
