The power cost of doing this widely would be staggering, surely?
Local models are not the same as the giant ones in data centers. They're in about the about the same ballpark as running a AAA game on max settings.
Sure, it's more computationally-expensive than running an HTML selection, but it's also not "staggering" by any reasonable stretch.
Depending on the model, I’m guessing it’s dwarfed by just about any electron app
Apple Intelligence? I mean, you're only processing for the time your "AI|browser|agent" is acting as a firewall between you and Meta. Cache locally after processing and filtering. Use alongside the accessibility API. LLMs can, in many cases, reliably solve CAPTCHAs. I find it difficult to imagine they cannot defeat Meta ad blocking countermeasures.
EFF: Adversarial Interoperability - https://www.eff.org/deeplinks/2019/10/adversarial-interopera...
Worth a white paper to see which costs less energy, using Apple's built in LLM, or downloading and displaying all the FB ads using radio, playback, and screen animation energy.