Run a 120B-parameter Mixture of Experts model on an Android mid-range phone CPU-only using llama.cpp.
Why it made the shelf
It is a developer tool and implementation leveraging llama.cpp to run large language models locally on mobile hardware.
Seen on
Behind the tool
People, backing, growth, and ownership. Every claim links to its source.
Source review queued
Public trail
1 reference from 1 publisher.
Find the right AI models that run on your specific hardware setup.