LocalEngine: Think Tank

Private on-device AI runtime

Free · In‑App Purchases

Now with MLX models: run Qwen3.5, Gemma 4, MiniCPM5 and LFM2.5 natively on Apple silicon. Private, on-device AI chat with no account and no cloud. LocalEngine runs open-source AI language models entirely on your iPhone and iPad. Everything happens on-device: your prompts, conversations, and images never leave your device and are never sent to any server. Built in pure Swift with SwiftUI, LocalEngine runs models through a Metal-accelerated llama.cpp engine, so you get fast, low-latency AI completely offline — no account, no sign-in, no cloud. WHAT YOU CAN DO • Chat — Hold multi-turn conversations with a local model. Replies stream token-by-token in real time. Start a new chat or stop a response at any time. • See images — Turn on vision in Settings, attach photos to your message, and ask about them. Images are understood entirely on-device by a vision-capable model. • Download models — Browse a curated catalog of open models and download the one that fits your device. Downloads run as resumable background jobs with live progress, and the screen can stay awake while a large model downloads. • Tune generation — Adjust the response length (max tokens) and the temperature to shape how the model replies. • Unlock larger models — LocalEngine is free to use with starter models. A one-time Pro upgrade unlocks larger catalog downloads on iPhone and iPad. CURATED MODELS LocalEngine is ready to chat out of the box with a small, fast default model, and its catalog includes compact vision-language models you can grow into: • Qwen3.5 — 0.8B and 2B free, with 4B available through Pro • Gemma 4 — E2B free, with E4B available through Pro Smaller models are quick and light on storage; larger Pro models are more capable on newer devices. Each vision model pairs with a multimodal projector so it can understand images — all on-device. PRIVATE BY DESIGN • 100% on-device inference — no cloud, no telemetry, no analytics. • No account and no sign-in required. • Runs fully offline once a model is downloaded. • Built to App Store privacy and security standards. REQUIREMENTS • iOS / iPadOS 17.0 or later • A recent iPhone or iPad is recommended, especially for the larger models • Free storage space for the models you choose to download LocalEngine puts the power of modern AI — text and vision — in your pocket, without giving up your privacy.

  • This app hasn’t received enough ratings or reviews to display an overview.

• MLX models are here — Qwen3.5, Gemma 4, MiniCPM5 and LFM2.5 now also come in MLX versions, built for Apple silicon. Find them next to the GGUF versions in Models. • Automatic engine switching — pick any model and LocalEngine uses the right engine for it, no settings to change. • Updated on-device AI engines for better stability and model support. Everything still runs 100% on-device — no account, no cloud, no telemetry.

The developer, GUOYU WANG, indicated that the app’s privacy practices may include handling of data as described below. For more information, see the developer’s privacy policy .

  • Data Not Collected

    The developer does not collect any data from this app.

    Privacy practices may vary, for example, based on the features you use or your age. Learn More

    The developer has not yet indicated which accessibility features this app supports. Learn More

    Seller
    GUOYU WANG
    Size
    44.6 MB
    Category
    Productivity
    Compatibility
    Requires iOS 17.0 or later.
    • iPhone
      Requires iOS 17.0 or later.
    • iPad
      Requires iPadOS 17.0 or later.
    • Mac
      Requires macOS 14.0 or later.
    • Apple Vision
      Requires visionOS 1.0 or later.
    Languages
    English and 9 more
    • English, French, German, Italian, Japanese, Korean, Portuguese, Simplified Chinese, Spanish, Traditional Chinese
    Age Rating
    13+
    In-App Purchases
    Yes
    • LocalEngine Pro $4.99
    • Local Engine Pro $2.99
    • Professional $9.99
    Copyright
    © 2026 GUOYU WANG