LLM Server

비즈니스

₩33,000 · iPad⁠용으로 디자인됨 · macOS⁠용으로는 확인되지 않음

Your Phone Is Now an AI Server LLM Server turns your iPhone or iPad into a private AI inference server. Run large language models entirely on-device, expose an OpenAI-compatible API to your local network, and chat with your model from any browser on any device — laptop, desktop, tablet, even another phone. No cloud, no subscription, no data leaving your hardware. Chat From Any Browser on Your Network Start the server and open the URL on any device sharing your Wi-Fi. The built-in web interface gives you a clean chat UI instantly — no client app to install, no account to create. Your laptop, your partner's tablet, a colleague's machine: all talking to a model running on your phone. OpenAI-Compatible API, Drop-In Ready A standard OpenAI-compatible API means LLM Server slots into any tool you already use — Continue, Open WebUI, LangChain, custom scripts, the lot. Chat completions, text completions, streaming (SSE), and model listing endpoints all supported. Ollama CLI commands work too. Fully Offline, Fully Yours - Download a GGUF model once and you're done with the cloud. Inference runs locally on Apple Metal GPU with no internet required. Your prompts, your conversations, your data — none of it leaves the device. Airplane mode works fine. Any GGUF Model From Hugging Face - Browse and download directly from Hugging Face with built-in search, or import your own files. LLaMA, Mistral, Phi, Gemma, Qwen, DeepSeek, and every other llama.cpp-supported architecture runs out of the box. Background downloads with progress tracking so you can keep working. Enterprise-Grade Security - TLS/HTTPS encryption — generate self-signed certificates or import your own chain and private key. - API key authentication — Bearer tokens with per-key management. Generate cryptographically secure keys or bring your own. - Bind control — lock to localhost, open to your LAN, or pin to a specific interface. - Optional Web Search Tools (currently supporting Perplexity, Brave, Exa, Tavily and Ollama API keys, more will be added later) Tune Every Knob Full control over generation: context size up to 32K, temperature, top-p, top-k, repeat penalty, frequency and presence penalties, max tokens, seed, GPU layer offloading, and thread count. Save presets globally or per model. Smart Resource Management Your phone stays responsive under load. Real-time thermal monitoring with automatic thread reduction under pressure and request rejection at critical temperatures. Memory-aware model loading with conservative budgeting. Configurable request queues and per-request timeouts. Built for Developers - Live API docs with copy-paste curl examples - Structured logging (debug, info, warning, error) - One-tap copy for server addresses and API keys What's Inside Dashboard with one-tap server control, model manager with download progress, complete settings hub (server, inference, security, API keys, developer tools), and guided onboarding for first-time setup.

  • 이 앱은 개요를 표시할 만큼 충분한 리뷰 또는 평가를 받지 않았습니다.

Optimized App UI for iPhone Duo

Linosec 개발자가 아래 설명된 데이터 처리 방식이 앱의 개인정보 처리방침에 포함되어 있을 수 있다고 표시했습니다. 자세한 내용은 개발자의 개인정보 처리방침 을 참조하십시오.

  • 데이터가 수집되지 않음

    개발자가 이 앱에서 데이터를 수집하지 않습니다.

    개인정보 처리방침은 사용하는 기능이나 사용자의 나이 등에 따라 달라질 수 있습니다. 더 알⁠아⁠보⁠기

    개발자가 이 앱이 지원하는 손쉬운 사용 기능을 아직 등록하지 않았습니다. 더 알아보기

    제공자
    • Linosec
    크기
    • 573.6 MB
    카테고리
    • 비즈니스
    호환성
    iOS 16.4 이상 필요
    • iPhone
      iOS 16.4 이상 필요
    • iPad
      iPadOS 16.4 이상 필요
    • Mac
      macOS 13.3 이상 및 Apple M1 칩 이상이 탑재된 Mac이 필요
    • Apple Vision
      visionOS 1.0 이상 필요
    언어
    • 영어
    연령 등급
    19+
    • 19+
    • 드물게 다음이 포함됨
      사실적인 폭력
      욕설 또는 노골적인 유머
      성적인 내용 또는 선정적인 테마
      잔혹/공포 테마
      의료 치료 정보
      음주, 흡연, 약물의 사용이나 언급
      성적인 내용 또는 노출
      총 또는 기타 무기

      다음이 포함됨
      사용자 생성 콘텐츠
      건강 또는 웰빙 주제
    저작권
    • © Linosec