DEPLOY FORWARD LLC
Run quantized large language models directly on your iPhone. No cloud, no internet required.
目前在 Apple 公開 App Store 頁面可見的項目,可能不等同於 App Store Connect 中的完整目錄。
Run quantized large language models directly on your iPhone. No cloud, no internet required.
Access state-of-the-art quantized AI models optimized for mobile hardware. Download GGUF-format models that compress billion-parameter networks into mobile-friendly sizes while maintaining performance.
COMPLETE MODEL SUITE
• Llama 3.2 1B/3B (Meta) - Q4/Q8 quantization
• Gemma 3 270M/2B/9B (Google) - IQ4_NL optimization
• Qwen 2.5 0.5B-7B (Alibaba) - Multiple quantization levels
• LLaVA 1.5/1.6 (Vision) - Multimodal image understanding
• Direct integration with Hugging Face model repository
TECHNICAL FEATURES
• GGML/llama.cpp inference engine
• Metal GPU acceleration on Apple Silicon
• Dynamic context window management (2K-8K tokens)
• Retrieval-Augmented Generation (RAG) with embeddings
• Real-time streaming with token/second metrics
• SQLite conversation storage with vector search
SYSTEM REQUIREMENTS
Models run efficiently when file size ≤ available RAM. Recommended minimum 6GB RAM for larger models. iPhone 15 Pro/Pro Max optimal. iOS26 for Apple foundation model.
Zero telemetry. Zero data transmission. Pure local AI computing.
根據最新觀測到的 App Store 資料快速解答。
最新觀測到的下載價格為 免費。請開啟 App Store 官方連結確認目前結帳價格。
Apple 公開頁面目前在此商店顯示 0 個購買項目,公開清單可能不完整。