๐ Ollama Local AI is an independent, privacy-first Ollama Android app for running, connecting, and routing Large Language Models (LLMs) from your phone.
โ ๏ธ Important notice: Ollama Local AI is not affiliated with, sponsored by, or endorsed by Ollama. Ollama is a trademark of its respective owner.
Turn your Android device into a secure local-network OpenAI-compatible LLM proxy for IDEs, coding assistants, scripts, and custom AI tools.
๐ฑ Connect Cursor, VS Code, Antigravity, Windsurf, or any OpenAI-compatible client to your phoneโs LAN IP. Manage local models, llama.cpp, cloud providers, and custom endpoints from one Android app instead of configuring multiple endpoints on your development machine.
๐ Local-first privacy
Run local AI models on-device with no internet required. Local model data, provider settings, and API keys are stored on your Android device. When you choose a cloud provider, requests go directly to the provider you configureโwithout a third-party relay from this app.
โก Key features
๐ง Local LLM inference: Run local models on your phone with llama.cpp engine support.
๐ฌ Chat mode: Chat with local models or configured cloud providers such as OpenAI and Claude, then switch providers when needed.
๐จ Canvas mode: Build and test websites inside chat while the AI constructs them live.
๐ WebX Preview: Preview websites created by Ollama or local AI from any browser on your network.
๐ฅ๏ธ Cross-device WebUI: Open the chat and management console from your phone, tablet, or PC.
๐ฅ๏ธ Zero-config local hosting: Use the embedded NanoHTTPD server directly on Android.
๐ OpenAI-compatible API: Use standard endpoints such as /v1/chat/completions, /v1/models, and /health with compatible IDEs, scripts, and tools.
๐ Provider switching: Save and switch between Ollama, llama.cpp, NVIDIA, OpenAI, Claude, Hugging Face, and custom API endpoints.
๐ฆ Smart proxy routing: Use provider timeouts, rate limits, session-stable routing, local pools, failover, and round-robin request distribution.
๐ก LAN master/worker pairing: Link Android phones and devices into a private local AI network.
๐ Multi-device clusters: Connect devices into a local LLM inference cluster.
๐ค Multi-agent orchestration: Coordinate agents, providers, tools, and tasks.
๐๏ธ Live agent stages: Follow agent progress and execution stages in real time.
๐ Traffic Observatory: Monitor request flow, routing, performance statistics, and proxy diagnostics.
๐งช Tool Lab: Test AI tools, provider connections, API requests, and workflows.
๐ฅ๏ธ Web Console: Manage routing, chat, providers, tools, and diagnostics in your browser.
๐จ Material 3 interface: Use a clean, responsive Android UI designed for local AI development.
๐ Efficient routing: Lazy token-bucket rate limiting helps control traffic and avoid unnecessary background work.
๐ Encrypted storage: Protect provider credentials with AES-256 encrypted EncryptedSharedPreferences.
๐งต Responsive performance: Coroutines and offloaded I/O keep the Compose interface fluid during routing and network work.
๐ ๏ธ Built for developers
Use Ollama Local AI with Cursor, VS Code, Antigravity, Windsurf, OpenAI-compatible clients, automation scripts, and custom AI workflows. Chat, code, route requests, preview web apps, connect devices, and observe AI traffic from one Android application.
๐ฃ๏ธ More in beta
Additional local AI integrations, provider support, developer tools, and performance improvements are in development.
๐ฅ Build a flexible local LLM workflow with Ollama Local AIโyour Android hub for local models, cloud routing, IDE connectivity, and private network AI development.
Ollama Local AI is an independent third-party application and is not affiliated with or endorsed by Ollama.