Where LM Studio Stores JSON
- Windows:
%USERPROFILE%/.lmstudio/config-presets - macOS / Linux:
~/.lmstudio/config-presets
Find the LM Studio profile .json file on your computer, then transfer it to your mobile device before importing into LMSA.

[NOTE]
Enable "Serve on local network" in LM Studio after starting your server. Make sure your device is on the same Wi‑Fi as the server computer.
Always load models from within LMSA, not from LM Studio directly. Loading models directly in LM Studio can cause multiple models to stack in memory instead of being replaced, leading to excessive resource usage on your server computer.
You can switch between different models directly in the app:
The app will automatically eject the current model and load the new one.
Save a favorite model as the default so it automatically loads when you start the app:
LMSA can import system prompts exported from LM Studio profile JSON files stored on your computer, then save them as reusable prompts in LMSA.
Where LM Studio Stores JSON
%USERPROFILE%/.lmstudio/config-presets~/.lmstudio/config-presetsFind the LM Studio profile .json file on your computer, then transfer it to your mobile device before importing into LMSA.
How to Import It
If the JSON contains a profile name, LMSA uses it as the prompt name; otherwise it uses the filename.
To connect LMSA to Ollama on your local network, configure Ollama to listen on all network interfaces. Follow the OS-specific steps below.
Manual configuration:
OLLAMA_HOST0.0.0.0Edit the Ollama systemd service:
sudo nano /etc/systemd/system/ollama.serviceIn the [Service] section add:
Environment="OLLAMA_HOST=0.0.0.0"OpenRouter is a cloud AI service that lets you access hosted models using your API key. When enabled, LMSA sends requests to OpenRouter instead of a local server.
[WARN]
When OpenRouter is enabled, your messages are sent to OpenRouter's servers and then forwarded to the model provider. Do not share sensitive personal information, passwords, or private data in your conversations. Your API key is stored locally on your device only.
OpenRouter also supports Zero Data Retention for certain models — see Zero Data Retention (ZDR) for details.
Verify your API key and re-enter it in Settings.
Ensure the OpenRouter toggle is on and a valid API key is entered.
Check your internet connection and account credits.
Turn the OpenRouter toggle off in Settings.
LMSA features an optional Web Search capability powered by Brave Search. This allows your AI models to access real-time information from the live web to provide more accurate and timely responses.
Brave Search is a privacy-focused search engine. It does not track users, build search profiles, or sell your data. Your web-enhanced chats remain anonymous.
All search queries are transmitted via secure, encrypted HTTPS channels directly to the Brave Search API.
[NOTE]
LMSA can generate high quality AI images directly inside your chat using the /image command.
/image followed by a description of the image you want.For better results, describe the subject, style, lighting, camera angle, colors, background, and mood. More specific prompts usually produce stronger images.
Images generated with /image can be saved and used for personal or commercial purposes.
*Unlimited image generation is subject to our Fair Usage Policy.
LMSA includes a built-in, privacy-first Voice Mode that lets you have hands-free, voice-in/voice-out conversations with your AI models. Speak naturally, and LMSA will transcribe your voice, query the model (local or cloud), and read the response back to you.
How to Enter Voice Mode
You have two simple options to start a voice chat:
Privacy-First Design
Once active, Voice Mode remains running for a continuous, back-and-forth conversation. The full loop works seamlessly:
Speak naturally
No need to tap send between turns.
Local Generation
Processed by your PC or OpenRouter.
Listen out loud
LMSA speaks the response back.
To get the best possible hands-free experience, consider these optimizations:
For a deeper look at the privacy architecture behind real-time voice chat with local LLM models, see our Privacy-First Voice Mode page.
When you first open LMSA, you'll see a Welcome Screen and an Onboarding Quick Guide that introduces you to the app. This guide helps you configure your essential preferences and get oriented.
What you need to know:
The sidebar contains all navigation:
Main Menu (Options section):
Active Template Indicator:
Chat History:
Bottom Links:
Connection Status (Sidebar Footer):
Message Tips:
When the AI includes code, it appears with:
Don't like a response? You can regenerate it:
Quick-suggestion buttons may appear below messages:
Tip: Save important responses as files to keep them even if you delete the chat.
AI Models are different versions of artificial intelligence with different strengths:
You can also switch connection modes directly from the sidebar footer:
This is the fastest way to move between server types without navigating multiple settings screens.
| Feature | Local | Cloud |
|---|---|---|
| Speed | Very Fast | Fast |
| Power | Moderate | Very High |
| Cost | Free | Pay-per-token |
| Internet Required | No | Yes |
| Setup Difficulty | Medium | Easy |
| Model Variety | Good | Excellent |
Method 1: Tap the hamburger menu (≡) in the header to open sidebar Method 2: In the sidebar Options section, tap "Settings" (gear icon) Method 3: From the welcome screen, tap the "Settings" button
LMSA has 20+ customizable options organized in categories:
Temperature
System Prompt
Max Tokens
App Language
Chat Bubble Font
Chat Bubble Font Size
Auto-Scroll
Chat Scrollbar
Hide Thinking
Auto-Generate Titles
Biometric Unlock
Enter Key Behavior
TTS Voice Selection
Smart Replies
Web Search (Optional)
Haptic Feedback
Light/Dark Theme
Saved Custom Endpoints
Connection Timeout
Personas are pre-configured AI personalities that change how the AI responds. Switch personalities instantly without retyping instructions.
LMSA includes 16 pre-made templates that you can activate instantly:
Write effective prompts:
"You are a Python expert. Provide code solutions with explanations.
Always suggest the most Pythonic approach. Include type hints."Include:
Avoid:
Premium is an optional one-time purchase ($14.99) that unlocks the complete LMSA experience with unlimited conversations, zero distractions, and full feature access forever.
| Feature | Free | Premium |
|---|---|---|
| Chat usage | Unlimited | Unlimited |
| Voice Mode | Unlimited | Unlimited |
| LM Studio support | ||
| Ollama support | ||
| File attachments | ||
| OpenRouter support | ||
| Web search | Unlimited | |
| Text-to-Image Generation (/image) | Limited | Unlimited |
| Text-to-Speech (TTS) | ||
| Custom endpoint usage | ||
| Adjust max tokens / context length | ||
| Templates & personas | ||
| Save custom endpoints | ||
| Offline mode | ||
| PIN Unlock | ||
| Biometric Unlock | ||
| Advanced settings | ||
| Edit system prompts | ||
| Import/Export Saved Chats | ||
| Ad-free experience | ||
| Email Support | ||
| Cost | $0 (Free forever) | $14.99 one-time |
To quickly verify premium status, check the button on the main welcome screen and in the side menu under the premium category:
Attach files to your messages for AI analysis:
How to attach files:
Supported formats:
Save any response as a file:
Hear responses read aloud:
Features:
Generate AI images directly inside your chats using the `/image` command:
/image followed by a description of what you want to create (e.g., /image a futuristic cyberpunk cat wearing a glowing neon visor).Limits:
Offline Mode is a Premium-only feature that lets you run LMSA entirely on your local network without any internet connection. You can chat with AI models running on your own devices (LM Studio or Ollama) while keeping all conversations completely private and local.
Premium Users Only — Offline Mode is exclusively available to users with a Lifetime Premium purchase ($14.99 one-time).
Setup:
During a chat:
✅ Available:
❌ NOT Available:
Offline Mode in Offline Mode:
When Internet Is Needed:
Local connections are typically unencrypted:
Offline Mode is perfect for:
"Can't connect to server"
"Slow responses"
"Lost connection"
Some AI models (like Claude) include a "thinking" process where the AI reasons through complex problems.
How it works:
Best for:
MCP (Model Context Protocol) allows AI models to use external tools and integrations like web lookup, weather data, code execution, and more.
mcp.json, e.g., mcp/playwright)Important: The Ephemeral URL option for MCP server URLs is currently disabled. LM Studio now rejects dynamic remote MCP URLs that resolve to a non-public address (this blocks LAN and Tailscale/VPN addresses). Use the Plugin ID method instead, which references a pre-configured entry in LM Studio's mcp.json file.
Web Search lets LMSA query live web sources via the Brave Search API before the model answers. This improves timeliness and can increase answer accuracy on topics that change often while maintaining a high standard of privacy.
When to turn it on:
When to keep it off:
How to use:
Privacy note: Web Search is opt-in and powered by Brave Search, a privacy-first search engine. When enabled, AI-generated search queries are sent to Brave to retrieve context. Brave does not track users or build search profiles, keeping your web-enhanced chats anonymous. See the Privacy Policy for more details.
Export all your chats:
Import chats:
Use case: Switch devices or backup important conversations
Create multiple personas for different tasks:
Tap between them as you switch tasks!
Better prompts = better answers. Examples:
❌ Weak: "Tell me about Python" ✅ Better: "Explain Python decorators with a practical example I can use in a web app"
❌ Weak: "Write code for me" ✅ Better: "Write a Python function that parses a CSV file and filters rows where age > 18, then saves to a new file"
Break complex tasks into steps:
This guides the AI and captures nuance better than one long question.
Batch Similar Tasks:
Use System Prompts Creatively:
Temperature Locking: Set your preferred temperature and lock it so it persists across chats. No more adjusting every time!
Custom Persona Library: Build personas for your specific needs:
✅ DO:
❌ DON'T:
Templates are pre-configured AI personas for different tasks (e.g., Math Tutor, Code Assistant).
[TIP]
Larger, instruction‑following models (13B+) will adhere to templates better. For best results use a capable model.
LMSA templates now support full v2 character cards. You can build cards in LMSA or import existing card files from other tools.
Basic vs Advanced v2
How to Import v2 Cards
Tip: If you edit prompts after import, save the template again so LMSA updates the generated runtime system prompt.
Change chat bubble font style and size in Settings:
[NOTE]
Changing Chat Bubble Font Size only affects chat bubbles (messages). It does not change text size in menus, settings, or other UI.
LMSA includes native multilingual support across the entire mobile experience, allowing you to use the app and interact with your AI models in your preferred language.
Supported Languages
Onboarding Quick Guide
When first launching LMSA on your device, the onboarding quick guide welcomes you and lets you choose your preferred language right away.
Selecting your language during onboarding instantly localizes the interface, navigation menus, default persona prompts, and voice mode.
You can switch languages at any time without resetting your data:
Clearing app cache: Safe—does not delete any data.
Clearing app storage, uninstalling, or reinstalling: Permanently deletes all data including saved chats, system prompts, templates, and settings. Exported chat files are unencrypted JSON — store them securely.
Connections to LM Studio and Ollama servers are not encrypted. Anyone monitoring network traffic could intercept messages. OpenRouter API and chat data are fully encrypted during transit via HTTPS.
[SECURITY]
Look for the green shield icon
In the Models menu, any OpenRouter model marked with a green shield icon supports Zero Data Retention (ZDR). For these models, LMSA automatically tells OpenRouter to route your messages only to provider endpoints that enforce a zero data retention policy — meaning your prompts and responses are not stored by the provider.
LMSA is built privacy-first, so ZDR enforcement for qualifying OpenRouter models is always on and cannot be disabled within the app. Users should be aware that not all OpenRouter models listed in LMSA qualify for ZDR — those that don't will not have the green shield icon and will be treated as normal, without any special privacy routing instructions. Always check the Models menu for the green shield icon to see whether a given model is covered by ZDR.
Keep in mind that privacy-first doesn't always mean the lowest cost, fastest, or first-available option. Restricting requests to ZDR-compliant endpoints can mean fewer providers are eligible to handle a given model. During periods of high traffic, this may occasionally result in an error — if that happens, simply try again. You should also keep an eye on your OpenRouter dashboard for accurate, live usage and pricing information.
ZDR is an OpenRouter feature, not an LMSA feature — LMSA simply requests it and respects OpenRouter's compatibility list. For full details, see the Zero Data Retention section of our Privacy Policy or OpenRouter's official ZDR documentation.
Background: Early LMSA releases were paid-only. Legacy paid users can reactivate premium access in the newer app using a legacy promo code.
Connect to your local LLM server from anywhere, securely—without opening ports or using VPNs. This guide walks you through setting up Tailscale with LMSA, step by step. 🔒
Note: LMSA does not offer official support for using Tailscale as a secure tunnel beyond the scope of this guide.
Tailscale is a modern VPN service built on top of WireGuard. It creates a secure, private network between your devices without needing to open ports on your home router or deal with complex VPN configurations.
When you install Tailscale on your LLM server and LMSA-enabled device, they both join your private Tailscale network. Even though they're on different networks (home vs. mobile, office vs. home, etc.), they can communicate securely as if they're on the same private network.
Run a local LLM server on your home computer and access it from LMSA on your phone or tablet while traveling.
Direct, encrypted connection between devices—faster than traditional VPNs.
Tailscale has a free tier perfect for home setups and small teams.
No port forwarding, firewall rules, or complex networking knowledge required.
Before you start, you'll need:
100.x.x.x)100.64.45.123On Windows/macOS:
100.x.x.x)On Linux:
tailscale ip -4This outputs your Tailscale IPv4 address directly.
ping commandsNow that Tailscale is set up, configure LMSA to use your remote LLM server:
http://100.x.x.x:YOUR_PORTWhere:
100.x.x.x = Your server's Tailscale IP (from Part 2)YOUR_PORT = The port your LLM server is running onCommon Examples:
http://100.64.45.123:11434http://100.64.45.123:1234http://100.64.45.123:8000http://100.64.45.123:8080> Note: If you're not sure what port your server uses, check your LLM server's documentation or startup logs. It usually displays: "Server listening on port XXXX"
See the Troubleshooting section below.
Possible Causes & Solutions:
100.x.x.x)Possible Causes & Solutions:
Possible Causes & Solutions:
Possible Causes & Solutions:
A: Yes! Tailscale's free tier is perfect for personal use and home projects. You get:
A: Absolutely! Install Tailscale on any device you want to connect to your LLM server:
A: Check these common defaults:
Otherwise:
A: This is normal. Here's how to get a permanent IP:
Alternatively, just update the IP in LMSA each time (or use a hostname if using Tailscale's DNS feature).
A: Yes, but it's not recommended:
A: Yes, Tailscale uses:
A: Yes! Tailscale works globally:
A: Depends on your Tailscale plan:
A: Brief outages only affect initial connection:
You're now ready to access your local LLM server securely from anywhere! Here's what you've learned:
✅ Installed & configured Tailscale on server and phone ✅ Found your server's Tailscale IP ✅ Configured LMSA with the server URL ✅ Tested the connection successfully ✅ Understand security best practices
$ lmsa connect --help
If you ran into any issues, our support team typically replies within one business day. We're here to help!
Contact Support