LM Studio
Run local LLMs offline with LM Studio's Bionic agent
LM Studio is the best desktop app for running local LLMs, and Bionic turns it into a true agentic workstation. The Free tier is generous, mobile access via Locally adds reach, and ZDR cloud credits are honest. But model format support is narrower than Ollama, and Bionic Pass pricing is still unknown. If you want a private local agent, LM Studio is your pick.
Verified 4d ago · liveness 81/100 · cite: rightaichoice.com/tools/lm-studio
- Developers needing a private local agent for coding, automation, and document tasks
- Privacy-conscious users who want offline LLM inference with zero data retention
- Agentic workflow experimentation with repeated long-context sessions
- Mobile AI: running large models on iPhone/iPad via Locally
- Users who need cloud-scale inference or model hosting as a service
- Those requiring support for proprietary models like GPT-4 or Claude
- Tinkerers wanting a vast model hub with every Hugging Face format
We scan live Reddit threads, YouTube comments, X posts, G2 reviews and other communities — and hand you an honest verdict in under a minute.
- Honest verdict, not marketing
- Real pros & cons from real users
- Attributed quotes with receipts
3 free scans · no card needed
Skip LM Studio if you need a vast model hub with every Hugging Face format, cloud-scale inference for large workloads, or support for proprietary models like GPT-4 or Claude; if you prefer a fully cloud-hosted assistant with minimal setup, LM Studio's local-first approach may not fit your workflow.
Going past the Free tier's web search limits may require logging in and could incur charges or restrictions on usage.
LM Studio's pricing fits privacy-conscious developers and indie hackers who want a free local agent with optional cloud credits. It's cheaper than cloud-only APIs like OpenAI for heavy local use, but cloud inference costs per token can rival other providers. For teams needing enterprise-scale cloud inference, Ollama's simpler pricing may be more predictable.
In short
LM Studio — Run local LLMs offline with LM Studio's Bionic agent. Best for Developers needing a private local agent for coding, automation, and document tasks, Privacy-conscious users who want offline LLM inference with zero data retention, Agentic workflow experimentation with repeated long-context sessions. Free to use.
What's new in LM Studio
Checked 4 days agoAcross the latest 5 updates: 1 feature update, 2 launches and 2 changelog entries.
Bionic 1.0.9
Bionic 1.0.9 improves PDF processing for larger embedded images, fixes skill-related bugs, and enhances markdown copying and shell output truncation.
Bionic now supports skills
LM Studio Bionic now supports Skills, allowing users to teach repeatable actions to the agent.
Bionic 1.0.8
Bionic 1.0.8 adds skill onboarding, /install-skill and /create-skill commands, @ mentions for external files, and clearer cloud plan details.
Run Muse Glimmer locally
Meta's 30B agentic open-source model Muse Glimmer is now available in LM Studio Bionic.
Introducing LM Studio Bionic
LM Studio Bionic is introduced as an AI agent built for open models, designed to get things done.
What people actually say about LM Studio — is it worth it?
We ran a structured research pass across product reviews, community discussions, and post-purchase forum threads to surface the patterns vendors won't publish themselves. Below: the recurring strengths, the hidden costs people mention most, and the cohort that consistently regrets adopting this tool.
105 mentions across 6 sources (Hacker News, YouTube, Product Hunt, Bluesky, Stack Overflow, Lemmy) · researched Jul 26, 2026.
- +Polished GUI makes local LLMs accessible to non-experts.
- +MLX engine with KV cache checkpointing speeds up agentic workflows.
- +MTP speculative decoding accelerates generation on supported hardware.
- +OpenAI-compatible API and SDKs simplify integration with existing tools.
- +Locally iPhone/iPad app extends local LLM use to mobile.
- −Closed-source license alienates open-source advocates.
- −Bionic update removed many advanced settings power users depend on.
- −Windows/Linux performance and stability fall behind Mac.
- −Model support limited to GGUF/MLX; fewer formats than Ollama.
- −No persistent learning; models forget corrections between sessions.
- • Pro/Enterprise pricing not publicly listed, causing uncertainty.
- • Running large models may require hardware upgrades (RAM/GPU).
Viability Score
How well maintained and how widely used is LM Studio? Built from what the vendor actually publishes (docs, changelog, tutorials, integrations, pricing), whether the site is live, and how much real users discuss it. How we calculate this
Last calculated: September 2026
How we score →Key Features
- Bionic agent for document editing, coding, automations, and computer control
- Real-time offline voice transcription in multiple languages
- Local LLM downloads via LM Studio Hub (GLM 5.2, Kimi K3, DeepSeek V4 Pro, etc.)
- MLX engine with KV cache checkpointing for long-context agentic workflows
- Multi-GPU support via tensor parallelism (CUDA, ROCm, Vulkan)
- MTP speculative decoding (stable)
- OpenAI-compatible API and REST API (beta)
- JavaScript SDK (@lmstudio/sdk)
- Python SDK (lmstudio)
- Command-line tool (lms) for scripting and integration
- LM Link for remote model access (up to 5 devices on Free)
- Headless CLI for Linux/cloud/CI
- Locally iPhone/iPad companion app
- Physical Batch Size option for finer control
- Diagram and chart creation (Bionic 1.0.7)
About LM Studio
LM Studio is a free desktop app for macOS and Windows that lets you download and run large language models entirely on your own hardware — no data ever leaves your device. Its headline feature, Bionic, is an AI agent for open models that handles document editing, coding, automations, and computer control, with real-time, offline voice transcription in multiple languages. The latest release, Bionic 1.0.7, adds diagram and chart creation, Simplified Chinese interface, and support for new models, while Bionic 1.0.5 introduced folder attachments and file attachments up to 100 MiB with auto context sizing for MLX. For developers, LM Studio offers an OpenAI-compatible API, a JavaScript SDK (@lmstudio/sdk), a Python SDK (lmstudio), and a command-line tool (lms) for scripting and integration. The app includes LM Link for secure remote connections (up to 5 devices on the Free plan) and a mobile companion app called Locally for iPhone and iPad. The engine supports MLX with KV cache checkpointing and multi-GPU tensor parallelism across CUDA, ROCm, and Vulkan, making long-context agentic workflows efficient. MTP speculative decoding is stable, and a Physical Batch Size option gives finer control over performance. On the cloud side, LM Studio offers pay-as-you-go credits for frontier open-source models like DeepSeek V4 Pro, GLM-5.2, and Kimi K3, with US-based inference and Zero Data Retention (ZDR) by default. Pricing per million tokens ranges from $0.13 input for DeepSeek V4 Flash to $3.00 input for Kimi K3, with output up to $15.00. The Free tier includes local LLM inference, offline voice transcription, web search (with ZDR), and LM Link for up to 5 devices. A Bionic Pass subscription is announced as 'coming soon.' Unlike cloud-only assistants, LM Studio puts you in control: your data stays local, and you can switch between local and cloud inference as needed. It's a strong choice for privacy-conscious developers and researchers who want agentic AI without subscription
Behind the Verdict
LM Studio stands out as a desktop-first agentic platform that prioritizes privacy and local control. The Bionic agent is deeply integrated into the app, handling document creation, coding, and automations with automatic saving and real-time voice transcription that stays on-device. The recent addition of skills (Bionic 1.0.8) and support for models like Kimi K3 and DeepSeek V4 Flash make it a compelling choice for developers who want to experiment with frontier open models without sending data to the cloud. Strengths include the generous free tier, the ability to run models locally with MLX and llama.cpp, and the option to use cloud credits with zero data retention for heavier tasks. The developer tooling (OpenAI-compatible API, JS/Python SDKs, CLI) makes it easy to integrate into existing workflows. The mobile companion 'Locally' extends reach to iPhone and iPad. Weaknesses: Bionic Pass pricing remains unknown, and the model hub is not as expansive as Ollama's, particularly for less common formats. Windows users may miss some MLX optimizations that are Mac-first. Where it fits: Privacy-conscious developers, researchers, and enterprise teams that need a local agent for coding and document work. It's less suitable for teams needing cloud-scale inference or support for proprietary models like GPT-4.
Researching LM Studio? Get your full AI stack in 60 seconds.
Free, no signup — tell us your goal and get tools matched to your budget & existing stack.
Real-world workflow fit
Concrete scenarios for the personas LM Studio actually fits — and what changes day-one when you adopt it.
A developer wants to use an AI coding assistant without sending code to the cloud. They download LM Studio, install a local model like DeepSeek V4 Flash, and use Bionic to review code, generate tests, and refactor, all offline.
Outcome: The developer gets instant, private coding assistance with no data leakage, improving productivity while maintaining confidentiality.
A researcher needs to analyze sensitive data and create reports. They use Bionic to turn research into polished Word documents and PowerPoint decks using Kimi K3, with all processing local or via ZDR cloud credits.
Outcome: The researcher saves time on document creation while ensuring data privacy, with the ability to switch to cloud models for heavy tasks.
An enthusiast wants to automate repetitive computer tasks. They use Bionic's computer control and skills to create custom automations, such as organizing files or generating charts, all with real-time voice commands.
Outcome: The user gains a powerful automation agent that can handle complex workflows, reducing manual effort and increasing efficiency.
Use Cases
- Run local LLMs for coding assistance and data analysis without internet
- Deploy private AI on headless servers for enterprise workflows
- Use LM Studio on iPhone/iPad to run large models on the go via Locally
- Experiment with agentic workflows using KV cache checkpointing
- Serve local models via OpenAI-compatible API for integration with other tools
- Create and edit documents with Bionic's agentic capabilities
- Automate tasks with computer control and coding agents
- Use skills to teach the agent repeatable actions
Models Under the Hood
as of 2026-08-31
Limitations
- Bionic Pass pricing and plan details are coming soon.
- The Free tier includes local LLMs and voice transcription, while cloud inference is available via pay-as-you-go credits with zero data retention.
- Web search tool has limits and requires login.
- All cloud services are zero data retention.
as of 2026-08-29
Verification history
We have re-verified LM Studio 17 times since . Each pass re-reads the vendor's own pages and re-checks every listed field against that evidence; passes where nothing had changed are marked as such.
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
Showing the 6 most recent of 17 verification passes.
Free to cite with attribution — this page re-verifies continuously.
12-month cost
Project the real annual outlay, including the implied monthly cost when only an annual tier is published.
Vendor list price only. Add-on usage, seat overages, and contract minimums are surfaced under Hidden costs & gotchas.
Plans compared
For each published LM Studio tier: who it actually fits, and what it adds vs. the previous tier. Cross-reference the cost calculator above for projected annual outlay.
Free
$0/mo
Ideal for
Privacy-conscious individuals and developers who want to run local LLMs and voice transcription on their own hardware without any cost.
What this tier adds
Starting tier: includes local LLM inference, offline voice transcription, LM Link for up to 5 devices, and Bionic Agent access, all at $0.
Cloud Credits
Pay as you go
Ideal for
Users who need occasional cloud inference for heavy tasks, such as running DeepSeek V4 Pro or Kimi K3, while keeping costs usage-based.
What this tier adds
Adds pay-as-you-go access to frontier models with US-based inference and ZDR, priced per million tokens (e.g., $0.13–$3.00 input).
Bionic Pass
Coming soon
Ideal for
Users anticipating a subscription for advanced agent features and expanded cloud usage; currently in announced as coming soon.
What this tier adds
Upcoming tier with pricing details to be announced; likely to add more capabilities or higher usage limits than Free.
Where the pricing makes sense
The company stage and team size where LM Studio's pricing actually pencils out — and where peers do it cheaper.
LM Studio's pricing fits privacy-conscious developers and indie hackers who want a free local agent with optional cloud credits. It's cheaper than cloud-only APIs like OpenAI for heavy local use, but cloud inference costs per token can rival other providers. For teams needing enterprise-scale cloud inference, Ollama's simpler pricing may be more predictable.
Setup time & first value
How long it actually takes to get something useful out of LM Studio — broken out by persona, not the marketing-page minute.
For a typical developer: 10-15 minutes to download and install the app, create an account, and start running a local model; additional time to configure MLX and GPU settings. For cloud credits: a few minutes to set up billing and generate an API key. Skills setup: ~10 minutes to create or install skills from the composer.
Switching to or from LM Studio
How to bring data in from common predecessors and how to get it back out — written for the switcher, not the buyer.
- →From Ollama: Clone existing models via Hugging Face or use LM Studio Hub; switch to LM Studio's runtime and Bionic agent for local workflows.
- →From ChatGPT/Claude web: Export conversations or use the OpenAI-compatible API to connect existing tools.
- ↗To Ollama: Export models via Hugging Face; use LM Studio's local API to serve models and then point tools to Ollama's endpoint.
- ↗To a cloud API: Use LM Studio's OpenAI-compatible API to prototype, then migrate to OpenAI/Anthropic by changing the base URL and API key.
Integrations
Resources & Guides
Tutorials & Learning
Official links
Tools that pair well with LM Studio
Common stack mates teams adopt alongside LM Studio, with the specific reason each pairing earns its keep.
Featured Head-to-Head Comparisons
Lm Studio vs Spider Cloud
If you need to run LLMs locally for privacy and agentic workflows, LM Studio is the free, polished choice with recent updates like multi-GPU tensor parallelism and MTP speculative decoding. If your priority is web data extraction for AI agents, Spider Cloud offers a fast, Rust-based API with natural language crawling and AI extraction. These tools are complementary, not competitive—choose based on whether you need local inference or cloud web scraping.
Lm Studio vs Praktika
Choose LM Studio if you need to run LLMs locally for development or data tasks—it's free, offline, and developer-friendly. Choose Praktika if you're an intermediate language learner seeking AI-powered speaking practice with real-time feedback; its freemium tier offers limited daily sessions.
Lm Studio vs Polycam
Choosing between LM Studio and Polycam is straightforward: they solve completely different problems. LM Studio is ideal for developers who need private, offline LLM inference on their own hardware, while Polycam is purpose-built for professionals capturing 3D scans of objects and spaces. Neither can substitute the other; your choice should be driven by whether your primary need is local AI computation or 3D reality capture.
Alternatives to LM Studio
View allAtomic Chat
Free local AI chat running 1000+ open-source models fully offline.
RWKV Runner
Open-source desktop app for running RWKV RNN LLMs locally with infinite context.
Frequently Asked Questions
Categories
Best-of guides
Topics
Used LM Studio? Help shape our editorial sentiment research.


