Local LLM Playground
Local LLM Playground is #67 in Developer Tools Paid in United States and charting in 3 of 24 countries we track.
Updated 27 Sep 2026 · rankings refresh through the day
DeveloperShigeru Hasunuma
CategoryDeveloper Tools
Price10.00 USD
Released
Rating★ 0.0 (0)
Version2.0(1)
Age rating4+
First seen by us
25 Sep 2026
Rating
0.0 ★
0 ratings
Measured
Best rank now
#67
Measured
Local LLM Playground on the App Store
App details
App ID
6760124117
Publisher
Shigeru Hasunuma
Content rating
4+
Where it ranks now
Every top chart it appears in across 24 countries.
| Country | Charts | Rank |
|---|---|---|
| Developer ToolsPaid | #67 | |
| Developer ToolsPaid | #176 | |
| Developer ToolsPaid | #135 |
Apps near Local LLM Playground on the chart
Developer Tools Paid, around #67.
Ratings
App Store, worldwide.
Rating
0.0 ★
0 ratings
Measured
Ratings
0
Measured
About Local LLM Playground
Local LLM Playground is an AI playground for running local language models directly on your iPhone.
The app supports both GGUF and MLX language models, and now also supports Apple Foundation Models. From entering prompts to viewing inference results, you can easily experiment with on-device AI entirely on your device.
The following AI models are bundled with the app, so you can start using local AI immediately after installation:
* Qwen2.5-0.5B-Instruct-Q4_K_M (GGUF)
* Qwen3-0.6B-MLX-4bit (MLX)
No cloud connection or API key is required. All AI processing is performed locally on your device, and your prompts and personal data are never transmitted to external servers.
In addition to prompt experimentation and model evaluation, the app provides detailed inference metrics, including:
* Token generation speed (tokens/sec)
* Inference time (prefill / decode)
* Token usage statistics
* Memory usage
* Context length (n_ctx)
You can also import your own GGUF and MLX models to experiment with a wide variety of local AI models.
Key Features:
* Load GGUF language models
* Load MLX language models
* Support for Apple Foundation Models
* Run AI inference from custom prompts
* Display token generation speed
* Visualize inference timing
* Monitor memory usage
* Configure context size
* Advanced inference parameter settings
* Import GGUF models
* Import MLX models
Highlights:
* Fully on-device AI
* Supports GGUF, MLX, and Apple Foundation Models
* No cloud services required
* No API keys required
* Privacy-focused by design
* Ideal for experimenting with and learning about local AI
Recommended For
* Anyone interested in local LLMs
* Developers exploring MLX models
* Users who want to experience Apple Foundation Models
* AI enthusiasts evaluating model behavior
* Prompt engineering and experimentation
* Comparing the performance of lightweight language models
Important Notice:
AI-generated responses may contain inaccuracies or incorrect information. Please independently verify any important information before relying on it.
The app supports both GGUF and MLX language models, and now also supports Apple Foundation Models. From entering prompts to viewing inference results, you can easily experiment with on-device AI entirely on your device.
The following AI models are bundled with the app, so you can start using local AI immediately after installation:
* Qwen2.5-0.5B-Instruct-Q4_K_M (GGUF)
* Qwen3-0.6B-MLX-4bit (MLX)
No cloud connection or API key is required. All AI processing is performed locally on your device, and your prompts and personal data are never transmitted to external servers.
In addition to prompt experimentation and model evaluation, the app provides detailed inference metrics, including:
* Token generation speed (tokens/sec)
* Inference time (prefill / decode)
* Token usage statistics
* Memory usage
* Context length (n_ctx)
You can also import your own GGUF and MLX models to experiment with a wide variety of local AI models.
Key Features:
* Load GGUF language models
* Load MLX language models
* Support for Apple Foundation Models
* Run AI inference from custom prompts
* Display token generation speed
* Visualize inference timing
* Monitor memory usage
* Configure context size
* Advanced inference parameter settings
* Import GGUF models
* Import MLX models
Highlights:
* Fully on-device AI
* Supports GGUF, MLX, and Apple Foundation Models
* No cloud services required
* No API keys required
* Privacy-focused by design
* Ideal for experimenting with and learning about local AI
Recommended For
* Anyone interested in local LLMs
* Developers exploring MLX models
* Users who want to experience Apple Foundation Models
* AI enthusiasts evaluating model behavior
* Prompt engineering and experimentation
* Comparing the performance of lightweight language models
Important Notice:
AI-generated responses may contain inaccuracies or incorrect information. Please independently verify any important information before relying on it.
Latest updates
2.0(1)
What’s New in This Version
* Added support for MLX inference engine
* Added support for Apple Foundation Models
* Bundled Qwen3-0.6B-MLX-4bit for instant MLX experience
* Continued support for Qwen2.5-0.5B-Instruct-Q4_K_M (GGUF)
* Improved local model import experience
* Updated app icon and UI refinements
* Performance improvements and bug fixes