All the power of the terminal agents, none of the terminal.
A harness is the layer between you and any model: over 280 built-in tools,
approval gates, memory, planning and the surfaces to use them - chat, voice, code, Studio and your phone.
It installs like an app and runs on your computer, not in someone's cloud.
Bring Claude, GPT, Gemini or Grok, run models on your own hardware with
Skales Local, or start with no key at all on
Skales IQ.
Installed in 30 seconds. Free for personal use.
No cloud · No terminal · No signup · No subscription · Source-available on
GitHub
/goalplan my week and book the meetings
Type a goal. Skales runs it on its own - across as many steps as it takes, even with the chat closed.
1,749GitHub stars
40.7K+installs worldwide
286built-in tools
26AI providers
12languages
100%local & private
Every AI assistant lives in someone else's cloud.
Your questions. Your files. Their servers.
Skales lives on yours.
Local models via OllamaOr your own API keysNo account. Ever.
Not a chatbot. A worker.
Skales · your model · local
On it. Checking Trattoria da Luigi availability and preparing the booking.
Frontier capability is already a commodity - you can buy it by the token from a dozen vendors, and next month there will be a better one. What decides whether an AI actually finishes your work is everything around the model. That layer is the harness, and it is the whole product here.
01
The model is a part you swap
26 providers are built in, plus any OpenAI-compatible endpoint. Switch mid-conversation, chain fallbacks for outages, or take the model out of the cloud entirely and run it on your own hardware. Nothing you build is tied to one vendor's roadmap.
02
The harness is what does the work
over 280 built-in tools - 47 for WordPress alone - plus approval gates, a file sandbox, memory that survives the session, planning, and goals that resume after a restart. Tools are loaded on demand rather than all at once, so the context stays lean and small models stay reliable.
03
And the surfaces you reach it through
One agent, many doors: chat, Iris the voice surface, Code, Studio and Flow, Cockpit for everything running in the background, the Desktop Buddy on your screen, and a native app on your phone that is an agent in its own right.
On the stores it is filed under AI agent & task, which is the
honest short version: an agent that takes tasks off your list and works them, not a window you chat into.
01The Goal Engine
Close the chat. The work continues.
Every other chat app stops when you stop typing. Skales takes on goals - and keeps working in the background until they are done.
01 · You type one line
/goal research suppliers and make a comparison sheet
Or just ask naturally - Skales recognizes a goal from how you phrase it.
A real plan with success criteria - visible, editable, yours.
03 · It works. You close the chat.
Browsing, extracting, writing files - in the background
Multiple goals run in parallel. Schedules repeat them. The app keeps going.
04 · It asks before anything risky
Approve / Decline - right in the chat or on your phone
Approval gates and a killswitch are on by default. You stay the boss.
05 · Done - and remembered
comparison-sheet.xlsx saved. Lessons folded into Memory.
Every finished goal makes the next one start ahead.
Goals, Workflows, schedules, an OODA self-correction loop and approval gates -
see the autonomy stack →
02Desktop Buddy
The most-used feature is alive.
A little companion lives on your screen - Skales the gecko,
Bubbles the blob or Capy the capybara.
It is a full agent: ask it something and it works through
every step - files, web, email, calendar - with Approve / Decline right in its
speech bubble. It speaks your language, remembers you, and streams live progress while it works.
Want a different look? Generate your own custom pixel skin right inside Skales.
Full agent loopApprove / Decline3 mascotsCustom pixel skinsOwn conversation thread
“this looks awesome! but after bonzi buddy went rogue.. i have ptsd”
- @aleksblackz on YouTube.
Fair. This one asks before it does anything.
Try it right now - the Buddy in the corner of this page is real. ↘
SkalesBubblesCapy
03Skales Studio
Make things, not just text.
Design, image, video, voice and music - one creative suite inside your agent.
Prompt-to-design returns production-ready HTML + Tailwind across 7 templates.
Images via FLUX, SDXL, DALL·E or local ComfyUI.
Video in 1080p via Veo, Kling, Runway and fal.ai - native 9:16 included.
A Brand Kit keeps everything on-brand.
Code Mode binds any chat to a folder - Plan investigates read-only, Code edits,
Auto runs the whole task. Skales Code is the window it grew into: a command
palette, search across the whole project, a change you review hunk by hunk, and git without leaving for a
terminal. And Lio is the friendly one - a 6-year-old shipped a Snake game
with it. All three work with Claude, GPT, Gemini or your local Ollama model.
Command paletteHunk-by-hunk reviewGit in the windowYour git identityApproval gatesAgent Skills import
Iris listens, she works, and she speaks. The wake word runs
on your device. And the part that matters:
she has the same tools a chat turn has - the whole palette,
not a shortlist - and the skills you switched on for the chat. That is the difference between a voice
assistant that sets a timer and an agent you happen to be talking to. When something needs your say-so
she asks out loud and takes yes or no for an answer. Results arrive as orbits instead of a flood of
pop-ups; Escape closes them again.
Wake word on deviceFull tool parity with chatYour skills, spokenOrbits, not pop-ups
Iris Orbit is becoming a product of its own in Q4 2026. You do not have to wait for it:
she already lives inside Skales and works there today,
with the full tool set of this machine behind her.
Local does not have to mean somebody else’s runtime.
Every local-first app on the market means the same thing by it: go install Ollama first. Skales brings its own inference server instead - it starts it, stops it and updates it, and there is nothing else to set up.
A catalog of 56 models
Text, vision, speech and image models, picked and tested rather than scraped. Every entry shows its size, its license and a SHA-256 you can verify before it ever runs - including whether commercial use is allowed.
Images on your own hardware
Image generation runs on the device through a bundled stable-diffusion runtime. No credits, no queue, no upload of the thing you are making.
Speech in and out, offline
Whisper listens and Piper speaks, both locally. Voice keeps working on a plane, in a basement, and on a machine that has never seen an API key.
Ollama and LM Studio still work exactly as before - they are a choice now, not a prerequisite.
See every way to run a model →
Skales IQ · 07
Start with no key at all.
The hardest part of any AI tool has never been the model. It is the twenty minutes between installing
something and having a working API key. Skales IQ removes them: it is the first provider in the list,
it is already on when you open the app for the first time, and it does real agent work - tools and images
included, not a crippled demo mode.
Ready on first launch, no key, no signup
Tool calling and image understanding included
The same agent, the same tools, the same gates
When the trial ends: paste your own key, or go local, and keep going free
Skales IQ is also the foundation of the paid tier that is coming - a managed way to run the agent for
people who would rather not think about providers at all. The free agent stays free either way: your own
key and your own local models are never touched by it.
Your vibe
One app. Three looks. Twelve languages.
Skales-X, Classic and Flat, each in light and dark and with an accent tone you pick - drag the handle to compare. The whole UI speaks 12 languages, chosen on first launch.
Dark modeLight mode
The honest comparison
All the power of the terminal monsters. None of the terminal.
OpenClaw needs curl, Node, a daemon and ideally a dedicated machine. Hermes wants a VPS. Agent Zero wants Docker. Skales wants a double-click.
Every claim dated and sourced. Last verified July 22, 2026.
08Skales Mobile
An agent, in your pocket.
The mobile app is a complete agent on its own:
111 tools that run on the phone, a model that runs
on the device with no signal at all, plus Flow,
Studio and voice. Pair it via QR over an
end-to-end encrypted relay and it also drives
your desktop's 286 tools from anywhere. Live on Google Play and the App Store.
Voices from GitHub, YouTube, Hacker News and the community - linked where the platform allows.
“From every tool I've tested in this space, I haven't found one that delivers intelligence without complexity, a companion instead of a tool, visualization without needing to write code, or value without hype. Skales has the foundation to tell that story. No one else in this landscape is close.”
“Bro, Skales is awesome. I use it for a week with local LLM and it's really work. Sure, here is some bugs, but I hope it will disappear in future version. 10/10. Thank you so much!”
“Good luck with Skales - the accessibility angle (no Docker, no CLI) is genuinely underserved.”
“This is honestly the best implementation of a local agent UI I have seen so far. The character interaction feels very fluid and much less robotic than other wrappers.”
“I love this bro. Yesterday installed and its great”
“The modular skill system is brilliant. It's the first time I've felt like I'm building a real assistant rather than just a chatbot. The codebase is surprisingly clean too.”
“The way it handles multi-step tasks is insane compared to standard chat interfaces. It actually does the work instead of just telling me how to do it.”
“This is the first project that actually makes 'Agents' accessible for non-coders. I set it up in 5 minutes and it's already organizing my downloads folder.”
“this looks awesome! but after bonzi buddy went rogue.. i have ptsd”
“I love the fact that it doesn't just sit in the browser. Having it as a desktop companion makes it feel like it's part of the OS. Plus, the privacy controls are top-tier.”
“ngl this is cooler than the 'just watches your terminal' ones, actually doing stuff makes it feel less gimmicky.”
Shipping weekly
What’s new in Skales
The app updates itself - and so does this section.
v12.9
Cockpit 2.0
One head, one action: an execution board in four columns - pending, in progress, completed, blocked - with goals, tasks, Autopilot and schedules as cards you can filter, drag and act on.
v12.9
Skales Code
The coding window grew up: command palette, project-wide search, review a change hunk by hunk, and git from inside - branches, push, stash, commits.
v12.9
Teams got rooms
Up to twelve computers in one end-to-end encrypted room, joined by code or QR, every message signed - and your own agents can sit in it as members.
v12.9
Memory refuses your secrets
A card number, an IBAN, a key or a password is no longer written into memory at all - because a memory is read back into the prompt of every later conversation.
Free for personal, education & internal business use
Skales is source-available under BSL 1.1 and automatically converts to
Apache 2.0 in 2030 - what you build today keeps working in five years. No rug pulls.
Embedding Skales in a product you sell? Commercial licensing is a friendly conversation:
dev@mariosimic.at
In the works · Waitlist open
Skales Business Suite
Persistent goals that survive restarts, CRM sync, newsletter automation, a social queue on your
schedule - same local-first story. No date promised until it is real. Get an early invite:
Opt-in only, unsubscribe anytime. Skales itself never needs an account.
FAQ
Questions, answered.
Everything you'd want to know before the download.
What is Skales?
Skales is an AI harness you install like an app - on the stores it is filed as an AI agent & task app. A harness is the layer between you and the model: 286 built-in tools, approval gates, memory, planning, and the surfaces to use them, on Windows, macOS, Linux, Android and iOS. The model itself is a part you swap: Claude, GPT, Gemini, Grok, DeepSeek, Mistral, or models running on your own hardware through Skales Local. It does real work - background goals, coding, email, calendar, browsing, media generation. Free for personal use, no account required.
What is an AI harness, and why does it matter more than the model?
Frontier model capability is close to a commodity now: several vendors sell something comparable, and the leader changes every few months. What decides whether an AI finishes your task is the harness around it - which tools it can reach, whether it asks before doing something irreversible, whether it remembers, whether it can plan and resume after a restart, and which surfaces you can reach it through. Skales is that layer, and the model plugs into it rather than the other way round.
Is Skales really free?
Yes. Skales is free for personal, educational and internal business use under the Business Source License 1.1, which automatically converts to Apache 2.0 in 2030. You only pay your AI provider for API usage - or pay nothing at all by running models on your own hardware with Skales Local, Ollama or LM Studio. Skales itself never charges a subscription.
How do I install Skales?
Download the installer for your OS from skales.app/download - Windows EXE, macOS DMG (signed), Linux .deb or AppImage, Android on Google Play, iOS on the App Store. Double-click, done in about 30 seconds. No Docker, no terminal, no admin rights, no account.
Is my data private?
Yes - Skales is local-first. All conversations, memories, files and settings live in ~/.skales-data on your machine. API calls go directly from your computer to your chosen AI provider (BYOK - bring your own key). No cloud backup, no middleman, no data sharing. With Skales Local it works entirely offline: the model, the image generation and the speech all run on your machine. Memory is deliberate about what it will not keep: a value shaped like a card number, an IBAN, an API key, a token or a password is not written down at all, because a memory is read back into the prompt of every later conversation.
Can I reach Skales from another device?
Yes, two ways. The Skales desktop serves its own interface, so with remote access on you can open that interface in a browser on a tablet, a second computer or a phone - one switch under Settings > Security, off by default, and turning it on means reachable over LAN or Tailscale AND token-protected as one inseparable setting, with the access URL, a QR code and a regenerate button on the same page. A browser session can additionally be asked for a six-digit code from an authenticator app, with ten recovery codes. The other way is the mobile app, paired by QR over an end-to-end encrypted relay, which is a native agent in its own right rather than a window onto the desktop.
How do I keep an autonomous agent from running up a bill?
Every conversation shows what it costs while it runs - a price beside the context meter, a price on each answer, and how much of the input came out of the provider's cache instead of being paid for twice. You can give a conversation a ceiling: at half of it and again at the ceiling Skales stops and asks whether to carry on, switch to a cheaper model or stop, and that question is answered by you rather than by the model, so being asked costs nothing. The ceiling holds for a turn typed on your phone as well, Telegram and WhatsApp count per conversation and per day, and the killswitch is still one click away.
Which AI models and providers does Skales support?
26 providers out of the box: Skales IQ (the built-in zero-key trial), OpenAI (GPT), Anthropic (Claude), Google (Gemini), xAI (Grok), DeepSeek, Mistral, Groq, Together, MiniMax, Moonshot AI (Kimi, with a switch between the international and China endpoints), GLM (z.ai), Qwen via DashScope, Hunyuan, GigaChat, AtlasCloud, Unsloth, Cloudflare Workers AI, NVIDIA NIM, Hugging Face and OpenRouter - plus local models through Skales Local, Ollama, LM Studio, KoboldCpp or any OpenAI-compatible endpoint. You can even sign in with your existing ChatGPT subscription, no API key needed. Every provider card also carries a "Can this model see images?" switch - Auto, Yes or No - so a wrong guess about a model's eyes is something you can correct.
Do I need Ollama to run Skales offline?
No. Skales Local ships its own inference server: Skales starts it, stops it and updates it, and a catalog of 56 models covers text, vision, speech and images - each with its size, its license and a SHA-256 you can check before it runs. Image generation happens on the device, speech comes in through Whisper and goes out through Piper, all without a key. The GPU badge is read from what the engine actually loaded rather than from what the build was capable of, so a model that ended up on the processor says so and says why. Ollama and LM Studio still work exactly as before; they are a choice now, not a prerequisite.
Can I talk to Skales instead of typing?
Yes. Iris is the voice surface and has her own place in the sidebar: the wake word runs on your device, she speaks back, and - the part that matters - she has the same full tool palette a chat turn gets, plus the skills you switched on for the chat, not a shortlist of voice commands. When something needs your permission she asks out loud and takes yes or no for an answer, while anything that sends, deletes, runs or deploys still wants your eyes on the screen. A call is the same mind in the same conversation rather than a smaller one of its own. Every chat also has plain voice in and out, with speech recognition and text-to-speech that can run entirely offline.
Is there an AI agent app for iOS and Android?
Yes. Skales Mobile is a real native app on the App Store and Google Play - not a web view and not a remote control that needs a computer. It runs 111 tools on the phone itself, can run a model on the device with no connection at all, and includes Flow, Studio, voice, agents and skills. Pairing it with a Skales desktop over an end-to-end encrypted relay is optional and adds that machine's full tool set - starting and stopping goals, the live plan, approving a step, reading the desktop workspace. No account, no subscription.
What can Skales actually do autonomously?
Type /goal and what you want. Skales plans the steps, keeps working in the background with the chat closed, runs recurring schedules, and parks with a decision card when it needs your approval. A run survives a restart: close, update or crash Skales mid-task and it comes back as an unfinished run that says how far it got, with one Continue to pick it up. Cockpit shows all of it on one execution board. Code Mode works in any folder like a coding agent, Workflows replay recorded tasks, and the Organization feature delegates work across specialized agents - with approval gates throughout.
Windows says "Windows protected your PC" - is Skales safe?
Yes. The Windows build is not code-signed yet, so SmartScreen shows a one-time notice. Click "More info", then "Run anyway". The published installer scan shows 0 detections across 50 antivirus engines on VirusTotal, Skales is source-available, and you can verify your download against the SHA-512 checksums published with every release. The macOS build is fully signed and notarized by Apple.
What is the Desktop Buddy?
A floating mascot that lives on your screen - Skales the gecko, Bubbles or Capy, plus thousands of custom pixel pets in the open Petdex format. The Buddy is a full agent: ask it anything and it works through multi-step tasks with approve/decline buttons right in its speech bubble.
What are the system requirements?
Windows 10/11 (64-bit), macOS 11 Big Sur or newer (Apple Silicon or Intel), or any x64 Linux distro (Ubuntu 20+, Fedora 36+, Debian via .deb). Skales uses about 300 MB of RAM in normal operation - no GPU required. Local models need whatever the model needs; the Skales Local catalog lists the size of each one before you download it.
Stop configuring. Start asking.
30 seconds from download to first task. No Docker, no terminal, no subscription -
and everything stays on your machine.