An open local AI workshop

Run it here.
See how.

KEYO Studio is a local LLM/SLM runner for developers. Inspect and own model execution on your machine, powered by our JavaScript CPU inference engine—not a wrapper around Ollama, llama.cpp, or vLLM.

Download size: about 158 MB for Windows. Both links deliver the same portable ZIP. If GitHub is slow, try the Hugging Face mirror; speed depends on your connection and location. The first test model is a separate download of about 1 GB. No models are bundled.

Slow or interrupted download? Use your browser’s Downloads menu to pause or resume when supported. If you switch hosts, start a fresh file—do not combine partial downloads. Extract All, keep the files together, then open KEYO Studio.exe. No separate Node.js installation is needed for the Windows app.

Developer alpha 0.1.0-alpha.5. Ugandan-built by KEYO Technologies. Unsigned Windows portable and Linux desktop archives are available with our own CPU engine and model-download library. Extract the entire archive before opening the app. Windows runtime execution is not certified by the Linux build host. CLI requires Node.js 20+; desktop packages include their runtime.

Explore the workspace interface preview →

Ugandan-built by KEYO Technologies, KEYO Studio aims to contribute to the AI revolution and add value to the development of AI and future superintelligence.

KEYO / execution pathlocal · alpha
execution stays on your machineNO TELEMETRY
JavaScript CPU inferenceLocal model filesVersioned localhost APINo paid APINo telemetry
Choose your platform

Windows, Linux and Mac—not iOS.

Desktop builds use our own local CPU engine. Models download separately. None of these archives installs an iPhone or iPad app.

PlatformDownload and instructionsVerification status
Windows x64Portable ZIP, about 158 MB. Windows instructions.Unsigned developer build; execution on Windows is unverified.
Linux x64 desktopArchive, about 123 MB. Linux instructions. Requires a graphical desktop—not ARM or a headless VPS.Packaged local model-download/chat smoke passed; not every distribution is certified.
macOS Apple Silicon / IntelSeparate arm64 / x64 previews, about 130 / 134 MB. Mac instructions. No separate Node.js installation.Must finish signing on a Mac. Not Developer ID signed or notarized; Mac GUI execution is unverified.
iOS iPhone / iPadNo mobile application yet.Desktop archives do not run on iOS. Safari offers only an interface preview—not local inference.
Linux / x64 desktop

Extract, launch, then load a model.

Download the Linux desktop archive — about 123 MB. Open Terminal and run these commands as your normal user—not with sudo.

cd ~/Downloads
tar -xzf KEYO-Studio-0.1.0-alpha.5-linux-x64.tar.gz
cd KEYO-Studio-0.1.0-alpha.5-linux-x64
./keyo-studio

Keep the extracted files together. In Model library, select Qwen2.5 0.5B Instruct → Download model → accept the licence and bandwidth confirmation → wait for verification → Load downloaded model → create a conversation. Try “2+2=? Reply with just the number.” The model downloads separately (about 1 GB); afterwards, generation works offline without an account or AI API credits.

Missing library or sandbox error? Report the exact message and Linux distribution/version. Do not disable the sandbox or run as root to bypass an error.

macOS / Developer preview

Choose your Mac’s processor.

macOS 13 or newer. Apple menu → About This Mac: an Apple M-series chip needs arm64 (about 130 MB); an Intel processor needs x64 (about 134 MB). Use the matching archive, not the Windows or Linux package.

Extract the entire ZIP and read READ-ME-FIRST.txt for the runtime’s minimum macOS version. These cross-built previews require a local signing step: open Terminal, change into the extracted KEYO-Studio folder, then run bash finish-on-mac.command. The script only applies local ad-hoc code signatures; it is not Apple notarization and does not bypass Gatekeeper. After it succeeds, open KEYO Studio.app and follow the same Model library → download → load → new conversation steps above.

Not a certified Mac release. If macOS blocks opening or the signing step fails, report the exact message and Mac processor/macOS version. Do not disable Gatekeeper, remove quarantine, or bypass security protections. A real Mac GUI test and Apple-signed/notarized distribution remain separate requirements.

Windows / First run

Download, open, then try a model.

Portable means there is no installer wizard. Use Windows x64. This is an unsigned developer alpha; execution on Windows is not yet certified.

Step 1 / Get the app

Choose the Windows ZIP.

Use either Windows app button above. The filename must be KEYO-Studio-0.1.0-alpha.5-win32-x64.zip (about 158 MB). Source code, source.tgz, github-upload and release-files ZIPs are for developers—not the runnable Windows app. Wait for the download to finish.

Step 2 / Extract everything

Right-click → Extract All.

In Downloads, right-click the ZIP, select Extract All, then Extract. Open the extracted folder and double-click KEYO Studio.exe. Do not run it inside the ZIP or move only the EXE. Keep the other files beside it. No separate Node.js installation is needed.

Step 3 / Get a small model

Open Model library.

Select Qwen2.5 0.5B Instruct, click Download model, and review the licence and bandwidth confirmation. Keep Internet connected until downloading and checksum verification finish. Allow about 1 GB for this separate model download. Start with 0.5B; 7B and 14B operation is not certified.

Step 4 / Load and chat

Load downloaded model.

Click Load downloaded model and wait for the loaded status. Create a new conversation, then send: “2+2=? Reply with just the number.” The expected answer is 4. CPU generation can be slow; allow it to finish or use the app’s cancellation control.

Step 5 / Test offline

Disconnect and ask again.

After downloading and loading, disconnect Wi-Fi and ask another short question. Generation runs locally without an AI API key or inference credits. Close and reopen the app to check saved conversations; load the downloaded model again if needed. No account is required for these local tests.

Troubleshooting / Stay safe

Check the exact message.

No EXE? Check the ZIP filename in step 1. Slow download? Try the alternate host or resume through your browser when supported. Windows warning, missing DLL or launch failure? Keep the exact message and report it with your Windows version. Do not disable antivirus or other security protections. Compare SHA-256 with the release checksum file if you need to verify the archive.

01 / What it is

A runner you can look inside.

KEYO loads supported model assets from a folder you provide, runs our transformer engine on your CPU, and exposes a versioned API bound to localhost with authentication. It is a small, inspectable starting point for local AI development.

Execution, not delegation

Your model. Your machine.

Prompts are processed locally. Explicit model downloads contact Hugging Face only after you accept the licence and bandwidth use. Downloads are pinned and checksum-verified; inference works offline afterwards. No paid inference API or telemetry is required.

A real alpha, with edges

Early and deliberately scoped.

The engine implements basic GPT-Neo, Qwen2 and Llama architectures. One small public Llama checkpoint has produced text locally; other families have mathematical fixture tests, not pretrained certification. Your existing Afro AI model is not yet certified.

02 / Quickstart

Download a model—or bring your own.

Desktop: extract the archive, open Model library, download Qwen2.5 0.5B, then load it and create a conversation. CLI: use keyo models and keyo download with explicit licence acceptance. Models are not included in the app download.

TERMINAL / commands to run locally
# Node.js 20 or newer
$ npm install -g ./afro-ai-keyo-studio-0.1.0-alpha.5.tgz
$ keyo inspect /path/to/model
$ keyo chat /path/to/model --prompt "Hello"
$ keyo serve /path/to/model
These are usage examples, not a live terminal. Download the source package first and run commands from your own environment.
RuntimeNode.js ≥ 20
ExecutionOur CPU engine
Model sourceLocal folder
03 / Input contract

Keep the model folder explicit.

The alpha expects these files to already exist together in a supported local model directory. Missing files or incompatible model configs can prevent a model from loading.

What you provide

A compatible model directory with a single safetensors checkpoint up to 4 GiB, or experimental indexed shards up to 32 GiB total (8 GiB per shard). KEYO reads matrix blocks from disk instead of decoding all weights into RAM. Its optional KEYO-owned native CPU kernel requires a local source build. Context stays limited to 512 tokens. GPU, GGUF and actual 7B/14B checkpoint certification remain unfinished. KEYO does not bundle private weights, automatically fetch them, or send prompts to a cloud service.

/path/to/model
├── config.json
├── tokenizer.json
└── model.safetensors

Required files · local only · not included in the package

04 / Scope & boundaries

Know what this release does—and doesn’t.

We’d rather name the edges clearly than imply broad compatibility. The alpha is a focused local execution experiment for developers.

In this alpha

Basic GPT-Neo, Qwen2 and Llama execution using our JavaScript CPU engine, with explicitly limited variants and model sizes.

Not included: GPU

GPU acceleration is roadmap work, not a capability of this release.

Local asset loading

Safetensors weights and BPE/tokenizer assets from a provided model folder.

Not included: GGUF or larger models

Quantization, sharded checkpoints and large-model certification remain future work.

Authenticated localhost API

A versioned API intended for local development and inspection.

Not included: fine-tuning

Training and fine-tuning are not part of this alpha.

No paid API, no telemetry

Execution stays local; there is no cloud prompt service or telemetry collection.

Unsigned developer desktop packages

Windows portable ZIP and Linux archive include our own engine and model library. Windows/macOS execution certification, signing, GPU acceleration and general 14B support remain unfinished. Never disable security protections to run a package.

Alpha caution: Local execution does not automatically make every workflow secure or production-ready. Review the source, keep the authenticated API on localhost, and evaluate the model and its license before use.
05 / Get the alpha

Choose the app or developer source.

To run on Windows, use a Windows app link above. Developer source and release-file bundles below are not Windows installers; they require Node.js 20 or newer.

KEYO Studio 0.1.0-alpha.5

Early developer alpha · Source package · Model weights not included

Published developer source: GitHub — KEYO Studio. SLM/LLM roadmap. Live browser interface preview: Hugging Face — KEYO Studio. The preview does not run models.