Independent software studio · Greece

Tools for people who run their own models.

We build open-source developer tools that make local AI models useful for real work, on hardware you already own. No cloud account, no telemetry, and your data stays on your machine.

3VS Code extensions
9Open-source repos
2Open Greek models
0Telemetry endpoints
Flagship · VS Code

Forge

An open-source coding agent for VS Code, built for people who run their own models.

Forge loads GGUF models through llama.cpp and manages them properly: start, keep loaded, share across editor windows. Cloud models and the Claude Code and Codex CLI agents are there when you choose to add them.

  • Local runtime, managed. Model loading, GPU placement and context limits are handled for you.
  • Tool calls that hold up. Tuned for the context and output limits that real local models have.
  • Every change reviewable. Inline diffs, per-turn checkpoints and Keep/Undo for each file.
  • Long sessions. Context compaction keeps work going past the window limit.
  • Plays well with others. MCP servers, delegation to CLI agents, and remote control from your phone.
VS CodeApache-2.0
Forge running an agent turn in the VS Code sidebar: a tool call, an inline diff of the edited file, and the Undo and Review buttons for the turn
The Forge model picker listing local llama.cpp GGUF models alongside Ollama local and cloud models
Flagship · Desktop

HalluScribe

A private, searchable memory of every AI coding session you have had.

HalluScribe reads the session files your AI tools already write, summarises each one with a local model, and keeps a readable Markdown archive on your disk. Nothing is uploaded and there is no API bill.

  • Many sources. Claude Code, Codex and Forge sessions, plus exported ChatGPT, Claude.ai, Gemini and Grok chats.
  • Local summaries. Gemma 4 by default, turning 200K-token logs into short, structured entries.
  • Runs on its own. A nightly sweep archives new sessions automatically.
  • Memory for agents. A read-only MCP server lets coding agents search what earlier sessions learned.
DesktopOpen source
The HalluScribe session archive: 3,059 summarised sessions from Claude Code, Codex, Forge, Grok, ChatGPT, Gemini and Claude.ai, with titles blurred for privacy
In progress

NVIDIA Nemotron, in Forge and in Greek

We are adding first-class support for NVIDIA's open Nemotron models to Forge, with a tuned tool-calling profile for local agentic coding on NVIDIA GPUs. Next, we plan to fine-tune Nemotron for Greek, using the same pipeline and native-speaker data behind our Greek Gemma 4 model. Greek-language benchmarks will be published in the Forge repository.

More from evolv

Smaller tools, same principles

VS Code

CacheWarden

Keeps assistant prompt caches warm while you step away, so you don't pay to rebuild them.

VS Code

Forge Relay

A coordination board for several AI agents at once: file claims, a shared event feed, and stop and pause for all.

Desktop

HalluMeter

A desktop ring that shows how full your model's context window is, and when quality is likely to drop.

DesktopOffline

Gemma4kids

An offline learning companion for children aged 6–11, answering in their own language. No internet, no subscription.

Open models

Greek, because nobody else was doing it

Greek is one of the EU's 24 official languages and one of the least served by open models. We release ours publicly, in formats that run on a single consumer GPU.

gemma-4-E4B-it-GR-v2

Gemma 4 E4B fine-tuned for Greek speech understanding and Greek text, trained on 3,217 native voice recordings across 17 categories and 2,476 curated question–answer pairs. Ships as a single GGUF for Ollama or llama.cpp.

JOY — Greek voice

A high-quality open Greek text-to-speech voice for Piper, recorded by a native speaker. Released under CC BY-NC 4.0.

About

A small studio with a narrow focus

evolv is an independent software studio in Greece. We work on one problem: making open-weight models genuinely useful for real work, on hardware people already own.

Raw inference speed is llama.cpp's job, and it does it well. The hard part is everything above the runtime: tool calls that don't break halfway, context that doesn't quietly overflow, file edits you can undo. That layer is what we build.

Efstathios Outas Founder · Software engineering
Chara Kaltsou Greek language & voice data · the voice of JOY
  • Local by defaultCloud providers work only when you configure them yourself.
  • No telemetryNo analytics, no usage pings, no auto-update calls.
  • Your keys stay yoursCredentials live in the editor's secret storage, never in a config file.
  • Reversible by designEvery file-changing action is visible, confirmable and undoable.

Get in touch

Questions, collaboration, or feedback on our tools.

hello@evolvlabs.dev