Intake
Ideas arrive as plain language, by a file or a chat command. Each becomes a goal with a stated "done when", reviewed by a person before it is active.
Direction is human. Labour is machine. Nothing counts until a machine can check it.
Artemis is a long-running project to turn a small, self-hosted computing estate into a self-directing engineering team. You describe an outcome in a sentence. Artemis works out who should act, what to change, where it lives and why it matters, then plans the work, carries it out on its own hardware with its own models, proves the result, and tells you what it did and why.
Humans set direction and approve the few things that deserve a human decision. The machine does the labour. This page is the top-down view; it leaves out the operating details on purpose.
Most automation runs a fixed script. Most "AI agents" run one conversation and forget it. Artemis is neither. It reads the goals it has been given, breaks them into tasks, assigns each task to the cheapest worker that can do it, checks the work mechanically before anyone sees it, lands it, measures what happened, and feeds the lesson back into the next round. Over time it should need the humans less, not because it hides its work, but because its record of being right grows.
People decide what to pursue and sign the rules the system runs under. The system does not activate its own goals or widen its own permissions.
A change is done when a test goes green that was first proven to go red. Where no mechanical check exists yet, the system asks rather than guesses, and the ceiling rises only when a new check is built.
The models that do the work run on hardware we own. Confidential material never leaves the premises and never touches a third-party model. A separate, sealed lane exists for work under contract, and it does not open until a rehearsal proves it cannot leak.
Each is a real component with a job, and each is small enough to understand on its own.
Ideas arrive as plain language, by a file or a chat command. Each becomes a goal with a stated "done when", reviewed by a person before it is active.
A planner reads active goals and the current state of each project, selects the next milestone, and writes small, verifiable tasks with explicit success criteria.
Work is routed to the least expensive capable worker first: local open-weight models on our own GPUs for most tasks, stronger models only when a task has failed locally and only within a budget.
Every task specification is read by a reviewer that can approve, fix, or refuse. Refusal exists for one question only: should this run at all.
Each task carries a mechanical check. If it does not pass in a clean environment, the task is not done.
Passing work is merged under rules that depend on how much damage a mistake in that project could do. Low-risk projects land automatically and report afterwards; high-risk ones ask first.
Every attempt, pass, retry and escalation is recorded, so the system can tell which tiers earn their keep and where its own estimates are wrong.
Outcomes become lessons; lessons become changes to the rules and the prompts; the map of the estate regenerates itself so the planner is never working from a stale picture.
The system decides which model runs on which cards for which task, leasing and releasing capacity rather than pinning one model to one card forever.
Alerts, approvals, digests and questions flow through one chat server with a small set of commands. Decisions that need two people need two presses. Messages carry names, identifiers and links, never the work itself.
What may leave the network is declared and proven, not assumed. The sealed lane for contract work has its own storage, its own models and no route out.
And its leak rehearsal.
So an idea typed in chat becomes a reviewed goal without a human relaying it.
From measurement to learning, so lessons change behaviour without a person editing a prompt.
One proven check at a time.
The usual failure of autonomous systems is not that they do too little. It is that they do the wrong thing confidently, and nobody can tell afterwards why. Artemis is built around the opposite bet: every action leaves a record a person can audit, every rule it follows was written by a person and can be read, and every increase in its freedom is paid for with a new way to prove it right.
Slower to start. Far easier to trust.
Progress notes will appear on this page as milestones land. The project is private while it is being built; the ideas are not, and we are glad to talk about them.
[email protected]