Goals in · finished work out

Send one goal.
Get finished work.

Mrrlin plans it, coordinates your agents, sends weak output back, and stops at every step that needs a human.

Map my first workflow
Director plan · 8 tasks88%

Launch our new marketing site, SEO base, and first growth loop before the next founder update.

Claude Code · local CLICodex CLI · local subscriptionGemini review · consensus
Runs on your Claude Code and Codex CLI subscriptions — no metered API bills.
Turn positioning into homepage messagingaccepted
Build landing page and product demo sectionin execution
Compare SEO titles across three modelsconsensus review
Package launch checklist and analytics noteshandoff ready
?Confirm final CTA before publish1 question
Launch packagesite + SEO ready

From teams running Mrrlin

Less coordination, more shipped
“It cut a week of coordination down to an afternoon. I stopped being the bottleneck between five agent tabs.”
Marta KovalenkoFounder, B2B SaaS
“Free tiers of Gemini and Grok now do the volume work. My paid agents only get the calls that actually need them.”
Daniel RothTechnical co-founder
“I can see which model did what, what got rejected, and why. That transparency is the reason I trust it with client work.”
Anna SilvaAgency owner
“More work ships, and it ships better — the review pass catches the things I used to catch at 1am.”
Tomás NevesHead of Growth
“Roughly half our tasks now close without anyone touching them. I only show up for the decisions.”
Jamie WuOps lead, 9-person team

Workspace

One goal. Every agent visible.

Plan, execution, review, and the approval inbox live in one place — not across five chat tabs.

Founder Growth/marketing site · SEO · outreachAutopilot on
Open12
Needs you2
In review5
Ready8

Active goals

G-01Launch conversion-ready marketing site88%
G-02Founder-led SaaS outreach74%
G-03Turn feedback into a product spec81%
G-04Client campaign pack57%
G-05Technical project audit69%

Approval inbox

Needs you · 10:44

Marketing site is ready to publish

Copy, SEO, schema, and mobile checks passed. Two claims were rewritten after review.

Approve launchSend back
Review · 08:35

6 weak personalization lines rejected

Memory · yesterday

SEO opportunity map saved

Chat tools vs Mrrlin

Not a chatbot. Not a task board. A director that finishes work.

Chat-based AI tools make you the project manager. Mrrlin runs the whole loop — and does it cheaper.

Chat-based AI tools
Mrrlin
Token cost
Full context re-sent with every prompt
Progressive context compression — the fewest tokens per task
Project memory
Forgets between sessions — you re-explain and re-instruct
Memory is captured continuously — context stays attached to the work
Work structure
One long chat thread, no visible progress
Spec-driven: epics, tasks, and a board — progress at a glance
Clarifications
Blocks the chat with one question at a time
Async inbox — answer questions across dozens of tasks when it suits you
Models & cost
One expensive model for everything
A pool of models — the Director puts cheap models on simple tasks

scroll to follow the loop →

Without Mrrlin

  1. You

    paste the context again

  2. One model, one chat

    one opinion, one pass

  3. “Done.”

    you are the only check

The chat ends and the context goes with it. Next time you paste it all again.

With Mrrlin

plan

  1. You

    A goal

    One sentence, in your own words.

  2. Mrrlin · Director

    Spec

    Context is pulled from memory, not re-explained by you.

  3. Gate · review panel

    Consensus

    Independent reviewers, rounds until they converge.

    no consensus → rerun

  4. Isolated

    Run

    Its own git worktree. Every artifact recorded as it goes.

  5. Gate · proof required

    Receipt

    A live address, a green build, or your own sign-off.

    no proof → stays open

Written back

Memory

What was decided, and why.

remember

Two gates the work cannot walk around, and a memory it cannot forget.

Execution layer

Two gates between
a goal and
done.

Everything that usually falls on you after the prompt — clarification, routing, review, retries, handoff. Six steps, and twice the work has to prove itself before it moves on.

01

Start with the outcome

Founder input

One plain-language business goal becomes scope, constraints, acceptance checks, and visible tasks.

02

Questions surface first

2 questions

If context is missing, Mrrlin asks before agents waste time on the wrong work.

03

Every task gets the right agent

Local CLI

Routine work goes to cheaper models; hard or code-heavy work routes to Claude Code, Codex CLI, Sonnet, DeepSeek, or local agents.

04

The first answer is not final

Consensus

Consensus review checks evidence, weak assumptions, brand fit, and acceptance criteria before delivery.

05

Weak outputs go back

Re-run

Rejected drafts, thin research, or shaky implementation are sent back with review notes instead of landing in your inbox.

06

Only finished work asks for you

Needs you

Sensitive claims, sends, publishing, deploys, and final deliverables wait for human approval.

What lands in your inbox

Marketing launch pack

Homepage, SEO, analytics, launch checklist

copy acceptedschema verifiedCTA needs sign-off

Founder outbound pack

50 leads, ICP scoring, personalized sequence

12 rejected leads3 models reviewedsend list waiting

Repo audit report

UX gaps, implementation risks, scoped fixes

risk rankedtests proposedclient report ready

Where to start

The work your team keeps postponing.

Founder-led SaaS

Launch a sharper marketing site, SEO base, analytics, and first growth campaign without becoming the AI project manager.

Launch package8 tasks · founder sign-off
Homepage positioningaccepted
SEO base + schemarunning
Growth loop QAconsensus
!CTA directionneeds you

OutputPublished site, SEO checklist, launch plan, reviewed copy

AI-native agency

Turn a client brief into campaign strategy, content, landing variants, and final QA across multiple agents.

Client pack7 tasks · multi-agent QA
Brief parsedaccepted
Landing variantsin execution
Claims + tone checkreviewing
Client deckhandoff

OutputCampaign brief, content pack, review notes, client-ready assets

Growth lead

Run acquisition and onboarding experiments end-to-end while keeping approvals and measurement visible.

Growth sprint6 tasks · metrics attached
Hypothesis rankedaccepted
Onboarding copybuilding
Event trackingverified
!Ship threshold1 question

OutputExperiment plan, copy, tracking plan, results memo

Technical agency

Convert a client request or repo audit into scoped work, fixes, tests, and a polished recommendation report.

Audit delivery9 tasks · repo context
Repo scancomplete
Fix planin execution
Risk rankingconsensus
Client reportready

OutputDelivery plan, implementation tasks, reviewed report

Agent harness

One Director routes work to the right agents.

Claude Code, Codex, Sonnet, DeepSeek, or your own local setup. Mrrlin plans, asks clarifying questions, sends important calls to consensus, then routes execution per task.

Fig 0.0 — how the harness is wired

The state lives in Mrrlin and outlives every session. The model runs on your machine and is yours to swap. Nothing crosses between them without a receipt.

Claude CodeCodex CLICursorWeb appPhone

Mrrlin — the state that outlives every session

41 tables · the rules are database constraints, not team habits

Tasks and epics

Status, dependencies, and how much autonomy each one is allowed.

Project memory

Specs and evidence, ranked by how current they are — not just by keyword.

Decisions and why

The reasoning is stored with the decision, so the next session inherits it.

Receipts

The proof each finished task produced, attached to the task itself.

Your machine — your models, your keys, your subscription

Mrrlin never holds a model credential and never resells tokens

  1. 01

    Intent

    What you actually meant, pinned down before anything gets built.

  2. 02

    Spec

    Architect, security and scope reviewers — each on its own model — have to converge.

  3. 03

    Run

    An isolated git worktree, never your checkout. Artifacts saved as it goes.

  4. 04

    Review

    A separate read — never the executor grading its own work.

  5. 05

    Receipt

    A live address, a green build, or your own sign-off.

No receiptThe task stays open. A model saying it finished is not evidence, and the database will not accept it as one.

A resumed run ships a delta, not the whole history — measured on one real task: 641 KB → 25 KB of prompt.

The outside — nothing reaches it that you did not grant

Your repository

Branches, pull requests, workflows.

Your site or CMS

Pages that return a real address.

Messaging

Handoffs that actually arrive.

Ad accounts

Read and change, per client.

Fig 0.1

Plan with your context

Claude Code reads the repo and docs. The Director writes the plan, acceptance checks, and open questions before work starts.

claude coderepo context2 questions

Fig 0.2

Consensus finds gaps

Important decisions go through independent model review. Missing evidence and weak assumptions become specs.

reviewer areviewer bspec gaps

Fig 0.3

Route by complexity

Simple work goes to cheaper models. Hard tasks route to Sonnet, DeepSeek, Codex, or specialized agents with visible checkpoints.

sonnetdeepseekcodex

Human control

Autonomous doesn't mean invisible.

Routine work continues without another prompt

Weak output is rejected and sent back for another pass

Evidence, reviewer notes, and model consensus stay attached

Publishing, sending, and client-facing claims wait for you

DirectorNeeds your call

Goal

Launch founder-led SaaS outreach without sending risky claims.

6 weak lead matches rejectedre-run
2 email claims need stronger evidenceblocked
CTA variant B has stronger ICP fitreviewed

Question

Use the competitor comparison angle, or switch to a safer productivity angle?

Approve angleAsk for safer version

FAQ

Before you start.

What makes this different from a chatbot?

A chatbot responds to prompts. Mrrlin turns a business goal into an AI workflow, plans the tasks, runs the work, remembers context, and keeps iterating until the outcome is ready.

Which teams is it for?

Founder-led teams, early-stage startups, small product and growth teams, and agencies or consultants that need repeatable business workflows.

Do we need to change our tools?

No. The goal is to work across the tools you already use, not force a migration to a new system.

How much control do humans keep?

Human review stays in the loop for final delivery and sensitive decisions, while the system handles the repetitive execution work.

Give Mrrlin the first goal.

Start with the site, the SEO base, the outreach, or the client work you keep postponing. Mrrlin will map the execution path.

No credit card · No setup · One goal

Mrrlin product film