Ungoverned AI
no limits- Acts on unverified data and missing documentation
- Blows past materiality limits and usage caps
- No gate between "uncertain" and "posted"
- The failure is silent — it looks like success
Aleq's autonomy is a confidence score that climbs as it proves itself on your books — inside limits you set. Below the limit it drafts and waits; above it, it posts and shows the receipt.
Everything Aleq does clears two bars: can it do the work, and has it earned the right to — with you, on your books. Not one or the other. Both, multiplied.
multiplication, not addition — anything times zero is zero. all the skill in the world, without earned trust, does nothing.
Overrule Aleq and turn out right — your word weighs more next time. Turn out wrong — it weighs less. Authority here tracks accuracy, not seniority. Not even the boss can shout the books into agreeing.
When Aleq checks with you instead of guessing, that counts as getting it right. It's never punished for asking — so it never bluffs to look confident.
It doesn't earn autonomy by being installed a long time. It earns it by being proven right — task by task, on your books, with your sign-offs as the evidence.
and some things don't bend: your preferences it learns in days · your controls it questions rarely · the rules that protect your books, never
The mode you see per task isn't a switch you flip — it's a read-out of how much Aleq has proven itself, gated by the limits and thresholds you've set.
New, unproven, or outside your limits. It looks things up and drafts the work. You make every call and post it yourself.
It prepares the entry and shows its work. One click to approve — or reject — before anything posts.
It does the work and posts it, logged. Reverse it any time with a standard reversing entry.
Today most of it is held together by hand. The rule that one person can't both approve a bill and pay it sits in a spreadsheet someone forgets to update. Approvals get chased over email. The week before the auditor shows up, someone builds the evidence by hand. And when they ask "why did this post?" there's no record — just memory. Aleq takes that work off your plate. Not by asking you to trust it, but by proving every move.
| Who can do what | tracked by hand in a sheet | stale |
| Approvals | chased over email | no trail |
| Audit evidence | built the week before | 38 hrs |
| "Why did this post?" | rebuilt from memory | no replay |
Aleq learns from what your team already does. What happens next teaches it. It earns your trust with proof — you don't hand it over up front.
the statement, the matches, the ledger, and how you've coded these before
you coded it this way 412 times and never reversed it
routine matches go through; two big ones wait for you
a logged entry you can reverse; what happens next updates the confidence score
When confidence on a task crosses your threshold, that task — and only that one — runs alone. Bank matching can run itself while every payment still waits for your OK. Each task has its own limit, and you set it.
runs on its own after 200+ clean matches · 18,420 runs · 98.7% right · routine cash only
always waits for you · drafts the wire, shows the proof, holds for your OK — never runs alone
gets the close ready · a person signs the month closed
| Task | Runs alone after | Confidence | Mode now |
|---|---|---|---|
| Matching the bank | 200 clean matches | 1.00 | Auto |
| Coding bills | 150 times · within 4% | 0.99 | Auto |
| Sending reminders | small ones, below your escalation line | 0.96 | Auto |
| Sending a payment | never alone · always waits for you | 0.92 | Assist |
| Closing the month | person only | — | Manual |
A reliable rule still stops when the amount, account, or risk is outside what you've approved. Then it asks a specific question instead of guessing.
The early-pay discount is worth taking, but this is over your $100k auto-payment limit.
Everything Aleq does gets a logged event you can reverse. It's not a screenshot of a dashboard. It's a real record your auditor can check on their own.
| Why | 247 lines tie to the bank to the penny |
| What set it off | difference = $0.00 · within the rules |
| Rule | P-RECON-DIFF-ZERO |
| What changed | posts JE-12491 · closes May |
Every entry Aleq posts can be reversed with a standard reversing entry, any time — the same mechanism your team already uses, not a special countdown window. You can also test any action on your real books first — it shows what would happen and changes nothing.
| approved | wire | $188,440 |
| paid early | discount captured | −$3,768 |
TAMi — The Aleq Mind — is the part of Aleq that tracks confidence per task from your team's own approvals and corrections. Above your threshold, it runs alone; below it, it drafts and asks. Every correction you make updates the score for next time.
| Rule | Confidence | Status |
|---|---|---|
| Stripe transfer ⇒ Stripe Clearing | 1.00 | runs alone |
| AWS bill in last 3 days, within 4% | 0.99 | runs alone |
| Kestrel early-pay discount always on | 0.92 | learning |
| Ironwood pays only after a reminder | 0.42 | asks first |
We ran 57 high-risk accounting tasks that leading AI models got wrong on their own. Then we ran the same 57 with Aleq in the loop — same models, the only change is what Aleq lets them do when the answer isn't certain.
From Aleq's internal eval runs — leading AI models (Claude, Codex, Gemini, Qwen) on the same tasks, with and without Aleq. Request the full methodology →
When a month closes, Aleq locks the period: postings, reversals, and edits to closed entries are blocked at the ledger level. Reopening a closed period is its own explicit, logged action — not something that happens by accident.
| Entries locked | 8,412 | locked |
| Closed by | Kurt | 14:02 |
| Status | closed | no edits |
Before any new version of Aleq goes live, we test it on real close, collections, and reconciliation cases — and run a set of safety checks. It has to score at least 0.90 out of 1.0. If it doesn't pass, it doesn't ship.
| month-end close | 0.96 | pass |
| collections | 0.94 | pass |
| reconciliation | 0.97 | pass |
| safety checks | — | enforced |
| must score 0.90+ | — | ships |
The same models power everything now. What differs is the permission model — autonomy someone configures on day one, or autonomy that's earned on your books and can be lost.
Asks about everything. Drafts, shows its work, posts nothing alone — exactly how you'd want a new controller to start.
Runs what it's proven — reconciliation alone while payments still ask — and every question it asks makes the next one rarer.
The board pack is drafted the morning after close. The reminder went out at day one past due. You stop assigning the work and start reviewing it.
Not blindly — trust should be earned per task, which is how Aleq is built. Autonomy is not a switch you flip; it is a confidence score each task earns from your team's own approvals and corrections. Below the threshold you set, Aleq drafts the work and waits; above it, it posts and shows the receipt. Every action is logged with the rule that triggered it, and every entry can be reversed with a standard reversing entry. You connect read-only, start everything in Assist mode, and let individual tasks move to Auto only after they have proven themselves on your books.
Connect read-only. Start everything in Assist. Let a task move to Auto only when it's earned it.