Proof stands in the light.

Grit builds AI systems for companies and individuals. Our product, RADAR, runs AI coding agents and closes a task only after it re-runs the acceptance test itself.

From Grit's own work

Observed, not a controlled comparison

  • 74 of 1,608times a coding agent reported a task as finished, Grit's re-run of the acceptance commands failed
  • 4.6%of those reports did not pass the re-run

A claim alone does not close a task. See the full record

Grit internal ledger, snapshot 28 Sep 2026 06:00 UTC

How we work

A task is finished when its test passes again.

  1. Every task gets its acceptance test before any work begins.

    Example. Nothing runs here.

    TaskFix the redirect after sign-in

    Pick the test that decides “done”

    Start stays locked until the task has a test anyone can re-run.

  2. The agent that did the work does not close it. The test is run again.

    Example. Nothing runs here.

    TaskFix the redirect after sign-inOpen

    Agent

    “Done. All tests pass.”

    Locked: the agent that did the work cannot close it.

    Checker

    Open: waiting for the checker.

  3. Every figure we publish comes from a record that names its source and date.

    A real figure, not an example

    From Grit's own work · Observed, not a controlled comparison

    4.6%

    In Grit's own work (20-28 Sep 2026), workers running one coding agent reported a task as finished 1,608 times (1,457 tasks); Grit's immediate re-run of the acceptance commands on the same server failed on 74 of those reports (4.6%). A claim alone does not close a task.

    Grit internal ledger, snapshot 28 Sep 2026 06:00 UTC

    See the full record

  4. The work is designed to run on more than one family of AI models.

    Example. Nothing runs here.

    TaskFix the redirect after sign-in

    AgentClaude Code

    Testpytest tests/test_login.py -k redirectSame test

    The agent, and the model family behind it, can change. The test that closes the task does not.

radar
sample data

radarsample data

Scroll or swipe to move around the screen.

RADAR

Our product

RADAR coordinates AI agents and checks what they deliver.

When an agent reports that a task is done, RADAR runs the task's acceptance test again. A passing test closes the task. Every step is written to a ledger you can read.

Example

Try it: can the agent close its own task?

Example tasks and output, written for this demo. Nothing runs here.

Task Open

Add a discount code to checkout

Acceptance test

pytest tests/test_checkout.py -k discount

Open: the agent is working on it.

Ledger

    Nothing recorded yet.

    Each task you close with proof lights one tile of the Grit disk.

    What we do

    Grit connects a company's data and automates its work, with systems that check their own results.

    • Websites

      Designed and built with you, from the first draft to the live site: accessible, quick to load, and in the languages your visitors read.

      For companies and individuals

    • AI automations

      Repetitive work handled by AI, the same steps done the same way each time. Every result is checked before it is used.

      For teams whose work repeats

    • Autonomous systems

      Agents that carry a piece of work from start to finish, built the way RADAR works: a task closes only when its test passes again.

      For work with a clear definition of done

    • Data

      Scattered data brought together in one place and cleaned, so it is reliable enough to work with and to decide on.

      For organisations whose data lives in many places

    The company

    The people who lead Grit.

    • İsmet AydınCo-founder
    • Mustafa TokerCo-founder
    • Ali Baha BerkalCo-founder

    Bring us the work you need done properly.

    Contact Grit