Lead-first
The team forms after understanding the project rather than from a rigid template.
OPEN LAB · FLAGSHIP
My open-source flagship investigates one practical question: when a software-agent team improves the outcome, and when it only adds coordination, cost and new ways to fail.
Build. Verify. Deliver.
01 · DESIGN
02 · BUILD
03 · VERIFICATION
Sample data
A Lead understands the objective, forms the team justified by risk and preserves issues, runs, costs, reviews and blockers in SQLite. Hard gates and evidence decide whether work is accepted, reopened or blocked. The repository is public, installation is verified on Windows and no stable release exists yet.
The team forms after understanding the project rather than from a rigid template.
Engineer, Reviewer, QA and other roles appear only when work and risk justify them.
Issues, runs, wakeups, interactions, costs and evidence survive in SQLite and can recover.
Runtime, authentication, health, authority and cost remain explicit for every agent.
A direct agent is compared with solo_lead, lead_quorum and full_team on the same case.
Inconclusive trials and cases where additional agents do not help remain part of the evidence.
The Lead translates the project into issues, dependencies, criteria and an exit condition.
→It hires the smallest structure capable of owning the work and its reviews.
→Heartbeats, runs and budgets keep activity and cost observable.
→Deterministic tests, review and evidence allow work to be accepted, reopened or blocked.
PUBLIC CODE + EVIDENCE
AI Teams 0.1.0 is public under Apache-2.0 with a reproducible Windows clean-room acceptance. Linux and macOS remain unverified; it does not promise full autonomy, guaranteed savings or full-team superiority. RogueBall remains a separate audited Godot runtime case.
Turn a prioritized opportunity into an owned, measurable and transferable system.
Explore this route↗Give the team the judgment and autonomy to operate, evaluate and extend what was built.
Explore this route↗You can inspect and experiment with the repository, but it remains active development without a stable release or verified multi-platform support.
Not necessarily. That is one of the hypotheses the benchmark compares against a direct agent.
Work, dependencies, runs, wakeups, interactions, costs, reviews and closure evidence in durable state.
No. It remains published as an audited Godot build, runtime and verification case with explicit provenance limits.
NEXT STEP
The flagship page exposes architecture, evaluation, code and limits. If your problem needs similar orchestration, we can start from its closure criteria.
This request is not open during launch. The information remains public to help you prepare the project; if you need consulting or business training, you can request a fit call.