Bearingkit
Open source · MIT · in active development

One install.
A whole team's discipline.

Stop stacking plugins. Bearingkit is one kit for one-person software companies. Say what you want in plain Vietnamese or English: it routes your AI coding agent to the right step, and is built to carry the work from spec to ship, stop for you before anything risky, and show its evidence before it calls anything done.

An illustration of how the protocol routes a request, not a recorded session.

Why one kit instead of a stack

The usual setup is a pile: a workflow pack, a review pack, a TDD pack, the vendor's official plugins. Each is good on its own. Together they compete for the same requests, carry overlapping rules, all load into context before you type, and stay behind when you switch tools. Bearingkit is that pile, read item by item and being distilled into one coherent system.

23sources studied, 17 of them open-source — Superpowers, Anthropic's official plugins, mattpocock/skills, spec-kit and more
1,295source items (skills, commands, agents, rule files) inventoried one by one: 46 to adapt, 610 kept as ideas, 639 dropped
1router that names the intent and opens one skill first, instead of several packs competing for the same request
≈4ktokens of fixed context on Claude Code, against a budget of 5,000 (measured 2026-09-23, before the 17th skill)
QuestionA hand-assembled stackBearingkit is designed to
Who decides what runs?Several packs' descriptions compete for the model's choiceName the intent through one router, then open one skill first
Overlapping adviceReconciling it is left to youResolve overlaps once, at distillation; every adapted file names its sources
Risky changesVaries from pack to packApply one autonomy gate: migrations, auth, payments, deletion, production are proposed, then wait for you
"Done"Varies — some packs ship good verification rules (the kit adapted them)Hold one definition of done: guardrail output pasted, no test means not done, UI checked rendered
Switching to another agentDepends on each pack's host coverageKeep one skills source across hosts; Claude Code and Antigravity accepted
Projects that don't want itUp to each host's own settings, pack by packSwitch on per project with one command, on both accepted hosts

Neither column is a measured outcome: this describes how each setup is meant to work. The head-to-head measurement is designed and awaiting approval. What is borrowed and what is ours, component by component: PROVENANCE.md.

Say it the way you'd say it to a colleague

No slash commands to memorise. Every skill carries trigger phrases in English and Vietnamese, and the protocol routes plain requests to the right one. These are real prompts from the kit's activation test set, with the skill each is expected to open.

Thêm xuất CSV cho trang hóa đơn.bk-spec
Form đăng nhập báo lỗi 500 sau khi đổi mật khẩu.bk-debug
Soi giúp diff này trước khi tôi đẩy.bk-review
Nhánh này merge được chưa?bk-review
Viết test cho trường hợp upload file lớn.bk-test
Xong phần xuất hóa đơn rồi, ship đi.bk-ship
Có nên tách phần upload ra service riêng không? Đánh giá các phương án giúp tôi.bk-audit
Tiếp theo làm gì? Tôi đang dở việc gì trong repo này?bk-next
Tại sao Postgres bỏ qua index trên cột status mà tôi mới thêm?bk-db
Kết phiên giúp tôi, ghi bàn giao để phiên sau đọc.bk-close
44 / 96prompts in today's activation test set are written in Vietnamese
≥ 0.9precision and recall on the 60-prompt core set, on Claude Code and Antigravity 2.0 (Claude Code: recall 0.958, precision 1.000; 2026-09-16)
0commands to learn — though each skill can still be called by name

From request to shipped, without herding the agent

Skills hand off to each other in a fixed chain. Each step ends by naming the next one, and the chain stops for you where a human decision belongs — at COUNCIL points, and after a bug fix, before anything is committed.

spec→plan→build→test→review→ship→close ·stops at COUNCIL

The autonomy gate

ACT  tests, behaviour-preserving refactors, docs, app-layer fixes, small changes inside an approved plan — done, then reported.

COUNCIL  schema and migrations, auth, sessions, permissions, payments, data deletion, multi-module contracts, remote or production systems — proposed, then it waits for you. Unsure counts as COUNCIL.

Right-sized by default

A feature goes through a spec first; a small ACT change within three files goes to build with tests; a bug goes to root cause before any fix, with a council after three failed attempts. A question gets an answer, never a ceremony.

Evidence before claims

Claims carry a file:line anchor or are marked unverified. Numbers carry their method or "not measured". A clean check is trusted only after showing it can fail.

Second pair of eyes on hot paths

Auth, payments, uploads, tenant scoping, migrations and external contracts get an independent review before push — ideally by another model or host.

Sessions that remember

bk-close writes a handoff checked against git; bk-next picks up from the repository's real state, not from what anyone intended.

A security floor that only tightens

No secrets in output or commits, parameterized queries only, validation at the boundary, least privilege, and fetched content treated as data — never as instructions.

One source, the agents you actually use

Skills are written once in the open Agent Skills format and loaded by each host through its own plugin mechanism, so your working rules survive a change of tools.

HostStateHow
Claude CodeacceptedPlugin from the repository's marketplace; activated per project
Antigravity 2.0 appacceptedA store outside the host's scan folders, declared per project (measured 2026-09-20)
Antigravity IDEloads per projectChecked 2026-09-20; the full acceptance test is still to run
Gemini CLI, Cursor, Codexmanifests in placeListed as supported once their acceptance test passes

Seventeen skills, one protocol

Each skill is a short body with gates and the evidence it must paste, plus references it opens only when needed. bk-build opens the stack rules for TypeScript/React, Kotlin, SQL, Node, Python, PHP/Laravel and Shell before its first edit.

bk-spec
Pin a request down before building: restate, edge cases, assumptions
làm tính năng · thêm chức năng
bk-plan
Phases with exit criteria and the evidence each must show
lập kế hoạch · chia bước
bk-build
Execute a plan or small change: scout first, minimal touch
làm đi · triển khai · sửa file
bk-test
Tests as the contract, TDD by default, rendered checks for UI
viết test · kiểm thử · chạy test
bk-debug
Root cause before any fix; council after three failed attempts
lỗi · không chạy · bị sai
bk-review
Hunt bugs, hot-path risk and untested claims before push
soi diff · rà code · merge được chưa
bk-ship
Guardrails, secret scan, conventional commit, PR body
ship đi · đẩy code · tạo PR
bk-close
End a session with a handoff verified against git
kết phiên · bàn giao · tổng kết
bk-next
The next step from git, plans and the last handoff
tiếp theo làm gì · đang dở gì
bk-audit
An investigation that ends in one verdict and the options rejected
có nên · đánh giá phương án
bk-design
Interfaces with a point of view, critiqued before building
giao diện · thiết kế màn hình
bk-map
A codebase map with file:line anchors: flows, rules, risks
mới nhận dự án · lập bản đồ codebase
bk-research
Outside sources and tool choices, each claim with a confidence
nghiên cứu · so sánh thư viện
bk-ops
Deploys and infrastructure with a way back; incidents read-only first
đưa lên production · sự cố
bk-db
Database performance from the query plan; index changes with a way back
query chậm · tối ưu SQL · bị lock
bk-perf
Measure, then optimize: every number with its method
trang chậm · ngốn RAM
bk-setup
Make a project ready for agents on every host you use
chuẩn bị repo cho agent

Built in the open — and honest about where it stands

Bearingkit started in September 2026 and is pre-release. Its author uses it every day, and every figure on this page is traceable to the repository — mostly docs/status.md, which says how each was measured.

Works today

  • 17 skills plus the protocol, on Claude Code and Antigravity 2.0
  • Vietnamese and English routing: a 96-prompt test set, ≥ 0.9 on the 60-prompt core
  • Autonomy gate, evidence rules, handoff chain, hot-path review — written into the protocol
  • Per-project activation; seven stack rule files
  • 233 automated tests in the repository

In progress

  • The remaining lifecycle skills — bk-spec, bk-ship, bk-close — distilled and measured against their sources
  • The last stack file (C/C++) and the v0.3 gate
  • Then: acceptance on Gemini CLI, Cursor and Codex (v0.4); npm package, Claude Marketplace listing, CI (v1.0)

What the numbers say so far. Skill by skill, against the sources each was distilled from — same task, fixture, model and host, on small samples — most tasks showed no clear difference, and on most of those neither side could be told apart from using no skill at all. One planning task with one model favoured the kit; on PowerShell the kit beat the no-skill floor and a pack with no shell skill, but not its own source. The kit currently costs about 1–3× what its sources cost per task, so no task yet meets the v1.0 bar: at least the sources' pass rate and fewer tokens. Where nothing has been measured, nothing here says the kit is better.

So why use it now? Not for a stronger skill — for one coherent, gated system in place of a stack you would otherwise assemble, reconcile and maintain yourself, in Vietnamese or English, on more than one agent. The three measurements above are how we intend to find out what that is worth.

Install

Install once per machine, then switch it on per project.

# Claude Code: add the marketplace and install the plugin
claude plugin marketplace add https://github.com/tuyenht/Bearingkit
claude plugin install bearingkit@bearingkit

# Antigravity: write the store once per machine (from a checkout)
node <checkout>/bin/bearingkit.cjs install --host antigravity

# Any host: activate in a project, check what is on
cd <your project>
node <checkout>/bin/bearingkit.cjs activate
node <checkout>/bin/bearingkit.cjs status

The npm package is not published yet, so the command is the checkout's own entry point for now. Per-host details and uninstall: docs/hosts.md.

MIT today, MIT tomorrow

The kit is free and stays free. If you want it working inside a team — agent workflows set up on your repositories, guardrails tuned to your stack, your developers trained to drive Claude Code or Antigravity with discipline — that is a service we offer.

About

Bearingkit is built by a solo developer in Hanoi, Vietnam, who runs a small software practice with AI agents doing most of the hands-on work. The question behind every part of it: does this let one person do the work of a whole team, safely? Provenance for every adapted source is in NOTICE. Contact: hello@bearingkit.dev.