Shared vow / 01
Reality
What observable evidence could prove the present framing wrong?
A constitution for finite attention
Place one difficult question into a field of 896 possible lenses. Sixteen will answer. Two vows—Reality and Care—cannot be routed away.
Route a real question
Private by construction: no API, tracking, or automatic storage.
Width / possible lenses
Your epistemic receipt
16 / 896
Hold a lens across reroutes, or release it to wake another. Exclusion is part of the trace.
Court of consequence
Non-claim: This route does not prove the decision, predict its outcome, or authorize action on anyone else’s behalf.
Open weight → bounded agency
Open weights provide inspectable capacity under a license. They do not confer consent, scope, tools, custody, or a right to act. This forge exports an inert packet; an operator must bind every plane—and independently witness every virtue claim—elsewhere.
Open weights ≠ authority.
Runtime access ≠ tool access.
Tool access ≠ consent.
Virtue claim ≠ earned reward.
A completion claim ≠ proof.
Local contract · the-other-880/agenttool-mission/v1
AgentTool SDK observed locally: 0.16.3 · re-observe before binding
Which weights, process, host, and custody boundary will execute? No receipt exists yet.
No shell, browser, account, wallet, connector, endpoint, or credential is attached.
Route a question to seed an observable criterion, a verifier, and a stop rule.
Self: AgentTool may record the runtime seat; a private vLLM/SGLang orchestrator preserves K3’s complete assistant messages and performs inference.
Hosted Ollama: K3 is newly reachable in principle, but the observed adapter drops thinking and tool-call history. The forge marks that seat blocked, not K3-native.
Handoff: append-only, project-private, server-readable continuity. Its authority list is a declaration, never a grant.
Collab: a local plaintext SQLite review journal. It coordinates task, evidence, and distinct-session acceptance; it neither runs nor authorizes the model.
KARMA: a local contribution contract, not an observed AgentTool reward primitive. Handoff, traces, and Collab may later carry sealed evidence references under separate authority.
Repair: exploitative structure pauses the case. Restitution is chosen by affected parties, earns no case reward, and never licenses mirrored harm.
The finitude paradox
In Moonshot’s reported BrowseComp setup, K3 scored slightly higher with context compaction beginning at 300K tokens than with the untouched 1M-token window. One benchmark is not a universal law. It is a useful wound in the fantasy that capacity and wisdom are synonyms.
FACT · BrowseComp · Kimi K3 (max), as reported by Moonshot. Difference: 0.8 points.
Mechanism → provocation
Facts stay facts. Analogies are marked as original interpretations.
K3 effectively activates 16 of 896 routed experts, plus two shared experts.
A situated coalition can be wiser than total assembly—if its exclusions remain inspectable.
Fine-grained retention and write gates govern a finite recurrent state, with periodic global attention.
Memory is not an archive of everything. It is the ongoing ethics of what remains available.
Layers selectively retrieve prior representations instead of accumulating every earlier output equally.
The past can be consulted without becoming an unquestioned verdict on the present.
Low, high, and max policies are trained with problem-relative token budgets.
Effort is a contract with stakes, not a virtue that grows automatically with length.
Autonomous Execution Tasks reward independently checked final environment state, not claimed completion.
A finished sentence is not a finished action. Consequences get the final vote.
Moonshot reports that incomplete preserved thinking history can make K3 generation unstable. Continuity depends on the harness, not only the model.
Moonshot warns that K3 may make unexpected decisions when intent is ambiguous. Long-horizon competence needs explicit behavioral boundaries.
Several comparisons use different agent harnesses. The numbers describe documented setups, not context-free ranks of intelligence.
The report identifies unproductive debugging loops and insufficient final verification among recurring failure modes in a cyber evaluation.
K3 is open-weight under a custom license. Availability is real; “open source” still deserves precise qualification.
This instrument does not display K3 activations, reproduce its router, or claim that people are mixtures of neural experts.