Human-agent operations
Human-agent operations: orchestration, runtime boundaries, memory, evaluation, security, and cost accounting. Token economics is a named sub-category.
Shared lane
This lane is carried by both hubs of this survey.
Entries are repositories — libraries, benchmarks, specifications — admitted as evidence, not as products.
Tasks in this category
- Agent identity & interoperability 0 of 91 entries
- Agent runtime & boundaries 0 of 91 entries
- Design agent organizations 0 of 91 entries
- Evaluate agents & skills 0 of 91 entries
- Manage agent memory & context 0 of 91 entries
- Operator & research consoles 0 of 91 entries
- Orchestrate multi-agent workflows 0 of 91 entries
- Package & port agent skills 0 of 91 entries
- Run a solo-operator agent firm 0 of 91 entries
- Secure the agent supply chain 0 of 91 entries
- Token economics & AI cost 10 of 91 entries
- Track tool & repo change 0 of 91 entries
- Allocate tasks across humans & agents Planned 0 of 91 entries
- Score orchestration Planned 0 of 91 entries
- MCP/tool interoperability Planned 0 of 91 entries
Token economics & AI cost
Campaigns placed on this task
Multiple runs of this task exist; each is its own sealed record — no supersession is implied.
-
2026-08-01__token-usage-observability-accounting
run type (the campaign's own designation) high-recall-map 10 of 91 entries ingested
- Ledger rows
- 41
- Screening rows
- 113
- Repos in ledger
- 14
-
2026-08-02__token-efficiency-routing-and-optimization
run type (the campaign's own designation) high-recall-map Not yet ingested
- Ledger rows
- 55
- Screening rows
- 160
- Repos in ledger
- 25
Entries — 10 of 91 on this hub
-
Human-agent operations Live
AgentOps-AI/agentops
Token economics & AI cost
Editorial draft Session-level agent observability in Python: token and cost tracking, tool calls, replay. Cited twice in the token-economics ledger.
Forks 612 Language Python Stars 5761
Snapshot · retrieved UTC
-
Human-agent operations Live
Arize-ai/openinference
Token economics & AI cost
Editorial draft OpenTelemetry-compatible tracing conventions and instrumentations for model applications, with masking support.
Forks 288 Language Python Stars 1138
Snapshot · retrieved UTC
-
Human-agent operations Live
BerriAI/litellm
Token economics & AI cost
Editorial draft A multi-provider model gateway; the ledger cites its spend-tracking dimensions and its public issue history as cost-accounting evidence.
Forks 10443 Language Python Stars 55954
Snapshot · retrieved UTC
-
Human-agent operations Live
crewAIInc/crewAI
Token economics & AI cost
Editorial draft A multi-agent framework in Python; two survey ledgers cite it — for usage-accounting fields and for role-and-process organization design.
Forks 8107 Language Python Stars 56858
Snapshot · retrieved UTC
-
Human-agent operations Live
future-agi/traceAI
Token economics & AI cost
Editorial draft OpenTelemetry tracing that instruments model calls, token counts and agent decisions.
Forks 38 Language Python Stars 210
Snapshot · retrieved UTC
-
Human-agent operations Live
lmnr-ai/lmnr
Token economics & AI cost
Editorial draft An OpenTelemetry-native tracing and evaluation platform for agents, in TypeScript.
Forks 220 Language TypeScript Stars 3155
Snapshot · retrieved UTC
-
Human-agent operations Live
open-telemetry/semantic-conventions
Token economics & AI cost
Editorial draft The OpenTelemetry semantic-conventions registry; kept as the source the GenAI conventions moved out of, with the successor listed separately.
Forks 377 Language Jinja Stars 627
Snapshot · retrieved UTC
-
Human-agent operations Live
open-telemetry/semantic-conventions-genai
Token economics & AI cost
Editorial draft Current home of the GenAI semantic conventions — the vocabulary that token and cost telemetry is standardizing on.
Forks 75 Language Python Stars 234
Snapshot · retrieved UTC
-
Human-agent operations Live
openlit/openlit
Token economics & AI cost
Editorial draft OpenTelemetry-native tracing and metrics with a model cost registry, in TypeScript.
Forks 350 Language TypeScript Stars 2676
Snapshot · retrieved UTC
-
Human-agent operations Live
traceloop/openllmetry
Token economics & AI cost
Editorial draft Apache-2.0 OpenTelemetry instrumentations spanning model and agent frameworks.
Forks 1047 Language Python Stars 7368
Snapshot · retrieved UTC
Tasks with no entry today
14 of 15 tasks in this category carry no entry today. Each is listed with its state: ingested with no repository evidence in its ledger, on the record and not yet ingested, or planned and not yet run. Opening a row shows the campaigns placed on it, with their ledger and screening rows as recorded.
-
Agent identity & interoperability Not yet ingested 2 campaigns on the record; one modern ledger has not been brought into the registry. The other predates the current standard; it carries flags, not counts.
Campaigns placed on this task
Multiple runs of this task exist; each is its own sealed record — no supersession is implied.
-
2026-07-27__agent-identity-authority-interoperability
run type (the campaign's own designation) pre-standard Not yet ingested
- Ledger rows
- —
- Screening rows
- —
- Repos in ledger
- 0
Coverage ceiling, quoted from the sealed record:
frozen SCOPE.md, scope_id wider-sota-04-scope-v1; coverage: effort-bounded, frozen before public discovery
-
2026-08-01__agent-identity-authority-interoperability-delta
run type (the campaign's own designation) high-recall-map Not yet ingested
- Ledger rows
- 20
- Screening rows
- 175
- Repos in ledger
- 13
-
-
Agent runtime & boundaries Not yet ingested 2 campaigns on the record; 2 modern ledgers have not been brought into the registry.
Campaigns placed on this task
Multiple runs of this task exist; each is its own sealed record — no supersession is implied.
-
2026-07-28__runtime-machinery
run type (the campaign's own designation) high-recall-map Not yet ingested
- Ledger rows
- 84
- Screening rows
- 327
- Repos in ledger
- 71
-
2026-08-01__runtime-machinery-delta
run type (the campaign's own designation) high-recall-map Not yet ingested
- Ledger rows
- 7
- Screening rows
- 43
- Repos in ledger
- 2
-
-
Design agent organizations Not yet ingested 2 campaigns on the record; one modern ledger has not been brought into the registry. The other predates the current standard; it carries flags, not counts.
Campaigns placed on this task
Multiple runs of this task exist; each is its own sealed record — no supersession is implied.
-
2026-07-27__agent-organization-design
run type (the campaign's own designation) pre-standard Not yet ingested
- Ledger rows
- —
- Screening rows
- —
- Repos in ledger
- 0
Coverage ceiling, quoted from the sealed record:
frozen expanded scope v2.0.0; coverage_target comprehensive-public-evidence-sweep, coverage_claim_ceiling systematic-effort-bounded
-
2026-08-01__agent-organization-design-v2
run type (the campaign's own designation) high-recall-map Not yet ingested
- Ledger rows
- 76
- Screening rows
- 217
- Repos in ledger
- 0
-
-
Evaluate agents & skills Not yet ingested 2 campaigns on the record; one modern ledger has not been brought into the registry. The other predates the current standard; it carries flags, not counts.
Campaigns placed on this task
Multiple runs of this task exist; each is its own sealed record — no supersession is implied.
-
2026-07-27__agent-evaluation-failures
run type (the campaign's own designation) pre-standard Not yet ingested
- Ledger rows
- —
- Screening rows
- —
- Repos in ledger
- 0
Coverage ceiling, quoted from the sealed record:
frozen SCOPE.md with post-discovery comprehensiveness addendum; coverage: effort-bounded, searched to a documented working plateau
-
2026-07-31__agent-skill-uplift-evaluation-and-ablation
run type (the campaign's own designation) high-recall-map Not yet ingested
- Ledger rows
- 41
- Screening rows
- 71
- Repos in ledger
- 10
-
-
Manage agent memory & context Pre-standard This campaign predates the current standard; it carries flags, not counts.
Campaigns placed on this task
-
2026-07-27__agent-context-memory-coordination
run type (the campaign's own designation) pre-standard Not yet ingested
- Ledger rows
- —
- Screening rows
- —
- Repos in ledger
- 0
Coverage ceiling, quoted from the sealed record:
SCOPE.md present but pre-JSONL; declares coverage: effort-bounded, candidate_bytes_acquired: false
-
-
Operator & research consoles Not yet ingested 2 campaigns on the record; one modern ledger has not been brought into the registry. The other predates the current standard; it carries flags, not counts.
Campaigns placed on this task
Multiple runs of this task exist; each is its own sealed record — no supersession is implied.
-
2026-07-26__control-console
run type (the campaign's own designation) pre-standard Not yet ingested
- Ledger rows
- —
- Screening rows
- —
- Repos in ledger
- 0
Pre-standard campaign. It predates the current standard and records no coverage qualifier, so none is shown.
-
2026-08-01__control-console-v2
run type (the campaign's own designation) high-recall-map Not yet ingested
- Ledger rows
- 37
- Screening rows
- 167
- Repos in ledger
- 52
-
-
Orchestrate multi-agent workflows Not yet ingested 1 campaign on the record; one modern ledger has not been brought into the registry.
Campaigns placed on this task
-
2026-07-30__opc-multi-agent-orchestration
run type (the campaign's own designation) high-recall-map Not yet ingested
- Ledger rows
- 33
- Screening rows
- 87
- Repos in ledger
- 42
-
-
Package & port agent skills Not yet ingested 1 campaign on the record; one modern ledger has not been brought into the registry.
Campaigns placed on this task
-
2026-07-31__skill-package-lifecycle-portability
run type (the campaign's own designation) high-recall-map Not yet ingested
- Ledger rows
- 40
- Screening rows
- 100
- Repos in ledger
- 14
-
-
Run a solo-operator agent firm Not yet ingested 2 campaigns on the record; one modern ledger has not been brought into the registry. The other predates the current standard; it carries flags, not counts.
Campaigns placed on this task
Multiple runs of this task exist; each is its own sealed record — no supersession is implied.
-
2026-07-27__agentic-firm-solo-operator
run type (the campaign's own designation) pre-standard Not yet ingested
- Ledger rows
- —
- Screening rows
- —
- Repos in ledger
- 0
Coverage ceiling, quoted from the sealed record:
SCOPE.md frozen before discovery; coverage: effort-bounded, candidate_bytes_acquired: false
-
2026-08-01__agentic-firm-solo-operator-v3
run type (the campaign's own designation) high-recall-map Not yet ingested
- Ledger rows
- 73
- Screening rows
- 248
- Repos in ledger
- 2
-
-
Secure the agent supply chain Not yet ingested 2 campaigns on the record; 2 modern ledgers have not been brought into the registry.
Campaigns placed on this task
Multiple runs of this task exist; each is its own sealed record — no supersession is implied.
-
2026-07-31__external-agent-skill-supply-chain-license-and-intake-security
run type (the campaign's own designation) high-recall-map Not yet ingested
- Ledger rows
- 77
- Screening rows
- 164
- Repos in ledger
- 12
-
2026-07-31__untrusted-content-agent-security
run type (the campaign's own designation) high-recall-map Not yet ingested
- Ledger rows
- 55
- Screening rows
- 105
- Repos in ledger
- 15
-
-
Track tool & repo change Not yet ingested 1 campaign on the record; one modern ledger has not been brought into the registry.
Campaigns placed on this task
-
2026-08-01__living-repository-change-intelligence
run type (the campaign's own designation) high-recall-map Not yet ingested
- Ledger rows
- 52
- Screening rows
- 104
- Repos in ledger
- 11
-
-
Allocate tasks across humans & agents Planned No campaign has run yet. Grounded in crosswalk row E01 (internal crosswalk id).
-
Score orchestration Planned No campaign has run yet. A grounding row is recorded, but naming it would name a course.
-
MCP/tool interoperability Planned No campaign has run yet. Grounded in crosswalk row N07 (internal crosswalk id).