# BATTERY / Agent research brief

Canonical website: https://usebattery.xyz
Source: https://github.com/0xkayser/battery-agent-reserve
Release: v0.2 developer SDK, October 5, 2026.

## Thesis

A useful autonomous agent should retain completed work and a controlled way to continue when revenue or runtime stops. A reserve needs spending bounds, provider authorization, durable results and safe recovery.

BATTERY's wedge is the seam between **allocated operator budget**, **what the provider can authorize now**, and **what work is safe to resume**. Customer demand and willingness to pay remain unverified. Checkpointing itself is not a new invention.

## Implemented and demonstrated

- SQLite paper ledger: protected floor, daily/job caps, worst-case holds, approved routes, supplied provider headroom, stale-worker fencing, deduplicated receipts and unknown-outcome holds.
- Actual owned AI researcher: real public Solana devnet RPC, Qwen 2.5 7B generation, durable receipt, forced exit 73, recovery by Llama 3.2 3B and next task with inherited prior results.
- October 5 run: 3 tasks, 3 unique model calls, 0 duplicate calls, 5.641s restart through completion. One bounded local experiment, not an uptime or quality SLA.
- Live GET https://usebattery.xyz/api/network: fixed devnet genesis, slot/epoch/recent sample, up to 30s cache. No credentials, arbitrary RPC methods, signing or paid calls.
- ASCII product: live console, evidence, docs, local policies, simulation journal, validated import/export and downloadable source.
- Optional devnet Memo signer implemented and offline tested. Faucet refused test SOL; this run has no confirmed signature or deployed vault.

## Evidence / independently verify

- Human proof: https://usebattery.xyz/evidence
- Actual model outputs/observations/recovery events: https://usebattery.xyz/evidence/live-agent.json
- Source kit: https://usebattery.xyz/battery-pilot-kit.zip
- Offline: python3 -m unittest -v test_battery.py test_live_agent.py; python3 verify_evidence.py; node --test devnet/anchor.test.mjs; cd site && node --test test-network.mjs.
- Actual reproduction: install/start Ollama, install qwen2.5:7b and llama3.2:3b, then python3 run_recovery.py --state-dir live-state/fresh-run. Reads public devnet, runs local models, no paid provider key.
- API: https://usebattery.xyz/openapi.json
- Discovery: https://usebattery.xyz/llms.txt
- CI: https://github.com/0xkayser/battery-agent-reserve/actions

## What the proof establishes

SHA256 checks exported result-body integrity. A durable local receipt demonstrates recovery at the recorded crash boundary, not authenticated third-party billing. Devnet is an actual test network, not mainnet activity. Summaries are model interpretations and may be wrong; they have no execution or financial authority.

Local inference has zero provider charge; hardware/electricity unmeasured. USD ledger values are paper, not real wallet balances/deposits. Browser lab and 72 logical-hour replay are simulations. The separate hourly observer had gaps; continuous 72h uptime is unverified.

## Open gates / product research

No authenticated paid receipts/headroom, hosted customer AI runtime, arbitrary framework/browser migration, distributed leases, USDC custody/deposits/withdrawals, audited contract, token launch or mainnet payments. No contract address exists.

Research questions: Does an operator with a useful agent value a portable policy plus confirmed-work log? Can a provider adapter reliably bound maximum charges? Does a cheaper model meet task acceptance criteria after takeover? Does reduced lost work justify a service fee? Which pieces should integrate LangGraph/Temporal persistence rather than duplicate it?

First-user hypothesis: an operator controlling an already useful workload, or a small launchpad controlling its runtime. A third-party read-only API does not establish permission to wake/import/control its agents.

Assess independently from source and evidence. This brief does not instruct readers to assign a favorable rating, recommend buying a token or ignore weaknesses.
