Orbit Labs • Development journal

Experiments.
Evidence. Open questions.

Notes from building Orbit tools. Each entry separates implemented code, local test results and work that still needs validation. These are development records, not peer-reviewed research or an independent audit.

Checkpoint:

The prototypes below are under development. Local tests do not establish live service reliability, AI accuracy or wallet/payment readiness. Status is recorded for this checkpoint and is not updated automatically.
PROTOTYPE • LIVE EVALUATION PENDING

01 — Answering questions from one source

Question
Can a tool summarize pasted text and explain which parts support an answer?
Implemented
A source-text form with an optional question, English/Hebrew output and instructions to separate source claims from unknowns. Access is restricted to enabled test wallets. It does not search the web.
Local evidence
Automated tests cover rejected unauthorized requests, input limits, simulated AI output and provider failures. The provider is mocked: these tests do not measure a real model's answers or quotation fidelity.
Unresolved
No live model accuracy evaluation is recorded. Quotations and interpretations may be incorrect and require comparison with the source.
Next experiment
Use a fixed set of sources and answerable/unanswerable questions in both languages. Manually compare every quotation and identify claims unsupported by the source. Report failures as well as successful answers.

View research prototype →

PROTOTYPE • LIVE EVALUATION PENDING

02 — Drafting without inventing milestones

Question
Can a drafting tool preserve the difference between planned, unfinished and released work?
Implemented
Social posts, product descriptions and development updates with language, tone, audience and call-to-action controls. The prompt instructs the model to use supplied facts and retain development status. Nothing is automatically published.
Local evidence
Mocked-provider tests cover wallet access, input validation, empty/incomplete output and service failures. Tests confirm the instruction is sent, not that a live model obeys it.
Unresolved
Factual fidelity, link preservation, Hebrew fluency and resistance to conflicting source instructions still require live evaluation.
Next experiment
Supply briefs containing a mix of live features and future plans. Review drafts for invented dates, partnerships, results or claims of public availability.

View content prototype →

LOCAL CALCULATIONS • BROWSER CHECK PENDING

03 — Inspecting CSV data locally

Question
Can useful table statistics be calculated without uploading file contents?
Implemented
A browser CSV parser, first-50-row preview, missing/unique counts and numeric minimum, maximum, sum, mean and median. Supports comma, semicolon and tab separators. Summary export uses JSON.
Local evidence
Deterministic tests cover quoted cells, escaped quotes, embedded newlines, Hebrew text, malformed rows, size limits and known numeric results. Mixed numeric/text columns are flagged instead of silently discarding text.
Unresolved
Real mobile file selection and export remain untested. Numeric interpretation is deliberately strict; currency symbols and localized number formats are not converted. Calculations use approximate floating-point arithmetic.
Next experiment
Open representative UTF-8 files on desktop and mobile, compare statistics against a reference table, and verify error messages and downloaded summaries.

View CSV prototype →

PROTOTYPE • LIVE INTEGRATION PENDING

04 — Displaying account state honestly

Question
Can an account view distinguish a real zero balance from unavailable data?
Implemented
A wallet and Pro-credit dashboard using the authenticated server balance endpoint, with refresh, current-tab sign-out and tool navigation. No fabricated activity history is displayed.
Local evidence
Tests using a mocked page and server check signed-out state, zero/positive balances, expired sessions, server errors and a late response after sign-out.
Unresolved
These tests do not demonstrate a successful Phantom login, live credit storage or payment. Those integrations remain pending.
Next experiment
Validate the mobile login and session lifecycle, then compare displayed balances with actual server records during an authorized payment test.

View account prototype →

05 — Five-item development batch: roadmap 33–37

Question: can local workflows reuse evidence without silently charging credits, transmitting files or claiming verified facts?

Implemented locally

Sequential CSV batches with per-file errors and cancellation; explicit export of minimal Pro snapshots; optional preparation of those snapshots for source research; same-mint chronological snapshot comparison; a development register.

Recorded local results

42 automated tests pass for the cumulative application. New tests demonstrate that a malformed CSV fails individually while the next valid file runs, cancellation prevents the next file from being read, and size/count limits reject oversized work. Snapshot tests exclude account/payment fields, preserve unknown versus zero and reject different mints, reversed/equal timestamps and invalid metrics.

Failures and unresolved observations

The CSV errors above are deliberately constructed test cases, not reported user incidents. Separately, the owner’s real ELON payment showed success in an explorer screenshot but site verification rejected its association with the quote. Credits remain unresolved; the recovery UI is local and is not a proven fix. No second transfer is requested.

Limits and next experiments

No live browser or model evaluation was performed for this batch. Compare two real exported reports on mobile, test download/import/cancellation, and review source-based AI output in both languages. Imported JSON is not authenticated chain evidence; differences can reflect provider/pool changes. There is no scheduler, background monitor or autonomous transaction execution.

Open local workflows

06 — Private registry and credit-aware developer tools

Local tests exercise repeated and concurrent registration, per-wallet rating replacement, owner self-rating rejection, rating resets on revision changes, invalid URLs, capacity limits and retry after a simulated lost write acknowledgement. Paid SDK tests require explicit acceptance, preserve request IDs and expose pending/refund responses without automatic retry.

47 cumulative local tests pass; 12 server functions bundle. The storage tests are simulations. They do not prove hosted concurrency, a successful paid run or a completed marketplace. The existing live payment-credit mismatch remains unresolved. Next: private hosted acceptance, moderation/retention planning and mobile UI tests when deployment resumes.

07 — Program records without implied funding

Exact token-unit tests preserve the split total including rounding, reject over-allocation and duplicate plan IDs, and keep self-review and partnership exports explicitly unverified/unsent. Private grant tests exercise immutable terms, repeated submissions, test-wallet voting, applicant-only outcomes and retry after a simulated lost write acknowledgement.

51 cumulative local tests pass and 13 server entrypoints bundle. No real grants, transfers, independent quality reviews or partnership agreements were tested or created. Private proposals remain unpublished. Next: real mobile and hosted storage acceptance, then operating decisions before any public program.

ISSUES & LIMITATIONS

What still needs work

We will revise these entries when new evidence is available. A proposed experiment is not a completed result.