An invasive speech neuroprosthesis decoded attempted speech at 62 words per minute with a 23.8% word error rate in one participant with ALS.
Build a more reliable mind and body.
Build steadier routines through small, measurable behavior experiments. Every recommendation shows its evidence and limits, while Phroneme's cognitive-autonomy research remains open to inspect.
- 90 days
- four phases, one behavior at a time
- 2 minutes
- for the daily reflection loop
- 1 experiment
- to review and adjust each week
Phroneme is developing a research instrument for a specific unresolved question: what happens to unaided reasoning, retention, transfer, and calibration as AI use changes?
Early evidence makes the question testable, not settled.
External studies suggest that assisted performance and later unaided performance may differ. Phroneme has not yet collected the participant data needed to estimate that difference or validate the Cognitive Autonomy Index.
The research opportunity is to build a versioned instrument, test its failure modes, and publish null as well as positive findings. The evidence chain is the product.
three tool conditions
Bar lengths are a qualitative interface schematic, not reported effect sizes.
In a 54-person preprint, the LLM-assisted group showed the weakest EEG connectivity of three tool conditions and struggled to quote its own work. The study does not establish long-term causal harm or population-wide effects.
Review the evidence record →Conceptual illustration, not a brain image, participant trajectory, or observed effect. The direction and magnitude remain research questions.
Do not outsource the capacity you want to protect.
The Cognitive Acceleration Lab is a local-first experiment prototype. It records short-form accuracy and confidence, while retention and transfer remain outcomes that a delayed, fresh-form study would need to establish.
Explore the learning lab →External research that defines the scientific neighborhood.
These examples are external evidence, not Phroneme-generated results. Each retained claim has a public evidence record with its classification and limitations.
Brain2Qwerty v2 reports 61% mean and 78% best-participant word accuracy on briefly memorized typed sentences decoded from MEG recordings.
In a 54-person preprint, the LLM-assisted group showed the weakest EEG connectivity of three tool conditions and struggled to quote its own work. The study does not establish long-term causal harm or population-wide effects.
Reads the brain's output, memorized sentences you imagine typing, at roughly one word in four wrong, from a $2M half-ton MEG scanner you sit motionless inside. A landmark result, and still a lab instrument.
Proposes a behavioral construct focused on unaided reasoning and related domains. The current browser demo is not a validated instrument and has no population norms.
Three proposed modules. One bounded estimate.
The current demo implements illustrative adaptive selection and 2PL scoring in the browser. Field-calibrated items, equated forms, norms, and a longitudinal trajectory model remain proposed work and must not be inferred from the demo.
Select the item with maximum information at the current ability estimate.
Two-parameter logistic model with Bayesian ability estimation and honest SEM.
Proposed: repeated equated forms, practice-effect controls, and uncertainty-aware change estimates.
The browser demo currently implements only adaptive selection and illustrative 2PL scoring. Inspect the current demo; the longitudinal layer remains proposed.
See what a trajectory looks like.
The intended product is repeated measurement with uncertainty, not one permanent label. Browse synthetic demonstration scenarios, then inspect the current assessment demo.
Generated example of a decline larger than the assumed measurement-noise threshold.
Under these assumed values, the change exceeds the demonstration threshold. This does not show that any person declined.
The displayed trajectories are synthetic demonstration scenarios used to explain measurement uncertainty. They are not participant observations, pilot results, norms, or evidence of change.
Run the assessment demo →Four deliberate moves, in one screen.
Brain-training apps optimize for entertainment. Standardized tests optimize for gatekeeping. Phroneme does neither.
The gamification arms race. This is the opposite of a dopamine loop.
Scope. The proposed instrument targets specific cognitive-autonomy domains rather than claiming to measure the whole mind.
Transparency. Evidence class, uncertainty, version, and limitations must travel with each material claim.
A versioned research platform that can test whether longitudinal cognitive-autonomy measurement is possible.
What each product category is designed to do.
This is a design comparison, not a performance benchmark. It shows intended product choices and the evidence still required before those choices can become claims.
| Design dimension | Brain-training apps | Standardized tests | Phroneme CAI |
|---|---|---|---|
| Primary optimization target | Engagement and repeat use | Achievement or selection | Proposed cognitive-autonomy construct |
| Longitudinal interpretation | Varies by product | Usually secondary | Proposed; not validated |
| Uncertainty shown to users | Product-specific | Method-specific | Present in demo; parameters illustrative |
| Current evidence status | Must be evaluated product by product | Established for specific instruments and uses | Awaiting pilot data |
Qualitative design comparison only. No numeric product-performance ranking is claimed, and competing products must be evaluated individually against their own evidence.
One instrument. Two jobs to be done.
What we have not solved yet.
We publish limits openly, because honesty is the moat incumbents cannot copy without indicting their own marketing.
Population norms do not exist
The demo runs real item-response math on illustrative item parameters. A production CAI needs thousands of respondents, drift monitoring, and differential-item-functioning screens before scores are comparable across people.
Validation architecture →One point is nearly meaningless
The browser demo shows an uncertainty band derived from illustrative parameters. Phroneme has not established that repeat measurements are comparable or that change can be separated from practice and noise.
See the demo →Category creation is slow and expensive
Institutions can substitute grades, existing tests, or doing nothing at zero marginal cost. Phroneme has not yet established willingness to pay or a validated institutional deliverable.
Five Forces analysis →Neural anchor is bounded, not magic
The CAI is proposed as a behavioral instrument with an optional neural convergent-validity study. No neural validation layer currently exists, and the product is not a brain-reading system.
Reverse inference paper →Everything currently available to inspect.
Draft papers, a proposed method, product demonstrations, sources, and an explicit public evidence-status ledger.
One measurement asset. Several uses.
If the construct survives simulation, expert review, feasibility work, psychometric validation, and independent replication, a versioned longitudinal measurement system could support individuals, researchers, and institutions. That outcome is not yet established.