Offload the analysis work without giving up reproducibility.

Research-ready pipelines for EEG and event-related potentials (ERP), speech, eye-tracking, surveys, corpora, and other quantitative or computational methods.

Scoped pilots, expert-reviewed — not production validation. Replication audits on published data are the proof and the lower-risk entry.

PhD in psycholinguistics (UNSW) · BA in linguistics (CUHK) — I've worked through the research lifecycle myself.

Get in touch Find your scenario — 30 seconds Email your paper or data link for a fit review. No call is needed to start. Replies within 24 hours on business days.

The proof: your own published numbers, reproduced

I re-run a published analysis from its available data and code, then provide a side-by-side comparison with the paper. The result may be an exact match, a close match, or a documented discrepancy; the comparison table shows which.

Published-data checkWhat was checkedResult
Public EEG reproductionSeven ERP components on public NEMAR dataFive cleared; two remain amber
L2 word-processing studyNineteen reported checksNineteen reproduced; four were close rather than exact
Metadiscourse corpus studyTwenty-four reported resultsTwenty exact; four discrepancies documented
Multimodal eye-tracking studyTwenty-one reported resultsTwenty matched; one mismatch documented
Public eye-tracking corpusFixation and reading-time measures on a stated complete-case subsetMeasures and effects reproduced for that subset

These are bounded audits, not claims that every analysis in each paper was rerun. Each Evidence Library record identifies the available data, the denominator used, the matched results, and the unresolved differences.

A trial applies the same process to one bounded question from your paper: recover the relevant result where possible, trace any discrepancy, and hand over the documented workflow.

Methods that can be scoped into a project

The appropriate deliverable depends on the data, design, validation standard, and what has already been demonstrated. The statuses below distinguish published-data reproduction from a tutorial or synthetic mechanism test.

If your research involves…A pipeline can cover…Current evidence
EEG / ERPFiltering, epochs, component measures, statistical checks, and documented handoverPublic EEG reproduction: five components cleared and two amberdetails →
Eye-trackingFixation and reading-time measures, exclusions, visualisation, and statistical modellingTwo public-data examples: one complete-case corpus reproduction and one study with 20 of 21 reported results matched
Speech or phoneticsTranscription, alignment, acoustic measures, error review, and reportingCurrent end-to-end speech pipeline evidence is synthetic; natural-audio validation remains necessary
Child speechMeasurement and error-analysis workflows for a defined research taskSynthetic mechanism test only; no real children's voices and no screening or diagnostic validation
Corpus and text researchCorpus preparation, coding, counts, classifiers, and comparison with published analysesPublished corpus reproductions are available; the sentiment tutorial is narrowly calibrated to English data
Experiments, surveys, coding, reviews, and translationA study-specific workflow with explicit human checks and validation criteriaCurrent examples include a mixture of published-data reproductions, real-data tutorials, and small synthetic tests; status is stated per record

See the Evidence Library for the data-status badge, denominator, limitations, and comparison table for each example.

No data yet? The pipeline starts before the data.

Literature and synthesis

Turn a reading list into a structured matrix of designs, samples, measures, and reported findings.

Current example: information extracted from eleven real-paper abstracts. This is an extraction tutorial, not a completed systematic review. Demo →

Instruments and experiments

Check questionnaire wording, translation choices, materials, and preregistration logic against stated sources and decision rules.

Current example: GAD-7 and PHQ-9 items compared with published Chinese versions, including a deliberately inserted error. This demonstrates the checking workflow, not general translation accuracy. Demo →

Grant proposals

Map a draft against the published assessment criteria and identify where the rationale, design, feasibility, or evidence needs clarification.

Current example: a criterion-by-criterion alignment exercise using the 2026/27 GRF2 criteria. It is not a funder score or a prediction of funding. Demo →

From fit review to documented handover

1 · Review the fit

Email the paper, data link, or research question. I will identify the bounded result or workflow that can be assessed first, together with the main limitations.

2 · Run a trial

Use one defined question from your published work to test result recovery, discrepancy tracing, and documentation before commissioning a larger package.

3 · Build the full workflow

Extend the agreed approach to the current dataset through fixed milestones, documented checks, and a handover that you can inspect and rerun.

For administrative, teaching, or private-assistant work, see the FAQ. For an isolated AI research-team setup, see Own an AI Team.

How the work is checked

I remain the accountable person for scoping, methodological decisions, review, and delivery. Specialist AI roles may support planning, analysis, drafting, and quality checks, but their outputs are reviewed against the source data, paper, code, or agreed criteria before handover.

The working sequence is: understand the data and research question → run basic integrity checks → build in an open, inspectable stack → compare the output with the paper or stated expectations → document the result, limitations, and handover.

If you prefer to operate the machinery yourself, an isolated AI research-team setup is available separately. Process & Pricing →

Contact

Email your paper, public data link, or bounded research question. No call is needed to start.

Pricing: One trial deliverable on your topic starts from US$1,000 (approximately HK$8,000), free if not useful.

Larger engagements are quoted as fixed-fee milestones rather than hourly work.

Email

Replies within 24 hours on business days.

Working arrangement

Fully remote, with scheduling across time zones

Fixed-fee pricing — no hourly meter.