DataLumio logo
Academic Research · Data Cleaning

DataLumio

DataLumio is a web-based AI data analysis platform that cleans spreadsheets, PDFs and survey exports, then runs qualitative and quantitative analysis side by side to produce interactive dashboards and Word reports without formulas, code or statistics software.

Active Subscription API available 18+ Verified by Guidaio
Overview

What is DataLumio?

DataLumio is a browser-based data analysis platform that sets out to replace the usual stack of spreadsheet, statistics package and slide deck with one connected workflow. The pitch is deliberately plain: clean, analyse, visualise and report on a dataset without writing a single formula, and without knowing Python, R, SQL or SPSS.

Six functions make up the product. Data Cleaning strips duplicates, blank rows, empty columns, missing values and inconsistent fields out of a spreadsheet. Quantitative Analysis returns descriptive statistics, frequency tables, outlier detection, chi-square and ANOVA output, regression-style summaries and clustering, each explained in everyday language. Qualitative Analysis works on documents: it identifies themes, pulls supporting verbatims, labels sentiment, codes and categorises responses, compares cases, and produces theme occurrence tables, word clouds and conclusions. The user can steer it with a research question or a list of imposed themes, or leave those fields empty for open exploration. PDF Analysis opens a long document in a side-by-side workspace, document on the left and chat on the right, keeps a saved conversation history per file, and can analyse a visual area selected inside the page, such as a chart or a table. The Data Visualization Dashboard builds bar, line, pie, scatter, area, heatmap, histogram and table views in five styles, one of which, AI Smart Mix, picks the combination itself. Data Integration connects Google Drive and Google Sheets so files can be analysed without being uploaded again, while fourteen further connectors are advertised as coming soon.

Accepted formats are PDF, CSV, XLS, XLSX, DOC and TSV, with a limit of four files per analysis. Results come back as structured Word reports, individually downloadable charts and interactive dashboards. A public API covers the quantitative side only, through one endpoint to generate a key and one to submit a file, and requires an active membership.

DataLumio is equally explicit about what it is not. Across its home page, about page, features page and FAQ it repeats that it assists rather than replaces: its output is a first pass to be checked against the source data, not a substitute for expert statistical judgement or for a researcher reading their own material.

What it does

  • Clean a messy spreadsheet by removing duplicates, blank rows, empty columns, missing values and inconsistent fields
  • Run quantitative analysis: descriptive statistics, frequency tables, histograms, chi-square, ANOVA, regression-style summaries and clustering
  • Code open-ended answers into themes and sentiment labels, with supporting verbatims and cross-case comparison
  • Chat with a long PDF in a side-by-side workspace, including analysis of a selected chart, table or figure
  • Build an interactive dashboard in one of five styles, from Executive Overview to AI Smart Mix
  • Export a structured Word report and download each chart individually
  • Connect Google Drive or Google Sheets, or automate quantitative analysis through the public API
Audience

When to use DataLumio / When not to

A quick filter to help you decide if DataLumio is the right fit.

When to use DataLumio

  • Academic researchers coding interview transcripts, field notes and survey datasets
  • Students working through course datasets, dissertations and thesis material
  • Market research and insights analysts handling mixed-method surveys with both figures and open comments
  • Business teams turning recurring spreadsheets into dashboards and written reports
  • Consultants and agencies producing first-draft analysis for recurring client deliverables

When not to use DataLumio

  • Organisations that need a data processing agreement and a named hosting country before uploading third-party personal data
  • Teams looking for a certified provider, since no SOC 2, ISO 27001 or equivalent audit is displayed
  • Analysts who need live database connections: Postgres, MySQL, Snowflake, BigQuery and Databricks are still listed as coming soon
  • Developers expecting full API coverage, as the public API handles quantitative analysis only
  • Anyone wanting a mobile or desktop application, or a free tier: DataLumio is web-only and every plan is paid
Get started

How to use DataLumio

A typical end-to-end flow, from setup to results.

  1. Create an account with an email address and password, or through a third-party provider such as Google
  2. Buy credits or take a plan, since the features sit in the paid tiers starting at USD 5 as a one-off payment
  3. Upload your files in CSV, XLS, XLSX, TSV, PDF or DOC, or connect Google Drive or Google Sheets, keeping to four files per analysis
  4. Run the file through Data Cleaning first if the spreadsheet carries duplicates, blank rows, empty columns or missing values
  5. Pick the function that matches the file: Quantitative for CSV and Excel, Qualitative for PDF and DOC, PDF Analysis for a long document, Dashboard for tabular data
  6. For qualitative work, type a research question or a list of themes, or leave both empty to let the tool explore on its own
  7. For a PDF, open the document in the reader, ask questions in plain language and select a chart or table on the page to have it analysed
  8. For a dashboard, choose a style among Executive Overview, Statistical Deep Dive, Time Series Focus, Group Comparison and AI Smart Mix
  9. Read the output back against the source file before using it, as the site itself insists on every page
  10. Export the report as a Word document, download the charts one by one, and delete files or reports from the dashboard when you are done
Quick read

Pros & Cons

Pros

  • Qualitative and quantitative analysis run on the same dataset in the same place, which is the tool's real differentiator
  • No coding or statistical software required: no Python, R, SQL, SPSS or advanced spreadsheet formulas
  • The whole chain is covered without switching tools, from cleaning through analysis and visualisation to the finished report
  • A shared credit balance across all functions rather than a separate price for every feature, with a very low USD 5 one-off entry point
  • An explicit no-training policy, repeated in two separate documents, alongside written user ownership of files and reports
  • No manual review of uploaded files by the team, and a retention table published data type by data type
  • Unusually candid documentation, with roughly thirty FAQ entries and a plain statement of the AI's limits

Cons

  • No explicit GDPR compliance statement and no certification such as SOC 2, ISO 27001 or HIPAA
  • No data processing agreement offered or even mentioned, and no published list of sub-processors, which blocks institutional use
  • No hosting country or region named: the policy only says servers may sit outside your country of residence
  • No Article 27 EU representative and no named DPO, although the platform openly targets European researchers
  • A single email address covers legal, privacy and support matters, with no chat and no phone line
  • No legal form published and a shared serviced-office address in London, on a domain registered only in May 2025
  • A narrow delivered scope: two live connectors out of sixteen, a quantitative-only API, four files per analysis on every plan, and no mobile, desktop or browser version
Pricing

Pricing & Plans

The published pricing table contains no free tier. The cheapest paid entry point is the Starter plan at USD 5, charged once and granting 25 credits. The subscriptions run at USD 10 per month for Standard (50 credits), USD 29 per month for Pro (150 credits) and USD 100 per month for Enterprise (unlimited credits, bought through Contact Sales). Every feature is available on every plan; only the credit allowance changes, with an analysis report, a cleaning job and a dashboard costing 10 credits each, an image 5 credits and a question on a PDF 1 credit. Payments are handled by Stripe. Readers should note that the site's headline invites visitors to start free while the table itself lists no free plan.

Plan 1
Starter
  • USD 5
  • one-off payment
  • 25 credits
  • presented for individuals and students getting started
Plan 3
Pro
  • USD 29 per month
  • 150 credits
  • presented for researchers and small teams and flagged as the Best Plan
Plan 4
Enterprise
  • USD 100 per month
  • unlimited credits
  • presented for organisations and institutions with high-volume needs
  • with priority support and an SLA
Prices and plans listed above may evolve. Always check the official pricing page before subscribing.
Trust & Privacy

Data, GDPR & hosting

A consolidated view of how DataLumio handles your data.

GDPR overview

DataLumio nowhere states in plain words that it is GDPR compliant, and displays no certification such as SOC 2 or ISO 27001. What it does publish is an alignment. Section 8 of the privacy policy grants seven rights, namely access, rectification, erasure, restriction, portability, objection and withdrawal of consent, and tells EU and UK users they may also lodge a complaint with their local supervisory authority. Verified requests are answered within thirty days, through the contact form or info@datalumio.co. California residents are addressed separately under the CCPA, with a statement that personal information is not sold. International transfers are said to rely on Standard Contractual Clauses approved by the European Commission. Absent, however, are any Article 27 EU representative, any named data protection officer, any offered data processing agreement and any published list of sub-processors.

Who owns the data?

Under the privacy policy and terms in force since 11 June 2026, everything a user uploads stays that user's property. DataLumio states that it claims no ownership over uploaded files and that both the source material and the reports generated from it remain the account holder's intellectual property. What the company keeps is what it built: the platform, its AI models, its software, its branding and its underlying technology. Uploaded material is processed only to deliver the analysis requested, is never read manually by the internal team, and can be deleted by the user at any time from the account dashboard, with permanent removal from the systems within thirty days.

Reuse rights

Because ownership of both the uploaded files and the generated reports stays with the user, nothing in the terms requires permission before reusing the output: reports, charts and dashboards can be quoted, published or folded into client deliverables freely. On DataLumio's side the declared use is narrow. Files are processed solely to run the analysis the user asked for, on legal bases listed as contract performance, legitimate interest, legal obligation and consent. Advertising use and any sale of personal information are ruled out in writing, the internal team does not manually open uploaded files, and third-party AI providers are contractually barred from using the data for model training. Sharing is limited to service providers for cloud hosting, payments through Stripe and email delivery, to AI infrastructure suppliers, and to cases of legal obligation or business transfer. Technical data such as IP address, approximate country and city, browser, device, pages viewed, referrer and session length is collected automatically, and the cookie policy covers strictly necessary, functional, analytics and security cookies only, with no advertising cookies, no advertising profiles and no cross-site tracking.

Data retention & training

Retention summary
Retention is published as a table in the AI and data usage policy. Uploaded files are kept for the active session and for the period they remain available in the dashboard, and are permanently removed within thirty days of a deletion request; the privacy policy adds that files are not kept beyond what is needed to complete and display the analysis session. Generated reports are kept until you delete them, and are removed within thirty days of an account closure. Analysis metadata is kept for twelve months for performance monitoring, and account information for the life of the account plus thirty days. Users can delete files and reports themselves from the account dashboard at any time. Verified rights requests are answered within thirty days. The policies took effect on 11 June 2026.
Trains on customer data
No
GDPR contact

Hosting summary

DataLumio names no hosting country and no hosting region anywhere on its site. The privacy policy says only that data may be processed and stored on servers located outside the user's country of residence. For users in the European Economic Area and the United Kingdom, the stated safeguard for such transfers is the use of Standard Contractual Clauses approved by the European Commission, which is a declared alignment rather than a disclosed location. Service providers are described by category, namely cloud hosting, payment processing and email delivery, without a named list. On the security side the site claims TLS encryption in transit, encryption at rest, need-to-know internal access controls and periodic security reviews. Note that the domain resolves to a Fastly anycast node in the United States: that is a content delivery edge, and it says nothing about where customer files are actually stored.

Watch-outs

Things to keep in mind

Risks and trade-offs to weigh before adopting DataLumio.

  • No explicit GDPR compliance claim and no security certification: check the ground rules before any institutional use
  • No data processing agreement is available, which is a blocker if you are the controller for other people's data
  • Research files often hold data from people who never agreed to it being uploaded, and the AI and data usage policy puts consent, ethics approval and legal rights squarely on the user
  • The hosting country is not disclosed, so there is no way to know where your files travel or rest
  • AI output must be read back against the source files, a warning the site itself repeats across four different pages
  • The Enterprise unlimited promise is subject to fair-use conditions that are not published, and the start free message is not supported by the pricing table
  • The fourteen coming soon connectors carry no date, and the API covers quantitative analysis only, with its exact endpoints kept private
Setup

Setup & Integrations

Technical difficulty

Very low for the web platform. Nothing is installed: you sign up with an email address and password or through a provider such as Google, then work in the browser. The site advertises setup in under a minute, no coding required, and a Google Drive connection made in minutes. The only genuinely technical route is the API, which means generating a key and posting a file as multipart form data with an Api-Key header. The site also notes that you still need to understand your dataset and choose a sensible analytical approach.

Deployment

Web appAPI

Integrations

Google Drive Google Sheets Stripe
Company

Behind DataLumio

Company name
DataLumio
Founded
15/07/2025
Country of origin
🇬🇧 United Kingdom
Headquarters
Level 1, Devonshire House, One Mayfair Place, London, UK
UBO
INFORMATION_NOT_FOUND
UBO country
INFORMATION_NOT_FOUND
Domain registrar country
🇺🇸 United States
Legal contact
Support contact

Social

Official links

Resources

All the official URLs gathered for verification and reference.

FAQ

Frequently asked questions

What does DataLumio actually do?
It is a web platform that cleans a data file, analyses it qualitatively and quantitatively, visualises the results in dashboards and turns them into a written report, without any formula or code.
Which file formats can I upload?
PDF, CSV, XLS, XLSX, DOC and TSV, with the accepted formats depending on the function chosen. Every plan allows four files per analysis.
Is my data used to train the AI?
No. The privacy policy and the AI and data usage policy both state that uploaded data is not used for training, fine-tuning or benchmarking, and that third-party AI providers are contractually barred from doing so.
Who owns the files I upload and the reports produced?
You do. The privacy policy states that uploaded research data remains your property and that you keep full ownership and intellectual property rights over both your files and the generated reports.
How long is my data kept?
Files are kept for the time needed to run and display the analysis, reports until you delete them, and analysis metadata for twelve months. Deletion requests and account closures are honoured within thirty days.
How much does it cost and how do credits work?
Starter costs USD 5 once for 25 credits, Standard USD 10 per month for 50, Pro USD 29 per month for 150 and Enterprise USD 100 per month for unlimited use. An analysis report, a cleaning job or a dashboard costs 10 credits, an image 5 and a PDF question 1.
Is there an API?
Yes, with two endpoints: one to generate a key and one to submit a file for quantitative analysis. It requires an active membership and does not cover qualitative analysis, dashboards or PDF chat.
Which integrations are live?
Google Drive and Google Sheets are in production. Fourteen further connectors, including Slack, Snowflake, BigQuery, Postgres and Databricks, are announced as coming soon without a date.
Is there a mobile app?
No. DataLumio is a web-only platform, with no iOS or Android application and no desktop version.
Does it replace a data analyst or a researcher?
No, and the site says so itself. The output is a first pass that must be read back against the source data, not a substitute for expert judgement. Users must also be at least 18 years old.
Conclusion

Should you pick DataLumio?

DataLumio is a young product, with a domain registered on 30 May 2025 and a first web archive capture dated 15 July 2025, but its proposition is genuinely useful. Very few tools put qualitative and quantitative analysis on the same dataset, in the same place, with no code and no statistics package, and almost none do it from a USD 5 one-off entry point. For a researcher facing a pile of interview transcripts, a student with a course dataset or a consultant assembling a first draft, that combination saves real hours on the first pass.

The data policy also reads well above average on paper: uploaded files and generated reports stay the user's property, nothing is used to train models, staff do not open files manually, and retention is documented type by type. What is missing is proof. There is no certification, no data processing agreement, no list of sub-processors and no named hosting country or region, only servers said to sit outside the user's country of residence. The publisher itself is hard to pin down: a brand with no stated legal form, a shared serviced-office address in London, and one email address covering legal, privacy and support alike.

The delivered scope is narrower than the marketing suggests too, with two live connectors out of sixteen, an API limited to quantitative analysis, four files per analysis on every plan including Enterprise, and a start free message the pricing table does not support.

Used for what it says it is, an assistant for the first pass whose output is re-read against the source files, DataLumio is a reasonable and inexpensive bet for individuals, small teams and consultants. For institutional, regulated or third-party personal data, the missing agreement and hosting disclosure should be settled with the vendor first.