How to use Dataset Profiler
- Paste the table into the box, or leave it empty and choose a CSV or TSV file.
- Set the delimiter only when the detected one looks wrong.
- Select Profile the columns and read the per-column summary of type, empty rate and unique values.
Example: Dataset Profiler
Check a small order file before importing it, where one quantity is empty and one customer name repeats.
Options
- Value types
- Each non-empty value gets the narrowest fitting type: integer, decimal, date, boolean or text. A column is only called numeric when every non-empty value is numeric; one text value makes the whole column mixed.
- Ignore spaces at the ends of a value
- On by default, so a cell holding only spaces counts as empty and a padded value counts the same as its trimmed form. Switching it off shows padding as a real difference in both the blank count and the unique count.
- Unique values
- The number of distinct non-empty values in a column. A column where that equals the row count has no repeated value, which is what makes it a candidate key; blank cells are never counted as a value.
- Range
- Shown only for a numeric column, as the smallest and largest number in it. The digits are read from the text, so 0012 appears in the samples as written but is compared as the number twelve for the range.
Supported inputs and limits
Where your input is processed
This tool processes your input in this browser. Your text and files are not uploaded to UseFreeTools. Check this tool's limits for anything it may save on your device.
Profiling before an import
Most import failures trace back to a column whose real contents differ from its label: a numeric field with a stray marker, a key with duplicate values, or a date written in more than one shape. Reading the empty rate, unique count and a few sample values first turns those into a decision rather than an error halfway through a load.
Why a guess is not a schema
The type in this report comes from the characters in the sample only. A column that holds only digits this month can hold a prefix next month, and a date written in short form may arrive in another format in the next export. The report is evidence about the file in front of you, so confirm the intended type against the data source before writing a schema.
Questions about Dataset Profiler
Why is my date column shown as text?
The type guess recognises only a few plain date shapes, such as 2026-09-29 and 29/09/2026, and keeps a written month as text because many formats are ambiguous. Check that column by hand before relying on it.
Why is my ID column called integer when it holds codes?
The guess describes the characters, not their meaning. A column of digits reads as a number even when the digits are a label. Treat the type as a hint and decide the real type for yourself.
Why is a numeric column called mixed?
At least one non-empty value is not numeric, such as n/a or a blank marker written as text. The column is reported as mixed so a numeric summary is never shown for data that partly is text.
Does the profiler change my file or upload it?
No. The table is read and summarised in this browser. Nothing is written, sent to a server, or changed on your device.
What is a candidate key?
A column with no empty cells and a different value in every row is named in the notes. It is a candidate only: a column can look unique in this sample and still repeat in the full dataset.