headers
Extracts field names from data files (CSV, JSON Lines, BSON, XML, Excel, and other iterabledata sources).
Default scan --limit is 10000. --format-out json (or --output ending in .json) writes JSON instead of text.
undatum headers data.jsonl
undatum headers data.csv --limit 50000
undatum headers data.csv --format-out json --output fields.json
undatum headers workbook.xlsx --table Sheet2
Reference
Reads: any readable format · Writes: field names (text or JSON) · Memory: streaming · Engines: python
undatum headers [OPTIONS] INPUT_FILE
| Argument | Description |
|---|---|
INPUT_FILE | Path to input file. (required) |
| Option | Description | Default |
|---|---|---|
-o, --output TEXT | Optional output file path. If not specified, prints to stdout. | |
-f, --fields TEXT | Field filter (kept for API compatibility, not currently used). | |
-d, --delimiter TEXT | CSV delimiter character (auto-detected when omitted). | |
--quotechar TEXT | CSV quote character (iterabledata default '"' when omitted). | |
--encoding TEXT | File encoding (e.g., 'utf8', 'latin1'). | |
-n, --limit INTEGER | Maximum number of records to scan for field detection. | 10000 |
--verbose / --no-verbose | Enable verbose logging output. | --no-verbose |
-F, --format-in TEXT | Override input file format detection (e.g., 'csv', 'jsonl', 'xml'). | |
-O, --format-out TEXT | Override output format (e.g., 'csv', 'json'). | |
--zipfile / --no-zipfile | Treat input file as a ZIP archive. | --no-zipfile |
--filter-expr TEXT | Filter expression (kept for API compatibility, not currently used). | |
--table, --sheet TEXT | Table or sheet name for multi-table sources (Excel, SQLite, lakehouse). | |
--start-page INTEGER | Sheet index (0-based) for Excel files. | 0 |
--trust | Acknowledge pickle deserialization risk when reading pickle sources. | |
--on-error TEXT | Parse-error policy: raise (default), skip, or warn. | |
--error-log TEXT | Append parse errors as JSONL (use with --on-error skip or warn). | |
--flatten-nested | Unfold nested dict / array-of-dict fields into dotted paths (e.g. city.lat). | |
--max-nested-depth INTEGER | With --flatten-nested, maximum nest depth to unfold (engine default 5). | |
--keep-nested-parents / --no-keep-nested-parents | With --flatten-nested, keep parent dict/array fields alongside dotted children. | --keep-nested-parents |
--json | Print the result as one JSON document (same as --format-out json). |
See also shared options.