Feather — Null-heavy Columns
Feather/Arrow IPC twin of the null-heavy table — grouped with the Parquet nulls fixture.
| id | name | score | note |
|---|---|---|---|
| 1 | Alpha | 90 | null |
| 2 | null | null | missing name |
| 3 | Gamma | 88.5 | null |
Specifications
- Format
- Apache Feather
- Edge
- null cells
- Rows
- 3
Testing contract
Expected to pass- Scenario
- Exercise Feather — Null-heavy Columns in its columnar workflow. Feather/Arrow IPC twin of the null-heavy table — grouped with the Parquet nulls fixture.
- Expected result
- 3 rows, 4 columns; fields: id: int64; name: string; score: double; note: string; column null counts=[0, 1, 1, 2]. Declared feature checks: edge=null cells.
What is a .feather file?
Feather (.feather) is a fast on-disk representation of Apache Arrow tables designed for zero-copy, language-agnostic data exchange between Python (pandas/pyarrow) and R. It preserves column types and structure exactly.
How to use this file
Use an example .feather file to test Arrow/Feather readers, round-trip type fidelity, and Feather-to-Parquet/CSV conversion.
How to use this file for testing
“Feather — Null-heavy Columns” is a deterministic Novus Examples fixture for Data import, Conversion testing. Realistic faker-generated datasets with documented schemas for testing import and ETL flows.
Documented properties for this file: 3 rows · Apache Feather. Compare results against paired or grouped companions on this page when present (clean↔damaged, searchable↔scanned, or format twins) so scores stay reproducible across runs.
Download the file once, keep the path stable in CI or local scripts, and treat the spec table as the contract: dimensions, seeds, field lists, and roles are intentional. Corrupt or invalid samples are labelled as such, expect parsers to fail loudly rather than silently accept them.
Data fixtures document their exact quirks (delimiters, encodings, null handling, schema, and row counts) in the spec table. Point your parser or importer at the file and assert it handles the documented edge cases; clean and deliberately-messy siblings make before/after diffs straightforward.
Related files
- featherFeather — Dictionary-encoded ColumnFeather file with dictionary-encoded strings — Arrow IPC edge-case fixture.

- parquetParquet — Decimal as String ColumnAmounts stored as strings in Parquet — common ingestion edge case for ETL parsers.

- parquetParquet — Dictionary-encoded ColumnParquet with dictionary-encoded string column — tests dictionary page decoding.

- parquetParquet — Duplicate Key ColumnParquet with repeated key values — tests join/aggregation edge cases.

- parquetParquet — Empty TableEmpty Parquet file with schema but no rows — edge case for readers.

- parquetParquet — List ColumnParquet with a list-of-string column — tests repeated-field decoding.

Generated by generation/data_wave_f.py. Free for any use, no attribution required, license.