This is an automated email from the ASF dual-hosted git repository.
Fokko pushed a commit to branch master
in repository https://gitbox.apache.org/repos/asf/parquet-testing.git
The following commit(s) were added to refs/heads/master by this push:
new 4b1ce45 Add file with an incompatible logical/physical type (#122)
4b1ce45 is described below
commit 4b1ce4502afff8d20c9b4bb08d07e04e21cdeff3
Author: Divjot Arora <[email protected]>
AuthorDate: Thu Sep 3 13:57:10 2026 -0400
Add file with an incompatible logical/physical type (#122)
---
data/README.md | 1 +
data/int32_with_uuid_logical_type.parquet | Bin 0 -> 353 bytes
2 files changed, 1 insertion(+)
diff --git a/data/README.md b/data/README.md
index bd3b4d2..036ecf6 100644
--- a/data/README.md
+++ b/data/README.md
@@ -58,6 +58,7 @@
| repeated_primitive_no_list.parquet | REPEATED INT32 and BYTE_ARRAY fields
without LIST annotation. See
[note](#REPEATED-primitive-fields-with-no-LIST-annotation) |
| map_no_value.parquet | MAP with null values, MAP with INT32 keys and no
values, and LIST<INT32> column with same values as the MAP keys. See
[map_no_value.md](map_no_value.md) |
| page_v2_empty_compressed.parquet | An INT32 column with DataPageV2, all
values are null, the zero-sized data is compressed using ZSTD. This is a valid
non-zero bytes ZSTD stream that uncompresses into 0 bytes. |
+| int32_with_uuid_logical_type.parquet | A single required INT32 column
`int32_uuid` (10 rows, values 0..9) annotated with the UUID logical type, which
is only applicable to FIXED_LEN_BYTE_ARRAY(16). This is an
unrecognized/incompatible logical-physical type combination; a reader should
tolerate it by ignoring the annotation (reading the column as its physical
INT32 type) and ignoring its statistics, rather than failing the whole file. |
| datapage_v2_empty_datapage.snappy.parquet | A compressed FLOAT column with
DataPageV2, a single row, value is null, the file uses Snappy compression, but
there is no data for uncompression (see [related
issue](https://github.com/apache/arrow-rs/issues/7388)). The zero bytes must
not be attempted to be uncompressed, as this is an invalid Snappy stream. |
| unknown-logical-type.parquet | A file containing a column annotated with a
LogicalType whose identifier has been set to an abitrary high value to check
the behaviour of an old reader reading a file written by a new writer
containing an unsupported type (see [related
issue](https://github.com/apache/arrow/issues/41764)). |
| int96_from_spark.parquet | Single column of (deprecated) int96 values that
originated as Apache Spark microsecond-resolution timestamps. Some values are
outside the range typically representable by 64-bit nanosecond-resolution
timestamps. See [int96_from_spark.md](int96_from_spark.md) for details. |
diff --git a/data/int32_with_uuid_logical_type.parquet
b/data/int32_with_uuid_logical_type.parquet
new file mode 100644
index 0000000..4001364
Binary files /dev/null and b/data/int32_with_uuid_logical_type.parquet differ