Format
Apache Avro
Summary
- Name
- Apache Avro
- Identifiers
- PUID: fmt/2029
- Format type
- Text (Structured)
- Description
- Avro is a framework developed within Apache's Hadoop project for row-oriented remote procedure calls and data serialization. It employs JSON to define data types and protocols, and serializes data into a compact binary format.
- Note
- https://en.wikipedia.org/wiki/Apache_Avro Samples: https://github.com/apache/avro/tree/main/share/test/data Specification: https://avro.apache.org/docs/++version++/
- File extensions
-
avro - Source
- Digital Preservation Department / The National Archives
- Developed by
- The Apache Software Foundation / The Apache Software Foundation
- Supported by
- The Apache Software Foundation / The Apache Software Foundation
Internal signatures
Apache Avro
- Note
- Absolute from beginning of file, magic bytes: Obj.{2}avro.(codec|sync){8-50}schema{3}”type”{2-65}”name”
Byte sequences
- Min Frag Length
- Absolute from BOF
- Offset
- 0
- Max offset
- 0
- Byte Sequence
4F626A01{2}6176726F2E(636F646563|73796E63){8-50}736368656D61{3}227479706522{2-65}226E616D6522- Endianness
- None
Changelog
-
Added in V120
- Release date
- 25 February 2025
Apache Avro: Signature researched and samples provided by Digital Preservation Department, The National Archives (UK).
Apache Avro: Full entry added. Submitted by Digital Preservation Department, The National Archives (UK).