GitHub - apache/arrow: Apache Arrow is the universal columnar format and multi-language toolbox for fast data interchange and in-memory analytics (original) (raw)

Apache Arrow

Fuzzing Status License Twitter Follow

Powering In-Memory Analytics

Apache Arrow is a universal columnar format and multi-language toolbox for fast data interchange and in-memory analytics. It contains a set of technologies that enable data systems to efficiently store, process, and move data.

Major components of the project include:

Arrow is an Apache Software Foundation project. Learn more atarrow.apache.org.

What's in the Arrow libraries?

The reference Arrow libraries contain many distinct software components:

Implementation status

The official Arrow libraries in this repository are in different stages of implementing the Arrow format and related features. See our currentfeature matrixon git main.

How to Contribute

Please read our latest project contribution guide.

Getting involved

Even if you do not plan to contribute to Apache Arrow itself or Arrow integrations in other projects, we'd be happy to have you involved: