# An example of Apache Arrow file?

**URL:** <https://discourse.julialang.org/t/an-example-of-apache-arrow-file/58299>\
**Category:** Data\
**Tags:** arrow\
**Created:** [March 31, 2021, 2:33pm UTC](https://discourse.julialang.org/t/an-example-of-apache-arrow-file/58299 "2021-03-31T14:33:31Z")\
**Posts on this page:** 1\
**Showing post:** 8

<div class="post-metadata">

**Author:** ![Sami](https://sea2.discourse-cdn.com/julialang/user_avatar/discourse.julialang.org/sami/32/20726_2.png) [@Sami](https://discourse.julialang.org/u/Sami)\
**Post date:** [April 22, 2021, 11:56am UTC](https://discourse.julialang.org/t/an-example-of-apache-arrow-file/58299/8 "2021-04-22T11:56:53Z")

</div>

I ended up doing the big arrow file with pyarrow, along with lines below:

```julia
with pa.output_stream("path/big.arrow") as sink:
    with pa.ipc.new_file(sink, schema) as writer:
        for arrowfile in glob.glob("path/to/files/*.arrow", recursive=False):
            with pa.input_stream(arrowfile) as source:
                with pa.ipc.open_file(source) as reader:
                    for i in range(0,reader.num_record_batches):
                        writer.write_batch(reader.get_batch(i))

```

That led to the another issue: [How well Apache Arrow’s zero copy methodology is supported?](https://discourse.julialang.org/t/how-well-apache-arrow-s-zero-copy-methodology-is-supported/59797)

---

_[View the full topic](https://discourse.julialang.org/t/an-example-of-apache-arrow-file/58299)._
