SmartQueryTools

Extract Head of Arrow Files Online

Extract the first N rows from Arrow files directly in your browser. Choose how many rows to keep and download the result — no upload required.

How to extract Head of Arrow files

  1. Drop your file onto the upload area. The whole file is loaded, the total row count is shown and the first 200 rows are previewed.
  2. Enter how many rows to keep in the Number of rows box. The default is 100.
  3. Click Extract First Rows. The preview switches to show only the rows you kept.
  4. Click Download to save them in the same format you uploaded, with _head added to the file name.

Your file is processed locally in your browser and is never uploaded. The free limit is 50 MB per file; larger files work if your device has the memory for them.

Worked example

A weather station logs a reading every ten minutes into one long file. An engineer wants a small slice to share in a bug report without sending the full year of data.

Input (Arrow)

Arrow IPC file (binary, columnar) — shown as a table with its schema

stationreading_timetemp_chumidity_pct
WS-142026-07-01 00:00:0011.482
WS-142026-07-01 00:10:0011.183
WS-142026-07-01 00:20:0010.985
WS-142026-07-01 00:30:0010.8NULL
WS-142026-07-01 00:40:0010.686

Schema: station VARCHAR, reading_time TIMESTAMP, temp_c DOUBLE, humidity_pct BIGINT

Settings

  • Number of rows: 3

Result

stationreading_timetemp_chumidity_pct
WS-142026-07-01 00:00:0011.482
WS-142026-07-01 00:10:0011.183
WS-142026-07-01 00:20:0010.985

The first three data rows are kept in their original order and the header is carried over. Nothing is sorted, so "first" means first in the file, not earliest by timestamp. Here the file was already in time order, so the result is also the three earliest readings.

Frequently Asked Questions

Does it read only the first N rows?

No. The whole file is loaded into the in-browser engine and counted first, then the first N rows are kept. Very large files are limited by the memory your browser can use, not by N.

What is the difference between Head and Sample?

Head keeps the first N rows in file order and gives the same result every time. Sample picks rows at random, which is better when the start of the file is not typical of the rest.

What if I ask for more rows than the file has?

You get every row. If the box is empty or not a number, the default of 100 is used. 0 and negative numbers are treated as 1, so you always get at least one row.

Related Tools