<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Utilities on SQLite Browser</title><link>https://nimh-dsst.github.io/sqlite-browser/docs/utilities/</link><description>Recent content in Utilities on SQLite Browser</description><generator>Hugo</generator><language>en</language><atom:link href="https://nimh-dsst.github.io/sqlite-browser/docs/utilities/index.xml" rel="self" type="application/rss+xml"/><item><title>Creating an SQLite file directly from a BIDS dataset</title><link>https://nimh-dsst.github.io/sqlite-browser/docs/utilities/bids2db/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>https://nimh-dsst.github.io/sqlite-browser/docs/utilities/bids2db/</guid><description>&lt;p&gt;&lt;code&gt;utilities/bids2db.py&lt;/code&gt; indexes a BIDS dataset with &lt;a href="https://github.com/childmindresearch/bids2table"&gt;bids2table&lt;/a&gt; (the same as &lt;code&gt;b2t2 index&lt;/code&gt;), adds each file&amp;rsquo;s JSON sidecar metadata, and writes the result to a SQLite file that SQLite Browser can open.&lt;/p&gt;&#10;&lt;p&gt;Use this instead of &lt;a href="https://nimh-dsst.github.io/sqlite-browser/docs/utilities/parquets2db/"&gt;parquets2db.py&lt;/a&gt; when you want sidecar metadata, because bids2table v2 Parquet files don&amp;rsquo;t include it.&lt;/p&gt;&#10;&lt;h2 id="prerequisites"&gt;Prerequisites&lt;/h2&gt;&#10;&lt;ul&gt;&#10;&lt;li&gt;SQLite Browser dependencies installed (&lt;code&gt;uv sync&lt;/code&gt; in the repo root), which includes bids2table&lt;/li&gt;&#10;&lt;li&gt;A BIDS dataset folder containing &lt;code&gt;dataset_description.json&lt;/code&gt; and &lt;code&gt;sub-*&lt;/code&gt; folders&lt;/li&gt;&#10;&lt;/ul&gt;&#10;&lt;h2 id="usage"&gt;Usage&lt;/h2&gt;&#10;&lt;div class="highlight"&gt;&lt;pre tabindex="0" style="background-color:#f7f7f7;-moz-tab-size:4;-o-tab-size:4;tab-size:4;-webkit-text-size-adjust:none;"&gt;&lt;code class="language-bash" data-lang="bash"&gt;&lt;span style="display:flex;"&gt;&lt;span&gt;uv run python utilities/bids2db.py PATH -o SQLITE_FILE &lt;span style="color:#0550ae"&gt;[&lt;/span&gt;options&lt;span style="color:#0550ae"&gt;]&lt;/span&gt;&#10;&lt;/span&gt;&lt;/span&gt;&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;table&gt;&#10;&#9;&lt;thead&gt;&#10;&#9;&#9;&#9;&lt;tr&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;th&gt;Argument&lt;/th&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;th&gt;Required&lt;/th&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;th&gt;Description&lt;/th&gt;&#10;&#9;&#9;&#9;&lt;/tr&gt;&#10;&#9;&lt;/thead&gt;&#10;&#9;&lt;tbody&gt;&#10;&#9;&#9;&#9;&lt;tr&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;&lt;code&gt;PATH&lt;/code&gt;&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;Yes&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;Root folder of the BIDS dataset&lt;/td&gt;&#10;&#9;&#9;&#9;&lt;/tr&gt;&#10;&#9;&#9;&#9;&lt;tr&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;&lt;code&gt;-o&lt;/code&gt; / &lt;code&gt;--output&lt;/code&gt;&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;Yes&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;Path for the output &lt;code&gt;.sqlite&lt;/code&gt; file (must end in &lt;code&gt;.sqlite&lt;/code&gt;)&lt;/td&gt;&#10;&#9;&#9;&#9;&lt;/tr&gt;&#10;&#9;&#9;&#9;&lt;tr&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;&lt;code&gt;--subjects&lt;/code&gt;&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;No&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;Subject folder names or glob patterns to include, with the &lt;code&gt;sub-&lt;/code&gt; prefix (e.g. &lt;code&gt;sub-01 'sub-1*'&lt;/code&gt;)&lt;/td&gt;&#10;&#9;&#9;&#9;&lt;/tr&gt;&#10;&#9;&#9;&#9;&lt;tr&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;&lt;code&gt;-j&lt;/code&gt; / &lt;code&gt;--workers&lt;/code&gt;&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;No&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;Worker processes for loading metadata (&lt;code&gt;0&lt;/code&gt; = main process (the default), &lt;code&gt;-1&lt;/code&gt; = all cores)&lt;/td&gt;&#10;&#9;&#9;&#9;&lt;/tr&gt;&#10;&#9;&#9;&#9;&lt;tr&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;&lt;code&gt;--use-threads&lt;/code&gt;&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;No&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;Use threads instead of processes when &lt;code&gt;--workers&lt;/code&gt; &amp;gt; 0&lt;/td&gt;&#10;&#9;&#9;&#9;&lt;/tr&gt;&#10;&#9;&#9;&#9;&lt;tr&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;&lt;code&gt;-q&lt;/code&gt; / &lt;code&gt;--no-progress&lt;/code&gt;&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;No&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;Disable the progress bar&lt;/td&gt;&#10;&#9;&#9;&#9;&lt;/tr&gt;&#10;&#9;&#9;&#9;&lt;tr&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;&lt;code&gt;-v&lt;/code&gt; / &lt;code&gt;--verbose&lt;/code&gt;&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;No&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;&lt;code&gt;-v&lt;/code&gt; for warnings, &lt;code&gt;-vv&lt;/code&gt; for more logging&lt;/td&gt;&#10;&#9;&#9;&#9;&lt;/tr&gt;&#10;&#9;&lt;/tbody&gt;&#10;&lt;/table&gt;&#10;&lt;p&gt;Example:&lt;/p&gt;</description></item><item><title>Concatenating SQLite files</title><link>https://nimh-dsst.github.io/sqlite-browser/docs/utilities/db-concatenate/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>https://nimh-dsst.github.io/sqlite-browser/docs/utilities/db-concatenate/</guid><description>&lt;p&gt;&lt;code&gt;utilities/db_concatenate.py&lt;/code&gt; appends the rows of one table (default &lt;code&gt;data&lt;/code&gt;) from several &lt;code&gt;.sqlite&lt;/code&gt; files into a single output &lt;code&gt;.sqlite&lt;/code&gt; file. This is useful when a dataset was indexed in pieces, for example one file per site produced by &lt;a href="https://nimh-dsst.github.io/sqlite-browser/docs/utilities/bids2db/"&gt;bids2db.py&lt;/a&gt; or &lt;a href="https://nimh-dsst.github.io/sqlite-browser/docs/utilities/parquets2db/"&gt;parquets2db.py&lt;/a&gt;.&lt;/p&gt;&#10;&lt;h2 id="usage"&gt;Usage&lt;/h2&gt;&#10;&lt;div class="highlight"&gt;&lt;pre tabindex="0" style="background-color:#f7f7f7;-moz-tab-size:4;-o-tab-size:4;tab-size:4;-webkit-text-size-adjust:none;"&gt;&lt;code class="language-bash" data-lang="bash"&gt;&lt;span style="display:flex;"&gt;&lt;span&gt;uv run python utilities/db_concatenate.py SQLITE_FILE &lt;span style="color:#0550ae"&gt;[&lt;/span&gt;SQLITE_FILE ...&lt;span style="color:#0550ae"&gt;]&lt;/span&gt; -o OUTPUT_FILE &lt;span style="color:#0550ae"&gt;[&lt;/span&gt;options&lt;span style="color:#0550ae"&gt;]&lt;/span&gt;&#10;&lt;/span&gt;&lt;/span&gt;&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;table&gt;&#10;&#9;&lt;thead&gt;&#10;&#9;&#9;&#9;&lt;tr&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;th&gt;Argument&lt;/th&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;th&gt;Default&lt;/th&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;th&gt;Description&lt;/th&gt;&#10;&#9;&#9;&#9;&lt;/tr&gt;&#10;&#9;&lt;/thead&gt;&#10;&#9;&lt;tbody&gt;&#10;&#9;&#9;&#9;&lt;tr&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;&lt;code&gt;SQLITE_FILE ...&lt;/code&gt;&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;required&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;Space-separated input files; each must end in &lt;code&gt;.sqlite&lt;/code&gt;&lt;/td&gt;&#10;&#9;&#9;&#9;&lt;/tr&gt;&#10;&#9;&#9;&#9;&lt;tr&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;&lt;code&gt;-o&lt;/code&gt; / &lt;code&gt;--output&lt;/code&gt;&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;required&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;Output file; must end in &lt;code&gt;.sqlite&lt;/code&gt; and must not be one of the inputs&lt;/td&gt;&#10;&#9;&#9;&#9;&lt;/tr&gt;&#10;&#9;&#9;&#9;&lt;tr&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;&lt;code&gt;-t&lt;/code&gt; / &lt;code&gt;--table&lt;/code&gt;&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;&lt;code&gt;data&lt;/code&gt;&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;Table to concatenate from every input&lt;/td&gt;&#10;&#9;&#9;&#9;&lt;/tr&gt;&#10;&#9;&lt;/tbody&gt;&#10;&lt;/table&gt;&#10;&lt;p&gt;Example:&lt;/p&gt;</description></item><item><title>Enriching an SQLite database with a TSV join</title><link>https://nimh-dsst.github.io/sqlite-browser/docs/utilities/db-join/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>https://nimh-dsst.github.io/sqlite-browser/docs/utilities/db-join/</guid><description>&lt;p&gt;After creating a SQLite file (see &lt;a href="https://nimh-dsst.github.io/sqlite-browser/docs/utilities/bids2db/"&gt;Indexing a BIDS dataset&lt;/a&gt; or &lt;a href="https://nimh-dsst.github.io/sqlite-browser/docs/utilities/parquets2db/"&gt;Creating a SQLite file from Parquet files&lt;/a&gt;), you may want to enrich the &lt;code&gt;data&lt;/code&gt; table with additional per-subject or per-session metadata stored in a separate TSV or CSV file — for example, a phenotype file or an ID-mapping file. &lt;code&gt;utilities/db_join.py&lt;/code&gt; performs that join and writes the result back into the database.&lt;/p&gt;&#10;&lt;h2 id="prerequisites"&gt;Prerequisites&lt;/h2&gt;&#10;&lt;ul&gt;&#10;&lt;li&gt;An existing &lt;code&gt;.sqlite&lt;/code&gt; database (e.g., produced by &lt;code&gt;bids2db.py&lt;/code&gt; or &lt;code&gt;parquets2db.py&lt;/code&gt;)&lt;/li&gt;&#10;&lt;li&gt;A TSV or CSV file with a column that shares values with a column in the database table&lt;/li&gt;&#10;&lt;/ul&gt;&#10;&lt;h2 id="usage"&gt;Usage&lt;/h2&gt;&#10;&lt;div class="highlight"&gt;&lt;pre tabindex="0" style="background-color:#f7f7f7;-moz-tab-size:4;-o-tab-size:4;tab-size:4;-webkit-text-size-adjust:none;"&gt;&lt;code class="language-bash" data-lang="bash"&gt;&lt;span style="display:flex;"&gt;&lt;span&gt;uv run python utilities/db_join.py DATABASE CSV_FILE -k DB_KEY -c CSV_KEY &lt;span style="color:#0550ae"&gt;[&lt;/span&gt;options&lt;span style="color:#0550ae"&gt;]&lt;/span&gt;&#10;&lt;/span&gt;&lt;/span&gt;&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;h3 id="required-arguments"&gt;Required arguments&lt;/h3&gt;&#10;&lt;table&gt;&#10;&#9;&lt;thead&gt;&#10;&#9;&#9;&#9;&lt;tr&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;th&gt;Argument&lt;/th&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;th&gt;Description&lt;/th&gt;&#10;&#9;&#9;&#9;&lt;/tr&gt;&#10;&#9;&lt;/thead&gt;&#10;&#9;&lt;tbody&gt;&#10;&#9;&#9;&#9;&lt;tr&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;&lt;code&gt;DATABASE&lt;/code&gt;&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;Path to the existing &lt;code&gt;.sqlite&lt;/code&gt; file&lt;/td&gt;&#10;&#9;&#9;&#9;&lt;/tr&gt;&#10;&#9;&#9;&#9;&lt;tr&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;&lt;code&gt;CSV_FILE&lt;/code&gt;&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;Path to the TSV or CSV file to join in&lt;/td&gt;&#10;&#9;&#9;&#9;&lt;/tr&gt;&#10;&#9;&#9;&#9;&lt;tr&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;&lt;code&gt;-k&lt;/code&gt; / &lt;code&gt;--db-key&lt;/code&gt;&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;Column name &lt;strong&gt;in the database table&lt;/strong&gt; to join on&lt;/td&gt;&#10;&#9;&#9;&#9;&lt;/tr&gt;&#10;&#9;&#9;&#9;&lt;tr&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;&lt;code&gt;-c&lt;/code&gt; / &lt;code&gt;--csv-key&lt;/code&gt;&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;Column name &lt;strong&gt;in the TSV/CSV file&lt;/strong&gt; to join on&lt;/td&gt;&#10;&#9;&#9;&#9;&lt;/tr&gt;&#10;&#9;&lt;/tbody&gt;&#10;&lt;/table&gt;&#10;&lt;h3 id="optional-arguments"&gt;Optional arguments&lt;/h3&gt;&#10;&lt;table&gt;&#10;&#9;&lt;thead&gt;&#10;&#9;&#9;&#9;&lt;tr&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;th&gt;Argument&lt;/th&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;th&gt;Default&lt;/th&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;th&gt;Description&lt;/th&gt;&#10;&#9;&#9;&#9;&lt;/tr&gt;&#10;&#9;&lt;/thead&gt;&#10;&#9;&lt;tbody&gt;&#10;&#9;&#9;&#9;&lt;tr&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;&lt;code&gt;-t&lt;/code&gt; / &lt;code&gt;--table&lt;/code&gt;&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;&lt;code&gt;data&lt;/code&gt;&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;Name of the database table to read from&lt;/td&gt;&#10;&#9;&#9;&#9;&lt;/tr&gt;&#10;&#9;&#9;&#9;&lt;tr&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;&lt;code&gt;-j&lt;/code&gt; / &lt;code&gt;--join-type&lt;/code&gt;&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;&lt;code&gt;outer&lt;/code&gt;&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;Join type: &lt;code&gt;left&lt;/code&gt;, &lt;code&gt;right&lt;/code&gt;, &lt;code&gt;inner&lt;/code&gt;, or &lt;code&gt;outer&lt;/code&gt;&lt;/td&gt;&#10;&#9;&#9;&#9;&lt;/tr&gt;&#10;&#9;&#9;&#9;&lt;tr&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;&lt;code&gt;-o&lt;/code&gt; / &lt;code&gt;--output-table&lt;/code&gt;&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;&lt;code&gt;data&lt;/code&gt;&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;Name of the table to write the joined result to&lt;/td&gt;&#10;&#9;&#9;&#9;&lt;/tr&gt;&#10;&#9;&#9;&#9;&lt;tr&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;&lt;code&gt;--replace&lt;/code&gt;&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;off&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;Replace the output table if it already exists&lt;/td&gt;&#10;&#9;&#9;&#9;&lt;/tr&gt;&#10;&#9;&#9;&#9;&lt;tr&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;&lt;code&gt;-s&lt;/code&gt; / &lt;code&gt;--separator&lt;/code&gt;&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;auto&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;Column delimiter; auto-detected from file extension (&lt;code&gt;.tsv&lt;/code&gt;/&lt;code&gt;.tab&lt;/code&gt; → tab, everything else → comma)&lt;/td&gt;&#10;&#9;&#9;&#9;&lt;/tr&gt;&#10;&#9;&lt;/tbody&gt;&#10;&lt;/table&gt;&#10;&lt;h2 id="step-by-step"&gt;Step-by-step&lt;/h2&gt;&#10;&lt;h3 id="1-identify-your-join-columns"&gt;1. Identify your join columns&lt;/h3&gt;&#10;&lt;p&gt;Open &lt;code&gt;db.sqlite&lt;/code&gt; in SQLite Browser and note the column name you want to join on. Then check your TSV or CSV file for the matching column name — the two do not need to have the same name.&lt;/p&gt;</description></item><item><title>Creating an SQLite file from Parquet files</title><link>https://nimh-dsst.github.io/sqlite-browser/docs/utilities/parquets2db/</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>https://nimh-dsst.github.io/sqlite-browser/docs/utilities/parquets2db/</guid><description>&lt;p&gt;&lt;code&gt;utilities/parquets2db.py&lt;/code&gt; collects one or more Parquet files, concatenates them into a single table, and writes the result to a SQLite file that can be opened directly in SQLite Browser. While it was originally built around &lt;a href="https://childmindresearch.github.io/bids2table/bids2table.html"&gt;bids2table&lt;/a&gt; outputs, it works with &lt;strong&gt;any&lt;/strong&gt; Parquet files.&lt;/p&gt;&#10;&lt;blockquote&gt;&#10;&lt;p&gt;&lt;strong&gt;Sidecar metadata:&lt;/strong&gt; bids2table v2 Parquet files don&amp;rsquo;t include JSON sidecar metadata. To include it, build the SQLite file directly from the BIDS dataset with &lt;a href="https://nimh-dsst.github.io/sqlite-browser/docs/utilities/bids2db/"&gt;bids2db.py&lt;/a&gt;.&lt;/p&gt;&#10;&lt;/blockquote&gt;&#10;&lt;h2 id="prerequisites"&gt;Prerequisites&lt;/h2&gt;&#10;&lt;ul&gt;&#10;&lt;li&gt;SQLite Browser dependencies installed (&lt;code&gt;uv sync&lt;/code&gt; in the repo root)&lt;/li&gt;&#10;&lt;li&gt;One or more &lt;code&gt;.parquet&lt;/code&gt; files to convert&lt;/li&gt;&#10;&lt;/ul&gt;&#10;&lt;h2 id="usage"&gt;Usage&lt;/h2&gt;&#10;&lt;div class="highlight"&gt;&lt;pre tabindex="0" style="background-color:#f7f7f7;-moz-tab-size:4;-o-tab-size:4;tab-size:4;-webkit-text-size-adjust:none;"&gt;&lt;code class="language-bash" data-lang="bash"&gt;&lt;span style="display:flex;"&gt;&lt;span&gt;uv run python utilities/parquets2db.py PARQUET_GLOB -o SQLITE_FILE&#10;&lt;/span&gt;&lt;/span&gt;&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;table&gt;&#10;&#9;&lt;thead&gt;&#10;&#9;&#9;&#9;&lt;tr&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;th&gt;Argument&lt;/th&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;th&gt;Required&lt;/th&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;th&gt;Description&lt;/th&gt;&#10;&#9;&#9;&#9;&lt;/tr&gt;&#10;&#9;&lt;/thead&gt;&#10;&#9;&lt;tbody&gt;&#10;&#9;&#9;&#9;&lt;tr&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;&lt;code&gt;PARQUET_GLOB&lt;/code&gt;&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;Yes&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;Python glob pattern matching all Parquet files to include&lt;/td&gt;&#10;&#9;&#9;&#9;&lt;/tr&gt;&#10;&#9;&#9;&#9;&lt;tr&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;&lt;code&gt;-o&lt;/code&gt; / &lt;code&gt;--output&lt;/code&gt;&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;Yes&lt;/td&gt;&#10;&#9;&#9;&#9;&#9;&#9;&lt;td&gt;Path for the output &lt;code&gt;.sqlite&lt;/code&gt; file (must end in &lt;code&gt;.sqlite&lt;/code&gt;)&lt;/td&gt;&#10;&#9;&#9;&#9;&lt;/tr&gt;&#10;&#9;&lt;/tbody&gt;&#10;&lt;/table&gt;&#10;&lt;h2 id="step-by-step"&gt;Step-by-step&lt;/h2&gt;&#10;&lt;h3 id="1-locate-your-parquet-files"&gt;1. Locate your Parquet files&lt;/h3&gt;&#10;&lt;p&gt;Here is an example of how Parquet files might look on the filesystem:&lt;/p&gt;</description></item></channel></rss>