/
/
JDBC

Parquet JDBC Driver

Easily connect live Parquet data with Java-based BI, ETL, reporting, AI/ML & custom apps.

Other Technologies
Decorative Icon Parquet Logo
Claude
Claude
GitHub Copilot
Gemini
New
CData Drivers now work with AI Coding tools

The Parquet JDBC Driver enables users to connect with live Parquet data, directly from any applications that support JDBC connectivity. Rapidly create and deploy powerful Java applications that integrate with Parquet.

JDBC architecture

Parquet JDBC Connectivity Features

  • SQL access to Apache Parquet data
  • Connect to live Apache Parquet data, for real-time data access with the Apache Parquet Python Connectors
  • Full support for data aggregation and complex JOINs in SQL queries
  • Secure connectivity through modern cryptography, including TLS 1.2, SHA-256, ECC, etc.
  • Generate table schema automatically based on existing Apache Parquet data or manually for greater control of the content you need
  • Seamless integration with leading BI, reporting, and ETL tools and with custom applications via the Parquet Connector.

Target Service, API

The driver provides SQL access to Apache Parquet files. Supports local files, network paths, and cloud storage locations (S3, Azure Blob, etc.).

Schema, Data Model

Automatically infers schema from Parquet file metadata. Supports complex types including nested structures and arrays.

Key Objects

Parquet files are exposed as tables. Each file or set of files in a directory becomes a queryable table. Supports partitioned datasets.

Operations

Read-only SQL queries on Parquet data. Supports filtering, projections, and aggregations with predicate pushdown for optimal performance.

Authentication

File system or cloud storage authentication. For cloud storage, supports AWS credentials, Azure credentials, or Google Cloud credentials.

Start a 30-day free trial today

JDBC access to Apache Parquet

Full-featured and consistent SQL access to any supported data source through JDBC


See what you can do with Parquet JDBC Driver

ETL, Data Warehouse
Data integration & ETL

Integrate Parquet into your systems and data warehouses through popular Java-based ETL/EAI tools. Supports both self-hosted environments and cloud service deployment.

BI, Reporting, Virtualization
BI, reporting & data virtualization

Connect to Parquet from any JDBC-compatible BI, reporting, and data virtualization platform. Provides seamless integration using SQL as the standard query interface across all tools.

Custom Java Applications
Custom Java applications

Use the Parquet JDBC Driver to rapidly deliver Java-based applications that connect with Parquet. Universal SQL-based interactivity simplifies integration and speeds time to market.

Connect to Parquet—empower every team

Management & Data Consumer
IT Department
ISVs & cloud service vendors

AI-assisted development with CData CLI

Build Parquet integrations faster with AI that understands your schema

Schema-aware AI

CData CLI gives AI coding tools access to your Parquet schema. No more guessing table names or column types—AI sees the same metadata in your JDBC Drivers.

Your AI knows SQL

How to find table names, column names, and how to generate SQL syntax are things that AI knows well from millions of training data. No need for customization, no hallucinations. Your AI acts like a domain specialist to Parquet.

More accurate, more token-efficient

With CData CLI's queryable schema detection and highly efficient queries with filters, aggregation, joins with correct pushdown, your AI will achieve more accuracy with less token usage.

Supported AI Coding Tools
Download CData CLI
FAQs

Frequently asked Parquet JDBC driver questions

Learn more about Parquet JDBC drivers for data and analytics integration

Ready to get started? Try CData for free today