Parquet JDBC Driver
Easily connect live Parquet data with Java-based BI, ETL, reporting, AI/ML & custom apps.
The Parquet JDBC Driver enables users to connect with live Parquet data, directly from any applications that support JDBC connectivity. Rapidly create and deploy powerful Java applications that integrate with Parquet.
Parquet JDBC Connectivity Features
- SQL access to Apache Parquet data
- Connect to live Apache Parquet data, for real-time data access with the Apache Parquet Python Connectors
- Full support for data aggregation and complex JOINs in SQL queries
- Secure connectivity through modern cryptography, including TLS 1.2, SHA-256, ECC, etc.
- Generate table schema automatically based on existing Apache Parquet data or manually for greater control of the content you need
- Seamless integration with leading BI, reporting, and ETL tools and with custom applications via the Parquet Connector.
Target Service, API
The driver provides SQL access to Apache Parquet files. Supports local files, network paths, and cloud storage locations (S3, Azure Blob, etc.).
Schema, Data Model
Automatically infers schema from Parquet file metadata. Supports complex types including nested structures and arrays.
Key Objects
Parquet files are exposed as tables. Each file or set of files in a directory becomes a queryable table. Supports partitioned datasets.
Operations
Read-only SQL queries on Parquet data. Supports filtering, projections, and aggregations with predicate pushdown for optimal performance.
Authentication
File system or cloud storage authentication. For cloud storage, supports AWS credentials, Azure credentials, or Google Cloud credentials.
Start a 30-day free trial today
JDBC access to Apache Parquet
Full-featured and consistent SQL access to any supported data source through JDBC
See what you can do with Parquet JDBC Driver
Integrate Parquet into your systems and data warehouses through popular Java-based ETL/EAI tools. Supports both self-hosted environments and cloud service deployment.
Connect to Parquet from any JDBC-compatible BI, reporting, and data virtualization platform. Provides seamless integration using SQL as the standard query interface across all tools.
Use the Parquet JDBC Driver to rapidly deliver Java-based applications that connect with Parquet. Universal SQL-based interactivity simplifies integration and speeds time to market.
Connect to Parquet—empower every team
AI-assisted development with CData CLI
Build Parquet integrations faster with AI that understands your schema
Schema-aware AI
CData CLI gives AI coding tools access to your Parquet schema. No more guessing table names or column types—AI sees the same metadata in your JDBC Drivers.
Your AI knows SQL
How to find table names, column names, and how to generate SQL syntax are things that AI knows well from millions of training data. No need for customization, no hallucinations. Your AI acts like a domain specialist to Parquet.
More accurate, more token-efficient
With CData CLI's queryable schema detection and highly efficient queries with filters, aggregation, joins with correct pushdown, your AI will achieve more accuracy with less token usage.
FAQs
Frequently asked Parquet JDBC driver questions
Learn more about Parquet JDBC drivers for data and analytics integration
Popular JDBC Videos: