AI - Parquet

Provides efficient columnar storage & simplified processing for large datasets via open-source Parquet tech. Committed to coding standards, community collaboration, and Apache licensing.

Logo of Parquet
Last Audited At

About Parquet

Parquet is an open-source data storage technology developed under the Apache Software Foundation. It focuses on delivering efficient, columnar storage solutions for large datasets. Parquet's primary goal is to make data analysis faster by providing optimized I/O and compression.

Parquet offers services that simplify data processing when using platforms like Hadoop or Spark. The company prides itself on strict adherence to coding standards, ensuring readability and maintainability of the codebase. Their key offerings include:

  1. Pig compatibility: Parquet's implementation of Pig Latin, the programming language for defining data flow graphs, allows for seamless integration with tools like Apache Pig and Hive.
  2. Schema conversion: Parquet offers automatic schema conversion for certain data storage methods to make the process easier for developers.
  3. Contributor community: The project boasts a dedicated community of authors and contributors that collaborate on enhancing and expanding Parquet's capabilities.
  4. Code of conduct: Parquet adheres to two codes of conduct - Apache Software Foundation's and Twitter's, ensuring a positive and inclusive development environment.
  5. Licensing: Parquet is licensed under the Apache License, Version 2.0.

To contribute to Parquet, one can send pull requests against the official Git repository at https://github.com/apache/parquet-mr. The project encourages contributions in various forms, including code improvements and issue reporting.

Was this page helpful?

More companies

KubeFlow

Empowering machine learning communities with a comprehensive platform for pipelines, training, and deployment through collaborative development and community involvement - KubeFlow.

Read more

Verily

Bridging the gap between research and care in precision health through Alphabet's AI expertise and Verily's clinical excellence, data privacy, and security.

Read more

Aerospike

Empowering businesses with high-performance, real-time data solutions through Aerospike's enterprise-grade databases, proven to outperform competitors and trusted by industry leaders.

Read more

Tell us about your project

Our Hubs

London, United Kingdom

A global AI hotspot, thrives on innovation, diverse talent, and a dynamic tech ecosystem, offering unparalleled opportunities for AI engineers.

Munich, Germany

A vibrant AI hub, merges cutting-edge technology with rich cultural experiences, creating an inspiring environment for AI engineers.