Skill 디렉토리

AI Agent를 위한 재사용 가능한 Skill을 찾으세요.

작업으로 실제 GitHub Skill을 검색하고 사용 전에 Stars, 신뢰, 감사, 카테고리, 설치 경로를 확인하세요.

모든 추천은 리포지토리, 감사, 설치 경로와 명확하게 연결됩니다.

검색 결과: parquet

영문 디렉토리

Open Lakehouse Format for Multimodal AI. Convert from Parquet in 2 lines of code for 100x faster random access, vector index, and data versioning. Compatible with Pandas, DuckDB, Polars, Pyarrow, and PyTorch with more integrations coming..

6.7K
Stars
86/100
신뢰
카테고리: ml-automation감사

pandas on AWS - Easy integration with Athena, Glue, Redshift, Timestream, Neptune, OpenSearch, QuickSight, Chime, CloudWatchLogs, DynamoDB, EMR, SecretManager, PostgreSQL, MySQL, SQLServer and S3 (Parquet, CSV, JSON and EXCEL).

4.1K
Stars
80/100
신뢰
카테고리: data-analysis감사

Chat with your database or your datalake (SQL, CSV, parquet). PandasAI makes data analysis conversational using LLMs and RAG.

24K
Stars
73/100
신뢰
카테고리: data-analysis감사

Specification for storing geospatial vector data (point, line, polygon) in Parquet

1.1K
Stars
84/100
신뢰
카테고리: geo-science감사

ADAM is a genomics analysis platform with specialized file formats built using Apache Avro, Apache Spark, and Apache Parquet. Apache 2 licensed.

1.1K
Stars
80/100
신뢰
카테고리: geo-science감사

Petastorm library enables single machine or distributed training and evaluation of deep learning models from datasets in Apache Parquet format. It supports ML frameworks such as Tensorflow, Pytorch, and PySpark and can be used from pure Python code.

1.9K
Stars
76/100
신뢰
카테고리: ml-automation감사
Dsq62

Commandline tool for running SQL queries against JSON, CSV, Excel, Parquet, and more.

3.9K
Stars
62/100
신뢰
카테고리: data-analysis감사

ETL framework for .NET (Parser / Writer for CSV, Flat, Xml, JSON, Key-Value, Parquet, Yaml, Avro formatted files)

859
Stars
72/100
신뢰
카테고리: data-analysis감사

EasyDB is a lightweight desktop app built with Tauri + Rust, powered by Apache DataFusion. Query local CSV, TSV, Text, NdJson, Excel, Parquet files and MySQL databases directly with SQL — no external database required. Handles datasets from hundreds of MB to several GB with ease. 让所有数据说同一种“语言”

625
Stars
70/100
신뢰
카테고리: data-analysis감사
Arc71

High-performance analytical database. 19.9M records/sec ingestion, 8.4M+ rows/sec queries. Ingestion, compaction, SQL, retention, continuous queries — one binary. Open Parquet on your storage. S3/Azure native. Air-gap ready. No vendor lock-in. AGPL-3.0.

609
Stars
71/100
신뢰
카테고리: devops감사

An experimental embedded SQL engine in C++20. Query Parquet, CSV, JSON, Arrow, Avro, SQLite, and Excel files directly with SQL, in-process. Early-stage.

448
Stars
67/100
신뢰
카테고리: data-analysis감사

sql driver for CSV, TSV, LTSV, JSON, Parquet, Excel with gzip, bzip2, xz, zstd support.

371
Stars
68/100
신뢰
카테고리: data-analysis감사