Google BigQuery support for Spark, Structured Streaming, SQL, and DataFrames with easy Databricks integration.
☆70May 8, 2023Updated 3 years ago
Alternatives and similar repositories for spark-bigquery
Users that are interested in spark-bigquery are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Google BigQuery support for Spark, SQL, and DataFrames☆156Dec 14, 2019Updated 6 years ago
- A handy Scala wrapper of Google BigQuery API 's Java Client Library.☆35Sep 29, 2018Updated 7 years ago
- ☆31Oct 17, 2018Updated 7 years ago
- IPython magics to work with DBT☆15Jul 22, 2022Updated 4 years ago
- An application that uses Cloud Dataflow and Cloud Build to copy/transfer BigQuery tables between locations/regions.☆14Mar 17, 2021Updated 5 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ScalikeJDBC extension for Google BigQuery☆18Mar 15, 2020Updated 6 years ago
- Spark pipelines that correspond to a series of Dataflow examples.☆27May 5, 2019Updated 7 years ago
- Java Swing in HTML5 Browser.☆29Aug 6, 2026Updated 3 weeks ago
- Recipes and examples for Apache Spark☆13Jan 21, 2015Updated 11 years ago
- ☆46Apr 27, 2020Updated 6 years ago
- Spark Structured Streaming State Tools☆35Jul 3, 2020Updated 6 years ago
- Minitime - a Java Time wrapper for Scala and Scala.js☆16Jan 17, 2020Updated 6 years ago
- hive_compared_bq compares/validates 2 (SQL like) tables, and graphically shows the rows/columns that are different.☆26Dec 13, 2017Updated 8 years ago
- Shapeless utilities for common data types☆68Updated this week
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Named Entity Extraction on Twitter Stream using Apache Spark Streaming and Stanford CoreNLP☆13Oct 12, 2016Updated 9 years ago
- Postgres extension drivers for quill☆14Oct 31, 2016Updated 9 years ago
- Hive Storage Handler for interoperability between BigQuery and Apache Hive☆19Jan 29, 2025Updated last year
- A SBT resolver and publisher for Google Cloud Storage☆23Dec 15, 2021Updated 4 years ago
- My entry to the Kaggle 2013 StumbleUpon competition. Ranked 4th on the final private leaderboard.☆15Apr 23, 2014Updated 12 years ago
- Using the Parquet file format (with Avro) to process data with Apache Flink☆14Aug 17, 2015Updated 11 years ago
- Spark data profiling utilities☆23Nov 24, 2018Updated 7 years ago
- Helm Chart for lyft/flinkk8soperator☆11Mar 10, 2020Updated 6 years ago
- Transfers inotify events from NFS server to client (for example, for MPD library auto-update)☆16Sep 16, 2013Updated 12 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Make your Spray applications simpler with Skinny components☆17Aug 20, 2016Updated 10 years ago
- Apache Calcite Adapter for Apache Kudu☆28Sep 26, 2025Updated 11 months ago
- ☆31Mar 7, 2025Updated last year
- ☆13Updated this week
- Python Environment Variables Manager☆11Jun 18, 2016Updated 10 years ago
- an example of integrating Spark Streaming with Google Pub/Sub and Google Datastore☆16Mar 22, 2017Updated 9 years ago
- Scripts to demonstrate VPC Service Controls between tenant and shared projects☆12Jun 11, 2019Updated 7 years ago
- Tutorials, Examples about Kubeflow Pipeline.☆13Nov 21, 2022Updated 3 years ago
- Test for SparkSQL ScalaPB☆14Jun 28, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Fuzzy matching function in spark (https://spark-packages.org/package/itspawanbhardwaj/spark-fuzzy-matching)☆24Dec 30, 2019Updated 6 years ago
- ☆67Aug 16, 2024Updated 2 years ago
- This repository is deprecated. All of its content and history has been moved to googleapis/google-cloud-node.☆12Jul 13, 2023Updated 3 years ago
- Introduction à la science des données et à l’intelligence artificielle☆18Jun 3, 2020Updated 6 years ago
- Connector between Spark and InfluxDB.☆23May 19, 2016Updated 10 years ago
- proof-of-concept implementation of Pig-on-Spark integrated at the logical node level☆29Jul 7, 2022Updated 4 years ago
- Learning Spark SQL, published by Packt☆43Jan 30, 2023Updated 3 years ago