Best way to get parquet data from AWS S3 bucket?


We have some parquet files being replicated to an AWS S3 bucket.
I've started to look to see if I can use Amazon Glue to crawl the bucket, Athena to query the Glue table, and then Domo to pull data from Athena. I'm running into a few issues (like the initial load file Glue picks up as a table, says it has rows, but Athena can't query any data from it) but I think I can get there.
However, before I go too far down the road, is there another approach that works?
Unfortunately the S3 connector doesn't read parquet files.
I could convert them to CSV and upload directly to Domo using something like https://stackoverflow.com/questions/62275672/converting-parquet-files-in-s3-to-csv-and-store-back-in-s3 but that seems ... cludgy?
If anyone has any suggestions, I'm all ears.
Answers
-
I think that would be your best bet currently. If you haven't upvoted this in ideas exchange to get parquet files supported in Domo I would do so https://dojo.domo.com/main/discussion/51684/ingesting-parquet-files
If I have answered your question, please click "Yes" on my comment option.
I also specialize in consumption consulting.0 -
Yep - I found that and upvoted it.
0 -
It looks like there is a parquet reader built into the Domo CLI tool.
0
Categories
- All Categories
- 2K Product Ideas
- 2K Ideas Exchange
- 1.6K Connect
- 1.3K Connectors
- 311 Workbench
- 6 Cloud Amplifier
- 9 Federated
- 3.8K Transform
- 657 Datasets
- 115 SQL DataFlows
- 2.2K Magic ETL
- 815 Beast Mode
- 3.3K Visualize
- 2.5K Charting
- 81 App Studio
- 45 Variables
- 775 Automate
- 190 Apps
- 481 APIs & Domo Developer
- 81 Workflows
- 23 Code Engine
- 40 AI and Machine Learning
- 20 AI Chat
- 1 AI Playground
- 1 AI Projects and Models
- 18 Jupyter Workspaces
- 410 Distribute
- 120 Domo Everywhere
- 280 Scheduled Reports
- 10 Software Integrations
- 144 Manage
- 140 Governance & Security
- 8 Domo Community Gallery
- 48 Product Releases
- 12 Domo University
- 5.4K Community Forums
- 41 Getting Started
- 31 Community Member Introductions
- 114 Community Announcements
- 4.8K Archive