# Dremio S3 Iceberg Catalog

**URL:** <https://community.dremio.com/t/dremio-s3-iceberg-catalog/11794>\
**Category:** Uncategorized\
**Created:** [May 6, 2024, 4:02pm UTC](https://community.dremio.com/t/dremio-s3-iceberg-catalog/11794 "2024-05-06T16:02:20Z")\
**Posts on this page:** 6\
**Page:** 1

<div class="post-metadata">

**Author:** ![sheinbergon](https://sea2.discourse-cdn.com/flex020/user_avatar/community.dremio.com/sheinbergon/32/2408_2.png) [@sheinbergon](https://community.dremio.com/u/sheinbergon)\
**Post date:** [May 6, 2024, 4:02pm UTC](https://community.dremio.com/t/dremio-s3-iceberg-catalog/11794/1 "2024-05-06T16:02:20Z")

</div>

Hello

We are using DBT + Dremio CE 24.3.0 to create materialized iceberg views of our queries.  
We instruct the the DBT to create the ICEBERG table on top of a specific S3 bucket, as it doesn’t support Glue integration.

We wish to be able to access these iceberg tables externally using PyIceberg.  
While static tables access is an option, we wish to be to do this using a standalone catalog.

What catalog does dremio use when Iceberg tables are created on top of S3 buckets? is it externally accessible? is it configurable?

---

<div class="post-metadata">

**Author:** ![balaji.ramaswamy](https://sea2.discourse-cdn.com/flex020/user_avatar/community.dremio.com/balaji.ramaswamy/32/843_2.png) [@balaji.ramaswamy](https://community.dremio.com/u/balaji.ramaswamy)\
**Post date:** [May 13, 2024, 6:57am UTC](https://community.dremio.com/t/dremio-s3-iceberg-catalog/11794/2 "2024-05-13T06:57:30Z")

</div>

@sheinbergon When you create ICeberg table on S3 source using Dremio it will default to a Hadoop catalog

> **[Apache Iceberg | Dremio Documentation](https://docs.dremio.com/current/sonar/query-manage/data-formats/apache-iceberg/#iceberg-catalogs-in-dremio)**
>
> Apache Iceberg is an open table format designed for gigantic, petabyte-scale tables and is rapidly becoming an industry standard for managing data in data lakes. A table format helps you manage, organize, and track all of the files that make up a...

When you create the iceberg catalog, you wil ldefine the catalog and Dremio supports catalogs listed in above link

---

<div class="post-metadata">

**Author:** ![rdkworld](https://avatars.discourse-cdn.com/v4/letter/r/6de8d8/32.png) [@rdkworld](https://community.dremio.com/u/rdkworld)\
**Post date:** [July 2, 2025, 2:01pm UTC](https://community.dremio.com/t/dremio-s3-iceberg-catalog/11794/3 "2025-07-02T14:01:25Z")

</div>

@sheinbergon @balaji.ramaswamy So, do we have a solution to this, how can iceberg tables with data in S3 (& default cataloging Hadoop) be able to be access externally using PyIceberg. Would like know if the default Hadoop catalog (on Dremio) is exposed via url, using Dremio 25.2.x

---

<div class="post-metadata">

**Author:** ![sheinbergon](https://sea2.discourse-cdn.com/flex020/user_avatar/community.dremio.com/sheinbergon/32/2408_2.png) [@sheinbergon](https://community.dremio.com/u/sheinbergon)\
**Post date:** [July 10, 2025, 7:22am UTC](https://community.dremio.com/t/dremio-s3-iceberg-catalog/11794/4 "2025-07-10T07:22:10Z")

</div>

Yes. I Just use dbt-dremio and create using Glue as the the table storage, not S3. Though not clearly documented in dbt-dremio adapter, it works perfectly OTB. I can then access these tables either through Dremio (Glue) or directly via PyIceberg’s Glue catalog spec

---

<div class="post-metadata">

**Author:** ![rdkworld](https://avatars.discourse-cdn.com/v4/letter/r/6de8d8/32.png) [@rdkworld](https://community.dremio.com/u/rdkworld)\
**Post date:** [July 15, 2025, 9:26pm UTC](https://community.dremio.com/t/dremio-s3-iceberg-catalog/11794/5 "2025-07-15T21:26:36Z")

</div>

@sheinbergon Thanks for the response, yes I was used dbt-dremio to create iceberg tables but it will use default hadoop catalog not Glue catalog. You mentioned after you inserted thru dbt, you were able to access via PyIceberg’s glue catalog, can you help me out here…how did you get the cataloging in Glue when it was in Hadoop (which Dremio used when inserting via dbt-dremio). From what I understand, cataloging and data are tied together, only query engine can differ (Dremio or pyiceberg). I want to access data inserted thru dbt-dremio using PyIceberg

---

<div class="post-metadata">

**Author:** ![sheinbergon](https://sea2.discourse-cdn.com/flex020/user_avatar/community.dremio.com/sheinbergon/32/2408_2.png) [@sheinbergon](https://community.dremio.com/u/sheinbergon)\
**Post date:** [July 31, 2025, 8:31am UTC](https://community.dremio.com/t/dremio-s3-iceberg-catalog/11794/6 "2025-07-31T08:31:09Z")

</div>

You just need to recreate the table on the glue catalog and insert the data there (via MERGE/INSERT). I’m not sure what’s gap here
