dataaaaa!
a platform to stack them all
489 Sharing resources collected and tagged on dataaaaa — 218 articles, 127 podcasts, 127 release notes, 12 events and 5 projects. The 30 most recent are listed below, newest first.
Focus on data distribution over storage or collection. Effective data architecture builds around primitives that preserve optionality rather than fixed answers, ensuring systems adapt as business needs shift. Treat data strategy as continuous work rather than a static roadmap.
JupyterGIS 0.16 adds support for GeoZarr, GeoPackage, and openEO layers. Data engineers can visualize remote processing pipelines as lazy, tile-based map layers directly from openEO backends without local materialization, and stream Xarray datasets via jupyter-tiler.
Snowflake published updates for Snowflake Data Clean Rooms alongside several data engineering features. Key platform enhancements include general availability for online modification of hybrid table constraints, external lineage tracking, and second generation Openflow runtimes and deployments.
Snowflake deployed Data Clean Rooms updates alongside external lineage support and data lineage for Cortex Agents. Releases also delivered an increased file limit for deployed dbt project objects, new Iceberg REST integrations, and CDC via Openflow Connector for SQL Server.
Snowflake released a preview of remote app operations for Snowflake Native Apps. Additional updates include remote development in VS Code, stable egress IP addresses on Azure, and Apache Iceberg catalog integrations to enhance data engineering workflows across platforms.
Traditional APIs limit data engineering by restricting queries, throughput, and cross-service joins. True interoperation requires streaming the entire dataset to build custom indexes, models, and shared databases rather than querying closed, rate-limited REST endpoints.
Snowflake released feature updates for Data Clean Rooms alongside improvements to the data platform. Recent additions include Cortex agent capabilities, multi-value tags, and enhanced data freshness for ORGANIZATION_USAGE views.
Snowflake introduced Request Access in Workspaces in preview, alongside Amazon S3 Tables Iceberg REST catalog integration. Recent data engineering updates also include CREATE OR ALTER support for dynamic Iceberg tables and new usage monitoring views.
Snowflake deployed updates for Snowflake Data Clean Rooms alongside general availability for AIEXTRACT and AIPARSE_DOCUMENT with client-side encrypted stages. Data engineers also gained new data movement policies and Horizon Catalog access to externally managed Apache Iceberg tables.
Snowflake enables life sciences M&A data integration without physical migration using Secure Data Sharing and federated architectures. For divestitures, Zero-Copy Cloning and row-level security isolate commingled datasets without breaking GxP compliance or duplicating storage costs.
Enterprises are adopting zero-copy data architectures using open table formats like Apache Iceberg to query source data directly, eliminating redundant ETL pipelines and cutting storage costs. Embedded AI models process unstructured files and run conversational queries directly where data resides.
Snowflake released updates for Data Clean Rooms, the Python API (v1.13.1), and ML Python (v1.51.0). Key data engineering features include Amazon S3 Tables Iceberg REST catalog integration, Openflow CDC for SQL Server, and adaptive or incremental refresh modes for dynamic tables.
Snowflake released workspace sharing in Declarative Native Apps as a public preview, enabling better collaboration during app development and deployment. The release improves how engineers manage app code directly within data workspaces.
Snowflake has officially launched General Availability for consuming Delta Shares within the Horizon Catalog. Data engineers can now directly access and integrate external open-format Delta Lake data provider shares alongside native governance features without copying or moving raw data assets
Snowflake announced feature updates for Data Clean Rooms alongside general availability for adaptive refresh mode and custom incremental dynamic tables. Data engineers also gain a new agenttype column in QUERYHISTORY views and SQL environment variable support for dbt Projects on Snowflake.
Snowflake introduced skill and plugin sharing in CoCo in public preview. Recent platform updates for data engineering workflows include custom incremental dynamic tables, support for private Git packages in dbt Projects, and altering multiple dynamic tables in a single statement.
Google introduced the borderless lakehouse architecture to connect data across AWS, Databricks, and Snowflake directly to AI agents. It simplifies cross-cloud data pipelines and unifies multi-cloud data engineering workflows without complex migrations.
Zero-copy architectures hide subtle network and compute tradeoffs. Catalog federation and sharing protocols move metadata or signed references, while federated queries pull remote rows per execution. The same work repeats on every execution, because nothing persists.
Snowflake released new updates for Snowflake Data Clean Rooms on July 23, 2026. This follows a series of recent platform enhancements, including the general availability of the SHOW CORTEX BASE MODELS command and support for refreshing multiple dynamic tables within a single ALTER statement.
To share unstructured files securely, you can use Snowflake directory tables and secure views to enforce row-level access control. Instead of managing multiple external stages, a single parent stage points to S3.
Snowflake has introduced several data engineering updates, including preview capabilities for DCM Projects. Notably, "VOLATILE scalar UDFs are supported with dynamic table incremental refresh" is now generally available, alongside general availability for cloning dynamic Iceberg tables.
Snowflake has released updates for its Data Clean Rooms. To ensure these new updates are applied to the user interface, clean room UI users must log out and log back in. New updates are applied to the UI automatically when a user starts a new session after a new update is deployed.
Snowflake Horizon Catalog supports bidirectional Iceberg REST interoperability, enabling full read and write operations on external tables. In contrast, Databricks Unity Catalog only supports inbound federation, restricting external tables to read-only access.
Team USA Bobsled/Skeleton partners with Snowflake CoWork to optimize athletic performance through data collaboration. By analyzing push starts and track curves, the team calculates the exact number of steps before loading.
ObsessionDB introduced Datashares for ClickHouse, enabling zero-copy data sharing via a shared-storage architecture. By decoupling stateless compute from immutable object storage, external or internal teams can query live tables on isolated compute without drifting replicas or cluster resource
Open Semantic Interchange is now Apache Ossie, an open specification for semantic layers and ontologies. It provides a vendor-neutral, machine-readable YAML format to define metrics, dimensions, and business rules across tools.
Snowflake released updates for its Data Clean Rooms. UI users receive automatic updates upon starting a new session, while API users can manage updates manually or automatically. API users can check their environment version by querying SELECT FROM SAMOOHABYSNOWFLAKELOCALDB.ADMIN.VERSION;*
Snowflake ML Jobs is now generally available, allowing data scientists to run distributed Python training and hyperparameter optimization directly within data clean rooms. This enables multiparty model training on combined datasets without raw records ever leaving the participants' accounts.
GitBiased enables data engineers to build customizable GitHub dashboards. It integrates live pull requests, CI checks, and DORA metrics. The tool focuses on providing weekly summaries and calm visibility into engineering workflows without the noise of standard notification systems.
Snowflake Data Clean Rooms now supports managing data offerings and templates directly in Snowsight. This update also allows users to drop or leave failed collaborations to maintain account hygiene.
See all 489 Sharing resources