What the meeting covers

Scaling the Data Lake for the AI Era will take place on 3 September from 17:30 to 19:30 Pacific Daylight Time in San Francisco. The event is intended for data engineers, machine-learning platform engineers and architects working with large analytics systems. Registration requires approval, and the organiser's page does not list a public price. [1 · Event page]

The central topic is petabyte-scale analytics in the context of TRM Labs. The discussion is intended to show how a data lake changes under AI workloads: it must simultaneously provide scale, data quality, governance, query performance and access to the datasets used by models. [1 · Event page]

Practical value

For an AI product, data infrastructure often becomes a constraint before model capabilities do. Difficulties arise around data updates and provenance, access control, reproducible samples and query execution costs. At petabyte scale, small architectural mistakes turn into substantial delays and expenses. [1 · Event page]

The event is useful for teams moving machine-learning workloads from experimentation into production. Reporting should avoid promising specific talks beyond the organiser's information: the source document identifies the general subject and audience, but does not contain a detailed programme or a confirmed speaker line-up. [1 · Event page]

Sources

  1. Event page — 3 September 2026, 17:30–19:30 Pacific Daylight Time; approval required