Skip to content

Details

Join us for an Apache Hudi Open Source Meetup, hosted by Onehouse at COWRKS Ecoworld, Bangalore! This edition focuses entirely on the next-generation Apache Hudi 1.2 architecture โ€” exploring how its major structural overhauls power modern AI/ML workloads. Ideal for data engineers, lakehouse architects, and AI/ML practitioners.

๐Ÿ‘‰ Registration: Open โ€” RSVP to secure your spot: https://forms.gle/UboRNekS8FQjcZjCA
๐Ÿ“… Date & Time: Wednesday, July 29, 2026 ยท 4:00 โ€“ 7:00 PM IST
๐Ÿ“ Venue: COWRKS Ecoworld, 10th Floor, Building 4D, ECOWORLD, Outer Ring Rd, Devarabisanahalli, Bellandur, Bengaluru, Karnataka 560103 โ€”[ https://maps.app.goo.gl/Q7tB1M6bYCtQBU529](https://maps.app.goo.gl/Q7tB1M6bYCtQBU529)

โฐ Agenda

  • 4:00 โ€“ 4:30 | Welcome & Introductions โ€” kickoff, opening remarks, community announcements
  • 4:30 โ€“ 6:45 | Hudi Tech Talk Sessions โ€” titles & speakers TBA
  • 6:45 โ€“ 7:00 | Open Q&A, Networking & Wrap-up

๐Ÿ’ก What you'll learn

  • Multimodal Data as a First-Class Citizen: How Hudi 1.2.0's new native logical types โ€” VECTOR, VARIANT, and BLOB โ€” plus the read_blob() and hudi_vector_search() Spark SQL functions bring embeddings, semi-structured payloads, and binary objects (images, audio) into the lakehouse alongside structured data, backed by RFC-99/100/102.
  • Optimized Storage for AI/ML workloads: Native Lance columnar format integration for AI/ML access patterns, and how Hudi's indexing primitives (record-level, secondary, expression indexes) now extend to accelerate vector search and point lookups โ€” no separate vector DB/document store required.
  • Streaming at Scale: Full Record Level Index (RLI) support landing on Flink โ€” global indexing, efficient record routing, and dynamic bucket scaling โ€” plus the new FLIP-27 Source V2 with resumable split assignment and stronger pushdown capabilities.

๐Ÿ‘ฅ Who should attend

  • Data, Platform, and Cloud Engineers building real-time lakehouses.
  • ML/AI engineers building high-throughput feature stores, RAG architectures, and vector-search layers.
  • Engineers working with Apache Spark, Flink, Trino, and cloud-native computing fabrics.
  • Open-source contributors interested in advanced transactional table formats.

๐Ÿ“ RSVP: Entry is free, but seats are limited by venue capacity โ€” please register in advance so we can pre-generate your visitor pass at the tech park security desk.

See you there โ€” let's build the AI-native lakehouse together!

Related topics

Big Data
Data Management
Database Professionals
Open Source
Software Development

You may also like