B.Tech CSE · PES University, Bengaluru
Apache Impala · Data Warehouse · Cloudera
I'm a Software Development Engineer I at Cloudera, on the Data Warehouse team — building and improving large-scale query engines and storage integrations that power modern analytics workloads.
As an Apache Impala contributor, I've worked on Apache Iceberg integration within Impala, CatalogD enhancements, and observability features, enabling performant, transactional analytics on large-scale data lakes.
I'm passionate about distributed systems and big data: query optimization, storage engines, and the systems that make analytical SQL fast at scale.
⚙️ At Work · SDE-I · Data Warehouse @ Cloudera
- Query engine internals in Apache Impala
- CatalogD, metadata & observability improvements
- Performance & correctness at warehouse scale
🔬 Deep Diving
- Distributed query execution & planning
- Open table formats — Iceberg, Hive
- Columnar storage & scan optimization
- Lakehouse architecture patterns
🧭 Learning & Exploring
- Storage–compute separation at scale
- Storage layer integrations for analytics
- Systems that make SQL fast at scale
Happy to chat over coffee ☕ about Impala internals, Iceberg, C++ systems, or careers in big data.
Apache Impala Contributor across the query engine backend, catalogd, and the metadata layer — Building features that improve analytical SQL performance and metrics instrumentation.
I've worked on Apache Iceberg v3 features integration in Impala, spanning metadata, schema evolution, and engine-side support for open table formats.
- 🔧 Backend & catalogd — Query engine, catalog service, and observability
- 📦 Metadata — Iceberg table format support across Impala's catalog and metadata path
- 🏗️ Iceberg v3 — Feature integration for open table formats in Impala
