
Job Overview
Location
Bay Area Office
Job Type
Full-time
Category
Software Engineering
Date Posted
August 8, 2026
Full Job Description
📋 Description
- • Build foundational data systems for AI as a Senior Software Engineer at Granica Inc.
- • Work on the core infrastructure behind Crunch, Granica’s continuous optimization product for enterprise lakehouse data, including systems for metadata management, table maintenance, file layout optimization, distributed compute, and workload-aware data reorganization across petabyte- and exabyte-scale environments.
- • Collaborate with Granica Research, led by Prof. Andrea Montanari at Stanford, to translate ideas from information theory, probabilistic modeling, compression, and learning efficiency into production systems.
- • Develop table-maintenance infrastructure for lakehouse formats such as Apache Iceberg, Delta Lake, and Apache Hudi.
- • Optimize file layout, clustering, compaction, data skipping, indexing, and metadata pruning.
- • Improve performance and cost efficiency across Spark, Flink, Trino, Presto, Databricks, Snowflake-adjacent, and cloud object storage environments.
- • Work with columnar formats such as Parquet and ORC, including encoding, compression, layout, and read-path optimization.
- • Build adaptive engines that learn from access patterns and workloads to reorganize data automatically.
- • Develop distributed compute pipelines that scale predictively and remain reliable under failure.
- • Debug performance bottlenecks across storage, metadata, query execution, network, and compute layers.
- • Implement research-driven algorithms in compression, representation, layout optimization, and data efficiency.
- • Contribute to open-source or publish research when appropriate.
🎯 Requirements
- • Strong engineering depth in distributed systems, storage systems, databases, or data infrastructure.
- • Production experience with modern data lake or lakehouse technologies such as Spark, Iceberg, Delta Lake, Hudi, Trino, Presto, Flink, Hive Metastore, Unity Catalog, or similar systems.
- • Hands-on experience with columnar formats such as Parquet or ORC.
- • Understanding of metadata-driven architectures, table formats, query planning, and physical data layout.
- • Strong programming skills in Rust, Go, C++, or similar systems-oriented languages.
🏖️ Benefits
- • Competitive salary, meaningful equity, and performance bonus for top performers.
- • 401(k) with company match, comprehensive health coverage, and unlimited PTO.
- • Daily catered meals in our Mountain View office.
- • Support for research, publication, and conference participation.
📍 Location
- • Bay Area Office
- • Mountain View, CA
- • On-site
Skills & Technologies
See exactly how your profile matches this role — strengths, skill gaps, and what to do about them.
About Granica Inc.
Granica builds an AI efficiency platform that compresses and secures petabyte-scale training data for cloud object stores. Its byte-granular deduplication and privacy filtering shrink S3 and GCS footprints, cutting storage and transfer costs while boosting downstream model accuracy. Designed for data scientists and MLOps teams, the service deploys as a transparent sidecar proxy, enforcing differential privacy and access policies without code changes. Founded in 2022 and headquartered in Palo Alto, the company targets enterprises running computer-vision and NLP workloads that need cheaper, safer data pipelines.
Subscribe to the weekly newsletter for similar remote roles and curated hiring updates.
Newsletter
Weekly remote jobs and featured talent.
No spam. Only curated remote roles and product updates. You can unsubscribe anytime.
Similar Opportunities

Horizon3.ai, Inc.
2 months ago

Horizon3.ai, Inc.
2 months ago

Primer Technologies, Inc.
28 days ago
22 days ago
