Software Engineer, Database Internals Intern

PLAID

Java

K8s

Redis

About this role

PLAID is developing an OLAP database that can return queries in under a second on more than a petabyte of user attribute and behavior data without pre-aggregation. The role covers work from high-level analytical algorithms through the design of columnar storage, query engines, and distributed-computing algorithms.

Responsibilities include developing ideas and implementing new capabilities for analytics functions, expanding infrastructure to support rapidly growing data volumes, removing bottlenecks by optimizing query engines and data formats, and integrating products that use the database.

The project addresses query-performance and cost challenges in the analysis infrastructure supporting KARTE, which collects large volumes of end-user behavior data. It uses knowledge of open-source OLAP databases and involves building a large-scale distributed system using object storage, columnar storage, and distributed query-engine architecture. The system is used by products supporting decision-making for hundreds of companies.

The named technology stack includes Java, Kubernetes, GCS, Spanner, Bigtable, Redis, BigQuery, Apache Arrow, SIMD, Columnar, and Roaring Bitmap.

The position is based at PLAID’s headquarters in Ginza, Tokyo. Work is hybrid, combining office attendance and remote work. Depending on company and team circumstances, a certain number of office days may be requested. The employment type is an internship.

This is an AI-generated summary of the employer's original posting — details can be incomplete, out of date or simply wrong. Always confirm everything on the official posting before applying.