Sr. Software Engineer - Ingestion Core team

Databricks

San Francisco, California

Deeply understanding what’s in the enterprise data has been a challenge that Databricks has been addressing by providing analytics and machine learning tools. From data warehousing with Databricks SQL to large-scale distributed processing with Spark and advanced ML tools for experimentation and model serving, we empower our customers to gain insights and drive innovation. To enable all of this on Databricks, making data ingestion seamless is crucial. That’s the mission of the Ingestion Core Team: to make the ingestion of all data—structured and unstructured—simple, reliable, and efficient. Simplifying the complex is hard, and that’s where you come in. This role requires building distributed platform systems to incrementally ingest high-volume, petabyte-scale data from diverse sources—including cloud storage (SQS, ADLS, GCS), databases (Oracle, SQL Server, MySQL, Postgres), and file sources (Google Drive, SharePoint)—at high throughput and low cost. The data includes structured formats (JSON, Parquet, CSV) as well as unstructured data (text, images, docs, PPTs, and blobs), all of which land in Delta Lake with schema evolution and change data capture (CDC) capabilities. Join us in ma

Ref. C6WOU

Info
-- / --
Info

Arrastra un PDF aquí, pega una URL o toca para seleccionar un archivo