<?xml version="1.0" encoding="UTF-8"?>
<source>
  <jobs>
    <job>
      <externalid>0c9961cb-26b</externalid>
      <title>Senior Data Engineer</title>
      <description><![CDATA[<p>We are seeking a Senior Data Engineer to join our team. As a Senior Data Engineer, you will be responsible for designing and implementing data integration solutions to ingest, transform, and integrate structured and unstructured data from various sources into a scalable data architecture platform.</p>
<p><strong>Job Responsibilities:</strong></p>
<ul>
<li>Design and implement data integration solutions to ingest, transform, and integrate structured and unstructured data from various sources into a scalable data architecture platform</li>
<li>Develop, construct, and maintain large-scale data processing systems that collect data from variety of structured and unstructured data sources</li>
<li>Store data in a scale-out data lake and prepare the data techniques in preparation for the data science data exploration and analytic modeling</li>
<li>Ingest, transform, and blend/integrate structured and unstructured data and deliver the data to a MDM, data warehouse, and/or data lake</li>
<li>Include both technical processes and business logic to transform data from disparate sources into cohesive meaningful data with quality, governance, and compliance considerations</li>
<li>Use technical and business processes to combine data from disparate sources into meaningful and valuable information</li>
<li>Design, build, and manage the information and big data infrastructure that helps analyze and process data that the organization requires and optimize systems to perform smoothly</li>
<li>Prepare the data for the data scientist&#39;s exploration and discovery process</li>
<li>Evaluate, compare, and improve the different approaches including design patterns innovation, data lifecycle design, data ontology alignment, annotated datasets, and elastic search approach</li>
</ul>
<p><strong>Job Requirements:</strong></p>
<ul>
<li>University degree in Computer Science, Engineering, and/or a technically oriented field</li>
<li>Over 5 years of experience with Flink/Spark, Databricks</li>
<li>Over 5 years of experience with Azure (DP200 and/or DP201, DP203 certification acts as a plus)</li>
<li>3+ years&#39; experience mining data as a data analyst</li>
<li>3+ years&#39; using programming languages such as Java, DAX, MDX, SQL, Python</li>
<li>3+ years of experience with PowerBI, Tableau, or Qlik</li>
<li>3+ years&#39; experience in the life insurance domain is required</li>
<li>Passionate about analytics, machine learning technology, and applications, and eager to learn</li>
<li>English communication</li>
</ul>
<p><strong>Knowledge and Skill:</strong></p>
<ul>
<li>Good knowledge of Big Data technologies, such as Spark, Hadoop/MapReduce</li>
<li>Good knowledge of Azure services like Storage Account, Azure DataBricks, etc.</li>
<li>Good knowledge of SQL and excellent coding skills</li>
<li>Strong knowledge of data modeling and data mining</li>
<li>Self-Development, communication, problem-Solving Skills</li>
<li>Open-minded, multi-tasking, teamwork, flexible, and interest to learn new things</li>
</ul>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>senior</experiencelevel>
      <workarrangement></workarrangement>
      <salaryrange></salaryrange>
      <skills>Big Data, Spark, Hadoop/MapReduce, Azure, Databricks, Flink, Java, Python, SQL, PowerBI, Tableau, Qlik, Machine Learning, Analytics</skills>
      <category>IT</category>
      <industry>Finance</industry>
      <employername>Prudential</employername>
      <employerlogo>https://logos.yubhub.co/prudential.com.png</employerlogo>
      <employerdescription>Prudential is a financial services company that provides protection and investment solutions to individuals and businesses.</employerdescription>
      <employerwebsite>https://prudential.com</employerwebsite>
      <compensationcurrency></compensationcurrency>
      <compensationmin></compensationmin>
      <compensationmax></compensationmax>
      <compensationinterval></compensationinterval>
      <applyto>https://prudential.wd3.myworkdayjobs.com/en-US/prudential/job/Thnh-ph-H-Ch-Minh/Senior-Data-Engineer--K-s-D-liu-Cp-cao-_26030503?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>Thành phố Hồ Chí Minh</location>
      <city>Thành phố Hồ Chí Minh</city>
      <state></state>
      <postalcode></postalcode>
      <country></country>
      <postedate>2026-09-04</postedate>
    </job>
    <job>
      <externalid>26d8e8b8-406</externalid>
      <title>Principal Engineer, Big Data Platform</title>
      <description><![CDATA[<p>We&#39;re looking for a principal software engineer to lead the next generation of data infrastructure at Pinterest, which powers mission-critical big data and AI applications.</p>
<p>You’ll be working on some of the most exciting big data and AI open-source technologies (Flink, Spark, Kubernetes, etc.), at the scale of exabytes of data to help Pinners discover and do what they love.</p>
<p>Responsibilities:</p>
<ul>
<li>Lead the strategy and technical direction of Pinterest’s data infrastructure for big data and AI applications</li>
<li>Build and scale data infra frameworks and infrastructure to process petabytes-scale datasets, including compute engines, job management, resource management, scheduling, and remote shuffling</li>
<li>Work with internal customers on critical business use cases that rely on big data</li>
<li>Provide thought leadership to the entire company on how data should be processed and stored more reliably, quickly, and efficiently at scale</li>
<li>Contribute to the team’s technical vision and long-term roadmap</li>
</ul>
<p>Requirements:</p>
<ul>
<li>12+ years of industry experience with a proven track record of technical excellence</li>
<li>8+ years of experience building and supporting large scalable Kubernetes or big data platforms</li>
<li>Deep knowledge of big data / ML technologies (e.g., Flink, Spark, Presto, Kubernetes, Ray, PyTorch/TensorFlow)</li>
<li>Proficiency in one or more programming languages (Java, Go, Scala, Python)</li>
<li>Experience with Kubernetes and AWS technologies</li>
<li>Exceptional collaboration skills with cross-functional partners, with the ability to navigate ambiguity, make trade-offs, and keep stakeholders aligned on priorities and progress.</li>
<li>Bachelor’s degree in Computer Science, a related technical field, or equivalent experience.</li>
</ul>
<p>Benefits:</p>
<ul>
<li>Salary range: $242,634-$499,541 USD</li>
<li>Eligible for equity</li>
</ul>
<p>In-office requirement: This role will need to be in the office for in-person collaboration 1-2 times/quarter.</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>senior</experiencelevel>
      <workarrangement>hybrid</workarrangement>
      <salaryrange>$242,634-$499,541 USD</salaryrange>
      <skills>Flink, Spark, Kubernetes, Java, Go, Scala, Python, PyTorch/TensorFlow, Ray, Presto</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>Pinterest</employername>
      <employerlogo>https://logos.yubhub.co/pinterest.com.png</employerlogo>
      <employerdescription>Pinterest is a platform where people find creative ideas, dream about new possibilities, and plan for memories.</employerdescription>
      <employerwebsite>https://pinterest.com</employerwebsite>
      <compensationcurrency>USD</compensationcurrency>
      <compensationmin>242634</compensationmin>
      <compensationmax>499541</compensationmax>
      <compensationinterval>yearly</compensationinterval>
      <applyto>https://job-boards.greenhouse.io/pinterest/jobs/7683981?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>San Francisco, CA</location>
      <city>San Francisco</city>
      <state>CA</state>
      <postalcode></postalcode>
      <country>US</country>
      <postedate>2026-09-03</postedate>
    </job>
    <job>
      <externalid>23da71a5-011</externalid>
      <title>Vice President - LME Data Platform</title>
      <description><![CDATA[<p>We are seeking a Vice President to lead the architecture, design, and technical strategy for our open-source data platform. You will make the end-to-end technical vision , from ingestion to analytics , while mentoring a team of data platform engineers. This is a hands-on leadership role responsible for designing blueprints and writing critical-path code.</p>
<p><strong>Responsibilities</strong></p>
<p><strong>Architecture &amp; Strategy</strong></p>
<ul>
<li>Define the overall data platform architecture on OpenShift, covering ingestion (Kafka/Redpanda), compute (Spark, Flink, Trino), storage (MinIO, Iceberg), catalog (Polaris/Gravitino), orchestration (Airflow/Dagster), and serving (StarRocks/ClickHouse).</li>
<li>Design the Data Lakehouse architecture: open table formats (Iceberg, Delta Lake), catalog federation, multi-engine interoperability.</li>
<li>Architect the real-time data path: Kafka topic design, schema registry (Apicurio/Confluent), stream processing (Flink/Spark Structured Streaming).</li>
<li>Design the analytical serving layer: OLAP engine selection (StarRocks/ClickHouse/Doris), materialized view strategy, query federation with Trino.</li>
<li>Design multi-tenancy, data security (Ranger/OpenPolicyAgent), RBAC, and governance (DataHub/Atlas) across all platform layers.</li>
<li>Set technical standards for Helm chart design, CI/CD pipelines, and GitOps (ArgoCD/Flux) workflows.</li>
</ul>
<p><strong>Technical Leadership</strong></p>
<ul>
<li>Lead a team of 3–6 data platform engineers; run code reviews, design reviews, and sprint planning.</li>
<li>Establish engineering best practices: testing, observability (OpenTelemetry, Prometheus, Grafana, Loki), incident response, runbooks.</li>
<li>Partner with data engineering, analytics, and business teams to translate their needs into platform capabilities.</li>
<li>Define platform SLOs/SLIs across freshness, latency, availability, and durability; drive the on-call rotation and incident post-mortems.</li>
</ul>
<p><strong>Hands-on Engineering</strong></p>
<ul>
<li>Build and maintain the core OpenShift infrastructure: operators, Helm charts, namespaces, RBAC, network policies.</li>
<li>Develop Spark and Flink job frameworks, tuning guides, workload scheduling (YuniKorn/Volcano).</li>
<li>Implement the data mesh or data product model: domain ownership, self-serve data infrastructure, federated governance.</li>
<li>Performance-optimize across the stack: Kafka throughput, Spark shuffle, Iceberg compaction, Trino query planning, OLAP cold reads.</li>
</ul>
<p><strong>Required Skills &amp; Experience</strong></p>
<ul>
<li>10+ years in data/platform engineering, with 3+ years as architect or tech lead.</li>
<li>Deep Kubernetes/OpenShift: operators, Helm, CRDs, admission webhooks, SCC, multi-tenancy at production scale.</li>
<li>Expert-level streaming: Kafka/Redpanda internals (partitioning, consumer groups, exactly-once semantics), schema registry, Kafka Connect.</li>
<li>Expert-level batch &amp; streaming compute: Spark internals (Catalyst, Tungsten, AQE, shuffle, dynamic allocation), Flink or Spark Structured Streaming.</li>
<li>Strong open table format knowledge: Iceberg and/or Delta Lake , table format internals, metadata evolution, partitioning transforms, maintenance, catalog integration (Polaris, Gravitino, Nessie, or Unity Catalog).</li>
<li>Hands-on with federated query engines: Trino (coordinator/worker architecture, connector ecosystem, fault-tolerant execution mode).</li>
<li>Solid OLAP knowledge: StarRocks, ClickHouse, or Apache Doris , primary key models, materialized views, query optimization.</li>
<li>Production MinIO/S3-compatible storage at scale: erasure coding, multi-site replication, IAM, tiering.</li>
<li>Pipeline orchestration: Airflow (DAG design, sensors, dynamic task mapping) or Dagster.</li>
<li>Infrastructure-as-code (Terraform/Pulumi/Crossplane) and GitOps (ArgoCD/Flux).</li>
<li>Fluent in at least one JVM language (Scala/Java) and Python.</li>
</ul>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>executive</experiencelevel>
      <workarrangement></workarrangement>
      <salaryrange></salaryrange>
      <skills>Kubernetes, OpenShift, Kafka, Redpanda, Spark, Flink, Trino, Iceberg, Delta Lake, Polaris, Gravitino, StarRocks, ClickHouse, Apache Doris, MinIO, Airflow, Dagster, Terraform, Pulumi, Crossplane, ArgoCD, Flux, Scala, Java, Python, dbt, DataHub, Apache Atlas, Apache Ranger, OpenPolicyAgent</skills>
      <category>IT</category>
      <industry>Technology</industry>
      <employername>ITD SZ (Hong Kong Exchanges and Clearing Limited&apos;s technology subsidiary)</employername>
      <employerlogo>https://logos.yubhub.co/hkex.com.png</employerlogo>
      <employerdescription>ITD SZ is a technology subsidiary of Hong Kong Exchanges and Clearing Limited, providing computer software, hardware, information systems, cloud storage, and other technology services.</employerdescription>
      <employerwebsite>https://hkex.com</employerwebsite>
      <compensationcurrency></compensationcurrency>
      <compensationmin></compensationmin>
      <compensationmax></compensationmax>
      <compensationinterval></compensationinterval>
      <applyto>https://hkex.wd3.myworkdayjobs.com/en-US/HKEXCareerPage/job/CN-Shenzhen-HyQ/Vice-President---LME-Market-Data_R004150?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>Shenzhen</location>
      <city>Shenzhen</city>
      <state></state>
      <postalcode></postalcode>
      <country></country>
      <postedate>2026-08-30</postedate>
    </job>
    <job>
      <externalid>cf02f854-3e7</externalid>
      <title>Staff Site Reliability Engineer, Ads</title>
      <description><![CDATA[<p>Reddit is a community of communities. It&#39;s built on shared interests, passion, and trust, and is home to the most open and authentic conversations on the internet.</p>
<p>The Ads organization powers Reddit&#39;s advertising platform, enabling advertisers to reach highly engaged communities while helping Reddit grow its business. The reliability of our Ads systems directly impacts advertiser success, revenue generation, and user experience.</p>
<p>We&#39;re looking for a Staff Site Reliability Engineer who will define and provide technical leadership for reliability initiatives across the Ads organization and help shape the future of Ads infrastructure at Reddit.</p>
<p>Responsibilities:</p>
<ul>
<li>Lead reliability initiatives across multiple Ads domains including ad serving, auctions, targeting, reporting, measurement, and billing.</li>
<li>Partner with engineering leadership to develop a roadmap to improve reliability, scalability, operational excellence, and engineering efficiency across the Ads organization.</li>
<li>Design and build platforms, tooling, and automation that improve reliability and developer productivity at scale.</li>
<li>Drive architecture reviews and influence technical decisions impacting critical revenue-generating systems.</li>
<li>Participate in on-call rotations, lead complex incident investigations and coordinate cross-functional response efforts during major production events.</li>
<li>Identify systemic reliability risks and drive long-term solutions that improve platform resilience.</li>
<li>Establish reliability metrics around advertiser-critical user journeys such as campaign creation, ad delivery, auction participation, reporting, attribution, and billing.</li>
<li>Mentor engineers and provide technical leadership across multiple teams.</li>
<li>Influence roadmap planning and ensure reliability considerations are incorporated into product and infrastructure investments.</li>
</ul>
<p>Requirements:</p>
<ul>
<li>8+ years of experience in Site Reliability Engineering, Infrastructure Engineering, or related roles operating large scale distributed systems.</li>
<li>Strong experience evolving high traffic, user-facing production environments.</li>
<li>Strong cross-functional collaborations skills to lead and influence projects driving operational excellence.</li>
<li>Deep expertise in modern distributed systems, scale engineering, and cloud-native architectures.</li>
<li>Experience designing highly-available systems with strong operational and reliability practices.</li>
<li>Strong software engineering skills in languages like general-purpose backend languages like Go.</li>
<li>Strong understanding of observability systems including metrics, logging, tracing, and alerting.</li>
<li>Experience improving reliability through SLOs, automation, incident management, and performance optimization.</li>
<li>Demonstrated ability to troubleshoot complex issues across a modern distributed system stack.</li>
<li>Strong collaboration and communication skills with the ability to influence technical direction across teams.</li>
</ul>
<p>Nice to Have:</p>
<ul>
<li>Experience supporting advertising technology platforms or other large-scale revenue-critical systems.</li>
<li>Deep understanding of reliability challenges associated with ad-serving, real-time auctions, budget pacing, campaign delivery, measurement, attribution, or billing systems.</li>
<li>Experience operating high-QPS, low-latency services where latency directly impacts business outcomes.</li>
<li>Experience establishing reliability programs that deliver meaningful, measurable business outcomes.</li>
<li>Experience with Kubernetes, cloud infrastructure, and large-scale distributed systems.</li>
<li>Familiarity with Kafka, ClickHouse, Spark, Flink, BigQuery, or similar large-scale data platforms.</li>
<li>Experience partnering with Product, Data Science, and Ads Engineering organizations.</li>
<li>Experience supporting machine learning inference or recommendation systems at scale.</li>
</ul>
<p>Benefits:</p>
<ul>
<li>Comprehensive Health benefits</li>
<li>401k Matching</li>
<li>Workspace benefits for your home office</li>
<li>Personal &amp; Professional development funds</li>
<li>Family Planning Support</li>
<li>Flexible Vacation &amp; Reddit Global Days Off</li>
<li>4+ months paid Parental Leave</li>
<li>Paid Volunteer time off</li>
</ul>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>staff</experiencelevel>
      <workarrangement>remote</workarrangement>
      <salaryrange>$217,000-$303,900 USD</salaryrange>
      <skills>Site Reliability Engineering, Infrastructure Engineering, distributed systems, cloud-native architectures, Go, observability systems, SLOs, automation, incident management, performance optimization, advertising technology platforms, ad-serving, real-time auctions, budget pacing, campaign delivery, measurement, attribution, billing systems, Kubernetes, cloud infrastructure, large-scale distributed systems, Kafka, ClickHouse, Spark, Flink, BigQuery, machine learning inference, recommendation systems</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>Reddit</employername>
      <employerlogo>https://logos.yubhub.co/redditinc.com.png</employerlogo>
      <employerdescription>Reddit is a community of communities with 100,000+ active communities and approximately 130 million daily active unique visitors.</employerdescription>
      <employerwebsite>https://www.redditinc.com</employerwebsite>
      <compensationcurrency>USD</compensationcurrency>
      <compensationmin>217000</compensationmin>
      <compensationmax>303900</compensationmax>
      <compensationinterval>yearly</compensationinterval>
      <applyto>https://job-boards.greenhouse.io/reddit/jobs/8090680?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>San Francisco, CA</location>
      <city>San Francisco</city>
      <state>CA</state>
      <postalcode></postalcode>
      <country>US</country>
      <postedate>2026-08-29</postedate>
    </job>
    <job>
      <externalid>06dfd0b0-8d8</externalid>
      <title>Member of Technical Staff - Data Flywheel Infra, Frontier Models</title>
      <description><![CDATA[<p>We are looking for a Data Flywheel Infrastructure Engineer to build the infrastructure that continuously turns 1P data, 3P data, model signals, evaluation results, and synthetic data into high-quality training data for frontier LLM and multimodal models.</p>
<p>This role owns the systems connecting Data Acquisition → Governance &amp; Compliance → Curation → Training → Evaluation → Failure Mining → Data Improvement.</p>
<p>A critical part of the role is enabling aggressive data iteration while ensuring that every dataset is secure, policy-compliant, rights-aware, traceable, and auditable.</p>
<p>Starting January 26, 2026, MAI employees are expected to work from a designated Microsoft office at least four days a week if they live within 50 miles (U.S.) or 25 miles (non-U.S., country-specific) of that location.</p>
<p>This role is part of Microsoft AI’s Superintelligence Team, created to push the boundaries of AI toward Humanist Superintelligence,ultra-capable systems that remain controllable, safety-aligned, and anchored to human values.</p>
<p><strong>Responsibilities</strong></p>
<ul>
<li>Build 1P &amp; 3P Data Flywheel Infrastructure: scalable systems for ingesting, processing, curating, versioning, and serving first-party and third-party data for pre-training and post-training.</li>
<li>Own Data Governance, Security &amp; Compliance Infrastructure: build governance and policy enforcement directly into the data platform.</li>
<li>Build Policy-Aware Data Acquisition &amp; Curation Systems: develop automated pipelines for 1P and 3P data ingestion, classification, filtering, deduplication, quality scoring, semantic enrichment, and dataset construction.</li>
<li>Build Evaluation-to-Data Feedback Loops: convert model evaluations and real-world failure signals into actionable data tasks.</li>
<li>Build Synthetic &amp; AI-Native Data Pipelines: use LLMs, VLMs, and Agents to automate data generation, labeling, filtering, quality validation, enrichment, and transformation.</li>
<li>Build Data Quality, Attribution &amp; Observability: develop metrics and infrastructure to measure dataset quality, coverage, diversity, contamination, duplication, policy compliance, and contribution to model capability improvements.</li>
</ul>
<p><strong>Qualifications</strong></p>
<ul>
<li>Master’s Degree in Computer Science, Math, Software Engineering, Computer Engineering, or related field AND 4+ years experience in business analytics, data science, software development, data modeling, or data engineering OR Bachelor’s Degree in Computer Science, Math, Software Engineering, Computer Engineering, or related field AND 6+ years experience in business analytics, data science, software development, data modeling, or data engineering OR equivalent experience.</li>
<li>Software Engineering experience using Python, SQL, Spark/Flink/Ray.</li>
</ul>
<p><strong>Preferred Qualifications</strong></p>
<ul>
<li>Experience managing third-party datasets, data partnerships, licensed content, or externally sourced data with complex contractual and usage restrictions.</li>
<li>Experience building privacy- and security-aware systems for first-party product or user data, including isolation, access controls, retention/deletion, and purpose limitation.</li>
<li>Experience with data clean rooms, privacy-preserving processing, de-identification, confidential computing, or secure data collaboration.</li>
<li>Experience building evaluation → failure mining → data generation → training feedback loops.</li>
<li>Experience with synthetic data, model graders, reward signals, hard-example mining, active learning, or data-mixture optimization.</li>
<li>Experience with multimodal or agentic datasets including text, image, video, audio, web, GUI, tool-use, or interaction trajectories.</li>
<li>Understanding of Modern LLM training workflows including Pre-training, SFT, RL/post-training, evaluation, and synthetic data.</li>
<li>Strong understanding of data governance, security, privacy, provenance, access control, and data lifecycle management.</li>
</ul>
<p><strong>Salary Information</strong></p>
<p>Data Engineering IC5 – The typical base pay range for this role across the U.S. is USD $142,800 – $274,800 per year. Data Engineering IC6 – The typical base pay range for this role across the U.S. is USD $165,600 – $296,400 per year.</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>senior</experiencelevel>
      <workarrangement>hybrid</workarrangement>
      <salaryrange>USD $142,800 – $274,800 per year</salaryrange>
      <skills>Python, SQL, Spark, Flink, Ray, Data Engineering, Data Governance, Security, Compliance, third-party datasets, data partnerships, licensed content, externally sourced data, privacy-preserving processing, de-identification, confidential computing, secure data collaboration, synthetic data, model graders, reward signals, hard-example mining, active learning, data-mixture optimization</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>Microsoft AI</employername>
      <employerlogo>https://logos.yubhub.co/microsoft.ai.png</employerlogo>
      <employerdescription>Microsoft AI is a startup-like team inside Microsoft, focused on pushing the boundaries of AI toward Humanist Superintelligence.</employerdescription>
      <employerwebsite>https://microsoft.ai</employerwebsite>
      <compensationcurrency>USD</compensationcurrency>
      <compensationmin>142800</compensationmin>
      <compensationmax>274800</compensationmax>
      <compensationinterval>yearly</compensationinterval>
      <applyto>https://microsoft.ai/job/member-of-technical-staff-data-flywheel-infra-frontier-models-2/?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location></location>
      <city></city>
      <state></state>
      <postalcode></postalcode>
      <country></country>
      <postedate>2026-08-28</postedate>
    </job>
    <job>
      <externalid>48627a05-84e</externalid>
      <title>Member of Technical Staff - Data Flywheel Infra, Frontier Models</title>
      <description><![CDATA[<p>We are looking for a Data Flywheel Infrastructure Engineer to build the infrastructure that continuously turns 1P data, 3P data, model signals, evaluation results, and synthetic data into high-quality training data for frontier LLM and multimodal models.</p>
<p>This role owns the systems connecting Data Acquisition → Governance &amp; Compliance → Curation → Training → Evaluation → Failure Mining → Data Improvement.</p>
<p>A critical part of the role is enabling aggressive data iteration while ensuring that every dataset is secure, policy-compliant, rights-aware, traceable, and auditable.</p>
<p>The MAIST is a startup-like team inside Microsoft AI, created to push the boundaries of AI toward Humanist Superintelligence,ultra-capable systems that remain controllable, safety-aligned, and anchored to human values.</p>
<p>Responsibilities:</p>
<ul>
<li>Build 1P &amp; 3P Data Flywheel Infrastructure</li>
<li>Own Data Governance, Security &amp; Compliance Infrastructure</li>
<li>Build Policy-Aware Data Acquisition &amp; Curation Systems</li>
<li>Build Evaluation-to-Data Feedback Loops</li>
<li>Build Synthetic &amp; AI-Native Data Pipelines</li>
<li>Build Data Quality, Attribution &amp; Observability</li>
</ul>
<p>Qualifications:</p>
<ul>
<li>Master&#39;s Degree in Computer Science, Math, Software Engineering, Computer Engineering, or related field AND 3+ years experience in business analytics, data science, software development, data modeling, or data engineering</li>
<li>Bachelor&#39;s Degree in Computer Science, Math, Software Engineering, Computer Engineering, or related field AND 4+ years experience in business analytics, data science, software development, data modeling, or data engineering</li>
<li>Software Engineering experience using Python, SQL, Spark/Flink/Ray</li>
</ul>
<p>Preferred Qualifications:</p>
<ul>
<li>Experience managing third-party datasets, data partnerships, licensed content, or externally sourced data with complex contractual and usage restrictions.</li>
<li>Experience building privacy- and security-aware systems for first-party product or user data, including isolation, access controls, retention/deletion, and purpose limitation.</li>
<li>Experience with data clean rooms, privacy-preserving processing, de-identification, confidential computing, or secure data collaboration.</li>
<li>Experience building evaluation → failure mining → data generation → training feedback loops.</li>
<li>Experience with synthetic data, model graders, reward signals, hard-example mining, active learning, or data-mixture optimization.</li>
<li>Experience with multimodal or agentic datasets including text, image, video, audio, web, GUI, tool-use, or interaction trajectories.</li>
<li>Understanding of Modern LLM training workflows including Pre-training, SFT, RL/post-training, evaluation, and synthetic data.</li>
<li>Strong understanding of data governance, security, privacy, provenance, access control, and data lifecycle management.</li>
</ul>
<p>Data Engineering IC4 – The typical base pay range for this role across the U.S. is USD $119,800 – $234,700 per year.</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>senior</experiencelevel>
      <workarrangement>hybrid</workarrangement>
      <salaryrange>USD $119,800 – $234,700 per year</salaryrange>
      <skills>Python, SQL, Spark, Flink, Ray, Data Engineering, Data Governance, Security, Compliance, third-party datasets, data partnerships, licensed content, externally sourced data, privacy-preserving processing, de-identification, confidential computing, secure data collaboration, synthetic data, model graders, reward signals, hard-example mining, active learning, data-mixture optimization, multimodal datasets, agentic datasets, LLM training workflows</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>Microsoft AI</employername>
      <employerlogo>https://logos.yubhub.co/microsoft.ai.png</employerlogo>
      <employerdescription>Microsoft AI is a startup-like team inside Microsoft, created to push the boundaries of AI toward Humanist Superintelligence.</employerdescription>
      <employerwebsite>https://microsoft.ai</employerwebsite>
      <compensationcurrency>USD</compensationcurrency>
      <compensationmin>119800</compensationmin>
      <compensationmax>234700</compensationmax>
      <compensationinterval>yearly</compensationinterval>
      <applyto>https://microsoft.ai/job/member-of-technical-staff-data-flywheel-infra-frontier-models/?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>Multiple Locations</location>
      <city>Multiple Locations</city>
      <state></state>
      <postalcode></postalcode>
      <country>US</country>
      <postedate>2026-08-27</postedate>
    </job>
    <job>
      <externalid>5e39092f-88c</externalid>
      <title>Sr. Staff Software Engineer, Big Data Platform</title>
      <description><![CDATA[<p>About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product.</p>
<p>We&#39;re looking for a senior staff software engineer to lead the next generation of data infrastructure at Pinterest which powers mission critical big data and AI applications. You’ll be working on some of the most exciting big data and AI open source technologies (Flink, Spark, Kubernetes, etc.), at the scale of exabytes of data to help Pinners discover and do what they love.</p>
<p>Responsibilities:</p>
<ul>
<li>Lead the strategy and technical direction of Pinterest’s data infrastructure for big data and AI applications</li>
<li>Build and scale data infra frameworks and infrastructure to process petabytes-scale datasets, including compute engines, job management, resource management, scheduling and remote shuffling</li>
<li>Work with internal customers on critical business use cases that rely on big data</li>
<li>Provide thought leadership to the entire company on how data should be processed and stored more reliably, quickly and efficiently at scale</li>
<li>Contribute to the team’s technical vision and long-term roadmap</li>
</ul>
<p>Requirements:</p>
<ul>
<li>10+ years of industry experience with a proven track record of technical excellence</li>
<li>5+ years of experience of building and support large scalable Kubernetes or big data platform</li>
<li>Deep knowledge of big data / ML technologies (e.g. Flink, Spark, Presto, Kubernetes, Ray, PyTorch/TensorFlow)</li>
<li>Proficiency in one or more programming languages (Java, Go, Scala, Python)</li>
<li>Experiences in Kubernetes and AWS technologies</li>
<li>Exceptional collaboration skills with cross-functional partners, with the ability to navigate ambiguity, make tradeoffs, and keep stakeholders aligned on priorities and progress.</li>
<li>Bachelor’s degree in Computer Science, a related technical field, or equivalent experience.</li>
</ul>
<p>In-Office Requirement Statement: We let the type of work you do guide the collaboration style. That means we&#39;re not always working in an office, but we continue to gather for key moments of collaboration and connection.</p>
<ul>
<li>This role will need to be in the office for in-person collaboration 1-2 times/quarter and therefore can be situated anywhere in the country.</li>
</ul>
<p>Relocation Statement: This position is not eligible for relocation assistance.</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>senior</experiencelevel>
      <workarrangement>hybrid</workarrangement>
      <salaryrange></salaryrange>
      <skills>Kubernetes, big data, AI, Flink, Spark, Presto, Ray, PyTorch/TensorFlow, Java, Go, Scala, Python, AWS</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>Pinterest</employername>
      <employerlogo>https://logos.yubhub.co/pinterest.com.png</employerlogo>
      <employerdescription>Pinterest is a platform where people find creative ideas, dream about new possibilities, and plan for memories. The company aims to bring everyone the inspiration to create a life they love.</employerdescription>
      <employerwebsite>https://www.pinterest.com</employerwebsite>
      <compensationcurrency></compensationcurrency>
      <compensationmin></compensationmin>
      <compensationmax></compensationmax>
      <compensationinterval></compensationinterval>
      <applyto>https://job-boards.greenhouse.io/pinterest/jobs/7494956?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>Palo Alto; Seattle, WA; New York, NY; San Francisco, CA, US; Remote, US</location>
      <city>Palo Alto; Seattle</city>
      <state></state>
      <postalcode></postalcode>
      <country></country>
      <postedate>2026-08-26</postedate>
    </job>
    <job>
      <externalid>21529588-114</externalid>
      <title>Senior Staff Machine Learning Systems Engineer, Ads ML Platform</title>
      <description><![CDATA[<p>Reddit is seeking a Senior Staff Machine Learning Systems Engineer to lead the technical strategy for the end-to-end Ads ML engineer lifecycle.</p>
<p>The Ads ML Platform team builds infrastructure that accelerates high-scale ML systems and tooling for Ads ML, while extending reusable capabilities to broader Reddit ML use cases where appropriate.</p>
<p>Responsibilities:</p>
<ul>
<li>Own the technical strategy for the end-to-end Ads ML engineer lifecycle, starting with feature development, training data, offline experimentation, and model iteration workflows.</li>
<li>Align Ads ML platform priorities with Reddit&#39;s broader ML Platform vision, translating Ads pain points into reusable platform capabilities where appropriate.</li>
<li>Define architecture and technical standards for ML feature and training-data systems across batch/streaming computation, backfills, lineage, quality, observability, and online/offline consistency.</li>
<li>Stay close to ML engineers and platform customers to identify high-leverage friction points and improve day-to-day development velocity.</li>
<li>Build platform abstractions and workflow automation that make ML development faster, safer, more reliable, and more self-service.</li>
<li>Over time, extend the platform strategy into serving and online experimentation workflows, creating a more seamless offline-to-online ML development experience.</li>
<li>Partner across Ads, ML Platform, Data Platform, modeling, product, and engineering teams to clarify ownership, resolve ambiguity, and drive durable execution.</li>
<li>Mentor Staff and senior engineers, raise the architecture and operational bar, and help grow the next generation of technical leaders.</li>
</ul>
<p>Requirements:</p>
<ul>
<li>8+ years of experience in infrastructure, distributed systems, ML platforms, data platforms, or large-scale backend systems.</li>
<li>4+ years building or operating production ML infrastructure, feature platforms, training data systems, experimentation systems, or large-scale data pipelines.</li>
<li>Experience leading broad, ambiguous, multi-team platform initiatives from strategy through adoption.</li>
<li>Experience building platforms used directly by ML engineers, data scientists, or product teams developing production ML systems.</li>
<li>Deep experience in ML platform, feature platform, training data, experimentation, developer infrastructure, or distributed data infrastructure.</li>
<li>Experience working with distributed data and compute systems such as Spark, Flink, Kafka, Ray, Airflow, Iceberg, Kubernetes, BigQuery, Snowflake, Databricks, or similar technologies.</li>
</ul>
<p>Benefits:</p>
<ul>
<li>Comprehensive Healthcare Benefits and Income Replacement Programs</li>
<li>401k with Employer Match</li>
<li>Global Benefit programs that fit your lifestyle, from workspace to professional development to caregiving support</li>
<li>Family Planning Support</li>
<li>Gender-Affirming Care</li>
<li>Mental Health &amp; Coaching Benefits</li>
<li>Flexible Vacation &amp; Paid Volunteer Time Off</li>
<li>Generous Paid Parental Leave</li>
</ul>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>senior</experiencelevel>
      <workarrangement>remote</workarrangement>
      <salaryrange>$292,500-$409,500 USD</salaryrange>
      <skills>ML platforms, distributed systems, infrastructure, data platforms, large-scale backend systems, Spark, Flink, Kafka, Ray, Airflow, Iceberg, Kubernetes, BigQuery, Snowflake, Databricks</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>Reddit</employername>
      <employerlogo>https://logos.yubhub.co/redditinc.com.png</employerlogo>
      <employerdescription>Reddit is a community of communities with 100,000+ active communities and approximately 130 million daily active unique visitors.</employerdescription>
      <employerwebsite>https://www.redditinc.com</employerwebsite>
      <compensationcurrency>USD</compensationcurrency>
      <compensationmin>292500</compensationmin>
      <compensationmax>409500</compensationmax>
      <compensationinterval>yearly</compensationinterval>
      <applyto>https://job-boards.greenhouse.io/reddit/jobs/8157275?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>Remote - United States</location>
      <city>Remote - United States</city>
      <state></state>
      <postalcode></postalcode>
      <country></country>
      <postedate>2026-08-26</postedate>
    </job>
    <job>
      <externalid>8237c942-838</externalid>
      <title>Staff Software Engineer, Data Warehouse Foundation</title>
      <description><![CDATA[<p>About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime.</p>
<p>At Pinterest, we&#39;re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product.</p>
<p>The Data Product Platform is mission-critical to accelerate data-driven decision-making at Pinterest on the foundation of 100s of thousands of tables and an exadata-scale data warehouse.</p>
<p>We are seeking a Staff Software Engineer to provide technical leadership and contribute to the development of solutions across both these pillars.</p>
<p>In this role, you will shape the vision and drive the execution of the solutions that handle increasing data volume, navigate business complexity, accelerate analytical velocity, and build the foundation for our next generation of agentic analytics products.</p>
<p>Responsibilities:</p>
<ul>
<li>Define and design Pinterest&#39;s data warehouse architecture, storage, and access patterns.</li>
<li>Lead ambiguous, highly challenging, and cross-functional initiatives across the data warehouse and agentic data analytics domains.</li>
<li>Mentor and elevate the technical bar for engineers, providing leadership on complex technical design, execution strategies, and critical trade-off decisions.</li>
<li>Collaborate with stakeholders across Product, Data Science, and Engineering teams to understand user needs, drive alignment and drive progress.</li>
</ul>
<p>Requirements:</p>
<ul>
<li>Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent experience.</li>
<li>8+ years of relevant industry experience with large scale data warehouses, data tools and platforms.</li>
<li>5+ years experience building data warehouses and tools around the data ecosystem with technologies such as Spark, Trino, Flink, Airflow, Querybook, Superset, DataHub, etc.</li>
<li>Ability to work with cross-functional partners across multiple organizations.</li>
<li>Hands-on experience building tools and data pipelines leveraging AI coding tools, e.g. Cursor, Claude Code, Codex, etc.</li>
<li>Hands-on experience building AI tools and platforms to accelerate data pipeline authoring, data analytics and engineering productivity.</li>
</ul>
<p>Benefits:</p>
<ul>
<li>Salary range: $177,185-$364,795 USD</li>
<li>Equity eligibility</li>
</ul>
<p>#LI-REMOTE #LI-AH2</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>staff</experiencelevel>
      <workarrangement>hybrid</workarrangement>
      <salaryrange>$177,185-$364,795 USD</salaryrange>
      <skills>data warehousing, data engineering, Spark, Trino, Flink, Airflow, Querybook, Superset, DataHub, AI coding tools</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>Pinterest</employername>
      <employerlogo>https://logos.yubhub.co/pinterest.com.png</employerlogo>
      <employerdescription>Pinterest is a platform where people find creative ideas, dream about new possibilities, and plan for memories.</employerdescription>
      <employerwebsite>https://pinterest.com</employerwebsite>
      <compensationcurrency>USD</compensationcurrency>
      <compensationmin>177185</compensationmin>
      <compensationmax>364795</compensationmax>
      <compensationinterval>yearly</compensationinterval>
      <applyto>https://job-boards.greenhouse.io/pinterest/jobs/8076015?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>Palo Alto, CA</location>
      <city>Palo Alto</city>
      <state>CA</state>
      <postalcode></postalcode>
      <country>US</country>
      <postedate>2026-08-25</postedate>
    </job>
    <job>
      <externalid>e25bc328-741</externalid>
      <title>Analytics Engineer - X</title>
      <description><![CDATA[<p>xAI is seeking a skilled Analytics Engineer to build and maintain robust data systems that enable high-impact quantitative analysis and business decision-making.</p>
<p>This role combines strong software engineering practices with expertise in large-scale data processing and advanced analytical methods to deliver reliable, scalable solutions across the organization.</p>
<p><strong>Responsibilities:</strong></p>
<ul>
<li>Design, implement, and optimize end-to-end data pipelines for processing high-volume datasets using tools such as Spark, Kafka, Flink, etc.</li>
<li>Develop quantitative models and statistical frameworks to support experimentation, forecasting, and performance measurement.</li>
<li>Build and maintain data infrastructure that ensures data quality, consistency, and accessibility for analytical workflows.</li>
<li>Collaborate with product engineering, product, and operations teams to translate business requirements into production-grade data systems and insights.</li>
<li>Conduct A/B tests, causal analysis, and performance evaluations to drive measurable improvements in key metrics.</li>
<li>Implement monitoring, alerting, and automation for data systems to support real-time decision support.</li>
<li>Mentor team members on best practices for scalable data engineering and quantitative problem-solving.</li>
</ul>
<p><strong>Basic Qualifications:</strong></p>
<ul>
<li>4+ years of experience building production data pipelines and infrastructure at scale.</li>
<li>Strong proficiency in Python, SQL, and distributed computing frameworks (e.g., Spark, Flink, Hadoop).</li>
<li>Demonstrated expertise in statistical methods, predictive modeling, hypothesis testing, and experimental design.</li>
<li>Solid understanding of cloud services for data storage, processing, and orchestration.</li>
<li>Bachelor&#39;s or Master&#39;s degree in Computer Science, Statistics, Applied Mathematics, or related quantitative field.</li>
<li>Excellent problem-solving skills with a focus on delivering business impact through reliable systems.</li>
</ul>
<p><strong>Preferred Skills and Experience:</strong></p>
<ul>
<li>Prior work in consumer technology, or social media domains.</li>
<li>Experience with real-time streaming systems and low-latency data processing.</li>
<li>Contributions to open-source data tools or publications on large-scale analytics systems.</li>
<li>Track record of reducing operational costs or improving system efficiency through data optimizations.</li>
<li>Ability to bridge engineering excellence with rigorous analytical approaches.</li>
</ul>
<p><strong>Compensation and Benefits:</strong></p>
<p>The salary range is $180,000 - $440,000 USD. The total rewards package at xAI also includes equity, comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short &amp; long-term disability insurance, life insurance, and various other discounts and perks.</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>senior</experiencelevel>
      <workarrangement></workarrangement>
      <salaryrange>$180,000 - $440,000 USD</salaryrange>
      <skills>Python, SQL, Spark, Flink, Hadoop, distributed computing, statistical methods, predictive modeling, hypothesis testing, experimental design, cloud services, real-time streaming systems, low-latency data processing, open-source data tools, large-scale analytics systems, data optimizations</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>xAI</employername>
      <employerlogo>https://logos.yubhub.co/xai.com.png</employerlogo>
      <employerdescription>xAI aims to create AI systems that accurately understand the universe and aid humanity in its pursuit of knowledge. The organization operates with a flat structure and focuses on engineering excellence.</employerdescription>
      <employerwebsite>https://xai.com</employerwebsite>
      <compensationcurrency>USD</compensationcurrency>
      <compensationmin>180000</compensationmin>
      <compensationmax>440000</compensationmax>
      <compensationinterval>yearly</compensationinterval>
      <applyto>https://job-boards.greenhouse.io/xai/jobs/5210564007?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>Palo Alto, California</location>
      <city>Palo Alto</city>
      <state>California</state>
      <postalcode></postalcode>
      <country>US</country>
      <postedate>2026-08-13</postedate>
    </job>
    <job>
      <externalid>a0c71b5a-41e</externalid>
      <title>Analytics Engineer - X</title>
      <description><![CDATA[<p>xAI&#39;s mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge.</p>
<p>We are seeking a skilled Analytics Engineer to build and maintain robust data systems that enable high-impact quantitative analysis and business decision-making.</p>
<p><strong>Responsibilities:</strong></p>
<ul>
<li>Design, implement, and optimize end-to-end data pipelines for processing high-volume datasets using tools such as Spark, Kafka, Flink, etc.</li>
<li>Develop quantitative models and statistical frameworks to support experimentation, forecasting, and performance measurement.</li>
<li>Build and maintain data infrastructure that ensures data quality, consistency, and accessibility for analytical workflows.</li>
<li>Collaborate with product engineering, product, and operations teams to translate business requirements into production-grade data systems and insights.</li>
<li>Conduct A/B tests, causal analysis, and performance evaluations to drive measurable improvements in key metrics.</li>
<li>Implement monitoring, alerting, and automation for data systems to support real-time decision support.</li>
<li>Mentor team members on best practices for scalable data engineering and quantitative problem-solving.</li>
</ul>
<p><strong>Basic Qualifications:</strong></p>
<ul>
<li>4+ years of experience building production data pipelines and infrastructure at scale.</li>
<li>Strong proficiency in Python, SQL, and distributed computing frameworks (e.g., Spark, Flink, Hadoop).</li>
<li>Demonstrated expertise in statistical methods, predictive modeling, hypothesis testing, and experimental design.</li>
<li>Solid understanding of cloud services for data storage, processing, and orchestration.</li>
<li>Bachelor&#39;s or Master&#39;s degree in Computer Science, Statistics, Applied Mathematics, or related quantitative field.</li>
<li>Excellent problem-solving skills with a focus on delivering business impact through reliable systems.</li>
</ul>
<p><strong>Preferred Skills and Experience:</strong></p>
<ul>
<li>Prior work in consumer technology or social media domains.</li>
<li>Experience with real-time streaming systems and low-latency data processing.</li>
<li>Contributions to open-source data tools or publications on large-scale analytics systems.</li>
<li>Track record of reducing operational costs or improving system efficiency through data optimizations.</li>
<li>Ability to bridge engineering excellence with rigorous analytical approaches.</li>
</ul>
<p><strong>Compensation and Benefits:</strong></p>
<p>The salary range is $180,000 - $440,000 USD. The total rewards package at xAI also includes equity, comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short &amp; long-term disability insurance, life insurance, and various other discounts and perks.</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>senior</experiencelevel>
      <workarrangement></workarrangement>
      <salaryrange>$180,000 - $440,000 USD</salaryrange>
      <skills>Python, SQL, Spark, Flink, Hadoop, distributed computing, statistical methods, predictive modeling, experimental design, cloud services, real-time streaming systems, low-latency data processing, open-source data tools, large-scale analytics systems, data optimizations</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>xAI</employername>
      <employerlogo>https://logos.yubhub.co/xai.com.png</employerlogo>
      <employerdescription>xAI aims to create AI systems that accurately understand the universe and aid humanity in its pursuit of knowledge. The team is small and highly motivated.</employerdescription>
      <employerwebsite>https://xai.com</employerwebsite>
      <compensationcurrency>USD</compensationcurrency>
      <compensationmin>180000</compensationmin>
      <compensationmax>440000</compensationmax>
      <compensationinterval>yearly</compensationinterval>
      <applyto>https://job-boards.greenhouse.io/xai/jobs/5210564007?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>Palo Alto, California</location>
      <city>Palo Alto</city>
      <state>California</state>
      <postalcode></postalcode>
      <country>US</country>
      <postedate>2026-08-12</postedate>
    </job>
    <job>
      <externalid>9d1f3b64-b80</externalid>
      <title>Staff Software Engineer, Data Quality and Governance</title>
      <description><![CDATA[<p><strong>Job Description</strong></p>
<p><strong>Who we are</strong></p>
<p>Stripe is a financial infrastructure platform for businesses. Millions of companies,from the world’s largest enterprises to the most ambitious startups,use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead.</p>
<p><strong>About the team</strong></p>
<p>Data Quality and Governance owns the infrastructure that makes Stripe&#39;s data trustworthy and findable - the Data Catalog, the Knowledge Graph Service, dataset tiering and governance standards, lineage tracking, and the quality scoring system (DQPD) that every engineering team reports against. They&#39;re the team that defines what &quot;good data&quot; means at Stripe and then builds the enforcement and measurement tools to drive adoption across the company.</p>
<p><strong>Responsibilities</strong></p>
<ul>
<li>Lead the technical outcomes for a team of ambitious, talented engineers, providing mentorship, guidance, and support to ensure their success</li>
<li>Build and operate large-scale data discovery, metadata, or catalog platform</li>
<li>Develop strong subject matter expertise and manage the SLAs of data pipelines and full stack web applications that support critical stakeholders</li>
<li>Collaborate with product managers and peers across the company to create/improve canonical datasets and data warehouses, use golden paths, and ensure Stripes and customers are using trustworthy data</li>
<li>Leverage AI/LLM and Agents at scale to produce and analyze high-quality data on ambiguous problems</li>
<li>Have the opportunity to drive the execution of key data initiatives for Stripe, overseeing the entire development lifecycle from planning to delivery while maintaining high standards of quality and timely completion</li>
<li>Foster a collaborative and inclusive work environment, promoting innovation, knowledge sharing, and continuous improvement within the team</li>
</ul>
<p><strong>Who you are</strong></p>
<p>We&#39;re looking for someone who meets the minimum requirements to be considered for the role. If you meet these requirements, you are encouraged to apply.</p>
<p><strong>Minimum requirements</strong></p>
<ul>
<li>This is a Staff-level role , that typically means 10+ years of experience building and operating data systems, pipelines, warehouses, infrastructure, and leading teams to deliver exceptional solutions</li>
<li>Strong distributed systems fundamentals are a must , this team&#39;s core services are high-availability infrastructure that the rest of Stripe&#39;s data tooling depends on</li>
<li>An inquisitive nature in diving into data inconsistencies to pinpoint issues, and resolve deep rooted data quality issues</li>
<li>Knowledge of a backend development language (such as Scala, Java, or Go) and strong SQL experience</li>
<li>Extreme customer focus, with a commitment to partnering with product, leaders across the business, and other Stripe engineers to understand their use cases</li>
<li>Effective cross-functional collaboration, with the ability to think rigorously, communicate clearly, and make or coordinate difficult decisions and trade-offs</li>
<li>Thrive with high autonomy and responsibility in an ambiguous environment</li>
<li>Ability to foster and work in a healthy, inclusive, challenging, and supportive work environment</li>
</ul>
<p><strong>Preferred qualifications</strong></p>
<ul>
<li>Our stack is made up of Iceberg, Kafka, Change Data Capture, Flink, Spark, Airflow, Hive Metastore, Pinot, Trino, and AWS Cloud - experience with all or some of these tools is a huge plus</li>
<li>Influencing open-source contributions</li>
<li>Experience creating and maintaining data marts / warehouses to power business reporting needs</li>
<li>Experience collaborating with Product, Go-To-Market, or Sales / Marketing teams</li>
<li>Genuine enjoyment of innovation and a deep interest in understanding how things work, with the ability to question and direct architectural decisions</li>
<li>Strong written and verbal communication skills for various audiences, including leadership, users, and company-wide</li>
</ul>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>staff</experiencelevel>
      <workarrangement></workarrangement>
      <salaryrange></salaryrange>
      <skills>distributed systems, data systems, pipelines, warehouses, infrastructure, backend development, SQL, collaboration, communication, Iceberg, Kafka, Change Data Capture, Flink, Spark, Airflow, Hive Metastore, Pinot, Trino, AWS Cloud, open-source contributions, data marts, Product, Go-To-Market, Sales, Marketing</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>Stripe</employername>
      <employerlogo>https://logos.yubhub.co/stripe.com.png</employerlogo>
      <employerdescription>Stripe is a financial infrastructure platform for businesses. Millions of companies use Stripe to accept payments, grow their revenue, and accelerate new business opportunities.</employerdescription>
      <employerwebsite>https://stripe.com</employerwebsite>
      <compensationcurrency></compensationcurrency>
      <compensationmin></compensationmin>
      <compensationmax></compensationmax>
      <compensationinterval></compensationinterval>
      <applyto>https://job-boards.greenhouse.io/stripe/jobs/8112042?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>Seattle</location>
      <city>Seattle</city>
      <state></state>
      <postalcode></postalcode>
      <country></country>
      <postedate>2026-08-06</postedate>
    </job>
    <job>
      <externalid>d79b6192-615</externalid>
      <title>Staff Software Engineer, Business Data</title>
      <description><![CDATA[<p><strong>Job Description</strong></p>
<p><strong>Who we are</strong></p>
<p>Stripe is a financial infrastructure platform for businesses. Millions of companies,from the world’s largest enterprises to the most ambitious startups,use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead.</p>
<p><strong>About the team</strong></p>
<p>We are experts in business data, building the canonical data foundation that powers decision-making across Finance &amp; Strategy, Business Data Science, GTM, and Marketing. We engineer scalable and efficient data pipelines and provide curated data products that give Stripe&#39;s teams a consistent, accurate, and fresh view of the business,from the top to the bottom of the funnel.</p>
<p><strong>What you’ll do</strong></p>
<p>We&#39;re looking for a person who could contribute to the team by solving high-impact, cutting-edge data problems. The ideal candidate will be someone that has built data pipelines for large scale volume, is deeply knowledgeable of key tools including Airflow/Spark/Kafka/Flink, is empathetic, excels at building strong relationships, and collaborates effectively with other Stripe teams to understand their use cases and unlock new capabilities.</p>
<p><strong>Responsibilities</strong></p>
<ul>
<li>Lead the technical outcomes for a team of ambitious, talented engineers, providing mentorship, guidance, and support to ensure their success</li>
<li>Partner with our recruiting team to attract and hire top talent</li>
<li>Deliver cutting-edge data pipelines that scale to users&#39; needs, focusing on reliability and efficiency</li>
<li>Develop strong subject matter expertise and manage the SLAs of data pipelines and full stack web applications that support critical stakeholders</li>
<li>Collaborate with product managers and peers across the company to create/improve canonical datasets and data warehouses, use golden paths, and ensure Stripes and customers are using trustworthy data</li>
<li>Leverage AI/LLM and Agents at scale to produce and analyze high-quality data on ambiguous problems</li>
<li>Have the opportunity to drive the execution of key data initiatives for Stripe, overseeing the entire development lifecycle from planning to delivery while maintaining high standards of quality and timely completion</li>
<li>Foster a collaborative and inclusive work environment, promoting innovation, knowledge sharing, and continuous improvement within the team</li>
</ul>
<p><strong>Who you are</strong></p>
<p><strong>Minimum requirements</strong></p>
<ul>
<li>This is a Staff-level role , that typically means 10+ years of experience building and operating data systems, pipelines, warehouses, infrastructure, and leading teams to deliver exceptional solutions</li>
<li>A strong engineering background and passion for data as well as prior experience with writing and debugging data pipelines using a distributed data framework</li>
<li>An inquisitive nature in diving into data inconsistencies to pinpoint issues, and resolve deep rooted data quality issues</li>
<li>Knowledge of a backend development language (such as Scala, Java, or Go) and strong SQL experience</li>
<li>Extreme customer focus, with a commitment to partnering with product, leaders across the business, and other Stripe engineers to understand their use cases</li>
<li>Effective cross-functional collaboration, with the ability to think rigorously, communicate clearly, and make or coordinate difficult decisions and trade-offs</li>
<li>Thrive with high autonomy and responsibility in an ambiguous environment</li>
<li>Ability to foster and work in a healthy, inclusive, challenging, and supportive work environment</li>
</ul>
<p><strong>Preferred qualifications</strong></p>
<ul>
<li>Our stack is made up of Iceberg, Kafka, Change Data Capture, Flink, Spark, Airflow, Hive Metastore, Pinot, Trino, and AWS Cloud - experience with all or some of these tools is a huge plus</li>
<li>Influencing open-source contributions</li>
<li>Experience creating and maintaining data marts / warehouses to power business reporting needs</li>
<li>Experience collaborating with Product, Go-To-Market, or Sales / Marketing teams</li>
<li>Genuine enjoyment of innovation and a deep interest in understanding how things work, with the ability to question and direct architectural decisions</li>
<li>Strong written and verbal communication skills for various audiences, including leadership, users, and company-wide</li>
</ul>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>staff</experiencelevel>
      <workarrangement></workarrangement>
      <salaryrange></salaryrange>
      <skills>data pipelines, distributed data framework, backend development language, SQL, Airflow, Spark, Kafka, Flink, Iceberg, Change Data Capture, Hive Metastore, Pinot, Trino, AWS Cloud, open-source contributions, data marts / warehouses</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>Stripe</employername>
      <employerlogo>https://logos.yubhub.co/stripe.com.png</employerlogo>
      <employerdescription>Stripe is a financial infrastructure platform for businesses. Millions of companies use Stripe to accept payments, grow their revenue, and accelerate new business opportunities.</employerdescription>
      <employerwebsite>https://stripe.com</employerwebsite>
      <compensationcurrency></compensationcurrency>
      <compensationmin></compensationmin>
      <compensationmax></compensationmax>
      <compensationinterval></compensationinterval>
      <applyto>https://job-boards.greenhouse.io/stripe/jobs/8112037?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>Seattle</location>
      <city>Seattle</city>
      <state></state>
      <postalcode></postalcode>
      <country></country>
      <postedate>2026-08-06</postedate>
    </job>
    <job>
      <externalid>ce0831a2-300</externalid>
      <title>Staff Software Engineer, Product Risk</title>
      <description><![CDATA[<p><strong>Job Description</strong></p>
<p>Stripe is a financial infrastructure platform for businesses. Millions of companies,from the world’s largest enterprises to the most ambitious startups,use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet.</p>
<p><strong>About the Team</strong></p>
<p>Product and Risk Data Engineering is Stripe&#39;s single source of truth data engineering layer for Payments, Risk, and Product , we enable Stripe to confidently run, measure, and grow the business by making accurate information easy to access.</p>
<p><strong>Responsibilities</strong></p>
<ul>
<li>Lead the technical outcomes for a team of ambitious, talented engineers, providing mentorship, guidance, and support to ensure their success</li>
<li>Partner with our recruiting team to attract and hire top talent</li>
<li>Deliver cutting-edge data pipelines that scale to users&#39; needs, focusing on reliability and efficiency</li>
<li>Develop strong subject matter expertise and manage the SLAs of data pipelines and full stack web applications that support critical stakeholders</li>
<li>Collaborate with product managers and peers across the company to create/improve canonical datasets and data warehouses, use golden paths, and ensure Stripes and customers are using trustworthy data</li>
<li>Leverage AI/LLM and Agents at scale to produce and analyze high-quality data on ambiguous problems</li>
<li>Have the opportunity to drive the execution of key data initiatives for Stripe, overseeing the entire development lifecycle from planning to delivery while maintaining high standards of quality and timely completion</li>
<li>Foster a collaborative and inclusive work environment, promoting innovation, knowledge sharing, and continuous improvement within the team</li>
</ul>
<p><strong>Requirements</strong></p>
<ul>
<li>This is a Staff-level role , that typically means 10+ years of experience building and operating data systems, pipelines, warehouses, infrastructure, and leading teams to deliver exceptional solutions</li>
<li>A strong engineering background and passion for data as well as prior experience with writing and debugging data pipelines using a distributed data framework</li>
<li>An inquisitive nature in diving into data inconsistencies to pinpoint issues, and resolve deep rooted data quality issues</li>
<li>Knowledge of a backend development language (such as Scala, Java, or Go) and strong SQL experience</li>
<li>Extreme customer focus, with a commitment to partnering with product, leaders across the business, and other Stripe engineers to understand their use cases</li>
<li>Effective cross-functional collaboration, with the ability to think rigorously, communicate clearly, and make or coordinate difficult decisions and trade-offs</li>
<li>Thrive with high autonomy and responsibility in an ambiguous environment</li>
<li>Ability to foster and work in a healthy, inclusive, challenging, and supportive work environment</li>
</ul>
<p><strong>Preferred Qualifications</strong></p>
<ul>
<li>Experience with Iceberg, Kafka, Change Data Capture, Flink, Spark, Airflow, Hive Metastore, Pinot, Trino, and AWS Cloud</li>
<li>Influencing open-source contributions</li>
<li>Experience creating and maintaining data marts / warehouses to power business reporting needs</li>
<li>Experience collaborating with Product, Go-To-Market, or Sales / Marketing teams</li>
<li>Genuine enjoyment of innovation and a deep interest in understanding how things work, with the ability to question and direct architectural decisions</li>
<li>Strong written and verbal communication skills for various audiences, including leadership, users, and company-wide</li>
</ul>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>staff</experiencelevel>
      <workarrangement></workarrangement>
      <salaryrange></salaryrange>
      <skills>data pipelines, distributed data framework, backend development language, SQL, cross-functional collaboration, Iceberg, Kafka, Change Data Capture, Flink, Spark, Airflow, Hive Metastore, Pinot, Trino, AWS Cloud, open-source contributions, data marts / warehouses</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>Stripe</employername>
      <employerlogo>https://logos.yubhub.co/stripe.com.png</employerlogo>
      <employerdescription>Stripe is a financial infrastructure platform for businesses. Millions of companies use Stripe to accept payments, grow their revenue, and accelerate new business opportunities.</employerdescription>
      <employerwebsite>https://stripe.com</employerwebsite>
      <compensationcurrency></compensationcurrency>
      <compensationmin></compensationmin>
      <compensationmax></compensationmax>
      <compensationinterval></compensationinterval>
      <applyto>https://job-boards.greenhouse.io/stripe/jobs/8112043?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>Toronto, ON</location>
      <city>Toronto</city>
      <state>ON</state>
      <postalcode></postalcode>
      <country>CA</country>
      <postedate>2026-08-06</postedate>
    </job>
    <job>
      <externalid>967fe4bc-d80</externalid>
      <title>Staff Software Reliability Engineer - Data Platform</title>
      <description><![CDATA[<p>Secure Every Identity, from AI to Human</p>
<p>Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organisations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence.</p>
<p>This is an opportunity to do career-defining work. We&#39;re all in on this mission. If you are too, let&#39;s talk.</p>
<p><strong>Staff Software Reliability Engineer - Data Platform</strong></p>
<p><strong>About the Team</strong></p>
<p>The Data Platform team is responsible for the foundational data services, systems, and data products for Okta that benefit our users. Today, the Data Platform team solves challenges and enables:</p>
<ul>
<li>Streaming analytics</li>
<li>Interactive end-user reporting</li>
<li>Data and ML platform for Okta to scale</li>
<li>Telemetry of our products and data</li>
</ul>
<p>Our elite team is fast, creative and flexible. We encourage ownership. We expect great things from our engineers and reward them with stimulating new projects, new technologies and the chance to have significant equity in a company. Okta is about to change the cloud computing landscape forever.</p>
<p><strong>About the Position</strong></p>
<p>This is an opportunity for experienced Software Reliability Engineers to join our fast growing Data Platform organisation that is passionate about scaling high volume, low-latency, distributed data-platform services &amp; data products. In this role, you will get to work with engineers throughout the organisation to build foundational infrastructure that allows Okta to scale for years to come. As a member of the Data Platform team, you will be responsible for designing, building, and deploying the systems that power our data analytics and ML. Our analytics infrastructure stack sits on top of many modern technologies, including Kinesis, Flink, ElasticSearch, and Snowflake.</p>
<p>We are looking for experienced Software Engineers who can help design and own the building, deploying and optimising the streaming infrastructure. This project has a directive from engineering leadership to make OKTA a leader in the use of data and machine learning to improve end-user security and to expand that core-competency across the rest of engineering. You will have a sizable impact on the direction, design &amp; implementation of the solutions to these problems.</p>
<p><strong>Job Duties and Responsibilities:</strong></p>
<ul>
<li>Design, implement and own data-intensive, high-performance, scalable platform components</li>
<li>Work with engineering teams, architects and cross functional partners on the development of projects, design, and implementation</li>
<li>Conduct and participate in design reviews, code reviews, analysis, and performance tuning</li>
<li>Coach and mentor engineers to help scale up the engineering organisation</li>
<li>Debug production issues across services and multiple levels of the stack</li>
<li>Participate in the on-call rotation, and incident management</li>
</ul>
<p><strong>Required Knowledge, Skills, and Abilities:</strong></p>
<ul>
<li>5+ years of industry experience</li>
<li>2+ years of experience in object-oriented language, preferably Java</li>
<li>Hands-on experience using a cloud-based distributed computing technologies including</li>
<li>Messaging systems such as Kinesis, Kafka</li>
<li>Data processing systems like Flink, Spark, Beam</li>
<li>Storage &amp; Compute systems such as Snowflake, Databricks, Hadoop</li>
<li>Coordinators and schedulers like the ones in Kubernetes, Hadoop, Mesos</li>
<li>Experience in developing and tuning highly scalable distributed systems</li>
<li>Excellent grasp of software engineering principles</li>
<li>Solid understanding of multithreading, garbage collection and memory management</li>
<li>Experience with reliability engineering specifically in areas such as data quality, data observability and incident management</li>
</ul>
<p><strong>Nice to have</strong></p>
<ul>
<li>Maintained security, encryption, identity management, or authentication infrastructure</li>
<li>Leveraged major public cloud providers to build mission-critical, high volume services</li>
<li>Hands-on experience in developing Data Integration applications for large scale (petabyte scale) environments with experience in both batch and online systems.</li>
<li>Contributed to the development of distributed systems or used one or more at high volume or criticality such as Kafka or Hadoop</li>
<li>Experience developing Kubernetes based services on AWS Stack</li>
</ul>
<p>The annual base salary range for this position for candidates located in Canada is between:$160,000-$220,000 CAD The Okta Experience</p>
<ul>
<li>Supporting Your Well-Being</li>
<li>Driving Social Impact</li>
<li>Developing Talent and Fostering Connection + Community</li>
</ul>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>staff</experiencelevel>
      <workarrangement>hybrid</workarrangement>
      <salaryrange>$160,000-$220,000 CAD</salaryrange>
      <skills>Java, Kinesis, Kafka, Flink, Spark, Beam, Snowflake, Databricks, Hadoop, Kubernetes, Mesos, distributed systems, software engineering, multithreading, garbage collection, memory management, reliability engineering, data quality, data observability, incident management, security, encryption, identity management, authentication infrastructure, public cloud providers, Data Integration applications, Kubernetes based services on AWS Stack</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>Okta</employername>
      <employerlogo>https://logos.yubhub.co/rewards.okta.com.png</employerlogo>
      <employerdescription>Okta builds identity and access management solutions for organisations.</employerdescription>
      <employerwebsite>https://rewards.okta.com/can</employerwebsite>
      <compensationcurrency>CAD</compensationcurrency>
      <compensationmin>160000</compensationmin>
      <compensationmax>220000</compensationmax>
      <compensationinterval>yearly</compensationinterval>
      <applyto>https://job-boards.greenhouse.io/okta/jobs/8082028?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>Toronto, ON</location>
      <city>Toronto</city>
      <state>ON</state>
      <postalcode></postalcode>
      <country>CA</country>
      <postedate>2026-07-29</postedate>
    </job>
    <job>
      <externalid>34c614e6-5e5</externalid>
      <title>Senior Software Engineer, Data Platform</title>
      <description><![CDATA[<p>Software is eating the world, but AI is eating software. We live in unprecedented times – AI has the potential to exponentially augment human intelligence.</p>
<p>At Scale, our products include the Generative AI Data Engine, SGP, Donovan, and others that power the most advanced LLMs and generative models in the world through world-class RLHF, human data generation, model evaluation, safety, and alignment.</p>
<p>In this role, you will lead the design and development of core data storage, streaming, caching, and indexing platforms and underlying systems. You’ll also get widespread exposure to the forefront of the AI race as Scale sees it in enterprises, startups, governments, and large tech companies.</p>
<p>Responsibilities:</p>
<ul>
<li>Drive the design, implementation, and reliability of our foundational data platforms and systems, working closely with stakeholders and internal customers to understand and refine requirements.</li>
<li>Collaborate with cross-functional teams to define, design, and deliver new features.</li>
<li>Proactively identify opportunities for, and drive improvements to, current programming practices, including process enhancements and tool upgrades.</li>
<li>Present technical information to teams and stakeholders, providing guidance and insight on development processes and technologies.</li>
</ul>
<p>Requirements:</p>
<ul>
<li>5+ years of full-time engineering experience, post-graduation with specialties in back-end systems, specifically related to building large-scale data storage, streaming, and warehousing systems.</li>
<li>Experience in various database technologies (MongoDB, Postgres), streaming/processing solutions (Kinesis, Flink, Spark), indexing/caching (ElasticSearch, Redis), and various data query engines (Trino, Presto, Snowflake, etc.).</li>
<li>Show a track record of mentoring and leading teams in successful projects.</li>
<li>Possess excellent communication and collaboration skills, and the ability to translate complex technical concepts to non-technical stakeholders.</li>
<li>Experience working fluently with standard containerization &amp; deployment technologies like Kubernetes and various public cloud offerings.</li>
<li>Extensive experience in software development and a deep understanding of distributed systems, cloud platforms and data systems.</li>
</ul>
<p>Nice to haves:</p>
<ul>
<li>Strong knowledge of software engineering best practices and CI/CD tooling (CircleCI).</li>
<li>Experience scaling products at hyper-growth startups.</li>
<li>Excitement to work with AI technologies.</li>
</ul>
<p>Benefits: No benefits mentioned.</p>
<p>About Us: At Scale, our mission is to develop reliable AI systems for the world&#39;s most important decisions. We work closely with industry leaders like Meta, Ernst &amp; Young, Mayo Clinic, Time Inc., the Government of Qatar, and U.S. government agencies including the Army and Air Force.</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>senior</experiencelevel>
      <workarrangement></workarrangement>
      <salaryrange></salaryrange>
      <skills>large-scale data storage, streaming, warehousing systems, database technologies, MongoDB, Postgres, Kinesis, Flink, Spark, ElasticSearch, Redis, Trino, Presto, Snowflake, Kubernetes, public cloud offerings, distributed systems, cloud platforms, data systems, software engineering best practices, CI/CD tooling, CircleCI, hyper-growth startups, AI technologies</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>Scale</employername>
      <employerlogo>https://logos.yubhub.co/scaleai.com.png</employerlogo>
      <employerdescription>Scale develops reliable AI systems for the world&apos;s most important decisions.</employerdescription>
      <employerwebsite>https://scaleai.com</employerwebsite>
      <compensationcurrency></compensationcurrency>
      <compensationmin></compensationmin>
      <compensationmax></compensationmax>
      <compensationinterval></compensationinterval>
      <applyto>https://job-boards.greenhouse.io/scaleai/jobs/4717101005?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>London</location>
      <city>London</city>
      <state></state>
      <postalcode></postalcode>
      <country>GB</country>
      <postedate>2026-07-22</postedate>
    </job>
    <job>
      <externalid>dafd169f-a53</externalid>
      <title>Staff Data Engineer - Data Engineering</title>
      <description><![CDATA[<p>CoreWeave is The Essential Cloud for AI. Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence.</p>
<p>The Data Engineering Team builds and operates the foundational data infrastructure powering analytics, AI, and operational decision-making across CoreWeave. We design resilient data pipelines, scalable lakehouse systems, and high-quality datasets that enable teams across Finance, HR, Operations, and Engineering to move faster and make smarter decisions.</p>
<p>We&#39;re seeking a Staff Data Engineer to define and drive the architecture of CoreWeave&#39;s enterprise data ecosystem. You will establish the modeling, semantic, governance, and platform standards that enable teams to build trusted and reusable data products at scale.</p>
<p>In this role, you will:</p>
<ul>
<li>Define the company-wide strategy for data modeling, semantic layers, enterprise metrics, and reusable data products.</li>
<li>Set architectural standards for CoreWeave&#39;s lakehouse, including data organization, schemas, metadata, governance, and serving patterns.</li>
<li>Design scalable frameworks for data quality, lineage, observability, access control, and policy enforcement.</li>
<li>Lead the architecture of complex, cross-domain data systems spanning operational, financial, product, people, and infrastructure data.</li>
<li>Establish durable data contracts, system boundaries, and reusable engineering patterns across teams.</li>
<li>Identify and resolve systemic performance, reliability, and scalability constraints across the data ecosystem.</li>
<li>Lead architecture reviews for high-impact initiatives and ensure alignment with the broader platform strategy.</li>
<li>Evaluate and standardize core data technologies across processing, orchestration, cataloging, modeling, and metadata management.</li>
<li>Mentor senior engineers and raise the quality of system design and technical decision-making across the organization.</li>
</ul>
<p>Who You Are:</p>
<ul>
<li>10+ years of experience in data engineering, software engineering, distributed systems, or data architecture roles.</li>
<li>Demonstrated experience setting technical direction for data systems spanning multiple teams, business domains, or platforms.</li>
<li>Deep expertise in enterprise analytical modeling, including dimensional, Data Vault, semantic modeling, and governed metrics.</li>
<li>Experience designing and operating large-scale lakehouse or streamhouse architectures using technologies such as Apache Iceberg, Delta Lake, Apache Hudi, Apache Paimon, or Apache Fluss.</li>
<li>Advanced knowledge of distributed OLAP, query, processing, and ingestion systems such as StarRocks, ClickHouse, Trino, Spark, Flink, or Kafka.</li>
<li>Expert-level SQL and strong programming expertise in Python, Scala, Java, or Rust, with experience building production-grade data systems.</li>
<li>Demonstrated experience implementing governance capabilities such as lineage, profiling, metadata management, data quality controls, access policies, or auditability.</li>
<li>Demonstrated ability to independently turn ambiguous problem statements into clear technical direction and execution plans.</li>
<li>Track record of leading cross-cutting architecture and establishing standards adopted across multiple engineering teams.</li>
</ul>
<p>The base salary range for this role is $207,000 to $275,000. The starting salary will be determined based on job-related knowledge, skills, experience, and market location. We strive for both market alignment and internal equity when determining compensation. In addition to base salary, our total rewards package includes a discretionary bonus, equity awards, and a comprehensive benefits program.</p>
<p>What We Offer:</p>
<ul>
<li>Medical, dental, and vision insurance - 100% paid for by CoreWeave</li>
<li>Company-paid Life Insurance</li>
<li>Voluntary supplemental life insurance</li>
<li>Short and long-term disability insurance</li>
<li>Flexible Spending Account</li>
<li>Health Savings Account</li>
<li>Tuition Reimbursement</li>
<li>Ability to Participate in Employee Stock Purchase Program (ESPP)</li>
<li>Mental Wellness Benefits through Spring Health</li>
<li>Family-Forming support provided by Carrot</li>
<li>Paid Parental Leave</li>
<li>Flexible, full-service childcare support with Kinside</li>
<li>401(k) with a generous employer match</li>
<li>Flexible PTO</li>
<li>Catered lunch each day in our office and data center locations</li>
<li>A casual work environment</li>
<li>A work culture focused on innovative disruption</li>
</ul>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>staff</experiencelevel>
      <workarrangement></workarrangement>
      <salaryrange>$207,000 to $275,000</salaryrange>
      <skills>Apache Iceberg, Delta Lake, Apache Hudi, Apache Paimon, Apache Fluss, StarRocks, ClickHouse, Trino, Spark, Flink, Kafka, SQL, Python, Scala, Java, Rust, event-driven architectures, AutoMQ, Pulsar, Kubernetes</skills>
      <category>IT</category>
      <industry>Technology</industry>
      <employername>CoreWeave</employername>
      <employerlogo>https://logos.yubhub.co/coreweave.com.png</employerlogo>
      <employerdescription>CoreWeave is a publicly traded company (Nasdaq: CRWV) that delivers a platform of technology, tools, and teams to enable innovators to build and scale AI with confidence.</employerdescription>
      <employerwebsite>https://www.coreweave.com</employerwebsite>
      <compensationcurrency>USD</compensationcurrency>
      <compensationmin>207000</compensationmin>
      <compensationmax>275000</compensationmax>
      <compensationinterval>yearly</compensationinterval>
      <applyto>https://job-boards.greenhouse.io/coreweave/jobs/4698414006?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>Livingston, NJ / New York, NY / Sunnyvale, CA / San Francisco, CA / Bellevue, WA</location>
      <city>Livingston</city>
      <state>WA</state>
      <postalcode></postalcode>
      <country>US</country>
      <postedate>2026-07-21</postedate>
    </job>
    <job>
      <externalid>b14109ed-290</externalid>
      <title>Software Engineer - X Data</title>
      <description><![CDATA[<p><strong>About the Role</strong></p>
<p>As a Software Engineer in X Data, you will play a key role in providing comprehensive data solutions that serve a wide range of stakeholders, including end-users and internal teams.</p>
<p><strong>Responsibilities</strong></p>
<p>You will help build and operate a distributed data platform that powers hundreds of real-time and batch pipelines processing billions of events per day.</p>
<ul>
<li>Design, build, and operate production-grade real-time and batch pipelines that ingest, process, validate, and deliver data powering user-behavior insights and product decisions.</li>
<li>Create shared datasets, fact tables, and internal data products that let other teams analyse, debug, and improve product performance.</li>
<li>Prototype and build tooling that automates and accelerates internal data workflows , backfills, dashboards, report generation, and self-serve access to data.</li>
<li>Own data correctness end to end: validate with output invariants, denominator reconciliation, and independent recomputation, and lead root-cause investigations when key metrics move unexpectedly.</li>
<li>Move fluidly across query engines and frameworks, choosing the right tool and adapting quickly to new infrastructure and environments.</li>
<li>Partner across product and business teams to surface where data gaps exist and prioritise the highest-impact opportunities for new data acquisition and improvement.</li>
<li>Iterate quickly on feedback, shipping the smallest useful increment with a strong bias toward efficient, accurate, and reliable solutions.</li>
</ul>
<p><strong>Basic Qualifications</strong></p>
<p>We are looking for an engineer with 3+ years of professional software engineering experience, ideally in data engineering or distributed systems.</p>
<ul>
<li>Hands-on expertise in Python, Rust, Scala, Go or Java, and data pipeline toolings and distributed systems.</li>
<li>Knowledge of real-time and batch data processing tools such as Spark/Kafka/Flink/SQL and various storage systems in RMDBs/NoSQL.</li>
<li>Experience solving large-scale problems and comfortable doing incremental quality work while building brand new systems to enable future quality improvements.</li>
<li>Proven records of interpreting product requirements into engineering implementation plans, and effectively communicating with different groups.</li>
</ul>
<p><strong>Compensation and Benefits</strong></p>
<p>$125,000 - $400,000 USD</p>
<p>Base salary is just one part of our total rewards package at xAI, which also includes equity, comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short &amp; long-term disability insurance, life insurance, and various other discounts and perks.</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>senior</experiencelevel>
      <workarrangement></workarrangement>
      <salaryrange>$125,000 - $400,000 USD</salaryrange>
      <skills>Python, Rust, Scala, Go, Java, Spark, Kafka, Flink, SQL, distributed systems</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>xAI</employername>
      <employerlogo>https://logos.yubhub.co/xai.com.png</employerlogo>
      <employerdescription>xAI creates AI systems to understand the universe and aid human knowledge pursuit. The team is small and highly motivated.</employerdescription>
      <employerwebsite>https://xai.com</employerwebsite>
      <compensationcurrency>USD</compensationcurrency>
      <compensationmin>125000</compensationmin>
      <compensationmax>400000</compensationmax>
      <compensationinterval>yearly</compensationinterval>
      <applyto>https://job-boards.greenhouse.io/xai/jobs/5182183007?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>Palo Alto, CA</location>
      <city>Palo Alto</city>
      <state>CA</state>
      <postalcode></postalcode>
      <country>US</country>
      <postedate>2026-07-13</postedate>
    </job>
    <job>
      <externalid>100128a3-41f</externalid>
      <title>Software Engineer - X Data</title>
      <description><![CDATA[<p><strong>About the Role</strong></p>
<p>As a Software Engineer in X Data, you will play a key role in providing comprehensive data solutions that serve a wide range of stakeholders, including end-users and internal teams.</p>
<p><strong>Responsibilities</strong></p>
<p>You will help build and operate a distributed data platform that powers hundreds of real-time and batch pipelines processing billions of events per day.</p>
<ul>
<li>Design, build, and operate production-grade real-time and batch pipelines that ingest, process, validate, and deliver data powering user-behavior insights and product decisions.</li>
<li>Create shared datasets, fact tables, and internal data products that let other teams analyse, debug, and improve product performance.</li>
<li>Prototype and build tooling that automates and accelerates internal data workflows , backfills, dashboards, report generation, and self-serve access to data.</li>
<li>Own data correctness end to end: validate with output invariants, denominator reconciliation, and independent recomputation, and lead root-cause investigations when key metrics move unexpectedly.</li>
<li>Move fluidly across query engines and frameworks (e.g., BigQuery, Trino, Clickhouse for analytics; Flink, Kafka, Spark/Scalding for streaming and batch), choosing the right tool and adapting quickly to new infrastructure and environments.</li>
<li>Partner across product and business teams to surface where data gaps exist and prioritise the highest-impact opportunities for new data acquisition and improvement.</li>
<li>Iterate quickly on feedback, shipping the smallest useful increment with a strong bias toward efficient, accurate, and reliable solutions.</li>
</ul>
<p><strong>Basic Qualifications</strong></p>
<ul>
<li>3+ years of professional software engineering experience, ideally in data engineering or distributed systems.</li>
<li>Hands-on expertise in Python, Rust, Scala, Go or Java, and data pipeline toolings and distributed systems.</li>
<li>Knowledge of real-time and batch data processing tools such as Spark/Kafka/Flink/SQL and various storage systems in RMDBs/NoSQL.</li>
<li>Experience solving large-scale problems and comfortable doing incremental quality work while building brand new systems to enable future quality improvements.</li>
<li>Proven records of interpreting product requirements into engineering implementation plans, and effectively communicating with different groups (AI, product, marketing/sales and engineering).</li>
</ul>
<p><strong>Compensation and Benefits</strong></p>
<p>$125,000 - $400,000 USD</p>
<p>Base salary is just one part of our total rewards package at xAI, which also includes equity, comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short &amp; long-term disability insurance, life insurance, and various other discounts and perks.</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>senior</experiencelevel>
      <workarrangement></workarrangement>
      <salaryrange>$125,000 - $400,000 USD</salaryrange>
      <skills>Python, Rust, Scala, Go, Java, Spark, Kafka, Flink, SQL, BigQuery, Trino, Clickhouse</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>xAI</employername>
      <employerlogo>https://logos.yubhub.co/xai.com.png</employerlogo>
      <employerdescription>xAI creates AI systems to understand the universe and aid human knowledge pursuit. The team is small and highly motivated.</employerdescription>
      <employerwebsite>https://xai.com</employerwebsite>
      <compensationcurrency>USD</compensationcurrency>
      <compensationmin>125000</compensationmin>
      <compensationmax>400000</compensationmax>
      <compensationinterval>yearly</compensationinterval>
      <applyto>https://job-boards.greenhouse.io/xai/jobs/5182183007?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>Palo Alto, CA</location>
      <city>Palo Alto</city>
      <state>CA</state>
      <postalcode></postalcode>
      <country>US</country>
      <postedate>2026-07-12</postedate>
    </job>
    <job>
      <externalid>e76ceb38-293</externalid>
      <title>Software Engineer</title>
      <description><![CDATA[<p>The Platform team at Spotify creates technology for rapid growth and scaling. The Financial Engineering organisation builds infrastructure and tooling for finance teams. The Oracles team is responsible for building data pipelines and backend services for financial infrastructure.</p>
<p>We are looking for a Software Engineer to join the team. The successful candidate will help shape the future of Spotify&#39;s financial data platform by designing robust data models, developing scalable backend services, and building high-volume data pipelines.</p>
<p><strong>Responsibilities</strong></p>
<ul>
<li>Build, operate, and evolve scalable Scio data pipelines on Google Cloud Platform that process high-volume financial data with accuracy and reliability.</li>
<li>Develop and enhance Java backend services that power integrations across Spotify&#39;s financial ecosystem.</li>
<li>Design and evolve data models and datasets that support accounting, reconciliation, financial reporting, and operational insights.</li>
<li>Build and improve integrations with external financial systems to bring critical financial data into Spotify&#39;s financial ecosystem.</li>
<li>Partner closely with Finance, Product, and Engineering teams to translate accounting requirements into reliable, scalable technical solutions.</li>
<li>Improve the scalability, observability, and resilience of financial data pipelines and backend services.</li>
<li>Contribute to engineering best practices through continuous delivery, automated testing, monitoring, and thoughtful code reviews.</li>
<li>Help drive technical decisions that improve the long-term maintainability of Spotify&#39;s financial engineering platform.</li>
</ul>
<p><strong>Requirements</strong></p>
<ul>
<li>3+ years of professional experience building backend systems and data-intensive applications.</li>
<li>Experience with Scala or Java and at least one distributed data processing framework such as Scio, Apache Beam, Spark, or Flink.</li>
<li>Knowledge of designing and building reliable backend services and scalable data pipelines that support business-critical workflows.</li>
<li>Experience designing data models that enable accurate, high-volume financial processes.</li>
<li>Experience working with workflow orchestration platforms such as Flyte, Airflow, or similar technologies.</li>
<li>Strong engineering fundamentals, including continuous delivery, automated testing, monitoring, and writing production-quality code.</li>
<li>Ability to collaborate with cross-functional partners and translate complex business requirements into elegant technical solutions.</li>
<li>Experience working within financial, accounting, or other regulated data domains is a plus.</li>
<li>A Bachelor&#39;s degree or higher in Computer Science or a related field is a plus.</li>
</ul>
<p><strong>Location and Benefits</strong></p>
<ul>
<li>This role is based in Toronto.</li>
<li>Flexible work arrangements available.</li>
<li>Salary range: CAD 103,321.00–147,602.00, plus equity.</li>
<li>Benefits include extended health and dental coverage, retirement savings plans, monthly meal allowance, 23 paid days off, 13 paid flexible holidays, and other benefits in accordance with Canadian employment standards.</li>
</ul>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>senior</experiencelevel>
      <workarrangement>hybrid</workarrangement>
      <salaryrange>CAD 103,321.00–147,602.00</salaryrange>
      <skills>Scala, Java, Scio, Apache Beam, Spark, Flink, Google Cloud Platform, backend services, data pipelines, workflow orchestration, financial, accounting, regulated data domains, Computer Science</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>Spotify</employername>
      <employerlogo>https://logos.yubhub.co/spotify.com.png</employerlogo>
      <employerdescription>Spotify is a music streaming service with a large user base. The Financial Engineering organisation builds infrastructure and tooling for finance teams.</employerdescription>
      <employerwebsite>https://www.spotify.com</employerwebsite>
      <compensationcurrency>CAD</compensationcurrency>
      <compensationmin>103321</compensationmin>
      <compensationmax>147602</compensationmax>
      <compensationinterval>yearly</compensationinterval>
      <applyto>https://jobs.lever.co/spotify/308d127c-e765-4895-9264-3765ddbfc620?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>Toronto, ON</location>
      <city>Toronto</city>
      <state>ON</state>
      <postalcode></postalcode>
      <country>CA</country>
      <postedate>2026-07-11</postedate>
    </job>
    <job>
      <externalid>2d283414-103</externalid>
      <title>Retail Platform Principal Engineer</title>
      <description><![CDATA[<p>We are currently seeking an experienced professional to join our team in the role of Retail Platform Principal Engineer in WPS Technology.</p>
<p>Responsibilities:</p>
<ul>
<li>Provide strategic technical leadership across engineering projects, ensuring alignment with business goals, architecture principles, and regulatory/compliance requirements.</li>
<li>Architect and design scalable, resilient, secure microservices and event-driven systems to support high-throughput financial services workloads.</li>
<li>Lead adoption and governance of engineering standards, patterns, and best practices (API design, domain-driven design, CI/CD).</li>
<li>Hands-on development and code review in core platforms; set standards for software craftsmanship across Modern Tech stacks and cloud-native services.</li>
<li>Mentor and coach senior engineers and engineering leads; run regular technical deep dives, architecture reviews, and cross-team workshops.</li>
<li>Lead and facilitate blameless post-mortems and structured root cause analysis (RCA), driving remediation plans and lessons learned into the development lifecycle.</li>
<li>Own the Tech Risk Control &#39;book of work&#39;: maintain backlog of control remediation tasks, coordinate with Risk/Compliance to define mitigations, track progress, and ensure audit readiness.</li>
<li>Drive the integration of AI/ML capabilities into production systems with appropriate controls for governance, model risk, explainability, and data privacy.</li>
<li>Collaborate with product, security, ops, and data teams to prioritise technical debt, reliability improvements, and platform investments.</li>
<li>Represent engineering in stakeholder forums, technical governance boards, and external partner engagements.</li>
</ul>
<p>What you will be doing:</p>
<ul>
<li>Act as the technical owner for core platform domains, driving architecture, design, and delivery across backend, data streaming, and cloud infrastructure; focus deeply on 2-3 areas (e.g., Java/Spring, Kafka/Stream Processing, Cloud/Kubernetes).</li>
<li>Lead targeted technical deep-dives (2-3 per quarter) to resolve architectural trade-offs, eliminate chronic pain points, and transfer knowledge across teams.</li>
<li>Own and deliver the Tech Risk Control book-of-work items for your area: define remediations, coordinate cross-functional delivery, produce evidence for audits, and close out controls.</li>
<li>Run structured root-cause analysis and blameless post-mortems for major incidents, translate findings into prioritised fixes, and ensure follow-through.</li>
<li>Provide hands-on technical direction: participate in design reviews, code reviews, prototyping, and proof-of-concepts to validate architecture decisions.</li>
<li>Mentor senior engineers and engineering leads; raise team capability through coaching, workshops, and documented standards/patterns.</li>
<li>Collaborate with product, security, data, and ops to align technical strategy with business objectives, regulatory requirements, and non-functional goals (scalability, resilience, security).</li>
<li>Drive adoption of cloud-native and DevOps practices (CI/CD, observability) and ensure production readiness for new services and AI/ML components.</li>
</ul>
<p>Qualifications Certifications &amp; Education:</p>
<ul>
<li>University degree in Computer Science, Engineering, or related discipline; MSc/PhD preferred.</li>
<li>Cloud certification(s) strongly preferred: AWS Certified Solutions Architect Professional or DevOps Engineer, Microsoft Certified: Azure Solutions Architect, or Google Professional Cloud Architect.</li>
<li>Desirable: Kubernetes (CKA/CKAD), security or data governance certificates.</li>
</ul>
<p>Preferred Qualifications:</p>
<ul>
<li>Experience in financial services or regulated industries with demonstrable knowledge of regulatory controls, audit preparation, and model risk governance.</li>
<li>Contributions to open source projects or academic/industry research; speaker at technical conferences is a plus.</li>
<li>Proven ability to lead cross-functional engineering teams through technological transformations, and to operationalise AI safely at scale.</li>
</ul>
<p>Deliverables &amp; Success Measures:</p>
<ul>
<li>Architecture and platform designs that meet non-functional requirements (scalability, resiliency, security).</li>
<li>Successful delivery of prioritised Tech Risk Control book-of-work items and remediation actions.</li>
<li>Reduction in incident mean-time-to-detect and mean-time-to-recover through improved observability, RCA, and runbooks.</li>
<li>Mature AI/ML deployment pipelines with documented governance, monitoring, and model lifecycle controls.</li>
<li>High team engagement, evidence of knowledge transfer via deep dive sessions, and measurable uplift in engineering capability.</li>
</ul>
<p>Technical Skills &amp; Experience:</p>
<ul>
<li>Strong Java expertise (Java 8/11/17/21/25), demonstrated through large-scale, production systems.</li>
<li>Deep experience with Spring ecosystem: Spring Boot, Spring Cloud, Spring Security, Spring Data, and related tooling.</li>
<li>Microservices and cloud-native architectures: containerisation (Docker), orchestration (Kubernetes), service mesh (Istio/Linkerd), and 12-factor app principles.</li>
<li>Event streaming and messaging: Apache Kafka (Kafka Streams, ksqlDB), Confluent tooling, and best practices for exactly-once semantics and fault-tolerant processing.</li>
<li>Data platform familiarity: batch/stream processing (Spark, Flink), data lakehouse concepts (Delta Lake), and data governance tools (e.g., Apache Atlas, Apache Falcon, or equivalent).</li>
<li>AI/ML/LLM productionisation: hands-on with MLOps and Model frameworks (LangGraph, Agent Development Kit, Microsoft Agent Framework); experience deploying and operating LLMs (e.g., Falcon family or other LLMs), embeddings, RAG architectures, and frameworks (LangChain, LlamaIndex).</li>
<li>Observability, reliability, and performance engineering: Prometheus/Grafana, ELK/EFK, distributed tracing (Jaeger/Zipkin), profiling, and tuning high-throughput systems.</li>
<li>DevOps, CI/CD: Terraform, Ansible, G3, and SonarQube.</li>
<li>Cloud platforms: proven experience on at least one major public cloud (AWS, Azure, or GCP) in production-scale environments.</li>
<li>Security, compliance &amp; risk controls: secure coding practices, threat modelling, encryption, key management, and OWASP mitigation patterns.</li>
</ul>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>senior</experiencelevel>
      <workarrangement></workarrangement>
      <salaryrange></salaryrange>
      <skills>Java, Spring, Kubernetes, Docker, Apache Kafka, Confluent, Spark, Flink, Delta Lake, Apache Atlas, LangGraph, Agent Development Kit, Microsoft Agent Framework, LangChain, LlamaIndex, Prometheus, Grafana, ELK/EFK, Jaeger, Zipkin, Terraform, Ansible, SonarQube, AWS, Azure, GCP, Cloud certification, Kubernetes certification, Security certification, Data governance certification</skills>
      <category>Engineering</category>
      <industry>Finance</industry>
      <employername>HSBC</employername>
      <employerlogo>https://logos.yubhub.co/portal.careers.hsbc.com.png</employerlogo>
      <employerdescription>HSBC is a global banking and financial services organisation.</employerdescription>
      <employerwebsite>https://portal.careers.hsbc.com</employerwebsite>
      <compensationcurrency></compensationcurrency>
      <compensationmin></compensationmin>
      <compensationmax></compensationmax>
      <compensationinterval></compensationinterval>
      <applyto>https://portal.careers.hsbc.com/careers/job/563774608800937?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>Guangzhou, Guangdong</location>
      <city>Guangzhou</city>
      <state>Guangdong</state>
      <postalcode></postalcode>
      <country>CN</country>
      <postedate>2026-07-09</postedate>
    </job>
    <job>
      <externalid>cba34f88-64a</externalid>
      <title>Senior Software Engineer, Data Enablement Platform</title>
      <description><![CDATA[<p>Brex is the intelligent finance platform that enables companies to spend smarter and move faster in more than 200 markets.</p>
<p>By combining global corporate cards and banking with intuitive spend management, bill pay, and travel software, Brex enables founders and finance teams to accelerate operations, gain real-time visibility, and control spend effortlessly.</p>
<p>The Data Enablement Platform team owns and operates the systems required to deliver data to clients, and partners with them to create impactful data-backed products.</p>
<p>As a DEP team member, you will help build Brex&#39;s Data Infrastructure and work across the entire data stack to help product teams build data-backed products.</p>
<p>Responsibilities:</p>
<ul>
<li>Build, improve, and maintain Brex&#39;s data platform and infrastructure: data warehouse &amp; analytics, data streaming, replication and materialization, data orchestration, data observability, access &amp; governance.</li>
<li>Partner with Brex&#39;s product teams, data science and analytics teams to help them launch superior products backed by data insights.</li>
<li>Level up the data competency at Brex through partnership and education.</li>
<li>Invest in building resilient data architectures, participate in on-call rotations, and continuously improve operational efficiency.</li>
</ul>
<p>Requirements:</p>
<ul>
<li>5+ years of professional experience in a data infra or data platform role</li>
<li>Experience with one or more of Snowflake, Flink, Airflow, dbt, CDC</li>
<li>Fluency in Kotlin or Kotlin-like development languages such as Python</li>
<li>Experience building and maintaining large-scale modern data stack</li>
<li>Experience with streaming infra like Kafka</li>
<li>Experience in backend engineering or full-stack development</li>
<li>Strong communication and collaboration skills, especially working across team boundaries and building XFN relationships</li>
</ul>
<p>The expected salary range for this role is $192,000-$240,000 USD.</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>senior</experiencelevel>
      <workarrangement>hybrid</workarrangement>
      <salaryrange>$192,000-$240,000 USD</salaryrange>
      <skills>Snowflake, Flink, Airflow, dbt, CDC, Kotlin, Python, Kafka, Experience collaborating with product teams or building full-stack data-backed products, Experience with data architecture and data strategy</skills>
      <category>Engineering</category>
      <industry>Finance</industry>
      <employername>Brex</employername>
      <employerlogo>https://logos.yubhub.co/brex.com.png</employerlogo>
      <employerdescription>Brex is an intelligent finance platform that enables companies to spend smarter and move faster in over 200 markets.</employerdescription>
      <employerwebsite>https://brex.com</employerwebsite>
      <compensationcurrency>USD</compensationcurrency>
      <compensationmin>192000</compensationmin>
      <compensationmax>240000</compensationmax>
      <compensationinterval>yearly</compensationinterval>
      <applyto>https://job-boards.greenhouse.io/brex/jobs/8605074002?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>Seattle, Washington</location>
      <city>Seattle</city>
      <state>Washington</state>
      <postalcode></postalcode>
      <country>US</country>
      <postedate>2026-06-24</postedate>
    </job>
    <job>
      <externalid>e21c120f-11b</externalid>
      <title>Senior Software Engineer, Data Enablement Platform</title>
      <description><![CDATA[<p>Brex is the intelligent finance platform that enables companies to spend smarter and move faster in more than 200 markets.</p>
<p>By combining global corporate cards and banking with intuitive spend management, bill pay, and travel software, Brex enables founders and finance teams to accelerate operations, gain real-time visibility, and control spend effortlessly.</p>
<p>The Data Enablement Platform team owns and operates the systems required to deliver data to clients, and partners with them to create impactful data-backed products.</p>
<p>As a DEP team member, you will help build Brex&#39;s Data Infrastructure and work across the entire data stack to help product teams build data-backed products.</p>
<p>Responsibilities:</p>
<ul>
<li>Build, improve, and maintain Brex&#39;s data platform and infrastructure: data warehouse &amp; analytics, data streaming, replication and materialization, data orchestration, data observability, access &amp; governance.</li>
<li>Partner with Brex&#39;s product teams, data science and analytics teams to help them launch superior products backed by data insights.</li>
<li>Level up the data competency at Brex through partnership and education.</li>
<li>Invest in building resilient data architectures, participate in on-call rotations, and continuously improve operational efficiency.</li>
</ul>
<p>Requirements:</p>
<ul>
<li>5+ years of professional experience in a data infra or data platform role.</li>
<li>Experience with one or more of Snowflake, Flink, Airflow, dbt, CDC.</li>
<li>Fluency in Kotlin or Kotlin-like development languages such as Python.</li>
<li>Experience building and maintaining large-scale modern data stacks.</li>
<li>Experience with streaming infra like Kafka.</li>
<li>Experience in backend engineering or full-stack development.</li>
<li>Strong communication and collaboration skills, especially working across team boundaries and building XFN relationships.</li>
</ul>
<p>Bonus points:</p>
<ul>
<li>Experience collaborating with product teams or building full-stack data-backed products.</li>
<li>Experience with data architecture and data strategy.</li>
</ul>
<p>The expected salary range for this role is $192,000-$240,000 USD.</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>senior</experiencelevel>
      <workarrangement>hybrid</workarrangement>
      <salaryrange>$192,000-$240,000 USD</salaryrange>
      <skills>Snowflake, Flink, Airflow, dbt, CDC, Kotlin, Python, Kafka, data architecture, data strategy, full-stack development</skills>
      <category>Engineering</category>
      <industry>Finance</industry>
      <employername>Brex</employername>
      <employerlogo>https://logos.yubhub.co/brex.com.png</employerlogo>
      <employerdescription>Brex is an intelligent finance platform that enables companies to spend smarter and move faster in over 200 markets.</employerdescription>
      <employerwebsite>https://brex.com</employerwebsite>
      <compensationcurrency>USD</compensationcurrency>
      <compensationmin>192000</compensationmin>
      <compensationmax>240000</compensationmax>
      <compensationinterval>yearly</compensationinterval>
      <applyto>https://job-boards.greenhouse.io/brex/jobs/8605062002?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>New York, New York</location>
      <city>New York</city>
      <state>New York</state>
      <postalcode></postalcode>
      <country>US</country>
      <postedate>2026-06-24</postedate>
    </job>
    <job>
      <externalid>9048fd9a-cf4</externalid>
      <title>Senior Software Engineer, Data Enablement Platform</title>
      <description><![CDATA[<p>We&#39;re looking for a Senior Software Engineer to join our Data Enablement Platform team at Brex. As a DEP team member, you will help build Brex&#39;s Data Infrastructure and work across the entire data stack to help product teams build data-backed products.</p>
<p>The Data Enablement Platform team owns and operates the systems required to deliver data to our clients and partners with them to create impactful data-backed products. You will be responsible for building, improving, and maintaining Brex&#39;s data platform and infrastructure, including data warehouse &amp; analytics, data streaming, replication and materialization, data orchestration, data observability, access &amp; governance.</p>
<p>You will partner with Brex&#39;s product teams, data science, and analytics teams to help them launch superior products backed by data insights. Additionally, you will level up the data competency at Brex through partnership and education, invest in building resilient data architectures, participate in on-call rotations, and continuously improve our operational efficiency.</p>
<p>The expected salary range for this role is $192,000-$240,000 USD. The starting base pay will depend on a number of factors, including the candidate&#39;s location, skills, experience, market demands, and internal pay parity.</p>
<p>Responsibilities:</p>
<ul>
<li>Build, improve, and maintain Brex&#39;s data platform and infrastructure</li>
<li>Partner with product teams, data science, and analytics teams to launch data-backed products</li>
<li>Level up data competency at Brex through partnership and education</li>
<li>Invest in building resilient data architectures and improve operational efficiency</li>
</ul>
<p>Requirements:</p>
<ul>
<li>5+ years of professional experience in a data infra or data platform role</li>
<li>Experience with one or more of Snowflake, Flink, Airflow, dbt, CDC</li>
<li>Fluency in Kotlin or Kotlin-like development languages such as Python</li>
<li>Experience building and maintaining large-scale modern data stacks</li>
<li>Experience with streaming infra like Kafka</li>
<li>Experience in backend engineering or full-stack development</li>
<li>Strong communication and collaboration skills</li>
</ul>
<p>Bonus points:</p>
<ul>
<li>Experience collaborating with product teams or building full-stack data-backed products</li>
<li>Experience with data architecture and data strategy</li>
</ul>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>senior</experiencelevel>
      <workarrangement>hybrid</workarrangement>
      <salaryrange>$192,000-$240,000 USD</salaryrange>
      <skills>Snowflake, Flink, Airflow, dbt, CDC, Kotlin, Python, Kafka, data architecture, data strategy, full-stack development</skills>
      <category>Engineering</category>
      <industry>Finance</industry>
      <employername>Brex</employername>
      <employerlogo>https://logos.yubhub.co/brex.com.png</employerlogo>
      <employerdescription>Brex is an intelligent finance platform that enables companies to spend smarter and move faster in over 200 markets. The platform combines global corporate cards and banking with spend management, bill pay, and travel software.</employerdescription>
      <employerwebsite>https://brex.com</employerwebsite>
      <compensationcurrency>USD</compensationcurrency>
      <compensationmin>192000</compensationmin>
      <compensationmax>240000</compensationmax>
      <compensationinterval>yearly</compensationinterval>
      <applyto>https://job-boards.greenhouse.io/brex/jobs/8605070002?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>San Francisco, California</location>
      <city>San Francisco</city>
      <state>California</state>
      <postalcode></postalcode>
      <country>US</country>
      <postedate>2026-06-24</postedate>
    </job>
    <job>
      <externalid>fe747376-33d</externalid>
      <title>Software Engineer, Data Orchestration</title>
      <description><![CDATA[<p><strong>Job Overview</strong></p>
<p>As a Software Engineer on the Data Orchestration team at Stripe, you will design, build, and maintain innovative data platform products. You will work on a wide range of technologies including Airflow, Spark, SQL, Kafka, Flink, Hive MetaStore, Trino, Pinot, Python, Java, Scala, S3, and Iceberg.</p>
<p><strong>Responsibilities</strong></p>
<ul>
<li>Design, build, and maintain key Data Platform products with a focus on usability, reliability, security, and efficiency.</li>
<li>Create ergonomic APIs and abstractions for internal Stripe users, enhancing their experience and that of millions of Stripe customers.</li>
<li>Ensure operational excellence and high availability of the Data Orchestration platform across batch workloads.</li>
<li>Collaborate with high-visibility teams to support their initiatives while building a robust platform benefiting all of Stripe.</li>
<li>Plan for Stripe&#39;s infrastructure growth by unblocking and supporting internal partners.</li>
</ul>
<p><strong>Requirements</strong></p>
<ul>
<li>8+ years of professional experience writing high-quality production-level code with an interest in Data Infrastructure.</li>
<li>Experience operating or enabling large-scale, high-availability data pipelines.</li>
<li>Expertise in Spark, Flink, Airflow, Python, Java, SQL, and API design is a plus.</li>
<li>Experience developing, maintaining, and debugging distributed systems built with open-source tools.</li>
<li>Strong collaboration and communication skills.</li>
</ul>
<p><strong>Preferred Qualifications</strong></p>
<ul>
<li>Experience writing production-level code in Scala, Spark, Flink, Airflow, Python, Java, and SQL.</li>
<li>Experience designing APIs or building developer platforms.</li>
<li>Experience optimizing the end-to-end performance of distributed systems.</li>
</ul>
<p><strong>Work Environment</strong></p>
<p>Office-assigned Stripes spend at least 50% of their time in their local office or with users.</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>senior</experiencelevel>
      <workarrangement>hybrid</workarrangement>
      <salaryrange></salaryrange>
      <skills>Airflow, Spark, SQL, Kafka, Flink, Python, Java, Scala, API design, Distributed systems, Distributed systems performance optimization</skills>
      <category>Engineering</category>
      <industry>Finance</industry>
      <employername>Stripe</employername>
      <employerlogo>https://logos.yubhub.co/stripe.com.png</employerlogo>
      <employerdescription>Stripe is a financial infrastructure platform for businesses. It serves millions of companies worldwide.</employerdescription>
      <employerwebsite>https://stripe.com</employerwebsite>
      <compensationcurrency></compensationcurrency>
      <compensationmin></compensationmin>
      <compensationmax></compensationmax>
      <compensationinterval></compensationinterval>
      <applyto>https://job-boards.greenhouse.io/stripe/jobs/7230670?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>N/A</location>
      <city>N/A</city>
      <state></state>
      <postalcode></postalcode>
      <country></country>
      <postedate>2026-06-18</postedate>
    </job>
    <job>
      <externalid>52b2edcb-c52</externalid>
      <title>Sr. Staff Software Engineer, Data Product Platform</title>
      <description><![CDATA[<p>We are seeking a Sr. Staff Software Engineer to provide technical leadership and direction across the Data Product Platform at Pinterest. The Data Product Platform is mission-critical to accelerating data-driven decision-making at Pinterest, which is built on the foundation of hundreds of thousands of tables and an exadata-scale data warehouse.</p>
<p>The successful candidate will spearhead the definition and design of Pinterest&#39;s data warehouse architecture, storage, governance, and usage. They will set the strategic direction for world-class data analytics tools, define and drive the adoption of comprehensive data governance policies and tooling, and lead ambiguous, highly challenging, and cross-functional initiatives across the data ecosystem.</p>
<p>Responsibilities:</p>
<ul>
<li>Spearhead the definition and design of Pinterest&#39;s data warehouse architecture, storage, governance, and usage.</li>
<li>Set the strategic direction for world-class data analytics tools, including experimentation platforms, ad hoc analysis systems, data visualization, data pipeline authoring, and AI-assisted data analytics capabilities.</li>
<li>Define, design, and drive the adoption of comprehensive data governance policies and tooling to ensure data quality, promote responsible data handling, and maximize efficiency.</li>
<li>Lead ambiguous, highly challenging, and cross-functional initiatives across the data ecosystem, expertly managing trade-offs among unique requirements and constraints to deliver a cohesive user experience.</li>
<li>Drive measurable adoption and significant business impact across the entire product portfolio through platform and tooling initiatives.</li>
<li>Mentor and elevate the technical bar for engineers, providing leadership on complex technical design, execution strategies, and critical trade-off decisions.</li>
<li>Collaborate with stakeholders across Product, Data Science, and Engineering teams to understand user needs, drive alignment, and drive progress.</li>
</ul>
<p>Requirements:</p>
<ul>
<li>Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent experience.</li>
<li>12+ years of relevant industry experience with large-scale data warehouses, data tools, and platforms.</li>
<li>5+ years of experience building data warehouses and tools around the data ecosystem with technologies such as Spark, Trino, Flink, Airflow, Querybook, Superset, DataHub, etc.</li>
<li>Ability to work with cross-functional partners across multiple organizations.</li>
<li>Hands-on experience building tools and data pipelines leveraging AI coding tools, e.g., Cursor, Claude Code, Codex, etc.</li>
<li>Hands-on experience building AI tools and platforms to accelerate data pipeline authoring, data analytics, and engineering productivity.</li>
</ul>
<p>Benefits:</p>
<ul>
<li>Salary range: $208,592-$429,454 USD</li>
<li>Equity eligibility</li>
<li>Flexible working model (PinFlex)</li>
<li>Opportunity to work in-office 1-2 times/quarter</li>
</ul>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>senior</experiencelevel>
      <workarrangement>hybrid</workarrangement>
      <salaryrange>$208,592-$429,454 USD</salaryrange>
      <skills>Spark, Trino, Flink, Airflow, Querybook, Superset, DataHub, Cursor, Claude Code, Codex</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>Pinterest</employername>
      <employerlogo>https://logos.yubhub.co/pinterest.com.png</employerlogo>
      <employerdescription>Pinterest is a platform where people find creative ideas, dream about new possibilities, and plan for memories. The company has millions of users worldwide.</employerdescription>
      <employerwebsite>https://pinterest.com</employerwebsite>
      <compensationcurrency>USD</compensationcurrency>
      <compensationmin>208592</compensationmin>
      <compensationmax>429454</compensationmax>
      <compensationinterval>yearly</compensationinterval>
      <applyto>https://job-boards.greenhouse.io/pinterest/jobs/7746946?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>San Francisco, CA</location>
      <city>San Francisco</city>
      <state>CA</state>
      <postalcode></postalcode>
      <country>US</country>
      <postedate>2026-06-18</postedate>
    </job>
    <job>
      <externalid>cbd44fe8-e23</externalid>
      <title>Software Engineer, Product Security Data Platforms</title>
      <description><![CDATA[<p><strong>About the Team</strong></p>
<p>The Product Security Data Platforms team is a newly established engineering team within Stripe Security. Our mission is to build the foundational infrastructure that provides our users with unprecedented visibility into the security posture of their Stripe integration. While Stripe is renowned for industry-leading payment protection, we are expanding our focus to provide a comprehensive security telemetry platform that helps businesses protect their entire digital ecosystem on Stripe.</p>
<p>As a founding member of this team, you will be architecting a large-scale customer-facing security data pipeline and presentation layer. Much like modern security observability platforms and data lakes that have transformed cloud infrastructure, we are building an API-first service that transforms massive streams of behavioral data into actionable security intelligence. This team operates at the intersection of high-throughput data engineering and cybersecurity, creating the systems that will allow the world’s most sophisticated companies to monitor, detect, and respond to threats in real-time.</p>
<p><strong>Responsibilities</strong></p>
<ul>
<li>Architect Scalable Foundations: Design and implement a highly available, low-latency pipeline capable of processing and augmenting millions of events per second into structured security telemetry.</li>
<li>Build API-First Products: Develop the core services and streaming APIs that enable enterprise customers to seamlessly ingest security signals into their own internal security operations centers and analytics tools.</li>
<li>Engineering Security Signals: Partner with security researchers and threat detection experts to build the logic that identifies anomalous behavior and surfaces high-fidelity security insights.</li>
<li>Define Technical Strategy: Lead the technical roadmap for the platform, making critical decisions on data modeling, storage strategies, and the abstraction layers that will support future security products.</li>
<li>Drive Engineering Excellence: As a senior leader, you will set the bar for code quality, system resilience, and operational maturity for a product that requires 99.99%+ availability.</li>
<li>Cross-functional Collaboration: Work closely with Stripe’s core platform and data teams to leverage global infrastructure while ensuring security data remains isolated and protected.</li>
</ul>
<p><strong>Requirements</strong></p>
<ul>
<li>8+ years of professional software development experience, with a history of shipping and maintaining complex, large-scale systems.</li>
<li>Expertise in Data Engineering &amp; Distributed Systems: Deep experience building and operating high-throughput data pipelines (e.g., Kafka, Flink, Spark, or similar streaming technologies).</li>
<li>Strong Programming Fundamentals: Significant experience in languages like Java, C++, or Rust.</li>
<li>System Design Leadership: A proven track record of designing robust, scalable architectures and leading cross-team technical initiatives from conception to launch.</li>
<li>Operational Mindset: Experience maintaining mission-critical services with high availability requirements, on-call rotation, and a strong focus on observability and debugging.</li>
<li>Communication &amp; Collaboration: Excellent technical writing skills for drafting design documents and the ability to mentor other engineers while collaborating with non-technical stakeholders.</li>
</ul>
<p><strong>Preferred Qualifications</strong></p>
<ul>
<li>Security Domain Knowledge: Prior experience building security analytics, threat detection systems, or observability platforms.</li>
<li>Platform Engineering: Experience building &quot;as-a-service&quot; infrastructure where the primary users are other engineers or external developers.</li>
<li>Cloud Native Infrastructure: Prior experience with AWS, Kubernetes, and infrastructure-as-code (Terraform).</li>
<li>Product Intuition: A desire to work on &quot;0-to-1&quot; initiatives where you help define the product requirements and user experience alongside engineering.</li>
<li>Front-end / Full Stack Experience: Prior React / TypeScript frontend experience is helpful</li>
</ul>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>senior</experiencelevel>
      <workarrangement></workarrangement>
      <salaryrange></salaryrange>
      <skills>Data Engineering, Distributed Systems, Java, C++, Rust, Kafka, Flink, Spark, System Design, Cloud Infrastructure, Security Analytics, Threat Detection, Observability Platforms, Platform Engineering, AWS, Kubernetes, Terraform, React, TypeScript</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>Stripe</employername>
      <employerlogo>https://logos.yubhub.co/stripe.com.png</employerlogo>
      <employerdescription>Stripe is a technology company that provides online payment processing systems.</employerdescription>
      <employerwebsite>https://stripe.com</employerwebsite>
      <compensationcurrency></compensationcurrency>
      <compensationmin></compensationmin>
      <compensationmax></compensationmax>
      <compensationinterval></compensationinterval>
      <applyto>https://job-boards.greenhouse.io/stripe/jobs/7761694?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>Seattle</location>
      <city>Seattle</city>
      <state></state>
      <postalcode></postalcode>
      <country></country>
      <postedate>2026-06-18</postedate>
    </job>
    <job>
      <externalid>6d70ce20-c16</externalid>
      <title>Staff Engineer — Data Platform</title>
      <description><![CDATA[<p>At Yuno, we are building the payment infrastructure that allows all companies to participate in the global market. Founded by seasoned experts from the payments and tech industries , including the team behind Rappi, one of Latin America&#39;s most ambitious tech companies , our technology provides access to leading payment capabilities, enabling companies to engage customers confidently and maintain global operations through seamless integrations.</p>
<p>We are orchestrating a high-performing data team that works with pace and enthusiasm! Yuno moves money across borders for companies that can&#39;t afford for payments to fail. Our data platform is what makes that visible , to our product teams, our clients, and ourselves.</p>
<p>You will play a pivotal role within the Data team that powers Yuno and its payment platform, while helping co-design and implement an architecture that enables the entire organization to operate on reliable, fast, and trustworthy data.</p>
<p><strong>Your Contribution Will Be</strong></p>
<p>The stack is modern: StarRocks as our primary analytical layer, Flink for processing, DBT for transformation, Airflow for orchestration and various tooling for surfacing insights. The hard work of making it super reliable is still in front of us , and that&#39;s exactly why this role exists.</p>
<p>Technical Leadership</p>
<ul>
<li>Define architecture within the data platform, structure and deliver projects and initiatives end-to-end.</li>
<li>Act as a technical reference point for the Data team, setting quality standards, testing, observability, data modeling, and documentation.</li>
<li>Lead the design and implementation of scalable, low-latency data pipelines that process high-volume payment transactions in real time.</li>
<li>Champion an AI-first engineering culture, establishing standards for AI-assisted development, automated data quality testing, and LLM-powered data workflows.</li>
</ul>
<p>Hands-On Engineering</p>
<ul>
<li>Design and build data pipelines for large volumes of payment data that are performant, reliable, and correct , not just fast.</li>
<li>Design scalable data models that support business-critical use cases: fraud detection, revenue analytics, payment success rate optimization, regulatory reporting.</li>
<li>Own platform reliability , SLAs, data quality, alerting, and incident response for data services.</li>
<li>Ensure secure data handling practices aligned with PCI-DSS, GDPR, and other compliance frameworks relevant to the payments industry.</li>
</ul>
<p>Cross-Functional Impact</p>
<ul>
<li>Partner with Product, AI, and Finance teams to translate business needs into scalable data solutions.</li>
<li>Contribute to the roadmap of the data platform and proactively identify opportunities to unlock new business value through data.</li>
<li>Mentor senior and mid-level engineers, raising the technical bar across the team through code reviews, design reviews, and knowledge-sharing sessions.</li>
<li>Collaborate with Data Consumers (analysts, data scientists, product managers) to ensure data products are reliable, well-documented, and fit for purpose.</li>
</ul>
<p><strong>Skills You Need</strong></p>
<p><strong>Minimum Qualifications</strong></p>
<ul>
<li>8+ years of experience in data engineering, software engineering, or a related field, with at least 2 years operating at a staff or principal level.</li>
<li>Deep expertise in designing and building large-scale data platforms , streaming, batch, or hybrid architectures.</li>
<li>Hands-on experience with Spark, Flink, Kafka, StarRocks, or equivalent.</li>
<li>Strong Python and SQL skills; comfort working across multiple languages and paradigms.</li>
<li>Solid understanding of data modeling techniques: dimensional modeling, Data Vault, or lakehouse patterns.</li>
<li>Experience with cloud data infrastructure (AWS, GCP, or Azure), including managed services for storage, compute, and orchestration.</li>
<li>Strong grasp of data quality, observability, and governance principles.</li>
<li>Proven ability to set standards and lead technical initiatives across multiple teams without direct authority.</li>
<li>Professional proficiency in English , written and spoken.</li>
</ul>
<p><strong>Preferred Qualifications</strong></p>
<ul>
<li>Experience in the payments or fintech industry.</li>
<li>Familiarity with dbt, Great Expectations, or similar.</li>
<li>Experience with event-driven services and data mesh approaches.</li>
<li>Exposure to ML platform design or feature store infrastructure.</li>
</ul>
<p><strong>What We Offer at Yuno</strong></p>
<ul>
<li>Competitive Compensation.</li>
<li>Remote Work – You can work from everywhere!</li>
<li>Home Office Bonus – A one-time allowance to help you create your ideal home office.</li>
<li>Work Equipment.</li>
<li>Stock Options.</li>
<li>Health Plan wherever you are.</li>
<li>Flexible Days Off.</li>
<li>Language, Professional, and Personal Growth courses.</li>
</ul>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>staff</experiencelevel>
      <workarrangement>remote</workarrangement>
      <salaryrange></salaryrange>
      <skills>data engineering, software engineering, StarRocks, Flink, DBT, Airflow, Python, SQL, data modeling, cloud data infrastructure, data quality, observability, governance, payments, fintech, Great Expectations, event-driven services, data mesh, ML platform design, feature store infrastructure</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>Yuno</employername>
      <employerlogo>https://logos.yubhub.co/yuno.com.png</employerlogo>
      <employerdescription>Yuno is building payment infrastructure for global companies. The company was founded by experts from the payments and tech industries.</employerdescription>
      <employerwebsite>https://yuno.com</employerwebsite>
      <compensationcurrency></compensationcurrency>
      <compensationmin></compensationmin>
      <compensationmax></compensationmax>
      <compensationinterval></compensationinterval>
      <applyto>https://jobs.lever.co/yuno/0a1e498f-b0af-4955-bf94-e23b48faec00?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>London</location>
      <city>London</city>
      <state></state>
      <postalcode></postalcode>
      <country></country>
      <postedate>2026-06-18</postedate>
    </job>
    <job>
      <externalid>7d1839cc-728</externalid>
      <title>Staff Software Engineer, Communication &amp; Connectivity</title>
      <description><![CDATA[<p>We connect Airbnb&#39;s community with the right information, in the right place, at the right time. As a Staff Software Engineer in the Communication and Connectivity (CnC) organization, you will lead key initiatives to design and build large-scale, distributed data systems - both batch and real-time processing. These systems will power machine learning models and unlock new product features.</p>
<p>Responsibilities:</p>
<ul>
<li>Shape the team&#39;s long-term vision and roadmap in close collaboration with cross-functional partners across Airbnb</li>
<li>Build strong relationships with partner engineering teams to drive aligned and impactful outcomes</li>
<li>Design, develop, and maintain reliable, scalable data pipelines - both batch and real-time - that collect, process, and serve data from diverse sources across Airbnb</li>
<li>Implement robust offline and online feature building processes to enable faster production of ML products</li>
<li>Architect and build ML infra and optimize for performance, scalability, and cost-effectiveness</li>
<li>Mentor and develop engineers on the team, while also contributing to and influencing the broader data engineering community at Airbnb</li>
</ul>
<p>Requirements:</p>
<ul>
<li>9+ years of relevant industry experience with a Bachelor&#39;s and/or Master&#39;s degree in CS/EE, or equivalent experience, or 6+ years of experience with a PhD</li>
<li>Strong CS fundamentals, and knowledge of architecture and common design patterns</li>
<li>Experience running data processing pipelines in production using distributed data processing frameworks like Apache Spark or Flink</li>
<li>Experience collaborating with client, backend, ml, analytics teams, product and business partners</li>
<li>Effectively work across team boundaries to establish overarching data architecture, data flow, and provide guidance to individual teams</li>
<li>Experience working on/with end-to-end Machine Learning products</li>
<li>Excellent communication skills, both written and verbal</li>
</ul>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>staff</experiencelevel>
      <workarrangement>remote</workarrangement>
      <salaryrange>$204,000-$255,000 USD</salaryrange>
      <skills>Apache Spark, Flink, Machine Learning, Data Engineering, Cloud Computing</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>Airbnb</employername>
      <employerlogo>https://logos.yubhub.co/airbnb.com.png</employerlogo>
      <employerdescription>Airbnb is a global online marketplace for short-term rentals and experiences, founded in 2007, with over 5 million hosts and 2 billion guest arrivals.</employerdescription>
      <employerwebsite>https://www.airbnb.com</employerwebsite>
      <compensationcurrency>USD</compensationcurrency>
      <compensationmin>204000</compensationmin>
      <compensationmax>255000</compensationmax>
      <compensationinterval>yearly</compensationinterval>
      <applyto>https://job-boards.greenhouse.io/airbnb/jobs/7421419?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>United States</location>
      <city>United States</city>
      <state></state>
      <postalcode></postalcode>
      <country></country>
      <postedate>2026-06-18</postedate>
    </job>
    <job>
      <externalid>96d6fe9b-093</externalid>
      <title>Staff Software Engineer, Big Data Storage</title>
      <description><![CDATA[<p>About Pinterest:</p>
<p>Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime.</p>
<p>At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work.</p>
<p>Creating a career you love? It’s Possible.</p>
<p>At Pinterest, AI isn&#39;t just a feature, it&#39;s a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think.</p>
<p>We’re looking for a Staff Software Engineer to help build the next generation of Pinterest’s big data storage platform. You’ll work on some of the most exciting big data open source technologies , especially Apache Iceberg , at exabyte scale to power the data infrastructure that helps Pinners discover and do what they love.</p>
<p>As a Staff Software Engineer, you’ll serve as a technical leader and hands-on contributor, designing and building highly scalable storage systems for Pinterest’s data lake. You’ll partner closely with teams across data, ML/AI, analytics, and infrastructure to evolve our storage and metadata management capabilities, enabling efficient, reliable, and governed access to data at massive scale.</p>
<p>What you’ll do:</p>
<ul>
<li>Design, implement, and optimize Pinterest’s exabyte-scale data lake storage platform.</li>
</ul>
<ul>
<li>Lead complex technical projects and initiatives for data lake storage and metadata management, driving them from architecture through execution.</li>
</ul>
<ul>
<li>Collaborate with stakeholders and partner teams across the organization to design storage and metadata layer technologies that unlock big data and ML/AI innovations.</li>
</ul>
<ul>
<li>Build storage capabilities that efficiently support large-scale ML/AI workloads, including high-throughput data access, schema evolution, and large-scale column backfills.</li>
</ul>
<ul>
<li>Shape the long-term technical direction for scalable, reliable, and efficient big data storage systems.</li>
</ul>
<ul>
<li>Engage with and contribute to open source communities such as Apache Iceberg, Spark, and Flink to help address Pinterest’s scaling challenges.</li>
</ul>
<p>What we’re looking for:</p>
<ul>
<li>8+ years of relevant industry experience designing and building large-scale production distributed systems.</li>
</ul>
<ul>
<li>Strong experience designing and maintaining scalable storage, metadata, or data lake infrastructure.</li>
</ul>
<ul>
<li>Experience building storage capabilities for large-scale ML/AI or analytics workloads, including high-throughput data access, schema evolution, and large-scale column backfills.</li>
</ul>
<ul>
<li>Deep knowledge with building distributed systems, data storage systems, and production infrastructure.</li>
</ul>
<ul>
<li>Experience with big data technologies such as Apache Iceberg, Spark, Flink, Presto/Trino, Hive, or similar systems.</li>
</ul>
<ul>
<li>Proficiency in programming languages like Java, Scala, or Python.</li>
</ul>
<ul>
<li>Proven ability to lead complex technical initiatives and influence architecture across teams.</li>
</ul>
<ul>
<li>Strong collaboration, communication, and problem-solving skills, with a drive for technical excellence and innovation.</li>
</ul>
<ul>
<li>Bachelor’s degree in a relevant field such as Computer Science, or equivalent experience</li>
</ul>
<p>In-Office Requirement Statement:</p>
<p>We let the type of work you do guide the collaboration style. That means we&#39;re not always working in an office, but we continue to gather for key moments of collaboration and connection.</p>
<p>This role will need to be in the office for in-person collaboration 1-2 times/quarter and therefore can be situated anywhere in the country.</p>
<p>Relocation Statement:</p>
<p>This position is not eligible for relocation assistance.</p>
<p>Visit our PinFlex page to learn more about our working model.</p>
<p>#LI-HYBRID #LI-AH2</p>
<p>At Pinterest we believe the workplace should be equitable, inclusive, and inspiring for every employee. In an effort to provide greater transparency, we are sharing the base salary range for this position. The position is also eligible for equity. Final salary is based on a number of factors including location, travel, relevant prior experience, or particular skills and expertise. Information regarding the culture at Pinterest and benefits available for this position can be found here.</p>
<p>US based applicants only</p>
<p>$177,185-$364,795 USD</p>
<p>Our Commitment to Inclusion:</p>
<p>Pinterest is an equal opportunity employer and makes employment decisions on the basis of merit. We want to have the best qualified people in every job. All qualified applicants will receive consideration for employment without regard to race, color, ancestry, national origin, religion or religious creed, sex (including pregnancy, childbirth, or related medical conditions), sexual orientation, gender, gender identity, gender expression, age, marital status, status as a protected veteran, physical or mental disability, medical condition, genetic information or characteristics (or those of a family member) or any other consideration made unlawful by applicable federal, state or local laws. We also consider qualified applicants regardless of criminal histories, consistent with legal requirements. If you require a medical or religious accommodation during the job application process, please complete this form for support.</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>staff</experiencelevel>
      <workarrangement>hybrid</workarrangement>
      <salaryrange>$177,185-$364,795 USD</salaryrange>
      <skills>Apache Iceberg, Spark, Flink, Java, Scala, Python, Distributed systems, Data storage systems, Production infrastructure</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>Pinterest</employername>
      <employerlogo>https://logos.yubhub.co/pinterest.com.png</employerlogo>
      <employerdescription>Pinterest is a social media platform that allows users to save and share images and videos. It has over 320 million monthly active users.</employerdescription>
      <employerwebsite>https://www.pinterest.com/</employerwebsite>
      <compensationcurrency>USD</compensationcurrency>
      <compensationmin>177185</compensationmin>
      <compensationmax>364795</compensationmax>
      <compensationinterval>yearly</compensationinterval>
      <applyto>https://job-boards.greenhouse.io/pinterest/jobs/7437356?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>Palo Alto, CA</location>
      <city>Palo Alto</city>
      <state>CA</state>
      <postalcode>94304</postalcode>
      <country>US</country>
      <postedate>2026-06-06</postedate>
    </job>
    <job>
      <externalid>7d8db69c-3f0</externalid>
      <title>Senior Analytics Engineer, People Data</title>
      <description><![CDATA[<p>This role is central to our People Data &amp; Analytics team, where you will be instrumental in building and maintaining the robust data infrastructure that powers our strategic insights. You&#39;ll own the full data lifecycle, from ensuring accurate ingestion and integration of diverse HR data sources, to designing, developing, and optimizing data models and pipelines. Your primary objective will be to transform raw, disparate information into clean, reliable, and analytics-ready datasets, empowering our People Analysts and business stakeholders to unlock deeper understanding of our workforce, enhance employee experience, and drive data-driven decision-making.</p>
<p>Design, build, and optimize robust ETL/ELT pipelines to reliably ingest, integrate, and transform diverse people data from various HR systems (HRIS, ATS, LMS, etc.) into our data platform.</p>
<p>Develop, maintain, and govern scalable and secure data models, schemas, and ontologies specifically for people analytics, ensuring data quality, consistency, and accessibility for downstream consumption.</p>
<p>Contribute to the strategic design, development, and evolution of our people data platform and tooling, advocating for engineering best practices, automation, and a scalable analytics ecosystem (e.g., leveraging SQLMesh, Iceberg, Flyte).</p>
<p>Partner closely with People Analysts, HR Business Partners, and other stakeholders to understand their analytical needs and translate them into robust data solutions, providing well-structured, documented, and reliable datasets.</p>
<p>Implement and monitor data quality checks, identify discrepancies, troubleshoot data issues, and ensure the reliability and integrity of people data across all systems.</p>
<p>Continuously monitor the performance of data pipelines and models, identifying bottlenecks and implementing solutions to ensure the efficiency and scalability of our people data infrastructure.</p>
<p>Create and maintain comprehensive documentation for data pipelines, models, and processes, and champion data engineering best practices (e.g., version control, testing, CI/CD) within the team.</p>
<p>Implement and enforce strict data security measures and ensure all data handling practices comply with internal policies and external regulations (e.g., GDPR, CCPA) related to employee data privacy.</p>
<p>Collaborate with broader enterprise analytics and data engineering teams to align on data architecture standards, integrate people data with other business domains, and contribute to the overall evolution of the company&#39;s data platform.</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>senior</experiencelevel>
      <workarrangement>onsite</workarrangement>
      <salaryrange>$166,000-$220,000 USD</salaryrange>
      <skills>SQL, Python, ETL/ELT, Data Models, Data Pipelines, Cloud-Based Data Platforms, Data Warehousing Solutions, Data Lake Technologies, Big Data Processing Frameworks, Advanced Data Warehousing Features, Apache Spark, Flink, Apache Iceberg, Delta Lake, Palantir Foundry, SQLMesh, Flyte, Terraform, CloudFormation, Docker, Kubernetes, Tableau, Power BI, Looker</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>Anduril Industries</employername>
      <employerlogo>https://logos.yubhub.co/anduril.com.png</employerlogo>
      <employerdescription>Anduril Industries is a defense technology company that transforms U.S. and allied military capabilities with advanced technology.</employerdescription>
      <employerwebsite>https://www.anduril.com/</employerwebsite>
      <compensationcurrency>USD</compensationcurrency>
      <compensationmin>166000</compensationmin>
      <compensationmax>220000</compensationmax>
      <compensationinterval>yearly</compensationinterval>
      <applyto>https://job-boards.greenhouse.io/andurilindustries/jobs/5147099007?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>Boston, Massachusetts, United States; Costa Mesa, California, United States</location>
      <city>Boston</city>
      <state></state>
      <postalcode></postalcode>
      <country></country>
      <postedate>2026-05-27</postedate>
    </job>
    <job>
      <externalid>af793978-ba1</externalid>
      <title>Senior Database Engineer</title>
      <description><![CDATA[<p>Join our global mission to revolutionize the way the world games. As a Senior Database Engineer at Razer, you willsparse your expertise in designing, implementing, and maintaining database systems to support our fast-paced gaming ecosystem.</p>
<p>Your primary responsibilities will include:</p>
<ul>
<li>Installing, configuring, and maintaining database management systems (DBMS) and related software</li>
<li>Monitoring and optimizing database performance to ensure efficient and reliable operations</li>
<li>Developing and implementing database backup and recovery plans, as well as disaster recovery plans</li>
<li>Managing and monitoring database security, ensuring that all sensitive data is protected</li>
<li>Troubleshooting and resolving database issues, including performance, connectivity, and data integrity issues</li>
<li>Developing and maintaining database documentation, including data models, database schemas, and system architecture diagrams</li>
</ul>
<p>In this role, you will have the opportunity to work closely with our software developers and other IT staff to design and develop database systems that meet the needs of our growing business.</p>
<p>If you are a motivated and experienced database professional looking for a new challenge, we encourage you to apply.</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>senior</experiencelevel>
      <workarrangement>onsite</workarrangement>
      <salaryrange></salaryrange>
      <skills>MariaDB, MySQL, SQL Server, Redshift, Timestream, Cassandra, Flink, Hadoop/HBase, Kafka, Redis</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>Razer</employername>
      <employerlogo>https://logos.yubhub.co/razer.com.png</employerlogo>
      <employerdescription>Razer is a leading gaming hardware and software company with a global presence across 5 continents.</employerdescription>
      <employerwebsite>https://www.razer.com/</employerwebsite>
      <compensationcurrency></compensationcurrency>
      <compensationmin></compensationmin>
      <compensationmax></compensationmax>
      <compensationinterval></compensationinterval>
      <applyto>https://razer.wd3.myworkdayjobs.com/en-US/Careers/job/Shah-Alam/Senior-Database-Engineer_JR2025006407?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>Shah Alam</location>
      <city>Shah Alam</city>
      <state></state>
      <postalcode></postalcode>
      <country></country>
      <postedate>2026-05-19</postedate>
    </job>
    <job>
      <externalid>63475652-dd3</externalid>
      <title>Data Engineer, People Innovation Labs</title>
      <description><![CDATA[<p><strong>Compensation</strong></p>
<p>$293K – $325K • Offers Equity</p>
<p>The base pay offered may vary depending on multiple individualized factors, including market location, job-related knowledge, skills, and experience. If the role is non-exempt, overtime pay will be provided consistent with applicable laws. In addition to the salary range listed above, total compensation also includes generous equity, performance-related bonus(es) for eligible employees, and the following benefits.</p>
<ul>
<li>Medical, dental, and vision insurance for you and your family, with employer contributions to Health Savings Accounts</li>
</ul>
<ul>
<li>Pre-tax accounts for Health FSA, Dependent Care FSA, and commuter expenses (parking and transit)</li>
</ul>
<ul>
<li>401(k) retirement plan with employer match</li>
</ul>
<ul>
<li>Paid parental leave (up to 24 weeks for birth parents and 20 weeks for non-birthing parents), plus paid medical and caregiver leave (up to 8 weeks)</li>
</ul>
<ul>
<li>Paid time off: flexible PTO for exempt employees and up to 15 days annually for non-exempt employees</li>
</ul>
<ul>
<li>13+ paid company holidays, and multiple paid coordinated company office closures throughout the year for focus and recharge, plus paid sick or safe time (1 hour per 30 hours worked, or more, as required by applicable state or local law)</li>
</ul>
<ul>
<li>Mental health and wellness support</li>
</ul>
<ul>
<li>Employer-paid basic life and disability coverage</li>
</ul>
<ul>
<li>Annual learning and development stipend to fuel your professional growth</li>
</ul>
<ul>
<li>Daily meals in our offices, and meal delivery credits as eligible</li>
</ul>
<ul>
<li>Relocation support for eligible employees</li>
</ul>
<ul>
<li>Additional taxable fringe benefits, such as charitable donation matching and wellness stipends, may also be provided.</li>
</ul>
<p><strong>About the Team</strong></p>
<p>At OpenAI, we’re building the connective tissue between our mission and our people. People Innovation Labs is a fast-moving engineering team embedded in the People organization, focused on rethinking how we find and retain the best talent and empower everyone to do their best work. From recruiting to culture, we’re designing systems that give our People Team a significant edge by infusing OpenAI’s models and first-principles thinking into every aspect of our work. Our projects range from greenfield 0-1 products like OpenHouse (our internal knowledge hub) to AI-powered automations and scalable recruiting tools. We’re defining the future of work at OpenAI, creating a blueprint for how AI can supercharge productivity, culture, and innovation.</p>
<p><strong>About the Role</strong></p>
<p>We are looking for a hands-on engineering manager to lead technical and product strategy and execution for People Innovation Labs’ OpenHouse pod. OpenHouse is our flagship employee-facing product, serving as a culture and communication hub and an organization-wide front door into all other aspects of People Innovation Labs’ work. The OpenHouse pod is composed of full stack product engineers who are deeply curious about culture, recruiting and people development, and want to know everything from the business strategy and metrics down through the code that gets us there. In this role, you will work with People Innovation Labs leadership to build and grow the team and OpenHouse product, innovating on how we apply LLMs along the way.</p>
<p>We’re seeking a Data Engineer to build data-intensive systems that will power People Innovation Labs’ internal products and enable the People Analytics function to do their best work. These data pipelines are crucial for our build-out of people products backed by business systems of record and for ongoing people data analytics.</p>
<p>One example of an employee-facing product you’ll help us build is OpenHouse, which serves as a culture and communication hub and an organization-wide front door into all other aspects of People Innovation Labs’ work. OpenHouse and other products in our portfolio are built by full stack product engineers who are deeply curious about culture, recruiting and people development, and want to know everything from the business strategy and metrics down through the code that gets us there. In this role, you will work with People Innovation Labs leadership and software engineers and the People Analytics team to build the data systems that enable this work.</p>
<p><strong>In this role, you will:</strong></p>
<ul>
<li>Design, build and manage people data pipelines, ensuring all data is seamlessly integrated into our Databricks warehouse.</li>
</ul>
<ul>
<li>Develop canonical datasets to track key people metrics and People Innovation Labs product metrics.</li>
</ul>
<ul>
<li>Work collaboratively with various teams, including, Data Platform, Data Science, People Analytics, and Compensation and Equity to understand their data needs and provide solutions.</li>
</ul>
<ul>
<li>Implement robust and fault-tolerant systems for data ingestion and processing.</li>
</ul>
<ul>
<li>Participate in data architecture and engineering decisions, bringing your strong experience and knowledge to bear as the primary data engineering expert on the team.</li>
</ul>
<ul>
<li>Ensure the security, integrity, and compliance of data according to industry and company standards.</li>
</ul>
<p><strong>Your background might look something like:</strong></p>
<ul>
<li>Have 3+ years of experience as a data engineer and 8+ years of any software engineering experience (including data engineering).</li>
</ul>
<ul>
<li>Proficiency in at least one programming language commonly used within Data Engineering, such as Python, Scala, or Java.</li>
</ul>
<ul>
<li>Experience with data warehousing technologies such as Databricks and Snowflake, and expertise with ETL schedulers such as Fivetran, Airflow, Dagster, Prefect, or similar.</li>
</ul>
<ul>
<li>Experience with distributed processing technologies and frameworks, such as Spark, Hadoop, Flink and distributed storage systems (e.g., HDFS, S3).</li>
</ul>
<p><strong>About OpenAI</strong></p>
<p>OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.</p>
<p>We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic.</p>
<p>For additional information, please see [OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement](https://cdn.openai.com/policies/eeo-policy-statement.pdf).</p>
<p>Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information techn</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>Full time</jobtype>
      <experiencelevel></experiencelevel>
      <workarrangement>hybrid</workarrangement>
      <salaryrange>$293K – $325K • Offers Equity</salaryrange>
      <skills>data engineering, Databricks, Snowflake, ETL schedulers, Fivetran, Airflow, Dagster, Prefect, Spark, Hadoop, Flink, distributed storage systems, HDFS, S3</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>OpenAI</employername>
      <employerlogo>https://logos.yubhub.co/openai.com.png</employerlogo>
      <employerdescription>OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity.</employerdescription>
      <employerwebsite>https://openai.com/</employerwebsite>
      <compensationcurrency>USD</compensationcurrency>
      <compensationmin>293000</compensationmin>
      <compensationmax>325000</compensationmax>
      <compensationinterval>yearly</compensationinterval>
      <applyto>https://jobs.ashbyhq.com/openai/579595e5-0b16-485e-b7c5-102fc7467def?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>San Francisco</location>
      <city>San Francisco</city>
      <state></state>
      <postalcode></postalcode>
      <country></country>
      <postedate>2026-05-13</postedate>
    </job>
    <job>
      <externalid>5a1f5eb4-c83</externalid>
      <title>Distributed Systems Engineer - Data Platform (Delivery, Database, Retrieval)</title>
      <description><![CDATA[<p>About Us</p>
<p>At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of websites and other Internet properties for customers ranging from individual bloggers to SMBs to Fortune 500 companies.</p>
<p>We protect and accelerate any Internet application online without adding hardware, installing software, or changing a line of code. Internet properties powered by Cloudflare all have web traffic routed through its intelligent global network, which gets smarter with every request. As a result, they see significant improvement in performance and a decrease in spam and other attacks.</p>
<p>Cloudflare was named to Entrepreneur Magazine’s Top Company Cultures list and ranked among the World’s Most Innovative Companies by Fast Company.</p>
<p><strong>About Role</strong></p>
<p>We are looking for experienced and highly motivated engineers to join our DATA Org and help build the future of data at Cloudflare. Our organisation is responsible for the entire data lifecycle - from ingestion and processing to storage and retrieval - powering the critical logs and analytics that provide our customers with real-time visibility into the health and performance of their online properties.</p>
<p>Our mission is to empower customers to leverage their data to drive better outcomes for their business. We build and maintain a suite of high-performance, scalable systems that handle more than a billion events in a second.</p>
<p>As an engineer in our organisation, you will have the opportunity to work on complex distributed systems challenges across different parts of our data stack.</p>
<p>Our Data Org is composed of several key teams, and you could contribute to any of the following areas:</p>
<ul>
<li>Data Delivery: You will build and operate our distributed data delivery pipeline, a high-throughput, low-latency system (primarily written in Go) responsible for ingesting, processing, and routing massive volumes of data from across Cloudflare&#39;s global network to multi-core destination.</li>
</ul>
<ul>
<li>Analytical Database Platform: Contribute to our core analytical platform powered by ClickHouse. This team builds and maintains a high-performance, scalable database platform optimised for the immense analytical workloads generated by our products and services.</li>
</ul>
<ul>
<li>Data Retrieval: Be responsible for building the customer-facing products that make data accessible and actionable. This includes developing our public GraphQL API, building robust log delivery solutions and integrations with customer destinations, and contributing to our alerting products, which empower users to configure and receive near real-time alerts based on the logs and metrics observed by our data platform.</li>
</ul>
<p><strong>Responsibilities</strong></p>
<p>As a Software Engineer in our Data Organisation depending on the team you join, you will focus on a subset of the following areas:</p>
<ul>
<li>Design, develop, and maintain scalable and reliable distributed systems across the entire data lifecycle.</li>
</ul>
<ul>
<li>Build and optimise key components of our high-throughput data delivery platform to ensure data integrity and low-latency delivery.</li>
</ul>
<ul>
<li>Develop new and improve existing components for the Cloudflare Analytical Platform to extend functionality and performance.</li>
</ul>
<ul>
<li>Scale, monitor, and maintain the performance of our large-scale database clusters to accommodate the growing volume of data.</li>
</ul>
<ul>
<li>Develop and enhance our customer-facing GraphQL APIs, log delivery, and alerting solutions, focusing on performance, reliability, and user experience.</li>
</ul>
<ul>
<li>Work to identify and remove bottlenecks across our data platforms, from streamlining data ingestion processes to optimising query performance.</li>
</ul>
<ul>
<li>Collaborate with other teams across Cloudflare to understand their data needs and build solutions that empower them to make data-driven decisions.</li>
</ul>
<ul>
<li>Collaborate with the ClickHouse open-source community to add new features and contribute to the upstream codebase.</li>
</ul>
<ul>
<li>Participate in the development of the next generation of our data platforms, including researching and evaluating new technologies and approaches.</li>
</ul>
<p><strong>Key Qualifications</strong></p>
<ul>
<li>3+ years of experience working in software development covering distributed systems and databases.</li>
</ul>
<ul>
<li>Strong programming skills (Golang is preferable), as well as a deep understanding of software development best practices and principles.</li>
</ul>
<ul>
<li>Hands-on experience with modern observability stacks, including Prometheus, Grafana, and a strong understanding of handling high-cardinality metrics at scale.</li>
</ul>
<ul>
<li>Strong knowledge of SQL and database internals, including experience with database design, optimisation, and performance tuning.</li>
</ul>
<ul>
<li>A solid foundation in computer science, including algorithms, data structures, distributed systems, and concurrency.</li>
</ul>
<ul>
<li>Strong analytical and problem-solving skills, with a willingness to debug, troubleshoot, and learn about complex problems at high scale.</li>
</ul>
<ul>
<li>Ability to work collaboratively in a team environment and communicate effectively with other teams across Cloudflare.</li>
</ul>
<ul>
<li>Experience with ClickHouse is a plus.</li>
</ul>
<ul>
<li>Experience with data streaming technologies (e.g., Kafka, Flink) is a plus.</li>
</ul>
<ul>
<li>Experience developing and scaling APIs, particularly GraphQL, is a plus.</li>
</ul>
<ul>
<li>Experience with Infrastructure as Code tools like SALT or Terraform is a plus.</li>
</ul>
<ul>
<li>Experience with Linux container technologies, such as Docker and Kubernetes, is a plus.</li>
</ul>
<p>If you&#39;re passionate about building scalable and performant data platforms using cutting-edge technologies and want to work with a world-class team of engineers, then we want to hear from you!</p>
<p>Join us in our mission to help build a better internet for everyone!</p>
<p>This role requires flexibility to be on-call outside of standard working hours to address technical issues as needed.</p>
<p><strong>What Makes Cloudflare Special?</strong></p>
<p>We’re not just a highly ambitious, large-scale technology company. We’re a highly ambitious, large-scale technology company with a soul.</p>
<p>Fundamental to our mission to help build a better Internet is protecting the free and open Internet.</p>
<p>Project Galileo: Since 2014, we&#39;ve equipped more than 2,400 journalism and civil society organisations in 111 countries with powerful tools to defend themselves against attacks that would otherwise censor their work, technology already used by Cloudflare’s enterprise customers--at no cost.</p>
<p>Athenian Project: In 2017, we created the Athenian Project to ensure that state and local governments have the highest level of protection and reliability for free, so that their constituents have access to election information and voter registration.</p>
<p>Since the project, we&#39;ve provided services to more than 425 local government election websites in 33 states.</p>
<p>1.1.1.1: We released 1.1.1.1 to help fix the foundation of the Internet by building a faster, more secure and privacy-centric public DNS resolver. This is available publicly for everyone to use - it is the first consumer-focused service Cloudflare has ever released.</p>
<p>Here’s the deal - we don’t store client IP addresses never, ever. We will continue to abide by our privacy commitment and ensure that</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>senior</experiencelevel>
      <workarrangement>hybrid</workarrangement>
      <salaryrange></salaryrange>
      <skills>Golang, Prometheus, Grafana, SQL, database internals, database design, optimisation, performance tuning, algorithms, data structures, distributed systems, concurrency, ClickHouse, Kafka, Flink, GraphQL, Infrastructure as Code, Linux container technologies, Docker, Kubernetes</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>Cloudflare</employername>
      <employerlogo>https://logos.yubhub.co/cloudflare.com.png</employerlogo>
      <employerdescription>Cloudflare is a technology company that helps build a better Internet by protecting and accelerating any Internet application online without adding hardware, installing software, or changing a line of code.</employerdescription>
      <employerwebsite>https://www.cloudflare.com/</employerwebsite>
      <compensationcurrency></compensationcurrency>
      <compensationmin></compensationmin>
      <compensationmax></compensationmax>
      <compensationinterval></compensationinterval>
      <applyto>https://job-boards.greenhouse.io/cloudflare/jobs/7462801?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>Hybrid</location>
      <city>Hybrid</city>
      <state></state>
      <postalcode></postalcode>
      <country></country>
      <postedate>2026-04-26</postedate>
    </job>
    <job>
      <externalid>c9e5f5a8-0e3</externalid>
      <title>Senior Backend Software Engineer - Infrastructure</title>
      <description><![CDATA[<p>A Senior Backend Software Engineer - Infrastructure at Palantir will contribute high-quality code to underpin Palantir Foundry and Gotham with performant, secure, and scalable building blocks, enabling products deployed to the most important institutions in the public and private sector.</p>
<p>You will build the foundational capabilities that power our products used by research scientists, aerospace engineers, intelligence analysts, and economic forecasters, in countries around the world.</p>
<p>We&#39;re hiring engineers who are passionate about solving real-world problems and empowering both developers and end-users to work optimally.</p>
<p>If you’re motivated to develop reliable, performant, and scalable systems, and to design robust APIs and primitives, this role offers the opportunity to make a significant impact on our products and the people who use them.</p>
<p>As a Senior Backend Software Engineer - Infrastructure, you will:</p>
<ul>
<li>Build a performant search and indexing ecosystem for complex granularly permissioned data</li>
<li>Contribute to open-source data processing libraries, integrating the latest innovations to achieve performance gains</li>
<li>Build the distributed systems that power large scale compute workloads, orchestrating and efficiently scheduling hundreds of thousands of containers every hour</li>
<li>Design architecture and opinionated APIs to keep application developers on the happy path</li>
<li>Trace and performance observability in high scale distributed microservice architectures</li>
<li>Build reliant, performant, and scalable systems for storage, auth, or asset serving to enable other product teams to build robust applications without deep domain expertise in the underlying systems</li>
<li>Automate the deployment, management, and operations of complex distributed systems like Cassandra, Elasticsearch, Kafka, and more across different environments</li>
</ul>
<p>We use different backend languages, including Java, Rust, and Go, and open-source technologies like Cassandra, ElasticSearch, Spark, Kafka, Kubernetes, Flink, and industry-standard build tooling, including Gradle and GitHub.</p>
<p>To succeed in this role, you will need to demonstrate:</p>
<ul>
<li>Demonstrated ability to collaborate and empathize with a variety of individuals</li>
<li>Ability to learn new technology and concepts, even without in-depth experience</li>
<li>Bias towards quality and thoughtful about edge cases (“anything that can go wrong will go wrong”); writes code that is defensive against all possibilities</li>
<li>Leading solutions and APIs with users in mind while maintaining a high engineering bar</li>
</ul>
<p>We require:</p>
<ul>
<li>6+ years of experience designing, building, and operating scalable and reliable infrastructure systems in a production environment</li>
<li>Engineering background in Computer Science, Mathematics, Software Engineering, Physics or similar field</li>
<li>Strong coding skills with demonstrated proficiency in programming languages, such as Java, C++, Python, Rust, or similar languages</li>
<li>Familiarity with storage and data processing systems, cloud infrastructure, and other technical tools</li>
<li>Strong written and verbal communication skills and ability to iterate quickly with teammates, incorporating feedback and holding a high bar for quality</li>
</ul>
<p>The estimated salary range for this position is $135,000 - $200,000/year.</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>senior</experiencelevel>
      <workarrangement>hybrid</workarrangement>
      <salaryrange>$135,000 - $200,000/year</salaryrange>
      <skills>Java, Rust, Go, Cassandra, ElasticSearch, Spark, Kafka, Kubernetes, Flink, Gradle, GitHub</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>Palantir</employername>
      <employerlogo>https://logos.yubhub.co/palantir.com.png</employerlogo>
      <employerdescription>Palantir builds software for data-driven decisions and operations, empowering partners to develop lifesaving drugs, forecast supply chain disruptions, and more.</employerdescription>
      <employerwebsite>https://www.palantir.com/</employerwebsite>
      <compensationcurrency>USD</compensationcurrency>
      <compensationmin>135000</compensationmin>
      <compensationmax>200000</compensationmax>
      <compensationinterval>yearly</compensationinterval>
      <applyto>https://jobs.lever.co/palantir/b5ad6660-8145-4be5-97e2-3799f2912f5b?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>New York</location>
      <city>New York</city>
      <state></state>
      <postalcode></postalcode>
      <country></country>
      <postedate>2026-04-25</postedate>
    </job>
    <job>
      <externalid>c1d78c04-4d5</externalid>
      <title>Senior Backend Software Engineer - Infrastructure</title>
      <description><![CDATA[<p>A Senior Backend Software Engineer - Infrastructure at Palantir will contribute high-quality code to underpin Palantir Foundry and Gotham with performant, secure, and scalable building blocks, enabling products deployed to the most important institutions in the public and private sector.</p>
<p>The role involves building a performant search and indexing ecosystem for complex granularly permissioned data, contributing to open-source data processing libraries, and designing architecture and opinionated APIs to keep application developers on the happy path.</p>
<p>As a member of the infrastructure team, you will work on building the foundational capabilities that power our products used by research scientists, aerospace engineers, intelligence analysts, and economic forecasters, in countries around the world.</p>
<p>We&#39;re hiring engineers who are passionate about solving real-world problems and empowering both developers and end-users to work optimally. If you’re motivated to develop reliable, performant, and scalable systems, and to design robust APIs and primitives, this role offers the opportunity to make a significant impact on our products and the people who use them.</p>
<p>Frontline Foundry Software Engineers may be offered the opportunity to Frontline, an exclusive program unlike any other. This unique, short-term assignment involves being embedded with customers, allowing you to work directly with users and gain firsthand insight into how our products are used and the challenges our customers face.</p>
<p>Some of our most successful products were built on the factory floor, addressing real-world problems for the world&#39;s most important institutions. These products were developed by some of our most successful product engineers, who began their careers in roles aligned with Frontline responsibilities, gaining a deep understanding of both our technology and our customers.</p>
<p>Core Responsibilities:</p>
<ul>
<li>Building a performant search and indexing ecosystem for complex granularly permissioned data</li>
<li>Contributing to open-source data processing libraries, integrating the latest innovations to achieve performance gains</li>
<li>Building the distributed systems that power large scale compute workloads, orchestrating and efficiently scheduling hundreds of thousands of containers every hour</li>
<li>Designing architecture and opinionated APIs to keep application developers on the happy path</li>
<li>Tracing and performance observability in high scale distributed microservice architectures</li>
<li>Building reliant, performant, and scalable systems for storage, auth, or asset serving to enable other product teams to build robust applications without deep domain expertise in the underlying systems</li>
<li>Automating the deployment, management, and operations of complex distributed systems like Cassandra, Elasticsearch, Kafka, and more across different environments</li>
</ul>
<p>Technologies We Use:</p>
<ul>
<li>Different backend languages, including Java, Rust, and Go</li>
<li>Open-source technologies like Cassandra, ElasticSearch, Spark, Kafka, Kubernetes, Flink</li>
<li>Industry-standard build tooling, including Gradle and GitHub</li>
</ul>
<p>What We Value:</p>
<ul>
<li>Demonstrated ability to collaborate and empathize with a variety of individuals. Able to iterate with users and non-technical stakeholders and understand how technical decisions impact them.</li>
<li>Ability to learn new technology and concepts, even without in-depth experience. Experience developing and managing highly-available distributed systems is beneficial, but not required.</li>
<li>Bias towards quality and thoughtful about edge cases (“anything that can go wrong will go wrong”); writes code that is defensive against all possibilities.</li>
<li>Leading solutions and APIs with users in mind while maintaining a high engineering bar. Seeks to centralize and abstract complexity away from our users in order to expose simple, powerful APIs for consumers.</li>
<li>Active UK Security clearance, or eligibility and willingness to obtain a UK Security clearance is beneficial, but not necessary.</li>
</ul>
<p>What We Require:</p>
<ul>
<li>6+ years of experience designing, building, and operating scalable and reliable infrastructure systems in a production environment</li>
<li>Engineering background in Computer Science, Mathematics, Software Engineering, Physics or similar field.</li>
<li>Strong coding skills with demonstrated proficiency in programming languages, such as Java, C++, Python, Rust, or similar languages.</li>
<li>Familiarity with storage and data processing systems, cloud infrastructure, and other technical tools.</li>
<li>Strong written and verbal communication skills and ability to iterate quickly with teammates, incorporating feedback and holding a high bar for quality.</li>
</ul>
<p>Additional Information:</p>
<p>Life at Palantir</p>
<p>We want every Palantirian to achieve their best outcomes, that’s why we celebrate individuals’ strengths, skills, and interests, from your first interview to your longterm growth, rather than rely on traditional career ladders. Paying attention to the needs of our community enables us to optimize our opportunities to grow and helps ensure many pathways to success at Palantir.</p>
<p>Promoting health and well-being across all areas of Palantirians’ lives is just one of the ways we’re investing in our community. Learn more at Life at Palantir and note that our offerings may vary by region.</p>
<p>In keeping consistent with Palantir’s values and culture, we believe employees are “better together” and in-person work affords the opportunity for more creative outcomes. Therefore, we encourage employees to work from our offices to foster connectivity and innovation. Many teams do offer hybrid options (WFH a day or two a week), allowing our employees to strike</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>senior</experiencelevel>
      <workarrangement>hybrid</workarrangement>
      <salaryrange></salaryrange>
      <skills>Java, Rust, Go, Cassandra, ElasticSearch, Spark, Kafka, Kubernetes, Flink, Gradle, GitHub, Computer Science, Mathematics, Software Engineering, Physics</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>Palantir</employername>
      <employerlogo>https://logos.yubhub.co/palantir.com.png</employerlogo>
      <employerdescription>Palantir builds software for data-driven decisions and operations, empowering partners to develop lifesaving drugs, forecast supply chain disruptions, and more.</employerdescription>
      <employerwebsite>https://www.palantir.com/</employerwebsite>
      <compensationcurrency></compensationcurrency>
      <compensationmin></compensationmin>
      <compensationmax></compensationmax>
      <compensationinterval></compensationinterval>
      <applyto>https://jobs.lever.co/palantir/2cd25c0b-088d-4a5c-9b96-1165a33fe652?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>London</location>
      <city>London</city>
      <state></state>
      <postalcode></postalcode>
      <country></country>
      <postedate>2026-04-25</postedate>
    </job>
    <job>
      <externalid>8a1c0b7d-ba9</externalid>
      <title>Senior Staff Software Engineer, Data Platform</title>
      <description><![CDATA[<p>Join us in building the future of finance.</p>
<p>Our mission is to democratize finance for all.</p>
<p>An estimated $124 trillion of assets will be inherited by younger generations in the next two decades. The largest transfer of wealth in human history.</p>
<p>If you’re ready to be at the epicenter of this historic cultural and financial shift, keep reading.</p>
<p><strong>About the team + role</strong></p>
<p>We are building an elite team, applying frontier technologies to the world’s biggest financial problems. We’re looking for bold thinkers. Sharp problem-solvers. Builders who are wired to make an impact.</p>
<p>Robinhood isn’t a place for complacency, it’s where ambitious people do the best work of their careers.</p>
<p>We’re a high-performing, fast-moving team with ethics at the center of everything we do. Expectations are high, and so are the rewards.</p>
<p>The Data Platform organization builds and operates the systems that power how data is stored, moved, and consumed across Robinhood.</p>
<p>This organization spans three core pillars: Storage (Postgres, DynamoDB, and caching systems), Streaming (real-time event infrastructure), and Data Lake (ingestion and compute systems built on Delta Lake).</p>
<p>Together, these platforms support transactional workloads, real-time data processing, and large-scale analytics that are critical to Robinhood’s products and operations.</p>
<p>The team owns the full lifecycle of data,from low-latency order path systems to near real-time and batch analytics,serving millions of users and internal teams across the company.</p>
<p>As a Senior Staff Software Engineer, you will serve as the technical lead across the Data Platform organization, shaping architecture and guiding execution across multiple teams.</p>
<p>You’ll work on complex distributed systems challenges such as database sharding and proxy architectures, real-time streaming and CDC systems, and large-scale data ingestion and compute platforms.</p>
<p>You’ll define and drive key technical bets, partner with engineering leaders to align platform capabilities with business needs, and lead 0→1 initiatives that introduce new capabilities across storage, streaming, and data systems.</p>
<p>This is a rare opportunity to influence multiple critical systems at once while raising the technical bar across an entire organization!</p>
<p><strong>What you’ll do</strong></p>
<p>Lead architectural direction across storage, streaming, and data lake platforms, connecting systems that handle transactional, real-time, and analytical workloads.</p>
<p>Design and guide implementation of distributed systems, including database sharding, proxy-based query routing, and real-time event processing pipelines.</p>
<p>Improve data freshness and latency by evolving streaming and ingestion systems toward near real-time processing goals.</p>
<p>Partner with engineering leaders and teams across Robinhood to define platform strategy, align roadmaps, and ensure systems meet reliability, scalability, and performance requirements.</p>
<p>Drive 0→1 initiatives that introduce new platform capabilities, including next-generation streaming, CDC, and data processing systems.</p>
<p><strong>What you bring</strong></p>
<p>Extensive experience building and scaling distributed systems, with deep expertise in at least two of the following areas: storage systems, streaming platforms, or data lake / large-scale data processing.</p>
<p>Strong understanding of database systems such as PostgreSQL and/or DynamoDB, including replication, sharding, and performance optimization.</p>
<p>Experience with streaming and event-driven architectures using technologies such as Kafka, Flink, or similar systems.</p>
<p>Familiarity with modern data platforms and compute engines such as Spark, Delta Lake, or equivalent large-scale data processing systems.</p>
<p>Proven ability to lead complex technical initiatives, define long-term architecture, and collaborate across multiple teams.</p>
<p><strong>What we offer</strong></p>
<p>Challenging, high-impact work to grow your career.</p>
<p>Performance driven compensation with multipliers for outsized impact, bonus programs, equity ownership, and 401(k) matching.</p>
<p>Best in class benefits to fuel your work, including 100% paid health insurance for employees with 90% coverage for dependents.</p>
<p>Lifestyle wallet – a highly flexible benefits spending account for wellness, learning, and more.</p>
<p>Employer-paid life &amp; disability insurance, fertility benefits, and mental health benefits.</p>
<p>Time off to recharge including company holidays, paid time off, sick time, parental leave, and more!</p>
<p>Exceptional office experience with catered meals, events, and comfortable workspaces.</p>
<p><strong>In addition to the base pay range listed below, this role is also eligible for bonus opportunities + equity + benefits.</strong></p>
<p>Base pay for the successful applicant will depend on a variety of job-related factors, which may include education, training, experience, location, business needs, or market demands.</p>
<p>The expected base pay range for this role is based on the location where the work will be performed and is aligned to one of 3 compensation zones.</p>
<p>For other locations not listed, compensation can be discussed with your recruiter during the interview process.</p>
<p>Base Pay Range:</p>
<p>Zone 1 (Menlo Park, CA; New York, NY; Bellevue, WA; Washington, DC)$264,000-$310,000 USD</p>
<p>Zone 2 (Denver, CO; Westlake, TX; Chicago, IL)$264,000-$310,000 USD</p>
<p>Zone 3 (Lake Mary, FL; Clearwater, FL; Gainesville, FL)$264,000-$310,000 USD</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>senior</experiencelevel>
      <workarrangement>onsite</workarrangement>
      <salaryrange>$264,000-$310,000 USD</salaryrange>
      <skills>distributed systems, database systems, streaming platforms, data lake / large-scale data processing, PostgreSQL, DynamoDB, Kafka, Flink, Spark, Delta Lake</skills>
      <category>Engineering</category>
      <industry>Finance</industry>
      <employername>Robinhood</employername>
      <employerlogo>https://logos.yubhub.co/robinhood.com.png</employerlogo>
      <employerdescription>Robinhood is a financial services company that provides a mobile app for buying and selling stocks, options, ETFs, and cryptocurrencies.</employerdescription>
      <employerwebsite>https://www.robinhood.com/</employerwebsite>
      <compensationcurrency>USD</compensationcurrency>
      <compensationmin>264000</compensationmin>
      <compensationmax>310000</compensationmax>
      <compensationinterval>yearly</compensationinterval>
      <applyto>https://job-boards.greenhouse.io/robinhood/jobs/7729014?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>Bellevue, WA</location>
      <city>Bellevue</city>
      <state>WA</state>
      <postalcode></postalcode>
      <country>US</country>
      <postedate>2026-04-25</postedate>
    </job>
    <job>
      <externalid>81d19b3d-95d</externalid>
      <title>Backend Software Engineer - Infrastructure, Foundations</title>
      <description><![CDATA[<p>A Backend Software Engineer at Palantir builds software at scale to transform how organisations use data. Collaborate closely with technical and non-technical teammates to understand customer problems and build products that solve them. Contribute high-quality code to underpin Palantir Foundry and Gotham with performant, secure, and scalable building blocks.</p>
<p><strong>Core Responsibilities</strong></p>
<ul>
<li>Building a performant search and indexing ecosystem for complex granularly permissioned data</li>
<li>Contributing to open-source data processing libraries, integrating the latest innovations to achieve performance gains</li>
<li>Building the distributed systems that power large scale compute workloads, orchestrating and efficiently scheduling hundreds of thousands of containers every hour</li>
<li>Designing architecture and opinionated APIs to keep application developers on the happy path</li>
<li>Tracing and performance observability in high scale distributed microservice architectures</li>
<li>Building reliant, performant, and scalable systems for storage, auth, or asset serving to enable other product teams to build robust applications without deep domain expertise in the underlying systems</li>
<li>Automating the deployment, management, and operations of complex distributed systems like Cassandra, Elasticsearch, Kafka, and more across different environments</li>
</ul>
<p><strong>Technologies We Use</strong></p>
<ul>
<li>Different backend languages, including Java, Rust, and Go</li>
<li>Open-source technologies like Cassandra, ElasticSearch, Spark, Kafka, Kubernetes, Flink</li>
<li>Industry-standard build tooling, including Gradle and GitHub</li>
</ul>
<p><strong>What We Value</strong></p>
<ul>
<li>Demonstrated ability to collaborate and empathize with a variety of individuals</li>
<li>Ability to learn new technology and concepts, even without in-depth experience</li>
<li>Bias towards quality and thoughtful about edge cases</li>
<li>Builds solutions and APIs with users in mind while maintaining a high engineering bar</li>
</ul>
<p><strong>What We Require</strong></p>
<ul>
<li>Engineering background in Computer Science, Mathematics, Software Engineering, Physics or similar field</li>
<li>Strong coding skills with demonstrated proficiency in programming languages, such as Java, C++, Python, Rust, or similar languages</li>
<li>Familiarity with storage and data processing systems, cloud infrastructure, and other technical tools</li>
<li>Strong written and verbal communication skills and ability to iterate quickly with teammates, incorporating feedback and holding a high bar for quality</li>
</ul>
<p><strong>Additional Information</strong></p>
<p>The estimated salary range for this position is estimated to be $135,000 - $200,000/year. Total compensation for this position may also include Restricted Stock units, sign-on bonus and other potential future incentives.</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>mid</experiencelevel>
      <workarrangement>hybrid</workarrangement>
      <salaryrange>$135,000 - $200,000/year</salaryrange>
      <skills>Java, Rust, Go, Cassandra, ElasticSearch, Spark, Kafka, Kubernetes, Flink, Gradle, GitHub</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>Palantir</employername>
      <employerlogo>https://logos.yubhub.co/palantir.com.png</employerlogo>
      <employerdescription>Palantir builds software for data-driven decisions and operations, empowering partners to develop lifesaving drugs, forecast supply chain disruptions, and more.</employerdescription>
      <employerwebsite>https://www.palantir.com/</employerwebsite>
      <compensationcurrency>USD</compensationcurrency>
      <compensationmin>135000</compensationmin>
      <compensationmax>200000</compensationmax>
      <compensationinterval>yearly</compensationinterval>
      <applyto>https://jobs.lever.co/palantir/fb2d3222-dbd8-4e03-8d39-47b820e9509c?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>New York</location>
      <city>New York</city>
      <state></state>
      <postalcode></postalcode>
      <country></country>
      <postedate>2026-04-25</postedate>
    </job>
    <job>
      <externalid>f913e2d7-8d5</externalid>
      <title>Backend Software Engineer - Infrastructure</title>
      <description><![CDATA[<p>A Backend Software Engineer at Palantir builds software at scale to transform how organisations use data. You will collaborate closely with technical and non-technical teammates to understand our customers&#39; problems and build products that solve them.</p>
<p>Our Software Engineers are involved throughout the product lifecycle, from idea generation, design, prototyping, and production delivery. We encourage movement across teams to share context, skills, and experience, so you&#39;ll learn about many different technologies and aspects of each product.</p>
<p>As a Software Engineer on infrastructure, you&#39;ll contribute high-quality code to underpin Palantir Foundry and Gotham with performant, secure, and scalable building blocks, enabling products deployed to the most important institutions in the public and private sector.</p>
<p>We&#39;re hiring engineers who are passionate about solving real-world problems and empowering both developers and end-users to work optimally. If you&#39;re motivated to develop reliable, performant, and scalable systems, and to design robust APIs and primitives, this role offers the opportunity to make a significant impact on our products and the people who use them.</p>
<p><strong>Core Responsibilities</strong></p>
<ul>
<li>Building a performant search and indexing ecosystem for complex granularly permissioned data</li>
<li>Contributing to open-source data processing libraries, integrating the latest innovations to achieve performance gains</li>
<li>Building the distributed systems that power large scale compute workloads, orchestrating and efficiently scheduling hundreds of thousands of containers every hour</li>
<li>Designing architecture and opinionated APIs to keep application developers on the happy path</li>
<li>Tracing and performance observability in high scale distributed microservice architectures</li>
<li>Building reliant, performant, and scalable systems for storage, auth, or asset serving to enable other product teams to build robust applications without deep domain expertise in the underlying systems</li>
<li>Automating the deployment, management, and operations of complex distributed systems like Cassandra, Elasticsearch, Kafka, and more across different environments</li>
</ul>
<p><strong>Technologies We Use</strong></p>
<ul>
<li>Different backend languages, including Java, Rust, and Go</li>
<li>Open-source technologies like Cassandra, ElasticSearch, Spark, Kafka, Kubernetes, Flink</li>
<li>Industry-standard build tooling, including Gradle and GitHub</li>
</ul>
<p><strong>What We Value</strong></p>
<ul>
<li>Demonstrated ability to collaborate and empathise with a variety of individuals. Able to iterate with users and non-technical stakeholders and understand how technical decisions impact them.</li>
<li>Ability to learn new technology and concepts, even without in-depth experience. Experience developing and managing highly-available distributed systems is beneficial, but not required.</li>
<li>Bias towards quality and thoughtful about edge cases (“anything that can go wrong will go wrong”); writes code that is defensive against all possibilities.</li>
<li>Builds solutions and APIs with users in mind while maintaining a high engineering bar. Seeks to centralise and abstract complexity away from our users in order to expose simple, powerful APIs for consumers.</li>
<li>Active UK Security clearance, or eligibility and willingness to obtain a UK Security clearance is beneficial, but not necessary.</li>
</ul>
<p><strong>What We Require</strong></p>
<ul>
<li>Engineering background in Computer Science, Mathematics, Software Engineering, Physics or similar field.</li>
<li>Strong coding skills with demonstrated proficiency in programming languages, such as Java, C++, Python, Rust, or similar languages.</li>
<li>Familiarity with storage and data processing systems, cloud infrastructure, and other technical tools.</li>
<li>Strong written and verbal communication skills and ability to iterate quickly with teammates, incorporating feedback and holding a high bar for quality.</li>
</ul>
<p><strong>Additional Information</strong></p>
<p>Life at Palantir</p>
<p>We want every Palantirian to achieve their best outcomes, that’s why we celebrate individuals’ strengths, skills, and interests, from your first interview to your longterm growth, rather than rely on traditional career ladders. Paying attention to the needs of our community enables us to optimize our opportunities to grow and helps ensure many pathways to success at Palantir.</p>
<p>Promoting health and well-being across all areas of Palantirians’ lives is just one of the ways we’re investing in our community. Learn more at Life at Palantir and note that our offerings may vary by region.</p>
<p>In keeping consistent with Palantir’s values and culture, we believe employees are “better together” and in-person work affords the opportunity for more creative outcomes. Therefore, we encourage employees to work from our offices to foster connectivity and innovation. Many teams do offer hybrid options (WFH a day or two a week), allowing our employees to strike the right trade-off for their personal productivity. Based on business need, there are a few roles that allow for “Remote” work on an exceptional basis. If you are</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>mid</experiencelevel>
      <workarrangement>hybrid</workarrangement>
      <salaryrange></salaryrange>
      <skills>Java, Rust, Go, Cassandra, ElasticSearch, Spark, Kafka, Kubernetes, Flink, Gradle, GitHub</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>Palantir</employername>
      <employerlogo>https://logos.yubhub.co/palantir.com.png</employerlogo>
      <employerdescription>Palantir builds software for data-driven decisions and operations. It empowers partners to develop lifesaving drugs, forecast supply chain disruptions, and locate missing children.</employerdescription>
      <employerwebsite>https://www.palantir.com/</employerwebsite>
      <compensationcurrency></compensationcurrency>
      <compensationmin></compensationmin>
      <compensationmax></compensationmax>
      <compensationinterval></compensationinterval>
      <applyto>https://jobs.lever.co/palantir/f70cdff7-c62f-4b73-a136-909e5e3d1891?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>London</location>
      <city>London</city>
      <state></state>
      <postalcode></postalcode>
      <country></country>
      <postedate>2026-04-25</postedate>
    </job>
    <job>
      <externalid>ce058b80-935</externalid>
      <title>Backend Software Engineer - Infrastructure</title>
      <description><![CDATA[<p>A Backend Software Engineer - Infrastructure at Palantir will contribute high-quality code to underpin Palantir Foundry and Gotham with performant, secure, and scalable building blocks. You&#39;ll build the foundational capabilities that power our products used by research scientists, aerospace engineers, intelligence analysts, and economic forecasters, in countries around the world.</p>
<p>As a Software Engineer on infrastructure working on our Foundry platform, you&#39;ll collaborate closely with technical and non-technical teammates to understand our customers&#39; problems and build products that solve them. You&#39;ll work autonomously and make decisions independently, within a community that will support and challenge you as you grow and develop, becoming a strong technical contributor and engineering leader.</p>
<p>Some of the key responsibilities of this role include:</p>
<ul>
<li>Building a performant search and indexing ecosystem for complex granularly permissioned data</li>
<li>Contributing to open-source data processing libraries, integrating the latest innovations to achieve performance gains</li>
<li>Building the distributed systems that power large scale compute workloads, orchestrating and efficiently scheduling hundreds of thousands of containers every hour</li>
<li>Designing architecture and opinionated APIs to keep application developers on the happy path</li>
<li>Tracing and performance observability in high scale distributed microservice architectures</li>
<li>Building reliant, performant, and scalable systems for storage, auth, or asset serving to enable other product teams to build robust applications without deep domain expertise in the underlying systems</li>
<li>Automating the deployment, management, and operations of complex distributed systems like Cassandra, Elasticsearch, Kafka, and more across different environments</li>
</ul>
<p>We&#39;re looking for engineers who are passionate about solving real-world problems and empowering both developers and end-users to work optimally. If you&#39;re motivated to develop reliable, performant, and scalable systems, and to design robust APIs and primitives, this role offers the opportunity to make a significant impact on our products and the people who use them.</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>mid</experiencelevel>
      <workarrangement>hybrid</workarrangement>
      <salaryrange>$135,000 - $200,000/year</salaryrange>
      <skills>Java, Rust, Go, Cassandra, ElasticSearch, Spark, Kafka, Kubernetes, Flink</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>Palantir</employername>
      <employerlogo>https://logos.yubhub.co/palantir.com.png</employerlogo>
      <employerdescription>Palantir builds software for data-driven decisions and operations, empowering partners to develop lifesaving drugs, forecast supply chain disruptions, and more.</employerdescription>
      <employerwebsite>https://www.palantir.com/</employerwebsite>
      <compensationcurrency>USD</compensationcurrency>
      <compensationmin>135000</compensationmin>
      <compensationmax>200000</compensationmax>
      <compensationinterval>yearly</compensationinterval>
      <applyto>https://jobs.lever.co/palantir/6fe5515f-f677-4d98-8ac2-1775a425f5e7?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>New York</location>
      <city>New York</city>
      <state></state>
      <postalcode></postalcode>
      <country></country>
      <postedate>2026-04-25</postedate>
    </job>
    <job>
      <externalid>22bcbb50-ef4</externalid>
      <title>Member of Technical Staff - Data Platform</title>
      <description><![CDATA[<p><strong>About the Role</strong></p>
<p>The Data Platform team at xAI builds and operates the infrastructure responsible for all large-scale data transport and processing across the company.</p>
<p>As a software engineer on the Data Platform team, you will design, build, and operate the distributed systems powering X&#39;s data movement and compute.</p>
<p><strong>Responsibilities</strong></p>
<ul>
<li>Design and implement high-throughput, low-latency data ingestion and transport systems.</li>
<li>Scale and optimise multi-tenant Kafka infrastructure supporting real-time workloads.</li>
<li>Extend and tune Spark, Flink, and Trino for demanding production pipelines.</li>
<li>Build interfaces, APIs, and pipelines enabling teams to query, process, and move data at petabyte scale.</li>
<li>Debug and optimise distributed systems, with a focus on reliability and performance under load.</li>
<li>Collaborate with ML, product, and infrastructure teams to unblock critical data workflows.</li>
</ul>
<p><strong>Basic Qualifications</strong></p>
<ul>
<li>Proven expertise in distributed systems, stream processing, or large-scale data platforms.</li>
<li>Proficiency in Rust, Go, Scala or similar systems languages.</li>
<li>Hands-on experience with Kafka, Flink, Spark, Trino, or Hadoop in production.</li>
<li>Strong debugging, profiling, and performance optimisation skills.</li>
<li>Track record of shipping and maintaining critical infrastructure.</li>
<li>Comfortable working in fast-moving, high-stakes environments with minimal guardrails.</li>
</ul>
<p><strong>Compensation and Benefits</strong></p>
<p>$180,000 - $440,000 USD</p>
<p>Base salary is just one part of our total rewards package at X, which also includes equity, comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short &amp; long-term disability insurance, life insurance, and various other discounts and perks.</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>staff</experiencelevel>
      <workarrangement>onsite</workarrangement>
      <salaryrange>$180,000 - $440,000 USD</salaryrange>
      <skills>Rust, Go, Scala, Kafka, Flink, Spark, Trino, Hadoop, distributed systems, stream processing, large-scale data platforms</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>xAI</employername>
      <employerlogo>https://logos.yubhub.co/x.com.png</employerlogo>
      <employerdescription>xAI creates AI systems to understand the universe and aid humanity in its pursuit of knowledge.</employerdescription>
      <employerwebsite>https://www.x.com/</employerwebsite>
      <compensationcurrency></compensationcurrency>
      <compensationmin></compensationmin>
      <compensationmax></compensationmax>
      <compensationinterval></compensationinterval>
      <applyto>https://job-boards.greenhouse.io/xai/jobs/4803862007?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>Palo Alto, CA</location>
      <city>Palo Alto</city>
      <state>CA</state>
      <postalcode></postalcode>
      <country>US</country>
      <postedate>2026-04-18</postedate>
    </job>
    <job>
      <externalid>f32fed2e-9ba</externalid>
      <title>Engineering Manager, Data Transformation</title>
      <description><![CDATA[<p>As an Engineering Manager of the Data Transformation team, you will lead a team of engineers, collaborate with infrastructure and product engineering orgs, and advance the Data Transformation roadmap and adoption at Stripe.</p>
<p>You will be driving critical workstreams for Stripe&#39;s topmost priorities around delivering high quality, materialized datasets Stripe products and AI agents.</p>
<p>Key responsibilities include:</p>
<ul>
<li>Delivering infrastructure and services that scale to our users&#39; needs with an eye on reliability and efficiency</li>
<li>Leading and managing a team of talented engineers on the team, providing mentorship, guidance, and support to ensure their success</li>
<li>Working with high-visibility teams and their stakeholders to support the Infrastructure&#39;s key engineering initiatives</li>
<li>Understanding user needs and pain points to prioritize engineering work and deliver high quality solutions that meet user needs</li>
<li>Driving the execution of projects, overseeing the entire development lifecycle from planning to delivery, while maintaining high standards of quality and timely completion</li>
</ul>
<p>You will also provide hands-on technical leadership (architecture/design, vision/direction/requirements setting, and incident response processes) for your reports, work with leaders across the company to create and drive toward the longer term vision of Stripe&#39;s Data Transformation roadmap, and foster a collaborative and inclusive work environment, promoting innovation, knowledge sharing, and continuous improvement within the team.</p>
<p>We&#39;re looking for someone who has 1-3 years of experience managing teams that shipped and operated data pipelines and critical distributed system infrastructure, successfully recruited and built great teams, and works effectively cross-functionally and is able to think rigorously, communicate effectively, and make or coordinate hard decisions and trade-offs.</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>mid</experiencelevel>
      <workarrangement>remote</workarrangement>
      <salaryrange></salaryrange>
      <skills>Kafka, Flink, Spark, Airflow, Python, SQL, API design</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>Stripe</employername>
      <employerlogo>https://logos.yubhub.co/stripe.com.png</employerlogo>
      <employerdescription>Stripe is a financial infrastructure platform for businesses, with millions of companies using its services.</employerdescription>
      <employerwebsite>https://stripe.com/</employerwebsite>
      <compensationcurrency></compensationcurrency>
      <compensationmin></compensationmin>
      <compensationmax></compensationmax>
      <compensationinterval></compensationinterval>
      <applyto>https://job-boards.greenhouse.io/stripe/jobs/7688358?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>N/A</location>
      <city>N/A</city>
      <state></state>
      <postalcode></postalcode>
      <country></country>
      <postedate>2026-04-18</postedate>
    </job>
    <job>
      <externalid>9b657c4e-8a1</externalid>
      <title>Member of Technical Staff - Data Platform</title>
      <description><![CDATA[<p><strong>About the Role</strong></p>
<p>As a software engineer on the Data Platform team, you will design, build, and operate the distributed systems powering X&#39;s data movement and compute. You will take ownership of infrastructure components that process trillions of events daily, driving the scalability, performance, and reliability of the systems that power product and ML workloads across the company.</p>
<p><strong>Responsibilities</strong></p>
<ul>
<li>Design and implement high-throughput, low-latency data ingestion and transport systems.</li>
<li>Scale and optimize multi-tenant Kafka infrastructure supporting real-time workloads.</li>
<li>Extend and tune Spark, Flink, and Trino for demanding production pipelines.</li>
<li>Build interfaces, APIs, and pipelines enabling teams to query, process, and move data at petabyte scale.</li>
<li>Debug and optimize distributed systems, with a focus on reliability and performance under load.</li>
<li>Collaborate with ML, product, and infrastructure teams to unblock critical data workflows.</li>
</ul>
<p><strong>Basic Qualifications</strong></p>
<ul>
<li>Proven expertise in distributed systems, stream processing, or large-scale data platforms.</li>
<li>Proficiency in Rust, Go, Scala or similar systems languages.</li>
<li>Hands-on experience with Kafka, Flink, Spark, Trino, or Hadoop in production.</li>
<li>Strong debugging, profiling, and performance optimization skills.</li>
<li>Track record of shipping and maintaining critical infrastructure.</li>
<li>Comfortable working in fast-moving, high-stakes environments with minimal guardrails.</li>
</ul>
<p><strong>Compensation and Benefits</strong></p>
<p>$180,000 - $440,000 USD</p>
<p>Base salary is just one part of our total rewards package at X, which also includes equity, comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short &amp; long-term disability insurance, life insurance, and various other discounts and perks.</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>staff</experiencelevel>
      <workarrangement>hybrid</workarrangement>
      <salaryrange>$180,000 - $440,000 USD</salaryrange>
      <skills>distributed systems, stream processing, large-scale data platforms, Rust, Go, Scala, Kafka, Flink, Spark, Trino, Hadoop, debugging, profiling, performance optimization</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>xAI</employername>
      <employerlogo>https://logos.yubhub.co/x.ai.png</employerlogo>
      <employerdescription>xAI creates AI systems to understand the universe and aid humanity in its pursuit of knowledge.</employerdescription>
      <employerwebsite>https://www.x.ai/</employerwebsite>
      <compensationcurrency></compensationcurrency>
      <compensationmin></compensationmin>
      <compensationmax></compensationmax>
      <compensationinterval></compensationinterval>
      <applyto>https://job-boards.greenhouse.io/xai/jobs/4803862007?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>Palo Alto, CA</location>
      <city>Palo Alto</city>
      <state>CA</state>
      <postalcode></postalcode>
      <country>US</country>
      <postedate>2026-04-18</postedate>
    </job>
    <job>
      <externalid>4b563c21-dd0</externalid>
      <title>Software Engineer, Data Infrastructure</title>
      <description><![CDATA[<p><strong>Software Engineer, Data Infrastructure</strong></p>
<p><strong>Location</strong></p>
<p>San Francisco</p>
<p><strong>Employment Type</strong></p>
<p>Full time</p>
<p><strong>Department</strong></p>
<p>Applied AI</p>
<p><strong>Compensation</strong></p>
<ul>
<li>$185K – $385K • Offers Equity</li>
</ul>
<p>The base pay offered may vary depending on multiple individualized factors, including market location, job-related knowledge, skills, and experience. If the role is non-exempt, overtime pay will be provided consistent with applicable laws. In addition to the salary range listed above, total compensation also includes generous equity, performance-related bonus(es) for eligible employees, and the following benefits.</p>
<p><strong>Benefits</strong></p>
<ul>
<li>Medical, dental, and vision insurance for you and your family, with employer contributions to Health Savings Accounts</li>
</ul>
<ul>
<li>Pre-tax accounts for Health FSA, Dependent Care FSA, and commuter expenses (parking and transit)</li>
</ul>
<ul>
<li>401(k) retirement plan with employer match</li>
</ul>
<ul>
<li>Paid parental leave (up to 24 weeks for birth parents and 20 weeks for non-birthing parents), plus paid medical and caregiver leave (up to 8 weeks)</li>
</ul>
<ul>
<li>Paid time off: flexible PTO for exempt employees and up to 15 days annually for non-exempt employees</li>
</ul>
<ul>
<li>13+ paid company holidays, and multiple paid coordinated company office closures throughout the year for focus and recharge, plus paid sick or safe time (1 hour per 30 hours worked, or more, as required by applicable state or local law)</li>
</ul>
<ul>
<li>Mental health and wellness support</li>
</ul>
<ul>
<li>Employer-paid basic life and disability coverage</li>
</ul>
<ul>
<li>Annual learning and development stipend to fuel your professional growth</li>
</ul>
<ul>
<li>Daily meals in our offices, and meal delivery credits as eligible</li>
</ul>
<ul>
<li>Relocation support for eligible employees</li>
</ul>
<ul>
<li>Additional taxable fringe benefits, such as charitable donation matching and wellness stipends, may also be provided.</li>
</ul>
<p><strong>About the Team</strong></p>
<p>Data Platform at OpenAI owns the foundational data stack powering critical product, research, and analytics workflows. We operate some of the largest Spark compute fleets in production; design, and build data lakes and metadata systems on Iceberg and Delta with a vision toward exabyte-scale architecture; run high throughput streaming platforms on Kafka and Flink; provide orchestration with Airflow; and support ML feature engineering tooling such as Chronon. Our mission is to deliver reliable, secure, and efficient data access at scale and accelerate intelligent, AI assisted data workflows.</p>
<p><strong>About the Role</strong></p>
<p>This role focuses on building and operating data infrastructure that supports massive compute fleets and storage systems, designed for high performance and scalability. You’ll help design, build, and operate the next generation of data infrastructure at OpenAI. You will scale and harden big data compute and storage platforms, build and support high-throughput streaming systems, build and operate low latency data ingestions, enable secure and governed data access for ML and analytics, and design for reliability and performance at extreme scale.</p>
<p>You will take full lifecycle ownership: architecture, implementation, production operations, and on-call participation.</p>
<p><strong>Responsibilities</strong></p>
<ul>
<li>Design, build, and maintain data infrastructure systems such as distributed compute, data orchestration, distributed storage, streaming infrastructure, machine learning infrastructure while ensuring scalability, reliability, and security</li>
</ul>
<ul>
<li>Ensure our data platform can scale by orders of magnitude while remaining reliable and efficient</li>
</ul>
<ul>
<li>Accelerate company productivity by empowering your fellow engineers &amp; teammates with excellent data tooling and systems</li>
</ul>
<ul>
<li>Collaborate with product, research and analytics teams to build the technical foundations capabilities that unlock new features and experiences</li>
</ul>
<ul>
<li>Own the reliability of the systems you build, including participation in an on-call rotation for critical incidents</li>
</ul>
<p><strong>Requirements</strong></p>
<ul>
<li>4+ years in data infrastructure engineering OR</li>
</ul>
<ul>
<li>4+ years in infrastructure engineering with a strong interest in data</li>
</ul>
<ul>
<li>Take pride in building and operating scalable, reliable, secure systems</li>
</ul>
<ul>
<li>Are comfortable with ambiguity and rapid change</li>
</ul>
<ul>
<li>Have an intrinsic desire to learn and fill in missing skills, and an equally strong talent for sharing learnings clearly and concisely with others</li>
</ul>
<p><strong>About OpenAI</strong></p>
<p>OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of human diversity.</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>mid</experiencelevel>
      <workarrangement>hybrid</workarrangement>
      <salaryrange>$185K – $385K • Offers Equity</salaryrange>
      <skills>data infrastructure engineering, infrastructure engineering, Spark, Kafka, Flink, Airflow, Chronon, Iceberg, Delta, Terraform, distributed systems, machine learning, data science, cloud computing, containerization, DevOps</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>OpenAI</employername>
      <employerlogo>https://logos.yubhub.co/openai.com.png</employerlogo>
      <employerdescription>OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products.</employerdescription>
      <employerwebsite>https://jobs.ashbyhq.com</employerwebsite>
      <compensationcurrency></compensationcurrency>
      <compensationmin></compensationmin>
      <compensationmax></compensationmax>
      <compensationinterval></compensationinterval>
      <applyto>https://jobs.ashbyhq.com/openai/f763c6b3-5167-4a67-b691-4c3fa2c44156?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>San Francisco</location>
      <city>San Francisco</city>
      <state></state>
      <postalcode></postalcode>
      <country></country>
      <postedate>2026-03-06</postedate>
    </job>
    <job>
      <externalid>c873a489-0dc</externalid>
      <title>Data Engineer, Analytics</title>
      <description><![CDATA[<p><strong>Data Engineer, Analytics</strong></p>
<p><strong>Location</strong></p>
<p>San Francisco</p>
<p><strong>Employment Type</strong></p>
<p>Full time</p>
<p><strong>Department</strong></p>
<p>Applied AI</p>
<p><strong>Compensation</strong></p>
<ul>
<li>$230K – $385K • Offers Equity</li>
</ul>
<p>The base pay offered may vary depending on multiple individualized factors, including market location, job-related knowledge, skills, and experience. If the role is non-exempt, overtime pay will be provided consistent with applicable laws. In addition to the salary range listed above, total compensation also includes generous equity, performance-related bonus(es) for eligible employees, and the following benefits.</p>
<p><strong>Benefits</strong></p>
<ul>
<li>Medical, dental, and vision insurance for you and your family, with employer contributions to Health Savings Accounts</li>
</ul>
<ul>
<li>Pre-tax accounts for Health FSA, Dependent Care FSA, and commuter expenses (parking and transit)</li>
</ul>
<ul>
<li>401(k) retirement plan with employer match</li>
</ul>
<ul>
<li>Paid parental leave (up to 24 weeks for birth parents and 20 weeks for non-birthing parents), plus paid medical and caregiver leave (up to 8 weeks)</li>
</ul>
<ul>
<li>Paid time off: flexible PTO for exempt employees and up to 15 days annually for non-exempt employees</li>
</ul>
<ul>
<li>13+ paid company holidays, and multiple paid coordinated company office closures throughout the year for focus and recharge, plus paid sick or safe time (1 hour per 30 hours worked, or more, as required by applicable state or local law)</li>
</ul>
<ul>
<li>Mental health and wellness support</li>
</ul>
<ul>
<li>Employer-paid basic life and disability coverage</li>
</ul>
<ul>
<li>Annual learning and development stipend to fuel your professional growth</li>
</ul>
<ul>
<li>Daily meals in our offices, and meal delivery credits as eligible</li>
</ul>
<ul>
<li>Relocation support for eligible employees</li>
</ul>
<ul>
<li>Additional taxable fringe benefits, such as charitable donation matching and wellness stipends, may also be provided.</li>
</ul>
<p><strong>About the team</strong></p>
<p>The Applied team works across research, engineering, product, and design to bring OpenAI’s technology to consumers and businesses.</p>
<p>We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. Safety is more important to us than unfettered growth.</p>
<p><strong>About the role</strong></p>
<p>We&#39;re seeking a Data Engineer to take the lead in building our data pipelines and core tables for OpenAI. These pipelines are crucial for powering analyses, safety systems that guide business decisions, product growth, and prevent bad actors. If you&#39;re passionate about working with data and are eager to create solutions with significant impact, we&#39;d love to hear from you. This role also provides the opportunity to collaborate closely with the researchers behind ChatGPT and help them train new models to deliver to users. As we continue our rapid growth, we value data-driven insights, and your contributions will play a pivotal role in our trajectory. Join us in shaping the future of OpenAI!</p>
<p><strong>In this role, you will:</strong></p>
<ul>
<li>Design, build and manage our data pipelines, ensuring all user event data is seamlessly integrated into our data warehouse.</li>
</ul>
<ul>
<li>Develop canonical datasets to track key product metrics including user growth, engagement, and revenue.</li>
</ul>
<ul>
<li>Work collaboratively with various teams, including, Infrastructure, Data Science, Product, Marketing, Finance, and Research to understand their data needs and provide solutions.</li>
</ul>
<ul>
<li>Implement robust and fault-tolerant systems for data ingestion and processing.</li>
</ul>
<ul>
<li>Participate in data architecture and engineering decisions, bringing your strong experience and knowledge to bear.</li>
</ul>
<ul>
<li>Ensure the security, integrity, and compliance of data according to industry and company standards.</li>
</ul>
<p><strong>You might thrive in this role if you:</strong></p>
<ul>
<li>Have 3+ years of experience as a data engineer and 8+ years of any software engineering experience(including data engineering).</li>
</ul>
<ul>
<li>Proficiency in at least one programming language commonly used within Data Engineering, such as Python, Scala, or Java.</li>
</ul>
<ul>
<li>Experience with distributed processing technologies and frameworks, such as Hadoop, Flink and distributed storage systems (e.g., HDFS, S3).</li>
</ul>
<ul>
<li>Expertise with any of ETL schedulers such as Airflow, Dagster, Prefect or similar frameworks.</li>
</ul>
<ul>
<li>Solid understanding of Spark and ability to write, debug and optimize Spark code.</li>
</ul>
<p><strong>About OpenAI</strong></p>
<p>OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.</p>
<p style="margin-top:24px;font-size:13px;color:#666;">XML job scraping automation by <a href="https://yubhub.co">YubHub</a></p>]]></description>
      <jobtype>full-time</jobtype>
      <experiencelevel>senior</experiencelevel>
      <workarrangement>onsite</workarrangement>
      <salaryrange>$230K – $385K • Offers Equity</salaryrange>
      <skills>Python, Scala, Java, Hadoop, Flink, HDFS, S3, Airflow, Dagster, Prefect, Spark</skills>
      <category>Engineering</category>
      <industry>Technology</industry>
      <employername>OpenAI</employername>
      <employerlogo>https://logos.yubhub.co/openai.com.png</employerlogo>
      <employerdescription>OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products.</employerdescription>
      <employerwebsite>https://jobs.ashbyhq.com</employerwebsite>
      <compensationcurrency></compensationcurrency>
      <compensationmin></compensationmin>
      <compensationmax></compensationmax>
      <compensationinterval></compensationinterval>
      <applyto>https://jobs.ashbyhq.com/openai/fc5bbc77-a30c-4e7a-9acc-8a2e748545b4?utm_source=yubhub.co&amp;utm_medium=jobs_feed&amp;utm_campaign=apply</applyto>
      <location>San Francisco</location>
      <city>San Francisco</city>
      <state></state>
      <postalcode></postalcode>
      <country></country>
      <postedate>2026-03-06</postedate>
    </job>
  </jobs>
</source>