Browse Data & ML Jobs

Search 150 curated tech job listings scraped in real-time from LinkedIn, Glassdoor, RemoteOK, and more. Filter by role, location, seniority, and source to find your next opportunity.

Filters:
× Clear filters

150 jobs found for "Spark"

Page 1 of 8

Lead Data Engineer (Python, AWS, Spark, Kafka, SQL, Snowflake, Databricks, GenAI)

Capital One Richmond, VA $215k – $246k

…Preferred Qualifications: Master's Degree 7+ years of experience in application development including Python, SQL, Spark, ETL tools, or AWS Glue 4+ years of experience with a public cloud (AWS, Microsoft Azure … Google Cloud) 4+ years experience with Distributed data/computing tools (MapReduce, Hadoop, Hive, EMR, Kafka, Spark, or MySQL) 4+ year experience working on real-time data and streaming applications 4+ years of experience…

LinkedIn

Lead Data Engineer (Python, AWS, Spark, Kafka, SQL, Snowflake, Databricks, GenAI)

Capital One Richmond, VA $215k – $246k

…Preferred Qualifications: Master's Degree 7+ years of experience in application development including Python, SQL, Spark, ETL tools, or AWS Glue 4+ years of experience with a public cloud (AWS, Microsoft Azure … Google Cloud) 4+ years experience with Distributed data/computing tools (MapReduce, Hadoop, Hive, EMR, Kafka, Spark, or MySQL) 4+ year experience working on real-time data and streaming applications 4+ years of experience…

LinkedIn

Lead Data Engineer (Python, Scala, Spark)

Capital One Richmond, VA $197k – $225k

…Azure, Google Cloud) 4+ years experience with Distributed data/computing tools (MapReduce, Hadoop, Hive, EMR, Kafka, Spark, Gurobi, or MySQL) 4+ year experience working on real-time data and streaming applications 4+ years…

LinkedIn

Senior Frontend Developer (+ Equity) at AI-native startup backed by YC and Spark Capital

TechTree San Francisco, CA $180k – $240k

…personal AI assistant for knowledge workers. They've raised $30M in seed funding from Spark Capital, Felicis, and Y Combinator, marking one of the largest seed rounds in YC history. The team…

LinkedIn

Python/PySpark Data Engineer

Synechron Dallas, TX $95k – $100k

…platforms and intelligent data solutions. The ideal candidate will have strong expertise in Python, Apache Spark, cloud data technologies, data engineering best practices, and the integration of AI/ML capabilities into enterprise data … ETL/ELT workflows for structured, semi-structured, and unstructured data. Develop data processing solutions using Apache Spark, Spark SQL, and DataFrames. Integrate data from APIs, databases, files, streaming platforms, and cloud services. Implement…

LinkedIn

Senior Developer - Python / PySpark

CGI Toronto, Ontario, Canada $80k – $130k

…market, trade, risk, and reference data. Develop high-performance data processing solutions using Apache Spark and Azure Databricks. Design and implement reliable data ingestion frameworks for internal and external data sources. Develop … reusable data transformation, validation, and reconciliation frameworks. Optimize Spark applications for scalability, reliability, and performance. Implement data quality controls, monitoring, and governance best practices. Troubleshoot production issues and perform root cause analysis…

LinkedIn

Senior Software Engineer - Data Platform

Samsara USA $131k – $220k

…specialized software engineering role focused on data infrastructure. You will work on systems such as Spark and Databricks infrastructure, Delta Lake on S3, data replication from primary data stores such … operational use cases. Improve the reliability, observability, scalability, security, and developer experience of Samsara’s Spark and Databricks-based data processing platform. Develop internal libraries, APIs, frameworks, and tooling in languages such…

Jobicy

Data Engineer

SoftStandard Solutions United States

…based applications, scripts, APIs, and automation solutions. Build and optimize data processing pipelines using PySpark/Apache Spark . Work with relational and NoSQL databases for data ingestion, transformation, and storage. Implement data quality, validation … APIs . Strong understanding of ETL/ELT, data warehousing, data modeling, and database concepts . Experience with Apache Spark and distributed data processing. Experience with at least one cloud platform: AWS, Azure, or GCP . Knowledge…

LinkedIn

Sr. Solutions Engineer New

databricks London

…public cloud platform (AWS, Azure, or GCP) Working knowledge of distributed data systems: Apache Spark™, Delta Lake, or equivalent (Hadoop, Kafka, Flink) Experience leading technical customer conversations — discovery, whiteboarding, architecture reviews Familiarity … Have: Databricks certification or experience with the Databricks Platform Experience with Unity Catalog, Lakeflow Spark Declarative Pipelines, or MLflow Background at a data/AI company, cloud provider, or technical consulting firm Interview Process…

Arbeitnow

Python Hadoop Engineer

Tata Consultancy Services Charlotte, NC $401k+

…distributed data processing jobs. Practical experience working with distributed data tools such as Apache Spark, Databricks, or Hadoop ecosystems. Experience building and managing datasets in relational and/or cloud-based data platforms (Teradata … observability for data pipelines (logs, metrics, health checks). Advanced experience with performance tuning of SQL, Spark, or distributed data workflows. Knowledge of data security practices (encryption, masking, PII handling). Experience supporting analytical…

LinkedIn

Sr. Solutions Engineer

databricks London

…public cloud platform (AWS, Azure, or GCP) Working knowledge of distributed data systems: Apache Spark™, Delta Lake, or equivalent (Hadoop, Kafka, Flink) Experience leading technical customer conversations — discovery, whiteboarding, architecture reviews Familiarity … Have: Databricks certification or experience with the Databricks Platform Experience with Unity Catalog, Lakeflow Spark Declarative Pipelines, or MLflow Background at a data/AI company, cloud provider, or technical consulting firm Interview Process…

Arbeitnow

Senior Data & Software Engineer

Accenture Federal Services McLean, VA $112k – $222k

…design patterns. What you'll need : Minimum of 5 years' experience with the following: Apache Spark & PySpark Using orchestration tools to deploy data pipelines, including configuring and updating Spark Jobs Advanced Python…

LinkedIn

Forward Deployed Engineer

databricks Berlin; Munich

…Azure, GCP) with expertise in at least one Deep experience with distributed computing with Apache Spark™ and knowledge of Spark runtime internals Familiarity with CI/CD for production deployments Working knowledge of MLOps…

Arbeitnow

Data Scientist II

Scribd, Inc. Seattle, WA CA$184k+

…PyTorch to third party LLM APIs Process massive amounts of data with Python, SQL and Spark Align with stakeholders through written and verbal communications methods on the approaches and results of projects … Hands-on experience building ML pipelines and working with distributed data processing frameworks like Apache Spark, Databricks, or similar. Intermediate level in at least three of these fields: classification algorithms, natural language…

LinkedIn

Data Scientist II

Scribd, Inc. Los Angeles, CA CA$184k+

…PyTorch to third party LLM APIs Process massive amounts of data with Python, SQL and Spark Align with stakeholders through written and verbal communications methods on the approaches and results of projects … Hands-on experience building ML pipelines and working with distributed data processing frameworks like Apache Spark, Databricks, or similar. Intermediate level in at least three of these fields: classification algorithms, natural language…

LinkedIn

Lead Data Engineer (Python, AWS, SQL, GenAI) (Enterprise Platforms Technology)

Capital One Chicago, IL $197k – $225k

…Azure, Google Cloud) Preferred Qualifications: 7+ years of experience in application development including Python, SQL, Spark, ETL tools, AWS Glue 4+ years of experience with a public cloud (AWS, Microsoft Azure, Google … Cloud) 4+ years of experience with Distributed data/computing tools (MapReduce, Hadoop, Hive, EMR, Kafka, Spark, or MySQL) 4+ years of experience working on real-time data and streaming applications 4+ years…

LinkedIn

Data Scientist II

Scribd, Inc. Atlanta, GA CA$184k+

…PyTorch to third party LLM APIs Process massive amounts of data with Python, SQL and Spark Align with stakeholders through written and verbal communications methods on the approaches and results of projects … Hands-on experience building ML pipelines and working with distributed data processing frameworks like Apache Spark, Databricks, or similar. Intermediate level in at least three of these fields: classification algorithms, natural language…

LinkedIn

Data Scientist II

Scribd, Inc. Houston, TX CA$184k+

…PyTorch to third party LLM APIs Process massive amounts of data with Python, SQL and Spark Align with stakeholders through written and verbal communications methods on the approaches and results of projects … Hands-on experience building ML pipelines and working with distributed data processing frameworks like Apache Spark, Databricks, or similar. Intermediate level in at least three of these fields: classification algorithms, natural language…

LinkedIn

Data Scientist II

Scribd, Inc. New York, NY CA$184k+

…PyTorch to third party LLM APIs Process massive amounts of data with Python, SQL and Spark Align with stakeholders through written and verbal communications methods on the approaches and results of projects … Hands-on experience building ML pipelines and working with distributed data processing frameworks like Apache Spark, Databricks, or similar. Intermediate level in at least three of these fields: classification algorithms, natural language…

LinkedIn

Data Scientist II

Scribd, Inc. Chicago, IL CA$184k+

…PyTorch to third party LLM APIs Process massive amounts of data with Python, SQL and Spark Align with stakeholders through written and verbal communications methods on the approaches and results of projects … Hands-on experience building ML pipelines and working with distributed data processing frameworks like Apache Spark, Databricks, or similar. Intermediate level in at least three of these fields: classification algorithms, natural language…

LinkedIn
Filters:
× Clear filters