Full-Time

Apache Spark Developer

Bright Vision Technologies

Bright Vision Technologies

51-200 employees

AI-powered talent intelligence and automation platform

Compensation Overview

$125k - $185k/yr

No H1B Sponsorship

Remote in USA

Remote

Category
Data & Analytics (1)
Required Skills
Datadog
Kubernetes
Microsoft Azure
Grafana
Airflow
Apache Flink
GitHub Actions
Git
Apache Spark
SQL
Apache Kafka
Postgres
MLflow
OpenTelemetry
Docker
AWS
Scala
Prometheus
Jenkins
Terraform
Apache Hive
Hadoop
Yarn
DevOps
Oracle
Databricks
Snowflake

Get referred to Bright Vision Technologies

See people who can refer or advise you

Requirements
  • Six or more years of professional software or data engineering experience.
  • Four or more years of hands-on Apache Spark development experience in enterprise production environments.
  • Strong proficiency in PySpark, Scala, or Spark SQL for distributed data processing.
  • Deep understanding of Apache Spark architecture, including RDDs, DataFrames, Datasets, Catalyst Optimizer, DAG execution, and Tungsten engine.
  • Strong experience with distributed computing concepts, including partitioning, shuffling, caching, broadcast joins, and fault tolerance.
  • Advanced SQL skills with databases such as SQL Server, Oracle, PostgreSQL, Snowflake, or Teradata.
  • Experience working with Hadoop ecosystem technologies, including Hive, HDFS, YARN, and Parquet.
  • Experience processing streaming data using Spark Structured Streaming, Apache Kafka, or Event Hubs.
  • Hands-on experience with cloud platforms including Azure Databricks, AWS EMR, AWS Glue, Azure Synapse Analytics, or Google Dataproc.
  • Experience integrating Spark applications with Delta Lake, Apache Iceberg, or Apache Hudi.
  • Strong understanding of data warehousing concepts, dimensional modeling, and data lake architecture.
  • Experience using Git, CI/CD pipelines, Azure DevOps, GitHub Actions, or Jenkins.
  • Strong debugging, troubleshooting, and Spark performance tuning skills.
  • Experience working in Agile Scrum development environments.
Responsibilities
  • Design, develop, and maintain high-performance distributed data processing applications using Apache Spark.
  • Build scalable batch and real-time ETL/ELT pipelines processing large volumes of enterprise data.
  • Develop Spark applications using PySpark, Scala, or Spark SQL for data transformation, aggregation, and analytics.
  • Optimize Spark jobs for memory utilization, partitioning strategies, shuffle performance, and execution efficiency.
  • Process structured, semi-structured, and streaming data from enterprise databases, APIs, Kafka, cloud storage, and data lakes.
  • Develop reusable Spark libraries, data processing frameworks, and metadata-driven ingestion pipelines.
  • Collaborate with cloud engineering teams to deploy Spark workloads on Databricks, EMR, Azure Synapse, or Kubernetes.
  • Implement data quality validation, reconciliation, monitoring, and automated error handling across distributed pipelines.
  • Integrate Spark applications with enterprise data warehouses, lakehouses, and reporting platforms.
  • Participate in architecture reviews, code reviews, technical design discussions, and Agile development activities.
  • Troubleshoot production issues involving distributed processing, cluster performance, resource utilization, and data quality.
  • Support cloud migration initiatives by modernizing legacy ETL workloads into Spark-based architectures.
Desired Qualifications
  • Experience building enterprise lakehouse architectures using Databricks or Delta Lake.
  • Familiarity with Apache Airflow, Azure Data Factory, AWS Step Functions, or Control-M for workflow orchestration.
  • Experience with machine learning workflows using Spark MLlib, MLflow, or feature engineering pipelines.
  • Knowledge of Kubernetes, Docker, and containerized Spark deployments.
  • Experience implementing data quality frameworks using Great Expectations or Deequ.
  • Familiarity with Apache NiFi, Apache Flink, Trino, or Presto.
  • Experience working with cloud object storage including Amazon S3, Azure Data Lake Storage Gen2, or Google Cloud Storage.
  • Knowledge of Infrastructure as Code using Terraform or ARM templates.
  • Experience with enterprise monitoring tools including Prometheus, Grafana, Datadog, or OpenTelemetry.
  • Cloud certifications in Azure, AWS, Databricks, or Apache Spark-related technologies.
Bright Vision Technologies

Bright Vision Technologies

View

Bright Vision Technologies offers Lumina, an AI-powered platform for talent intelligence and enterprise automation, along with consulting and staffing services. Lumina analyzes unstructured data with generative AI, Large Language Models orchestrated via LangChain, and Retrieval-Augmented Generation to support sourcing and screening candidates, integrated with CRM, ERP, and HRMS on a cloud-native microservices architecture across AWS, Azure, and Google Cloud. The company differentiates itself by combining a sophisticated AI product with hands-on consulting and staffing expertise, plus its minority-owned status and a dual model that helps clients implement technology and solve broader business challenges. Its goal is to streamline and automate complex workflows in IT talent acquisition and management to enable digital transformation and efficient enterprise operations.

Company Size

51-200

Company Stage

N/A

Total Funding

N/A

Headquarters

N/A

Founded

2020

Get referred to Bright Vision Technologies

See people who can refer or advise you

Simplify Jobs

Simplify's Take

What believers are saying

  • April 2026 LinkedIn launch signals active product commercialization and go-to-market acceleration.
  • August 2026 site updates show hiring across Bridgewater, Princeton, and Chennai.
  • The company’s minority-owned status and U.S.-India footprint strengthen enterprise procurement positioning.

What critics are saying

  • No disclosed funding, customers, or revenue makes Lumina's traction impossible to verify.
  • The June 2026 pivot from staffing to product risks channel conflict and execution drag.
  • Enterprise AI recruiting faces entrenched rivals like Workday, LinkedIn, and Eightfold by 2027.

What makes Bright Vision Technologies unique

  • April 2026 launch of Lumina combines talent intelligence, automation, and hybrid cloud.
  • Lumina integrates RAG, LLM orchestration, semantic matching, and blockchain credential verification.
  • Bright Vision serves staffing and consulting clients, enabling implementation alongside the product.

Help us improve and share your feedback! Did you find this helpful?

Benefits

Remote Work Options