Hadoop 解决方案开发工程师
Hadoop Solutions Developer
Hadoop 解决方案开发人员 - 远程办公
Bright Vision Technologies 是一家技术咨询和软件开发公司,为美国各地提供云、人工智能、数据和企业解决方案。
这是一个加入一家知名且受人尊敬的组织的绝佳机会,提供巨大的职业发展潜力。
职位名称:Hadoop 解决方案开发人员
工作地点:100% 远程(美国)
职位类型:全职,直接 W2
薪资范围:每年 10 万至 15 万美元
所需经验:6 年以上
赞助:美国公民、绿卡持有者、EAD 持有者以及 H-1B 转移候选人欢迎申请。我们无法为该职位的新 H-1B 签证申请提供赞助。
职位简介
我们正在寻找一位经验丰富的 Hadoop 开发人员,在 Hadoop 及相关大数据生态系统上设计、构建和运营大规模数据处理管道和分析平台。在该职位中,您将负责摄取、转换和分析大量结构化和非结构化数据,以支持企业分析、机器学习和报告工作负载。理想的候选人应具备 Hadoop 生态系统中的深厚技术专长,同时拥有扎实的软件工程基础,并清楚了解如何在生产环境中交付可靠、高性能且成本效益高的数据平台。
主要职责
· 在 Hadoop 上设计、开发和运营端到端的大数据管道,从关系型、基于文件、流式和 API 驱动的多种数据源中摄取数据。
- 使用 Apache Spark、Hive、Pig 和 Sqoop 构建稳健的 ETL/ELT 工作流,注重数据质量、幂等性、错误处理和可恢复性。
- 使用 Kafka、Spark Streaming 或 Flink 开发高吞吐量的流数据管道,并将其与下游分析和运营系统集成。
- 通过仔细调整分区、内存、序列化和数据倾斜处理来优化 Spark 和 MapReduce 作业,以最低成本满足严格的 SLA 要求。
- 在 HDFS、Hive、HBase 和现代湖仓格式(Parquet、ORC、Delta、Iceberg、Hudi)上设计和维护数据模型和存储布局,平衡灵活性和性能。
- 与数据治理和安全团队合作,实施数据治理、血缘关系和质量控制。
- 为大数据管道构建稳健的监控、警报和日志策略,包括作业级别的 SLA 和主动故障检测。
- 配合数据工程师和数据科学家进行数据产品设计和实现,确保数据质量和一致性。
- 参与架构评审和技术决策,推动最佳实践和标准化流程。
查看英文原文
Hadoop Solutions Developer - Remote
Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.
This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.
Job Title: Hadoop Solutions Developer
Location: 100% Remote (U.S.)
Position Type: Full-time, Direct W2
Salary Range: $100,000–$150,000 Annually
Experience Required: 6+ years
Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.
Job Summary
We are seeking an experienced Hadoop Developer to design, build, and operate large-scale data processing pipelines and analytics platforms on Hadoop and related big-data ecosystems. In this role you will be responsible for ingesting, transforming, and analyzing massive volumes of structured and unstructured data to support enterprise analytics, machine learning, and reporting workloads. The ideal candidate will combine deep technical expertise across the Hadoop ecosystem with strong software engineering fundamentals and a clear understanding of how to deliver reliable, performant, and cost-effective data platforms in production environments.
Key Responsibilities
· Design, develop, and operate end-to-end big-data pipelines on Hadoop, ingesting data from a diverse mix of relational, file-based, streaming, and API-driven sources.
- Build robust ETL/ELT workflows using Apache Spark, Hive, Pig, and Sqoop, with strong attention to data quality, idempotency, error handling, and recoverability.
- Develop high-throughput streaming data pipelines using Kafka, Spark Streaming, or Flink, and integrate them with downstream analytical and operational systems.
- Optimize Spark and MapReduce jobs through careful tuning of partitioning, memory, serialization, and skew handling to meet demanding SLAs at minimal cost.
- Design and maintain data models and storage layouts on HDFS, Hive, HBase, and modern lakehouse formats (Parquet, ORC, Delta, Iceberg, Hudi) to balance flexibility and performance.
- Implement data governance, lineage, and quality controls in collaboration with data governance and security teams.
- Build robust monitoring, alerting, and logging strategies for big-data pipelines, including job-level SLAs and proactive failure detection.
- Partner with data scientists and analysts to deliver curated, reliable, and well-documented datasets that accelerate their work.
- Automate pipeline orchestration using Airflow, Oozie, or similar workflow engines, with clean dependency management and clear ownership boundaries.
- Continuously evaluate and adopt new technologies in the big-data and cloud ecosystem (EMR, Databricks, Snowflake, BigQuery) where they offer meaningful improvements.
- Lead performance reviews and architecture audits of existing pipelines, proposing concrete refactoring and optimization initiatives.
- Document data architectures, schemas, pipeline behaviors, and operational runbooks in a way that makes the platform supportable as the team scales.
- Mentor junior engineers and contribute to the team’s engineering standards and best practices.
Required Qualifications
· Bachelor’s degree in Computer Science, Engineering, or a related technical discipline.
- Five or more years of professional experience designing and operating big-data pipelines on Hadoop.
- Strong hands-on expertise with Apache Spark (Scala, Python, or Java) in production environments.
- Solid experience with Hive, HDFS, Sqoop, HBase, and the broader Hadoop ecosystem.
- Hands-on experience with streaming data platforms such as Kafka, Spark Streaming, or Flink.
- Strong SQL skills and experience working with both relational and NoSQL data stores.
- Experience with workflow orchestration tools such as Airflow or Oozie.
- Solid understanding of distributed systems concepts, including partitioning, replication, and fault tolerance.
- Strong scripting skills in Python or Shell.
- Excellent troubleshooting, debugging, and documentation skills.
Preferred Qualifications
· Experience operating Hadoop on cloud platforms such as AWS EMR, Azure HDInsight, or Databricks.
- Familiarity with modern lakehouse formats (Delta, Iceberg, Hudi).
- Exposure to data governance tooling such as Apache Atlas or Collibra.
- Experience with Kubernetes-based data platforms (Spark-on-K8s, Trino).
- Hands-on experience with CI/CD and infrastructure-as-code in data engineering workflows.
How to Apply
Would you like to know more about this opportunity? For immediate consideration, please send your resume to or contact us at (908) 505-3544. Learn more about Bright Vision Technologies at .
Bright Vision Technologies is an Equal Opportunity Employer.
Equal Employment Opportunity (EEO) Statement
Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall.
BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.
Originally posted on Himalayas