NextRaiseNextRaiseFind jobs
Sign inSign up free
Jobs / Data Engineer in United States of America
12 days ago
Apply with autofill
Apply with autofill
Tp-link-usa-corp·12 days ago
12 days ago

Big Data Engineer

Irvine, United States of AmericaFull-timeMid · 2-4 yearsData Engineer

Sign up free to see how well your resume matches this role.

Boost your chances at tp-link-usa-corp

How you compare FREE

?
Your scoreYour score: not yet known
→
67
Top 10%Top 10%: 67 out of 100

Top 10% of NextRaise users matched against Data Engineer roles in United States.

Must-have skills for this role

  • sql
  • python
  • spark
  • airflow

PDF or DOCX · no account needed

Apply faster with autofill FREEtp-link-usa-corp uses Workable - autofill it instead of retyping.careers.example.com/applyAutofillingFull namePriya SharmaEmailpriya.sharma@example.comPhone+49 30 1234567LocationBerlGet the extension

What you'll do

  • Build, run, and own ETL pipelines on EMR (Spark) orchestrated in Airflow — including their monitoring, alerting, and recovery.
  • Write SQL and Python for data ingestion, transformation, and the datasets that analysts and dashboards depend on.
  • Own data quality for what you build — run the checks before you ship and add new ones where they're missing; confirm the results are correct, not only that the job completed.
  • Build and extend the data models for your area: fact and dimension tables following the team's layering conventions.
  • Keep your jobs efficient — watch runtime and cost, and raise slow or expensive jobs rather than living with them.
  • Use and extend the team's shared patterns and templates, and turn work you find yourself repeating into something automated or reusable.
  • Work with analysts and business stakeholders to turn requests into datasets that actually get used.

What they're looking for

  • 2–4 years of hands-on data development in a production environment — pipelines that run on a schedule with real downstream consumers, that you were responsible for including when they broke.
  • Strong SQL: window functions, complex multi-table joins, incremental loads, and the ability to work out why a query is slow.
  • Python for production ETL and tooling (PySpark, pandas, boto3) — code that runs on a schedule, not only notebooks.
  • Hands-on experience with Spark on a cloud platform: AWS EMR, Databricks, Glue, or equivalent.
  • Production experience with a scheduler, Airflow preferred: DAG design, dependencies, retries, and reruns that are safe to repeat.
  • Working understanding of dimensional modeling: fact and dimension tables, star schema, and warehouse layering.
  • Comfortable working on AWS (S3 with Parquet / ORC), Linux, and Git — and comfortable picking up new tools as the platform evolves.
  • Effective use of AI to solve data problems — using it to move faster on SQL, debugging, and unfamiliar schemas, with the judgment to catch output that looks right but isn't.
  • Bachelor's degree in Computer Science, Information Systems, or a related field, or equivalent practical experience.

Nice to have

  • Data ingestion or CDC tooling: DataX, Sqoop, Debezium, Fivetran, or similar.
  • An OLAP / MPP engine: StarRocks, Doris, ClickHouse, Redshift, or similar.
  • Spark performance work: partitioning, shuffle, skew, and memory tuning.
  • AWS cost optimization: EMR instance sizing and Spot strategy, S3 lifecycle policies.
  • Lakehouse formats (Iceberg, Hudi, Delta), streaming (Kafka, Flink), dbt, or a data quality framework.
  • QuickSight or another BI tool: dataset and permission design.
  • Self-directed learning: a project, an open-source contribution, or a tool you picked up on your own and put to real use

Summarised by NextRaise from the employer’s description, which follows in full below.

Full description from employer

ABOUT US:

Headquartered in the United States, TP-Link Systems Inc. is a global provider of reliable networking devices and smart home products, consistently ranked as the world’s top provider of Wi-Fi devices. The company is committed to delivering innovative products that enhance people’s lives through faster, more reliable connectivity. With a commitment to excellence, TP-Link serves customers in over 170 countries and continues to grow its global footprint.

We believe technology changes the world for the better! At TP-Link Systems Inc, we are committed to crafting dependable, high-performance products to connect users worldwide with the wonders of technology. 

Embracing professionalism, innovation, excellence, and simplicity, we aim to assist our clients in achieving remarkable global performance and enable consumers to enjoy a seamless, effortless lifestyle. 

KEY RESPONSIBILITIES

  • Build, run, and own ETL pipelines on EMR (Spark) orchestrated in Airflow — including their monitoring, alerting, and recovery.
  • Write SQL and Python for data ingestion, transformation, and the datasets that analysts and dashboards depend on.
  • Own data quality for what you build — run the checks before you ship and add new ones where they're missing; confirm the results are correct, not only that the job completed.
  • Build and extend the data models for your area: fact and dimension tables following the team's layering conventions.
  • Keep your jobs efficient — watch runtime and cost, and raise slow or expensive jobs rather than living with them.
  • Use and extend the team's shared patterns and templates, and turn work you find yourself repeating into something automated or reusable.
  • Work with analysts and business stakeholders to turn requests into datasets that actually get used.

Requirements

REQUIRED QUALIFICATIONS

  • 2–4 years of hands-on data development in a production environment — pipelines that run on a schedule with real downstream consumers, that you were responsible for including when they broke.
  • Strong SQL: window functions, complex multi-table joins, incremental loads, and the ability to work out why a query is slow.
  • Python for production ETL and tooling (PySpark, pandas, boto3) — code that runs on a schedule, not only notebooks.
  • Hands-on experience with Spark on a cloud platform: AWS EMR, Databricks, Glue, or equivalent.
  • Production experience with a scheduler, Airflow preferred: DAG design, dependencies, retries, and reruns that are safe to repeat.
  • Working understanding of dimensional modeling: fact and dimension tables, star schema, and warehouse layering.
  • Comfortable working on AWS (S3 with Parquet / ORC), Linux, and Git — and comfortable picking up new tools as the platform evolves.
  • Effective use of AI to solve data problems — using it to move faster on SQL, debugging, and unfamiliar schemas, with the judgment to catch output that looks right but isn't.
  • Bachelor's degree in Computer Science, Information Systems, or a related field, or equivalent practical experience.

PREFERRED QUALIFICATIONS

  • Data ingestion or CDC tooling: DataX, Sqoop, Debezium, Fivetran, or similar.
  • An OLAP / MPP engine: StarRocks, Doris, ClickHouse, Redshift, or similar.
  • Spark performance work: partitioning, shuffle, skew, and memory tuning.
  • AWS cost optimization: EMR instance sizing and Spot strategy, S3 lifecycle policies.
  • Lakehouse formats (Iceberg, Hudi, Delta), streaming (Kafka, Flink), dbt, or a data quality framework.
  • QuickSight or another BI tool: dataset and permission design.
  • Self-directed learning: a project, an open-source contribution, or a tool you picked up on your own and put to real use

Benefits

Base Salary Range: $100 - 120K

  • Free snacks and drinks
  • Fully paid medical, dental, and vision insurance (partial coverage for dependents)
  • Contributions to 401K funds
  • Bi-annual reviews, and annual pay increases
  • Health and wellness benefits, including free gym membership
  • Quarterly team-building events

At TP-Link Systems Inc., we are continually searching for ambitious individuals who are passionate about their work. We believe that diversity fuels innovation, collaboration, and drives our entrepreneurial spirit. As a global company, we highly value diverse perspectives and are committed to cultivating an environment where all voices are heard, respected, and valued. We are dedicated to providing equal employment opportunities to all employees and applicants, and we prohibit discrimination and harassment of any kind based on race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state, or local laws. Beyond compliance, we strive to create a supportive and growth-oriented workplace for everyone. If you share our passion and connection to this mission, we welcome you to apply and join us in building a vibrant and inclusive team at TP-Link Systems Inc.

Please, no third-party agency inquiries, and we are unable to offer visa sponsorships at this time.

Company

Tp-link-usa-corp
Irvine, United States of America

Company facts come from this company's own listings. We only show what the postings themselves carry.

Sourced from Tp Link Usa Corp's careers site·first seen 15 Sept 2026·last verified 15 Sept 2026·How we source jobs

Similar jobs

  • Sr. Data Engineer at nbcuniversal3New York, United States of America–match not yet calculated
  • Senior Synthetic Data Engineer - Autonomous Driving at NVIDIASanta Clara, United States of America–match not yet calculated
  • NCIS Data Engineer | Active Secret clearance at gditUSA VA Quantico–match not yet calculated
  • Senior Data Engineer, Selling Partner Insights and Analytics at AmazonSeattle, United States of America–match not yet calculated
  • Senior Data Engineer (Scala) - Remote (USA) at icfReston, United States of America–match not yet calculated

Browse more jobs

  • Data Engineer jobs in United States
  • Data Scientist jobs in United States
  • Data Analyst jobs in United States
  • Business Intelligence Analyst jobs in United States
  • Data Engineer jobs in India
  • Data Engineer jobs in United Kingdom