NextRaiseNextRaiseFind jobs
Sign inSign up free
Jobs / Data Engineer in India
5 days ago
Apply with autofill
Apply with autofill
Botvfx·5 days ago
5 days ago

Data Engineer (ETL Specialist)

Chennai, IndiaFull-timeMid · 3-5 yearsData Engineer

Sign up free to see how well your resume matches this role.

Boost your chances at botvfx

How you compare FREE

?
Your scoreYour score: not yet known
→
82
Top 10%Top 10%: 82 out of 100

Top 10% of NextRaise users matched against Data Engineer roles in India.

Must-have skills for this role

  • python
  • apache spark
  • etl
  • elt

PDF or DOCX · no account needed

Apply faster with autofill FREEThe NextRaise extension autofills your application in one click.careers.example.com/applyAutofillingFull namePriya SharmaEmailpriya.sharma@example.comPhone+49 30 1234567LocationBerlGet the extension

What you'll do

  • Lead the architecture, implementation, and automation of scalable ETL/ELT pipelines capable of moving and processing tens of terabytes of unstructured media data daily.
  • Implement automated "Purge & Keep" business rulesets to systematically clean directory trees, wiping intermediate render caches, temporary scratch files, and duplicates while preserving core visual assets.
  • Build high-throughput processing pipelines to transcode, compress, and restructure massive image sequences (linear EXR, DPX) and video formats into AI-friendly data structures (e.g., WebDataset shards).
  • Integrate and normalize lightweight metadata scrapings (camera telemetry, tracking curves, geometry tags) extracted from VFX workfiles (Silhouette, Nuke) and bind them deterministically to their corresponding visual assets.
  • Manage high-velocity data movement across tiered storage layers (Hot NVMe/SSD staging, Warm Object Storage, and Deep Cloud/Tape Archives), ensuring the pipeline never bottlenecks network bandwidth or staging limits.
  • Establish robust data quality assurance (QA) validation loops within the pipeline to catch corrupted frames, missing frame ranges, and accidental inclusion of restricted IP or sensitive assets.

What they're looking for

  • At least 3–5 years of experience as a Data Engineer building large-scale, automated ETL/ELT pipelines, with a proven track record of handling unstructured data, heavy media assets, or computer vision datasets.
  • Strong technical grounding in pipeline development using Python (and bash/shell scripting) along with heavy data processing frameworks (e.g., Apache Spark, Ray, Dask).
  • Hands-on experience with modern data architectures, containerization (Docker, Kubernetes), and cloud/on-prem block/object storage systems.
  • Solid problem-solving capabilities focused on maximizing throughput under strict hardware constraints (balancing egress speeds with staging capacities).
  • Exceptional communication skills with the ability to collaborate effectively across multidisciplinary teams, including VFX technical directors, systems engineers, and AI research teams.

Nice to have

  • Basic familiarity with VFX file structures, directory archetypes, image sequence formatting, or open-source media utilities (FFmpeg, OpenEXR toolset).

Summarised by NextRaise from the employer’s description, which follows in full below.

Full description from employer

POSITION SUMMARY

BOT VFX, a visual effects services company with global clients, is looking for a Data Engineer (ETL Specialist) as a full-time contract role. This is strictly an on location role in Chennai, India.
In this role, you will be responsible for transforming petabytes of unstructured, legacy VFX data (including heavy video plates, multi-layered OpenEXR sequences, and proprietary DCC workfiles) into highly structured, sanitized, and compressed datasets.
The ideal candidate brings deep expertise in large-scale data engineering and ETL pipeline architecture alongside an understanding of physical infrastructure limits. You will bridge legacy storage/tape systems, VFX pipeline scripts, and Creative Innovation needs. Reporting to the CIO, the role works closely with IT/IO and Creative Innovation teams.

POSITION RESPONSIBILITY
  • Lead the architecture, implementation, and automation of scalable ETL/ELT pipelines capable of moving and processing tens of terabytes of unstructured media data daily.
  • Implement automated "Purge & Keep" business rulesets to systematically clean directory trees, wiping intermediate render caches, temporary scratch files, and duplicates while preserving core visual assets.
  • Build high-throughput processing pipelines to transcode, compress, and restructure massive image sequences (linear EXR, DPX) and video formats into AI-friendly data structures (e.g., WebDataset shards).
  • Integrate and normalize lightweight metadata scrapings (camera telemetry, tracking curves, geometry tags) extracted from VFX workfiles (Silhouette, Nuke) and bind them deterministically to their corresponding visual assets.
  • Manage high-velocity data movement across tiered storage layers (Hot NVMe/SSD staging, Warm Object Storage, and Deep Cloud/Tape Archives), ensuring the pipeline never bottlenecks network bandwidth or staging limits.
  • Establish robust data quality assurance (QA) validation loops within the pipeline to catch corrupted frames, missing frame ranges, and accidental inclusion of restricted IP or sensitive assets.

REQUIRED SKILLS
  • At least 3–5 years of experience as a Data Engineer building large-scale, automated ETL/ELT pipelines, with a proven track record of handling unstructured data, heavy media assets, or computer vision datasets.
  • Strong technical grounding in pipeline development using Python (and bash/shell scripting) along with heavy data processing frameworks (e.g., Apache Spark, Ray, Dask).
  • Hands-on experience with modern data architectures, containerization (Docker, Kubernetes), and cloud/on-prem block/object storage systems.
  • Solid problem-solving capabilities focused on maximizing throughput under strict hardware constraints (balancing egress speeds with staging capacities).
  • Exceptional communication skills with the ability to collaborate effectively across multidisciplinary teams, including VFX technical directors, systems engineers, and AI research teams.


Preferred Skills
  • Basic familiarity with VFX file structures, directory archetypes, image sequence formatting, or open-source media utilities (FFmpeg, OpenEXR toolset).

Company

Botvfx
Chennai, India

Company facts come from this company's own listings. We only show what the postings themselves carry.

Sourced from Botvfx's careers site·first seen 15 Sept 2026·last verified 15 Sept 2026·How we source jobs

Similar jobs

  • Cloud Data Engineer - Snowflake, DBT, Airflow and AWS. at synechronBengaluru, India–match not yet calculated
  • Assoc Director- Platform & Data Engineer at NovartisHyderabad, India–match not yet calculated
  • BODS Data Engineer at weekdayworksBengaluru, India–match not yet calculated
  • Data Engineer at coretek-servicesKondapur, India–match not yet calculated
  • Data Engineer - BODS at Weekday AIPune, India–match not yet calculated

Browse more jobs

  • Data Engineer jobs in India
  • Data Analyst jobs in India
  • Data Scientist jobs in India
  • Business Intelligence Analyst jobs in India
  • Data Engineer jobs in United States
  • Data Engineer jobs in United Kingdom