Apply NowSave Job

Data Engineer

Wynd labs
RemoteRemote
Full Time
Posted Yesterday
Celery, Kafka, RabbitMQ*Python*

Role Overview

We are seeking a Data Engineer to support and improve large-scale data pipelines and infrastructure. You’ll work across data collection, processing, transformation, validation, and delivery, with a focus on scalability, reliability, and performance.

This is a hands-on role where you’ll work with distributed systems, large datasets, web scraping infrastructure, and production data workloads.

Please note: This role requires a work schedule that overlaps sufficiently with EST business hours to collaborate effectively with the team.

Responsibilities

What You'll Be Doing:

  • Maintain, optimize, and troubleshoot database queries and related data systems to support efficient data access, processing, and reliability.
  • Assist in creating, maintaining, and improving data pipelines used to collect, process, transform, validate, and deliver large-scale datasets.
  • Support web scraping and data collection initiatives, including developing, testing, and maintaining scripts or tools used to gather publicly available data in accordance with Company requirements.
  • Monitor and troubleshoot data pipeline issues, identify data quality concerns, and help implement timely fixes to maintain data accuracy and operational continuity.
  • Document engineering work, including database queries, pipeline processes, scraping workflows, technical decisions, issues encountered, and resolutions implemented.
  • Participate in research and development projects to improve the Company’s data products and workflows.

Why Work With Us:

  • Opportunity. We are at the forefront of developing a web-scale crawler and knowledge graph that improves access to public web data and extends the value of AI to the people.
  • Culture. We're a lean team with a high bar. We come to work not to be comfortable, but to find out what we're capable of and to do work that matters. We're not calling for people who keep things moving. We're calling for people who make everyone around them better.
  • We prioritize low ego and high output. This is a fully remote team.
  • Compensation. You’ll receive a competitive salary & benefits.


Requirements

Who You Are:

  • Bachelor’s degree or equivalent work experience
  • Python (advanced) — strong grasp of async programming, multiprocessing, and writing production-grade code for long-running data jobs
  • Web scraping at scale — hands-on experience with high-volume scraping (proxies, rate limiting, anti-bot evasion). Experience with platform APIs and large media/metadata datasets (video platforms, social media)
  • Distributed data pipelines — experience designing and operating pipelines across many workers/servers using task queues (Celery, Kafka, RabbitMQ, or similar)
  • Data warehousing — practical experience with columnar/analytical warehouses; Databend, ClickHouse, or BigQuery strongly preferred; comfortable with complex analytical queries, partitioning strategies, cost-aware querying on cloud warehouses
  • Docker & Kubernetes — containerizing workloads, writing Helm charts/manifests, managing deployments, autoscaling scraping/processing workloads
  • Linux & bare-metal ops — comfortable managing services on Linux servers, debugging performance issues (disk I/O, network, memory) without managed-cloud abstractions
  • CI/CD for data workflows (GitHub Actions, ArgoCD)
  • Writing Scalable API


Contact Wynd labs

Job Details

LocationRemote
Job TypeFull Time
Experience LevelMid Level
EducationBachelor's Degree
PostedOctober 8, 2026 at 06:12 AM

Ready to Apply?

Don't miss out on this opportunity. Apply now and take the next step in your career.

Apply Now
Data Engineer at Wynd labs | TechiHub Jobs | TechiHub