Back to Jobs
Wynd Labs
Development 1d ago

Senior Backend Engineer (Data Pipeline Infrastructure)

Wynd Labs
United StatesUnited States
Full-time
Not Disclosed
Senior-Level

Job Description

Key Skills Required

Master these to land this role

Python2h 41mFree Trial ✨
Start 10-Day Free Trial
Backend42mFree Trial ✨
Start 10-Day Free Trial
GolangKubernetesDistributed SystemsData Pipeline Infrastructure

Want to know if you're a match for this job?

Calculate My Match Score

Who We Are:

We build infrastructure that delivers massive amounts of web data to the companies training the world’s most powerful AI models.

We're the team that helps to power and support Grass, a bandwidth-sharing network that lets us operate a massive distributed crawler, giving us unique access to high-quality public web data at global scale. On top of that, we’ve built pipelines for ingesting, segmenting, and annotating billions of videos, transcripts, and audio files, powering dataset creation for frontier labs.

We’re lean, technical, and move fast. No red tape, no slow decision-making; just a team of builders pushing to expand what’s possible for open web data and AI.

The Role:

We are seeking a Senior Backend Engineer with deep expertise in backend development, particularly in data pipeline infrastructure. In this role, you will join a small, high-performing team to design and architect scalable systems, drive technical excellence, and help advance Grass’s role in shaping the future of the internet.

Note: This role requires a work schedule that sufficiently overlaps with EST business hours to collaborate effectively with the team.

Who You Are:

  • A candidate with 7+ years of experience in software development, with a track record of writing high-quality, maintainable, and robust code.

  • A Bachelor’s or Master’s degree in a STEM field, or equivalent practical experience.

  • Strong experience building and scaling large, distributed systems.

  • Deep expertise in designing, troubleshooting, and optimizing complex, live software systems.

  • Hands-on experience with modern development practices, including continuous integration and continuous deployment.

  • Strong analytical and problem-solving skills, particularly in diagnosing and resolving data flow and system performance issues.

  • Professional or native English proficiency.

What You'll Be Doing:

  • Design, build, and optimize scalable data pipeline infrastructure for real-time and batch data processing.

  • Develop new backend features and system improvements, focusing on reliability and performance.

  • Write clear, well-tested, and well-documented code.

  • Create and review technical designs, code, and documentation to maintain high engineering standards.

  • Contribute to Wynd’s infrastructure across mobile, desktop, and server-side systems.

Technologies Used:

  • Golang, Python, Redis Clusters, Kubernetes.

Why Work With Us:

  • Opportunity to work at the forefront of developing a web-scale crawler and knowledge graph that improves access to public web data and extends the value of AI to the people.

  • Culture of a lean team with a high bar, prioritizing low ego and high output. Fully remote team.

  • Competitive salary, benefits, and equity package.

How would you rate this job post?

See what other professionals think about this role.

banner

Wynd Labs is a premier, enterprise-grade data infrastructure platform engineered to orchestrate massive-scale public web data ecosystems and intelligent artificial intelligence training workflows. Operating as a highly integrated decentralized data hub, the company eliminates the operational friction of traditional localized web scraping by seamlessly deploying advanced distributed crawling telemetry, rigorous multimodal data pipelines, and cohesive residential proxy architectures. Moving beyond rigid legacy dataset providers, Wynd Labs empowers frontier AI labs, elite research teams, and data-driven enterprises to dynamically synchronize their machine learning models with instantaneous, internet-scale data ingestion. Under the hood, their sophisticated backend infrastructure natively handles complex high-throughput routing, scalable real-time search extraction, and seamless petabyte-scale multimedia annotation, ensuring frictionless data accessibility and uncompromising model training readiness. What sets Wynd Labs apart is its uncompromising dedication to frictionless data orchestration; by bridging the gap between decentralized bandwidth sharing and rigorous artificial intelligence development, the platform empowers organizations to radically accelerate their algorithmic velocity, optimize data acquisition, and build an unassailable foundation for continuous AI dominance in the modern computational landscape.

Safety First

  • Never pay for a job application.
  • Do not share sensitive bank info.
  • Verify the client before starting work.
Learn More