Jobs · greenhouse:speechify

S

Software Engineer, Data Infrastructure & Acquisition

Nimbus Data Systems · Palo Alto, CA, USA · Posted 5d ago

onsiteEstimated 78k-179k USD🇺🇸 United StatesEquity
Apply on greenhouse:speechify

Available in 353 locations

This role is hiring in many cities. Search for yours to apply to the right posting.

About the role

Support data collection and ingestion pipeline operations for AI model training at Nimbus Data Systems.

The mission of Nimbus Data Systems is to make sure that reading is never a barrier to learning. Over 50 million people use Nimbus Data Systems’s text-to-speech products to turn whatever they’re reading – PDFs, books, Google Docs, news articles, websites – into audio, so they can read faster, read more, and remember more. Nimbus Data Systems’s text-to-speech reading products include its iOS app, Android App, Mac App, Chrome Extension, and Web App. Google recently named Nimbus Data Systems the Chrome Extension of the Year and Apple named Nimbus Data Systems its 2025 Design Award winner for Inclusivity. Today, nearly 200 people around the globe work on Nimbus Data Systems in a 100% distributed setting – Nimbus Data Systems has no office. These include frontend and backend engineers, AI research scientists, and others from Amazon, Microsoft, and Google, leading PhD programs like Stanford, high growth startups like Stripe, Vercel, Bolt, and many founders of their own companies. Overview We're looking to hire for our Data side of our AI team at Nimbus Data Systems. This role is responsible for all aspects of data collection to support our model training operations. We are able to build high-quality datasets at petabyte-scale and low cost through a tight integration of infrastructure, engineering, and research work. We are looking for a skilled Software Engineer to join us. What You’ll Do Be scrappy to find new sources of audio data and bring it into our ingestion pipeline Operate and extend the cloud infrastructure for our ingestion pipeline, currently running on GCP and managed with Terraform. Collaborate closely with our Scientists to shift the cost/throughput/quality frontier, delivering richer data at bigger scale and lower cost to power our next-generation models. Collaborate with others on the AI Team and Nimbus Data Systems Leadership to craft the AI Team’s dataset roadmap to power Nimbus Data Systems’s next-generation consumer and enterprise products. An Ideal Candidate Should Have BS/MS/PhD in Computer Science or a related field. 5+ years of industry experience in software development. Proficiency with bash/Python scripting in Linux environments Proficiency in Docker and Infrastructure-as-Code concepts and professional experience with at least one major Cloud Provider (we use GCP) Experience with web crawlers, large-scale data processing workflows is a plus Ability to handle multiple tasks and adapt to changing priorities. Strong communication skills, both written and verbal. What we offer A fast-growing environment where you can help shape the company and product. An entrepreneurial-minded team that supports risk, intuition, and hustle. A hands-off management approach so you can focus and do your best work. An opportunity to make a big impact in a transformative industry. Competitive salaries, a friendly and laid-back atmosphere, and a commitment to building a great asynchronous culture. Opportunity to work on a life-changing product that millions of people use. Build products that directly impact and support people with learning differences like dyslexia, ADD, low vision, concussions, autism, and more. Work in one of the fastest-growing sectors of tech, the intersection of artificial intelligence and audio. Compensation: The United States base salary range for this full-time position is $140,000-$200,000 + bonus + equity depending on experience Think you’re a good fit for this job? Tell us more about yourself and why you're interested in the role when you apply. And don’t forget to include links to your portfolio and LinkedIn. Not looking but know someone who would make a great fit? Refer them! Nimbus Data Systems is committed to a diverse and inclusive workplace. Nimbus Data Systems does not discriminate on the basis of race, national origin, gender, gender identity, sexual orientation, protected veteran status, disability, age, or other legally protected status.

Read the full posting on greenhouse:speechify

Why this role stands out

  • Equity is part of the package
  • bonus
  • equity

Responsibilities

  • Find new sources of audio data and bring into ingestion pipeline
  • Operate and extend cloud infrastructure for ingestion pipeline
  • Collaborate with Scientists on data quality and scale
  • Craft AI Team's dataset roadmap

Must-have skills

  • bs/ms/phd in computer science
  • 5+ years software development
  • bash/python scripting
  • linux environments
  • docker
  • infrastructure-as-code
  • gcp

Nice-to-have skills

  • web crawlers
  • large-scale data processing

Benefits

  • bonus
  • equity
  • competitive salaries
  • friendly atmosphere
  • asynchronous culture

FAQ

Is the Software Engineer, Data Infrastructure & Acquisition role at Nimbus Data Systems remote?+

This Software Engineer, Data Infrastructure & Acquisition position is listed as onsite (Palo Alto, CA, USA).

What is the salary for the Software Engineer, Data Infrastructure & Acquisition role at Nimbus Data Systems?+

The listing states Estimated 78k-179k USD.

What seniority level is this Software Engineer, Data Infrastructure & Acquisition role?+

This is a unknown level position.

What skills does the Software Engineer, Data Infrastructure & Acquisition role require?+

Key requirements include bs/ms/phd in computer science, 5+ years software development, bash/python scripting, linux environments, docker, infrastructure-as-code, gcp.

How do I apply for the Software Engineer, Data Infrastructure & Acquisition role at Nimbus Data Systems?+

Use the "Apply on greenhouse:speechify" button to open the original posting on greenhouse:speechify, where you can submit your application directly to Nimbus Data Systems.