{"jobs":[{"id":"e1415ae0-801b-45ec-bab6-624d71f07b44","title":"Web Scraping Specialist","department":"Analytics","team":"Analytics","employmentType":"FullTime","location":"Remote","shouldDisplayCompensationOnJobPostings":false,"secondaryLocations":[],"publishedAt":"2026-09-04T21:54:10.951+00:00","isListed":true,"isRemote":null,"workplaceType":null,"address":{"postalAddress":{"addressCountry":"United States","addressLocality":""}},"jobUrl":"https://jobs.ashbyhq.com/wynd-labs/e1415ae0-801b-45ec-bab6-624d71f07b44","applyUrl":"https://jobs.ashbyhq.com/wynd-labs/e1415ae0-801b-45ec-bab6-624d71f07b44/application","descriptionHtml":"<p style=\"min-height:1.5em\"><strong>Who We Are:</strong></p><p style=\"min-height:1.5em\">We build infrastructure that delivers massive amounts of web data to the companies training the world’s most powerful AI models.</p><p style=\"min-height:1.5em\">We're the team that helps to power and support Grass, a bandwidth-sharing network that lets us operate a massive distributed crawler, giving us unique access to high-quality public web data at global scale. On top of that, we’ve built pipelines for ingesting, segmenting, and annotating billions of videos, transcripts, and audio files, powering dataset creation for frontier labs.</p><p style=\"min-height:1.5em\">We’re lean, technical, and move fast. No red tape, no slow decision-making; just a team of builders pushing to expand what’s possible for open web data and AI.</p><p style=\"min-height:1.5em\"><strong>The Role.</strong></p><p style=\"min-height:1.5em\">We are seeking a Web Scraping Specialist who is proficient and brings significant experience in data extraction and web scraping techniques. You will join a small, specialized team and lead efforts to gather and analyze data, optimize scraping processes, and support our vision for a future where Grass plays a crucial role in transforming internet data accessibility.</p><p style=\"min-height:1.5em\"><strong>Please note: </strong><em><strong>This role requires a work schedule that sufficiently overlaps with EST business hours to collaborate effectively with the team.</strong></em></p><p style=\"min-height:1.5em\"><strong>Who You Are.</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Demonstrated ability to extract data from complex websites with minimal supervision, with a portfolio or examples of past projects.</p></li><li><p style=\"min-height:1.5em\">Proficiency in languages such as Python or JavaScript, with strong skills in libraries and frameworks like BeautifulSoup, Scrapy, or Selenium.</p></li><li><p style=\"min-height:1.5em\">Knowledge of asynchronous programming, multithreading, and distributed scraping.</p></li><li><p style=\"min-height:1.5em\">In-depth knowledge of HTML, CSS, JavaScript, and the Document Object Model (DOM).</p></li><li><p style=\"min-height:1.5em\">Experience with NoSQL databases (MongoDB, Cassandra), capable of designing efficient storage solutions and managing data integrity.</p></li><li><p style=\"min-height:1.5em\">Ability to apply machine learning algorithms for data cleaning, categorization, or predictive analysis adds significant value.</p></li><li><p style=\"min-height:1.5em\">Experience with cloud services (AWS, Google Cloud, Azure) for deploying and managing scraping jobs at scale.</p></li><li><p style=\"min-height:1.5em\">Active participation in open-source projects related to web scraping, data processing, or similar fields.</p></li></ul><p style=\"min-height:1.5em\"><strong>What You'll Be Doing.</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Write, test, and refine code that extracts data from various online sources, ensuring reliability and efficiency.</p></li><li><p style=\"min-height:1.5em\">Perform data retrieval tasks, handling complexities such as pagination and dynamic content loaded with AJAX.</p></li><li><p style=\"min-height:1.5em\">Clean and format extracted data, ensuring it meets quality standards for further analysis or processing.</p></li><li><p style=\"min-height:1.5em\">Database management: Store and manage the scraped data in appropriate databases, optimizing for access speed and data integrity.</p></li><li><p style=\"min-height:1.5em\">Regularly monitor the scraping processes, identify and resolve any issues to maintain continuous data flow.</p></li></ul><p style=\"min-height:1.5em\"><strong>Why Work With Us:</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Opportunity. We are at the forefront of developing a web-scale crawler and knowledge graph that improves access to public web data and extends the value of AI to the people.</p></li><li><p style=\"min-height:1.5em\">Culture. We're a lean team with a high bar. We come to work not to be comfortable, but to find out what we're capable of and to do work that matters. We're not calling for people who keep things moving. We're calling for people who make everyone around them better. <br />We prioritize low ego and high output. This is a fully remote team.</p></li><li><p style=\"min-height:1.5em\">Compensation. You’ll receive a competitive salary, benefits and equity package.</p></li></ul>","descriptionPlain":"Who We Are:\n\nWe build infrastructure that delivers massive amounts of web data to the companies training the world’s most powerful AI models.\n\nWe're the team that helps to power and support Grass, a bandwidth-sharing network that lets us operate a massive distributed crawler, giving us unique access to high-quality public web data at global scale. On top of that, we’ve built pipelines for ingesting, segmenting, and annotating billions of videos, transcripts, and audio files, powering dataset creation for frontier labs.\n\nWe’re lean, technical, and move fast. No red tape, no slow decision-making; just a team of builders pushing to expand what’s possible for open web data and AI.\n\nThe Role.\n\nWe are seeking a Web Scraping Specialist who is proficient and brings significant experience in data extraction and web scraping techniques. You will join a small, specialized team and lead efforts to gather and analyze data, optimize scraping processes, and support our vision for a future where Grass plays a crucial role in transforming internet data accessibility.\n\nPlease note: This role requires a work schedule that sufficiently overlaps with EST business hours to collaborate effectively with the team.\n\nWho You Are.\n\n - Demonstrated ability to extract data from complex websites with minimal supervision, with a portfolio or examples of past projects.\n\n - Proficiency in languages such as Python or JavaScript, with strong skills in libraries and frameworks like BeautifulSoup, Scrapy, or Selenium.\n\n - Knowledge of asynchronous programming, multithreading, and distributed scraping.\n\n - In-depth knowledge of HTML, CSS, JavaScript, and the Document Object Model (DOM).\n\n - Experience with NoSQL databases (MongoDB, Cassandra), capable of designing efficient storage solutions and managing data integrity.\n\n - Ability to apply machine learning algorithms for data cleaning, categorization, or predictive analysis adds significant value.\n\n - Experience with cloud services (AWS, Google Cloud, Azure) for deploying and managing scraping jobs at scale.\n\n - Active participation in open-source projects related to web scraping, data processing, or similar fields.\n\nWhat You'll Be Doing.\n\n - Write, test, and refine code that extracts data from various online sources, ensuring reliability and efficiency.\n\n - Perform data retrieval tasks, handling complexities such as pagination and dynamic content loaded with AJAX.\n\n - Clean and format extracted data, ensuring it meets quality standards for further analysis or processing.\n\n - Database management: Store and manage the scraped data in appropriate databases, optimizing for access speed and data integrity.\n\n - Regularly monitor the scraping processes, identify and resolve any issues to maintain continuous data flow.\n\nWhy Work With Us:\n\n - Opportunity. We are at the forefront of developing a web-scale crawler and knowledge graph that improves access to public web data and extends the value of AI to the people.\n\n - Culture. We're a lean team with a high bar. We come to work not to be comfortable, but to find out what we're capable of and to do work that matters. We're not calling for people who keep things moving. We're calling for people who make everyone around them better. \n   We prioritize low ego and high output. This is a fully remote team.\n\n - Compensation. You’ll receive a competitive salary, benefits and equity package.","compensation":{"compensationTierSummary":null,"scrapeableCompensationSalarySummary":null,"compensationTiers":[],"summaryComponents":[]}},{"id":"6a11b221-5dc2-4bd7-9ec5-70e5c11c1e33","title":"Network Infrastructure Engineer (DevOps)","department":"Engineering","team":"Engineering","employmentType":"FullTime","location":"Ashburn ","shouldDisplayCompensationOnJobPostings":false,"secondaryLocations":[],"publishedAt":"2025-11-05T17:13:15.219+00:00","isListed":true,"isRemote":true,"workplaceType":"Hybrid","address":{"postalAddress":{"addressRegion":"D.C.","addressCountry":"United States","addressLocality":"Washington "}},"jobUrl":"https://jobs.ashbyhq.com/wynd-labs/6a11b221-5dc2-4bd7-9ec5-70e5c11c1e33","applyUrl":"https://jobs.ashbyhq.com/wynd-labs/6a11b221-5dc2-4bd7-9ec5-70e5c11c1e33/application","descriptionHtml":"<p style=\"min-height:1.5em\"><strong>Who We Are:</strong></p><p style=\"min-height:1.5em\">We build infrastructure that delivers massive amounts of web data to the companies training the world’s most powerful AI models.</p><p style=\"min-height:1.5em\">We're the team that helps to power and support Grass, a bandwidth-sharing network that lets us operate a massive distributed crawler, giving us unique access to high-quality public web data at global scale. On top of that, we’ve built pipelines for ingesting, segmenting, and annotating billions of videos, transcripts, and audio files, powering dataset creation for frontier labs.</p><p style=\"min-height:1.5em\">We’re lean, technical, and move fast. No red tape, no slow decision-making; just a team of builders pushing to expand what’s possible for open web data and AI.</p><p style=\"min-height:1.5em\"><strong>The Role.</strong></p><p style=\"min-height:1.5em\">We are looking for a talented Network Infrastructure Engineer to join our team with experience working with infrastructure and network architecture design . In this role, you will be responsible for building and maintaining the infrastructure that powers our AI-driven platform. You will work closely with our engineering and product teams to deploy physical infrastructure at our data centre to ensure that our systems are scalable, secure, and reliable. Your work will directly impact the efficiency and performance of our platform, enabling us to deliver cutting-edge solutions at scale.</p><p style=\"min-height:1.5em\"><strong>Who You Are.</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Proven experience designing networks and maintaining physical infrastructure.</p></li><li><p style=\"min-height:1.5em\">Possesses a deep understanding of routers/switches and establishing connectivity between servers.</p></li><li><p style=\"min-height:1.5em\">Strong proficiency with infrastructure as code (IaC) tools such as Terraform, Ansible, or similar.</p></li><li><p style=\"min-height:1.5em\">Experience with cloud platforms like AWS, GCP, or Azure, including networking, security, and automation.</p></li><li><p style=\"min-height:1.5em\">Experience with high-performance storage solutions, namely MinIO. </p></li><li><p style=\"min-height:1.5em\">Proficiency in containerization technologies (e.g., Docker) and orchestration tools like Kubernetes.</p></li><li><p style=\"min-height:1.5em\">Solid understanding of monitoring and logging tools such as Prometheus, Grafana, or ELK stack.</p></li><li><p style=\"min-height:1.5em\">Strong problem-solving skills with a focus on performance optimization and automation</p></li></ul><p style=\"min-height:1.5em\"><strong>What You'll Be Doing.</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\"><strong>Infrastructure Design and Implementation</strong>:</p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Oversee design, configuration, and deployment of networks, servers, network switches, systems, firewalls, and applications.</p></li><li><p style=\"min-height:1.5em\">Oversee the configuration of servers with virtualization and the troubleshooting of application issues</p></li><li><p style=\"min-height:1.5em\">Ensure infrastructure solutions meet business needs, including scalability and security.</p></li></ul></li><li><p style=\"min-height:1.5em\"><strong>System Maintenance and Support</strong>:</p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Ensure regular maintenance of servers, networks, and other infrastructure components to ensure uptime and reliability.</p></li><li><p style=\"min-height:1.5em\">Troubleshoot and resolve issues related to infrastructure components.</p></li></ul></li><li><p style=\"min-height:1.5em\"><strong>Network Management</strong>:</p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Ensure correct configuration of network components, including switches, routers, firewalls, and VPNs.</p></li><li><p style=\"min-height:1.5em\">Monitor system performance and implement solutions to enhance reliability, security, and scalability.</p></li><li><p style=\"min-height:1.5em\">Collaborate with engineering teams to integrate DevOps best practices into the development lifecycle.</p></li><li><p style=\"min-height:1.5em\">Ensure that all systems are secure and compliant with industry standards.</p></li></ul></li></ul><p style=\"min-height:1.5em\"><strong>Why Work With Us:</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Opportunity. We are at the forefront of developing a web-scale crawler and knowledge graph that improves access to public web data and extends the value of AI to the people.</p></li><li><p style=\"min-height:1.5em\">Culture. We're a lean team with a high bar. We come to work not to be comfortable, but to find out what we're capable of and to do work that matters. We're not calling for people who keep things moving. We're calling for people who make everyone around them better. <br />We prioritize low ego and high output. This is a fully remote team.</p></li><li><p style=\"min-height:1.5em\">Compensation. You’ll receive a competitive salary, benefits and equity package.</p></li></ul>","descriptionPlain":"Who We Are:\n\nWe build infrastructure that delivers massive amounts of web data to the companies training the world’s most powerful AI models.\n\nWe're the team that helps to power and support Grass, a bandwidth-sharing network that lets us operate a massive distributed crawler, giving us unique access to high-quality public web data at global scale. On top of that, we’ve built pipelines for ingesting, segmenting, and annotating billions of videos, transcripts, and audio files, powering dataset creation for frontier labs.\n\nWe’re lean, technical, and move fast. No red tape, no slow decision-making; just a team of builders pushing to expand what’s possible for open web data and AI.\n\nThe Role.\n\nWe are looking for a talented Network Infrastructure Engineer to join our team with experience working with infrastructure and network architecture design . In this role, you will be responsible for building and maintaining the infrastructure that powers our AI-driven platform. You will work closely with our engineering and product teams to deploy physical infrastructure at our data centre to ensure that our systems are scalable, secure, and reliable. Your work will directly impact the efficiency and performance of our platform, enabling us to deliver cutting-edge solutions at scale.\n\nWho You Are.\n\n - Proven experience designing networks and maintaining physical infrastructure.\n\n - Possesses a deep understanding of routers/switches and establishing connectivity between servers.\n\n - Strong proficiency with infrastructure as code (IaC) tools such as Terraform, Ansible, or similar.\n\n - Experience with cloud platforms like AWS, GCP, or Azure, including networking, security, and automation.\n\n - Experience with high-performance storage solutions, namely MinIO. \n\n - Proficiency in containerization technologies (e.g., Docker) and orchestration tools like Kubernetes.\n\n - Solid understanding of monitoring and logging tools such as Prometheus, Grafana, or ELK stack.\n\n - Strong problem-solving skills with a focus on performance optimization and automation\n\nWhat You'll Be Doing.\n\n - Infrastructure Design and Implementation:\n   \n   - Oversee design, configuration, and deployment of networks, servers, network switches, systems, firewalls, and applications.\n   \n   - Oversee the configuration of servers with virtualization and the troubleshooting of application issues\n   \n   - Ensure infrastructure solutions meet business needs, including scalability and security.\n\n - System Maintenance and Support:\n   \n   - Ensure regular maintenance of servers, networks, and other infrastructure components to ensure uptime and reliability.\n   \n   - Troubleshoot and resolve issues related to infrastructure components.\n\n - Network Management:\n   \n   - Ensure correct configuration of network components, including switches, routers, firewalls, and VPNs.\n   \n   - Monitor system performance and implement solutions to enhance reliability, security, and scalability.\n   \n   - Collaborate with engineering teams to integrate DevOps best practices into the development lifecycle.\n   \n   - Ensure that all systems are secure and compliant with industry standards.\n\nWhy Work With Us:\n\n - Opportunity. We are at the forefront of developing a web-scale crawler and knowledge graph that improves access to public web data and extends the value of AI to the people.\n\n - Culture. We're a lean team with a high bar. We come to work not to be comfortable, but to find out what we're capable of and to do work that matters. We're not calling for people who keep things moving. We're calling for people who make everyone around them better. \n   We prioritize low ego and high output. This is a fully remote team.\n\n - Compensation. You’ll receive a competitive salary, benefits and equity package.","compensation":{"compensationTierSummary":null,"scrapeableCompensationSalarySummary":null,"compensationTiers":[],"summaryComponents":[]}},{"id":"a148c710-7685-4184-9be8-c4d46a0a8f04","title":"Machine Learning Engineer","department":"Engineering","team":"Engineering","employmentType":"FullTime","location":"Remote","shouldDisplayCompensationOnJobPostings":false,"secondaryLocations":[],"publishedAt":"2026-09-04T22:43:32.306+00:00","isListed":true,"isRemote":null,"workplaceType":null,"address":{"postalAddress":{"addressCountry":"United States","addressLocality":""}},"jobUrl":"https://jobs.ashbyhq.com/wynd-labs/a148c710-7685-4184-9be8-c4d46a0a8f04","applyUrl":"https://jobs.ashbyhq.com/wynd-labs/a148c710-7685-4184-9be8-c4d46a0a8f04/application","descriptionHtml":"<p style=\"min-height:1.5em\"><strong>Who We Are:</strong></p><p style=\"min-height:1.5em\">We build infrastructure that delivers massive amounts of web data to the companies training the world’s most powerful AI models.</p><p style=\"min-height:1.5em\">We're the team that helps to power and support Grass, a bandwidth-sharing network that lets us operate a massive distributed crawler, giving us unique access to high-quality public web data at global scale. On top of that, we’ve built pipelines for ingesting, segmenting, and annotating billions of videos, transcripts, and audio files, powering dataset creation for frontier labs.</p><p style=\"min-height:1.5em\">We’re lean, technical, and move fast. No red tape, no slow decision-making; just a team of builders pushing to expand what’s possible for open web data and AI.</p><p style=\"min-height:1.5em\"><strong>The Role:</strong></p><p style=\"min-height:1.5em\">We are looking for a Machine Learning Engineer with strong skills and significant experience developing machine learning models. You will join a small, innovative team and lead efforts to advance our capabilities, drive model development, and support our vision for a future where Grass is transformative in the internet's evolution.</p><p style=\"min-height:1.5em\"><strong>Please note: </strong><em><strong>This role requires a work schedule that sufficiently overlaps with EST business hours to collaborate effectively with the team.</strong></em></p><p style=\"min-height:1.5em\"><strong>Who You Are:</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Bachelor’s, Master’s, or Doctoral degree in Data Science, Computer Science, Statistics, or a related field.</p></li><li><p style=\"min-height:1.5em\">A minimum of 3 years of work or research experience dealing with large datasets.</p></li><li><p style=\"min-height:1.5em\">Experience working with large-scale text datasets, NLP pipelines, or data preparation for LLM training is highly preferred.</p></li><li><p style=\"min-height:1.5em\">Strong coding skills in Python or other object-oriented programming languages.</p></li><li><p style=\"min-height:1.5em\">Experience with text deduplication, dataset filtering, corpus curation, or data distillation is a strong plus.</p></li><li><p style=\"min-height:1.5em\">Graduate-level knowledge of statistics, including but not limited to hypothesis testing, regression analysis, and probability.</p></li><li><p style=\"min-height:1.5em\">Excellent work ethic and the ability to thrive in a fast-paced startup environment.</p></li><li><p style=\"min-height:1.5em\">Strong problem-solving skills and attention to detail.</p></li><li><p style=\"min-height:1.5em\">Good communication skills, with the ability to articulate complex data concepts to non-technical stakeholders.</p></li><li><p style=\"min-height:1.5em\">Experience working in a high-output team.</p></li></ul><p style=\"min-height:1.5em\"><strong>What You'll Be Doing:</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Developing data processing pipelines and machine learning solutions for large-scale NLP and LLM applications, including improving the quality, filtering, and preparation of training datasets.</p></li><li><p style=\"min-height:1.5em\">Designing and implementing pipelines for processing and analyzing large datasets.</p></li><li><p style=\"min-height:1.5em\">Analyzing and interpreting complex time series data to provide actionable insights and solutions.</p></li><li><p style=\"min-height:1.5em\">Designing, implementing, and maintaining data-driven models and algorithms.</p></li><li><p style=\"min-height:1.5em\">Developing techniques for dataset curation to improve the quality and efficiency of AI training data.</p></li><li><p style=\"min-height:1.5em\">Building scalable pipelines for filtering, deduplicating, and improving large-scale text datasets used for LLM training.</p></li><li><p style=\"min-height:1.5em\">Collaborating with cross-functional teams to understand data needs and deliver timely solutions.</p></li><li><p style=\"min-height:1.5em\">Ensuring data quality and integrity throughout all processes.</p></li><li><p style=\"min-height:1.5em\">Utilizing Optical Character Recognition (OCR) technology to convert different types of documents into editable and searchable data.</p></li><li><p style=\"min-height:1.5em\">Continuously researching and implementing best practices in data science and machine learning.</p></li><li><p style=\"min-height:1.5em\">Contributing to the development and improvement of internal data processing tools and infrastructure.</p></li></ul><p style=\"min-height:1.5em\"><strong>Why Work With Us:</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Opportunity. We are at the forefront of developing a web-scale crawler and knowledge graph that improves access to public web data and extends the value of AI to the people.</p></li><li><p style=\"min-height:1.5em\">Culture. We're a lean team with a high bar. We come to work not to be comfortable, but to find out what we're capable of and to do work that matters. We're not calling for people who keep things moving. We're calling for people who make everyone around them better. <br />We prioritize low ego and high output. This is a fully remote team.</p></li><li><p style=\"min-height:1.5em\">Compensation. You’ll receive a competitive salary, benefits and equity package.</p></li></ul>","descriptionPlain":"Who We Are:\n\nWe build infrastructure that delivers massive amounts of web data to the companies training the world’s most powerful AI models.\n\nWe're the team that helps to power and support Grass, a bandwidth-sharing network that lets us operate a massive distributed crawler, giving us unique access to high-quality public web data at global scale. On top of that, we’ve built pipelines for ingesting, segmenting, and annotating billions of videos, transcripts, and audio files, powering dataset creation for frontier labs.\n\nWe’re lean, technical, and move fast. No red tape, no slow decision-making; just a team of builders pushing to expand what’s possible for open web data and AI.\n\nThe Role:\n\nWe are looking for a Machine Learning Engineer with strong skills and significant experience developing machine learning models. You will join a small, innovative team and lead efforts to advance our capabilities, drive model development, and support our vision for a future where Grass is transformative in the internet's evolution.\n\nPlease note: This role requires a work schedule that sufficiently overlaps with EST business hours to collaborate effectively with the team.\n\nWho You Are:\n\n - Bachelor’s, Master’s, or Doctoral degree in Data Science, Computer Science, Statistics, or a related field.\n\n - A minimum of 3 years of work or research experience dealing with large datasets.\n\n - Experience working with large-scale text datasets, NLP pipelines, or data preparation for LLM training is highly preferred.\n\n - Strong coding skills in Python or other object-oriented programming languages.\n\n - Experience with text deduplication, dataset filtering, corpus curation, or data distillation is a strong plus.\n\n - Graduate-level knowledge of statistics, including but not limited to hypothesis testing, regression analysis, and probability.\n\n - Excellent work ethic and the ability to thrive in a fast-paced startup environment.\n\n - Strong problem-solving skills and attention to detail.\n\n - Good communication skills, with the ability to articulate complex data concepts to non-technical stakeholders.\n\n - Experience working in a high-output team.\n\nWhat You'll Be Doing:\n\n - Developing data processing pipelines and machine learning solutions for large-scale NLP and LLM applications, including improving the quality, filtering, and preparation of training datasets.\n\n - Designing and implementing pipelines for processing and analyzing large datasets.\n\n - Analyzing and interpreting complex time series data to provide actionable insights and solutions.\n\n - Designing, implementing, and maintaining data-driven models and algorithms.\n\n - Developing techniques for dataset curation to improve the quality and efficiency of AI training data.\n\n - Building scalable pipelines for filtering, deduplicating, and improving large-scale text datasets used for LLM training.\n\n - Collaborating with cross-functional teams to understand data needs and deliver timely solutions.\n\n - Ensuring data quality and integrity throughout all processes.\n\n - Utilizing Optical Character Recognition (OCR) technology to convert different types of documents into editable and searchable data.\n\n - Continuously researching and implementing best practices in data science and machine learning.\n\n - Contributing to the development and improvement of internal data processing tools and infrastructure.\n\nWhy Work With Us:\n\n - Opportunity. We are at the forefront of developing a web-scale crawler and knowledge graph that improves access to public web data and extends the value of AI to the people.\n\n - Culture. We're a lean team with a high bar. We come to work not to be comfortable, but to find out what we're capable of and to do work that matters. We're not calling for people who keep things moving. We're calling for people who make everyone around them better. \n   We prioritize low ego and high output. This is a fully remote team.\n\n - Compensation. You’ll receive a competitive salary, benefits and equity package.","compensation":{"compensationTierSummary":null,"scrapeableCompensationSalarySummary":null,"compensationTiers":[],"summaryComponents":[]}},{"id":"dad06add-f092-4204-8eca-453097b59118","title":"Growth & Communications Lead","department":"Growth/ Marketing","team":"Growth/ Marketing","employmentType":"FullTime","location":"Remote","shouldDisplayCompensationOnJobPostings":false,"secondaryLocations":[],"publishedAt":"2026-09-05T01:50:29.756+00:00","isListed":true,"isRemote":null,"workplaceType":null,"address":{"postalAddress":{"addressCountry":"United States","addressLocality":""}},"jobUrl":"https://jobs.ashbyhq.com/wynd-labs/dad06add-f092-4204-8eca-453097b59118","applyUrl":"https://jobs.ashbyhq.com/wynd-labs/dad06add-f092-4204-8eca-453097b59118/application","descriptionHtml":"<p style=\"min-height:1.5em\"><strong>Who We Are:</strong></p><p style=\"min-height:1.5em\">We build infrastructure that delivers massive amounts of web data to the companies training the world’s most powerful AI models.</p><p style=\"min-height:1.5em\">We're the team that helps to power and support Grass, a bandwidth-sharing network that lets us operate a massive distributed crawler, giving us unique access to high-quality public web data at global scale. On top of that, we’ve built pipelines for ingesting, segmenting, and annotating billions of videos, transcripts, and audio files, powering dataset creation for frontier labs.</p><p style=\"min-height:1.5em\">We’re lean, technical, and move fast. No red tape, no slow decision-making; just a team of builders pushing to expand what’s possible for open web data and AI.</p><p style=\"min-height:1.5em\"><strong>The Role:</strong></p><p style=\"min-height:1.5em\">We’re seeking a <strong>Growth &amp; Communications Lead</strong> to tackle the challenge of shrinking online attention spans. You will help identify, develop, and execute strategies to drive user acquisition and shape our external narrative. This role demands the ability to craft clear, compelling, and on-brand messaging that cuts through the noise across all channels. Success hinges on refined judgment for tone and resonance combined with rapid experimentation to scale our network while also articulating our mission.</p><p style=\"min-height:1.5em\"><strong>Please note: This role requires a work schedule that sufficiently overlaps with EST business hours to collaborate effectively with the team</strong><em><strong>.</strong></em></p><p style=\"min-height:1.5em\"><strong>Who You Are:</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">A strong, versatile copywriter who can turn complex ideas into clear, engaging, and persuasive language.</p></li><li><p style=\"min-height:1.5em\">Curious and passionate about growth, experimentation, and user behavior.</p></li><li><p style=\"min-height:1.5em\">Embraces feedback as a tool for continuous improvement.</p></li><li><p style=\"min-height:1.5em\">Innovative thinker who thrives in fast-paced environments.</p></li><li><p style=\"min-height:1.5em\">Creative problem solver and strong communicator.</p></li><li><p style=\"min-height:1.5em\">Comfortable balancing short-term wins with long-term strategy.</p></li><li><p style=\"min-height:1.5em\">Persistent and resourceful in solving challenges.</p></li><li><p style=\"min-height:1.5em\">High integrity and seeks out responsibility.</p></li><li><p style=\"min-height:1.5em\">Resilient, motivated to get things done, and eager to learn.</p></li><li><p style=\"min-height:1.5em\">Values team success over personal recognition; organized, detail-oriented, and process driven.</p></li></ul><p style=\"min-height:1.5em\"><strong>What You'll Be Doing:</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Owning copywriting across key channels (email, landing pages, in-product copy, social, blogs, and campaigns) to drive user acquisition, activation, and engagement.</p></li><li><p style=\"min-height:1.5em\">Identifying and analyzing growth opportunities across user acquisition, and engagement channels.</p></li><li><p style=\"min-height:1.5em\">Building relationships with influencers, creators, and various internet communities to amplify brand visibility and drive adoption, including writing briefs and suggested copy.</p></li><li><p style=\"min-height:1.5em\">Managing and optimizing paid acquisition campaigns across major ad platforms (e.g., Meta, Google Ads) to scale growth efficiently.</p></li><li><p style=\"min-height:1.5em\">Tracking and analyzing KPIs (CTR, conversion rate, engagement, etc.) to measure the impact of copy and inform decisions.</p></li><li><p style=\"min-height:1.5em\">Developing and framing the company's brand position, narrative, and tone across various social and digital platforms.</p></li><li><p style=\"min-height:1.5em\">Shaping how we show up online through consistent, thoughtful, and on-brand messaging.</p></li><li><p style=\"min-height:1.5em\">Blending deep technical understanding with creative storytelling to explain our mission, products, and business model.</p></li><li><p style=\"min-height:1.5em\">Creating unexpected ways to showcase our work, including our open source initiatives and research.</p></li><li><p style=\"min-height:1.5em\">Creating multi-format educational content (short-form, long-form, visual-supporting copy) for a variety of audiences and depth levels.</p></li></ul><p style=\"min-height:1.5em\"><strong>Skills, Requirements and Qualifications:</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Bachelor’s degree or equivalent work experience</p></li><li><p style=\"min-height:1.5em\">Minimum of 2 years of experience in a growth, marketing, communications, or creative role <strong>with a primary focus on copywriting</strong></p></li><li><p style=\"min-height:1.5em\">A strong portfolio demonstrating clear, persuasive, and results-driven copy across multiple formats (email, web, social, product, campaigns).</p></li><li><p style=\"min-height:1.5em\">Exceptional written communication skills; you are an excellent writer and editor with high attention to detail, tone, and clarity.</p></li><li><p style=\"min-height:1.5em\">Strong analytical skills with experience using data to test, measure, and iterate on copy and campaigns.</p></li><li><p style=\"min-height:1.5em\">Ability to manage multiple projects, deadlines, and priorities simultaneously.</p></li><li><p style=\"min-height:1.5em\">Uses first principles and systems thinking to understand and solve problems.</p></li><li><p style=\"min-height:1.5em\">Strong interpersonal skills; you are personable and able to manage expectations across many stakeholders and multiple ongoing relationships.</p></li><li><p style=\"min-height:1.5em\">Ability to work under pressure, meet deadlines, and adapt quickly based on feedback and performance data.</p></li><li><p style=\"min-height:1.5em\">Strong strategic thinking and problem-solving skills; comfortable moving between high-level narrative and tactical execution.</p></li></ul><p style=\"min-height:1.5em\"><strong>Why Work With Us:</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Opportunity. We are at the forefront of developing a web-scale crawler and knowledge graph that improves access to public web data and extends the value of AI to the people.</p></li><li><p style=\"min-height:1.5em\">Culture. We're a lean team with a high bar. We come to work not to be comfortable, but to find out what we're capable of and to do work that matters. We're not calling for people who keep things moving. We're calling for people who make everyone around them better. <br />We prioritize low ego and high output. This is a fully remote team.</p></li><li><p style=\"min-height:1.5em\">Compensation. You’ll receive a competitive salary, benefits and equity package.</p></li></ul>","descriptionPlain":"Who We Are:\n\nWe build infrastructure that delivers massive amounts of web data to the companies training the world’s most powerful AI models.\n\nWe're the team that helps to power and support Grass, a bandwidth-sharing network that lets us operate a massive distributed crawler, giving us unique access to high-quality public web data at global scale. On top of that, we’ve built pipelines for ingesting, segmenting, and annotating billions of videos, transcripts, and audio files, powering dataset creation for frontier labs.\n\nWe’re lean, technical, and move fast. No red tape, no slow decision-making; just a team of builders pushing to expand what’s possible for open web data and AI.\n\nThe Role:\n\nWe’re seeking a Growth & Communications Lead to tackle the challenge of shrinking online attention spans. You will help identify, develop, and execute strategies to drive user acquisition and shape our external narrative. This role demands the ability to craft clear, compelling, and on-brand messaging that cuts through the noise across all channels. Success hinges on refined judgment for tone and resonance combined with rapid experimentation to scale our network while also articulating our mission.\n\nPlease note: This role requires a work schedule that sufficiently overlaps with EST business hours to collaborate effectively with the team.\n\nWho You Are:\n\n - A strong, versatile copywriter who can turn complex ideas into clear, engaging, and persuasive language.\n\n - Curious and passionate about growth, experimentation, and user behavior.\n\n - Embraces feedback as a tool for continuous improvement.\n\n - Innovative thinker who thrives in fast-paced environments.\n\n - Creative problem solver and strong communicator.\n\n - Comfortable balancing short-term wins with long-term strategy.\n\n - Persistent and resourceful in solving challenges.\n\n - High integrity and seeks out responsibility.\n\n - Resilient, motivated to get things done, and eager to learn.\n\n - Values team success over personal recognition; organized, detail-oriented, and process driven.\n\nWhat You'll Be Doing:\n\n - Owning copywriting across key channels (email, landing pages, in-product copy, social, blogs, and campaigns) to drive user acquisition, activation, and engagement.\n\n - Identifying and analyzing growth opportunities across user acquisition, and engagement channels.\n\n - Building relationships with influencers, creators, and various internet communities to amplify brand visibility and drive adoption, including writing briefs and suggested copy.\n\n - Managing and optimizing paid acquisition campaigns across major ad platforms (e.g., Meta, Google Ads) to scale growth efficiently.\n\n - Tracking and analyzing KPIs (CTR, conversion rate, engagement, etc.) to measure the impact of copy and inform decisions.\n\n - Developing and framing the company's brand position, narrative, and tone across various social and digital platforms.\n\n - Shaping how we show up online through consistent, thoughtful, and on-brand messaging.\n\n - Blending deep technical understanding with creative storytelling to explain our mission, products, and business model.\n\n - Creating unexpected ways to showcase our work, including our open source initiatives and research.\n\n - Creating multi-format educational content (short-form, long-form, visual-supporting copy) for a variety of audiences and depth levels.\n\nSkills, Requirements and Qualifications:\n\n - Bachelor’s degree or equivalent work experience\n\n - Minimum of 2 years of experience in a growth, marketing, communications, or creative role with a primary focus on copywriting\n\n - A strong portfolio demonstrating clear, persuasive, and results-driven copy across multiple formats (email, web, social, product, campaigns).\n\n - Exceptional written communication skills; you are an excellent writer and editor with high attention to detail, tone, and clarity.\n\n - Strong analytical skills with experience using data to test, measure, and iterate on copy and campaigns.\n\n - Ability to manage multiple projects, deadlines, and priorities simultaneously.\n\n - Uses first principles and systems thinking to understand and solve problems.\n\n - Strong interpersonal skills; you are personable and able to manage expectations across many stakeholders and multiple ongoing relationships.\n\n - Ability to work under pressure, meet deadlines, and adapt quickly based on feedback and performance data.\n\n - Strong strategic thinking and problem-solving skills; comfortable moving between high-level narrative and tactical execution.\n\nWhy Work With Us:\n\n - Opportunity. We are at the forefront of developing a web-scale crawler and knowledge graph that improves access to public web data and extends the value of AI to the people.\n\n - Culture. We're a lean team with a high bar. We come to work not to be comfortable, but to find out what we're capable of and to do work that matters. We're not calling for people who keep things moving. We're calling for people who make everyone around them better. \n   We prioritize low ego and high output. This is a fully remote team.\n\n - Compensation. You’ll receive a competitive salary, benefits and equity package.","compensation":{"compensationTierSummary":null,"scrapeableCompensationSalarySummary":null,"compensationTiers":[],"summaryComponents":[]}},{"id":"8aa370f3-8e80-44d4-aa9f-c28296d43efd","title":"Research Crawling Engineer","department":"Engineering","team":"Engineering","employmentType":"FullTime","location":"Remote","shouldDisplayCompensationOnJobPostings":false,"secondaryLocations":[],"publishedAt":"2026-09-05T01:48:14.825+00:00","isListed":true,"isRemote":null,"workplaceType":null,"address":{"postalAddress":{"addressCountry":"United States","addressLocality":""}},"jobUrl":"https://jobs.ashbyhq.com/wynd-labs/8aa370f3-8e80-44d4-aa9f-c28296d43efd","applyUrl":"https://jobs.ashbyhq.com/wynd-labs/8aa370f3-8e80-44d4-aa9f-c28296d43efd/application","descriptionHtml":"<p style=\"min-height:1.5em\"><strong>Who We Are:</strong></p><p style=\"min-height:1.5em\">We build infrastructure that delivers massive amounts of web data to the companies training the world’s most powerful AI models.</p><p style=\"min-height:1.5em\">We're the team that helps to power and support Grass, a bandwidth-sharing network that lets us operate a massive distributed crawler, giving us unique access to high-quality public web data at global scale. On top of that, we’ve built pipelines for ingesting, segmenting, and annotating billions of videos, transcripts, and audio files, powering dataset creation for frontier labs.</p><p style=\"min-height:1.5em\">We’re lean, technical, and move fast. No red tape, no slow decision-making; just a team of builders pushing to expand what’s possible for open web data and AI.</p><p style=\"min-height:1.5em\"><strong><u>Overview:</u></strong><br />As a Research Crawling Engineer, you will design and operate large-scale web data acquisition systems for research and model development. Your work will span distributed systems, scraping infrastructure, and data pipelines.</p><p style=\"min-height:1.5em\"><strong>Please note: This role requires a work schedule that sufficiently overlaps with EST business hours to collaborate effectively with the team</strong><em><strong>.  </strong></em><br /><br /><strong><u>Responsibilities:</u></strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Build and maintain large-scale web crawlers across diverse domains</p></li><li><p style=\"min-height:1.5em\">Design high-throughput, fault-tolerant systems for data collection (millions to billions of URLs/day)</p></li><li><p style=\"min-height:1.5em\">Handle anti-bot systems, rate limits, and dynamic/JS-heavy sites</p></li><li><p style=\"min-height:1.5em\">Develop pipelines for cleaning, deduplication, filtering, and normalization</p></li><li><p style=\"min-height:1.5em\">Construct and maintain datasets for research and model training</p></li><li><p style=\"min-height:1.5em\">Monitor crawl performance, coverage, and data quality; iterate quickly</p></li><li><p style=\"min-height:1.5em\">Collaborate with research teams to align data collection with modeling needs</p></li><li><p style=\"min-height:1.5em\">Optimize infrastructure for cost, latency, and reliability</p></li></ul><p style=\"min-height:1.5em\"><strong><u>Requirements:</u></strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Strong programming experience in one or more of: Go, Rust, Python, Java, or C++</p></li><li><p style=\"min-height:1.5em\">Experience building web crawlers or large-scale data pipelines</p></li><li><p style=\"min-height:1.5em\">Solid understanding of HTTP, networking, and browser behavior</p></li><li><p style=\"min-height:1.5em\">Familiarity with distributed systems and parallel processing</p></li><li><p style=\"min-height:1.5em\">Experience working with large datasets (TB–PB scale preferred)</p></li><li><p style=\"min-height:1.5em\">Ability to debug unstable or adversarial environments</p></li></ul><p style=\"min-height:1.5em\"><strong><u>Preferred / Bonus:</u></strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Experience with NLP pipelines or dataset curation for ML</p></li><li><p style=\"min-height:1.5em\">Familiarity with LLM pretraining data or retrieval systems</p></li><li><p style=\"min-height:1.5em\">Experience with headless browsers (e.g., Chrome DevTools Protocol, Playwright, Puppeteer)</p></li><li><p style=\"min-height:1.5em\">Knowledge of proxy systems, IP rotation, and large-scale request orchestration</p></li><li><p style=\"min-height:1.5em\">Background in data quality evaluation or benchmarking</p></li><li><p style=\"min-height:1.5em\">Experience running workloads on cloud or bare-metal infrastructure</p></li></ul><p style=\"min-height:1.5em\"><strong><u>What This Role Involves:</u></strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Operating at the boundary of scale and reliability</p></li><li><p style=\"min-height:1.5em\">Adapting to constantly changing web environments</p></li><li><p style=\"min-height:1.5em\">Balancing throughput, coverage, and data quality</p></li><li><p style=\"min-height:1.5em\">Owning end-to-end data acquisition pipelines</p></li></ul><p style=\"min-height:1.5em\"><strong><u>Evaluation Criteria:</u></strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Ability to design systems that scale without degrading quality</p></li><li><p style=\"min-height:1.5em\">Practical problem-solving under real-world constraints</p></li><li><p style=\"min-height:1.5em\">Speed of iteration and ownership</p></li><li><p style=\"min-height:1.5em\">Measurable improvements in data coverage, quality, or efficiency</p></li></ul><p style=\"min-height:1.5em\"><strong><u>Compensation:</u></strong></p><p style=\"min-height:1.5em\">Based on experience and demonstrated ability to operate at scale<br /><br /><strong><u>Example Projects:</u></strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Build a distributed crawler for a continuously updated, high-quality web project</p></li><li><p style=\"min-height:1.5em\">Design a system to classify and filter billions of pages for pretraining</p></li><li><p style=\"min-height:1.5em\">Extract structured data from dynamic, JS-heavy sites at scale</p></li><li><p style=\"min-height:1.5em\">Improve deduplication and quality scoring across multimodal datasets</p></li></ul><p style=\"min-height:1.5em\"><strong>Why Work With Us:</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Opportunity. We are at the forefront of developing a web-scale crawler and knowledge graph that improves access to public web data and extends the value of AI to the people.</p></li><li><p style=\"min-height:1.5em\">Culture. We're a lean team with a high bar. We come to work not to be comfortable, but to find out what we're capable of and to do work that matters. We're not calling for people who keep things moving. We're calling for people who make everyone around them better. <br />We prioritize low ego and high output. This is a fully remote team.</p></li><li><p style=\"min-height:1.5em\">Compensation. You’ll receive a competitive salary, benefits and equity package.</p></li></ul>","descriptionPlain":"Who We Are:\n\nWe build infrastructure that delivers massive amounts of web data to the companies training the world’s most powerful AI models.\n\nWe're the team that helps to power and support Grass, a bandwidth-sharing network that lets us operate a massive distributed crawler, giving us unique access to high-quality public web data at global scale. On top of that, we’ve built pipelines for ingesting, segmenting, and annotating billions of videos, transcripts, and audio files, powering dataset creation for frontier labs.\n\nWe’re lean, technical, and move fast. No red tape, no slow decision-making; just a team of builders pushing to expand what’s possible for open web data and AI.\n\nOverview:\nAs a Research Crawling Engineer, you will design and operate large-scale web data acquisition systems for research and model development. Your work will span distributed systems, scraping infrastructure, and data pipelines.\n\nPlease note: This role requires a work schedule that sufficiently overlaps with EST business hours to collaborate effectively with the team.  \n\nResponsibilities:\n\n - Build and maintain large-scale web crawlers across diverse domains\n\n - Design high-throughput, fault-tolerant systems for data collection (millions to billions of URLs/day)\n\n - Handle anti-bot systems, rate limits, and dynamic/JS-heavy sites\n\n - Develop pipelines for cleaning, deduplication, filtering, and normalization\n\n - Construct and maintain datasets for research and model training\n\n - Monitor crawl performance, coverage, and data quality; iterate quickly\n\n - Collaborate with research teams to align data collection with modeling needs\n\n - Optimize infrastructure for cost, latency, and reliability\n\nRequirements:\n\n - Strong programming experience in one or more of: Go, Rust, Python, Java, or C++\n\n - Experience building web crawlers or large-scale data pipelines\n\n - Solid understanding of HTTP, networking, and browser behavior\n\n - Familiarity with distributed systems and parallel processing\n\n - Experience working with large datasets (TB–PB scale preferred)\n\n - Ability to debug unstable or adversarial environments\n\nPreferred / Bonus:\n\n - Experience with NLP pipelines or dataset curation for ML\n\n - Familiarity with LLM pretraining data or retrieval systems\n\n - Experience with headless browsers (e.g., Chrome DevTools Protocol, Playwright, Puppeteer)\n\n - Knowledge of proxy systems, IP rotation, and large-scale request orchestration\n\n - Background in data quality evaluation or benchmarking\n\n - Experience running workloads on cloud or bare-metal infrastructure\n\nWhat This Role Involves:\n\n - Operating at the boundary of scale and reliability\n\n - Adapting to constantly changing web environments\n\n - Balancing throughput, coverage, and data quality\n\n - Owning end-to-end data acquisition pipelines\n\nEvaluation Criteria:\n\n - Ability to design systems that scale without degrading quality\n\n - Practical problem-solving under real-world constraints\n\n - Speed of iteration and ownership\n\n - Measurable improvements in data coverage, quality, or efficiency\n\nCompensation:\n\nBased on experience and demonstrated ability to operate at scale\n\nExample Projects:\n\n - Build a distributed crawler for a continuously updated, high-quality web project\n\n - Design a system to classify and filter billions of pages for pretraining\n\n - Extract structured data from dynamic, JS-heavy sites at scale\n\n - Improve deduplication and quality scoring across multimodal datasets\n\nWhy Work With Us:\n\n - Opportunity. We are at the forefront of developing a web-scale crawler and knowledge graph that improves access to public web data and extends the value of AI to the people.\n\n - Culture. We're a lean team with a high bar. We come to work not to be comfortable, but to find out what we're capable of and to do work that matters. We're not calling for people who keep things moving. We're calling for people who make everyone around them better. \n   We prioritize low ego and high output. This is a fully remote team.\n\n - Compensation. You’ll receive a competitive salary, benefits and equity package.","compensation":{"compensationTierSummary":null,"scrapeableCompensationSalarySummary":null,"compensationTiers":[],"summaryComponents":[]}},{"id":"1e0ceb41-435e-48f7-b70d-4e6b8ed2df0a","title":"Senior Software Engineer (Backend)","department":"Engineering","team":"Engineering","employmentType":"FullTime","location":"Remote","shouldDisplayCompensationOnJobPostings":false,"secondaryLocations":[],"publishedAt":"2026-09-05T01:46:57.315+00:00","isListed":true,"isRemote":null,"workplaceType":null,"address":{"postalAddress":{"addressCountry":"United States","addressLocality":""}},"jobUrl":"https://jobs.ashbyhq.com/wynd-labs/1e0ceb41-435e-48f7-b70d-4e6b8ed2df0a","applyUrl":"https://jobs.ashbyhq.com/wynd-labs/1e0ceb41-435e-48f7-b70d-4e6b8ed2df0a/application","descriptionHtml":"<p style=\"min-height:1.5em\"><strong>Who We Are:</strong></p><p style=\"min-height:1.5em\">We build infrastructure that delivers massive amounts of web data to the companies training the world’s most powerful AI models.</p><p style=\"min-height:1.5em\">We're the team that helps to power and support Grass, a bandwidth-sharing network that lets us operate a massive distributed crawler, giving us unique access to high-quality public web data at global scale. On top of that, we’ve built pipelines for ingesting, segmenting, and annotating billions of videos, transcripts, and audio files, powering dataset creation for frontier labs.</p><p style=\"min-height:1.5em\">We’re lean, technical, and move fast. No red tape, no slow decision-making; just a team of builders pushing to expand what’s possible for open web data and AI.</p><p style=\"min-height:1.5em\"><strong>The Role:</strong></p><p style=\"min-height:1.5em\">We are seeking a Senior Backend Engineer with deep expertise in backend development, particularly in data pipeline infrastructure. In this role, you will join a small, high-performing team to design and architect scalable systems, drive technical excellence, and help advance Grass’s role in shaping the future of the internet.</p><p style=\"min-height:1.5em\"><strong>Please note: This role requires a work schedule that sufficiently overlaps with EST business hours to collaborate effectively with the team</strong><em><strong>.</strong></em></p><p style=\"min-height:1.5em\"><strong>Who You Are:</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">7+ years of experience in software development, with a track record of writing high-quality, maintainable, and robust code.</p></li><li><p style=\"min-height:1.5em\">Bachelor’s or Master’s degree in a STEM field, or equivalent practical experience.</p></li><li><p style=\"min-height:1.5em\">Strong experience building and scaling large, distributed systems.</p></li><li><p style=\"min-height:1.5em\">Deep expertise in designing, troubleshooting, and optimizing complex, live software systems.</p></li><li><p style=\"min-height:1.5em\">Hands-on experience with modern development practices, including continuous integration and continuous deployment.</p></li><li><p style=\"min-height:1.5em\">Strong analytical and problem-solving skills, particularly in diagnosing and resolving data flow and system performance issues.</p></li><li><p style=\"min-height:1.5em\">Professional or native English proficiency.</p></li></ul><p style=\"min-height:1.5em\"><strong>What You'll Be Doing:</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Design, build, and optimize scalable data pipeline infrastructure for real-time and batch data processing.</p></li><li><p style=\"min-height:1.5em\">Develop new backend features and system improvements, focusing on reliability and performance.</p></li><li><p style=\"min-height:1.5em\">Write clear, well-tested, and well-documented code.</p></li><li><p style=\"min-height:1.5em\">Create and review technical designs, code, and documentation to maintain high engineering standards.</p></li><li><p style=\"min-height:1.5em\">Contribute to Wynd’s infrastructure across mobile, desktop, and server-side systems.</p></li></ul><p style=\"min-height:1.5em\"><strong>Technologies Used:</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Golang, Python, Redis Clusters, Kubernetes.</p></li></ul><p style=\"min-height:1.5em\"><strong>Why Work With Us:</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Opportunity. We are at the forefront of developing a web-scale crawler and knowledge graph that improves access to public web data and extends the value of AI to the people.</p></li><li><p style=\"min-height:1.5em\">Culture. We're a lean team with a high bar. We come to work not to be comfortable, but to find out what we're capable of and to do work that matters. We're not calling for people who keep things moving. We're calling for people who make everyone around them better. <br />We prioritize low ego and high output. This is a fully remote team.</p></li><li><p style=\"min-height:1.5em\">Compensation. You’ll receive a competitive salary, benefits and equity package.</p></li></ul>","descriptionPlain":"Who We Are:\n\nWe build infrastructure that delivers massive amounts of web data to the companies training the world’s most powerful AI models.\n\nWe're the team that helps to power and support Grass, a bandwidth-sharing network that lets us operate a massive distributed crawler, giving us unique access to high-quality public web data at global scale. On top of that, we’ve built pipelines for ingesting, segmenting, and annotating billions of videos, transcripts, and audio files, powering dataset creation for frontier labs.\n\nWe’re lean, technical, and move fast. No red tape, no slow decision-making; just a team of builders pushing to expand what’s possible for open web data and AI.\n\nThe Role:\n\nWe are seeking a Senior Backend Engineer with deep expertise in backend development, particularly in data pipeline infrastructure. In this role, you will join a small, high-performing team to design and architect scalable systems, drive technical excellence, and help advance Grass’s role in shaping the future of the internet.\n\nPlease note: This role requires a work schedule that sufficiently overlaps with EST business hours to collaborate effectively with the team.\n\nWho You Are:\n\n - 7+ years of experience in software development, with a track record of writing high-quality, maintainable, and robust code.\n\n - Bachelor’s or Master’s degree in a STEM field, or equivalent practical experience.\n\n - Strong experience building and scaling large, distributed systems.\n\n - Deep expertise in designing, troubleshooting, and optimizing complex, live software systems.\n\n - Hands-on experience with modern development practices, including continuous integration and continuous deployment.\n\n - Strong analytical and problem-solving skills, particularly in diagnosing and resolving data flow and system performance issues.\n\n - Professional or native English proficiency.\n\nWhat You'll Be Doing:\n\n - Design, build, and optimize scalable data pipeline infrastructure for real-time and batch data processing.\n\n - Develop new backend features and system improvements, focusing on reliability and performance.\n\n - Write clear, well-tested, and well-documented code.\n\n - Create and review technical designs, code, and documentation to maintain high engineering standards.\n\n - Contribute to Wynd’s infrastructure across mobile, desktop, and server-side systems.\n\nTechnologies Used:\n\n - Golang, Python, Redis Clusters, Kubernetes.\n\nWhy Work With Us:\n\n - Opportunity. We are at the forefront of developing a web-scale crawler and knowledge graph that improves access to public web data and extends the value of AI to the people.\n\n - Culture. We're a lean team with a high bar. We come to work not to be comfortable, but to find out what we're capable of and to do work that matters. We're not calling for people who keep things moving. We're calling for people who make everyone around them better. \n   We prioritize low ego and high output. This is a fully remote team.\n\n - Compensation. You’ll receive a competitive salary, benefits and equity package.","compensation":{"compensationTierSummary":null,"scrapeableCompensationSalarySummary":null,"compensationTiers":[],"summaryComponents":[]}},{"id":"e7ffd2fb-23cb-4d21-ac89-06b846a2f728","title":"Junior Software Engineer","department":"Engineering","team":"Engineering","employmentType":"FullTime","location":"Remote","shouldDisplayCompensationOnJobPostings":false,"secondaryLocations":[],"publishedAt":"2026-09-05T01:54:39.260+00:00","isListed":true,"isRemote":null,"workplaceType":null,"address":{"postalAddress":{"addressCountry":"United States","addressLocality":""}},"jobUrl":"https://jobs.ashbyhq.com/wynd-labs/e7ffd2fb-23cb-4d21-ac89-06b846a2f728","applyUrl":"https://jobs.ashbyhq.com/wynd-labs/e7ffd2fb-23cb-4d21-ac89-06b846a2f728/application","descriptionHtml":"<p style=\"min-height:1.5em\"><strong>Who We Are:</strong></p><p style=\"min-height:1.5em\">We build infrastructure that delivers massive amounts of web data to the companies training the world’s most powerful AI models.</p><p style=\"min-height:1.5em\">We're the team that helps to power and support Grass, a bandwidth-sharing network that lets us operate a massive distributed crawler, giving us unique access to high-quality public web data at global scale. On top of that, we’ve built pipelines for ingesting, segmenting, and annotating billions of videos, transcripts, and audio files, powering dataset creation for frontier labs.</p><p style=\"min-height:1.5em\">We’re lean, technical, and move fast. No red tape, no slow decision-making; just a team of builders pushing to expand what’s possible for open web data and AI.</p><p style=\"min-height:1.5em\"><strong>The Role:</strong></p><p style=\"min-height:1.5em\">We're looking for an early-career software engineer to help review our applications and ensure they meet compliance requirements before release. You'll be trusted to investigate application behavior independently, exercise sound judgment when identifying potential compliance issues, and take ownership of problems from discovery through resolution.</p><p style=\"min-height:1.5em\"><br />This role sits at the intersection of engineering, quality assurance, and compliance validation. You'll review product changes, investigate application behavior, identify potential compliance issues, document findings clearly, and work closely with engineering and operations teams to drive issues through resolution and verify that they’ve been properly addressed before deployment.</p><p style=\"min-height:1.5em\"></p><p style=\"min-height:1.5em\"><strong>Please Note: This role requires a work schedule that overlaps sufficiently with EST business hours (online until 3:00 p.m. EST) to collaborate effectively with the team.</strong></p><p style=\"min-height:1.5em\"><strong>Who You Are:</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">You have hands-on programming experience with Rust and/or Java — even if limited.</p></li><li><p style=\"min-height:1.5em\">You're a fast learner who can pick up unfamiliar systems, applications, and requirements quickly.</p></li><li><p style=\"min-height:1.5em\">You have strong analytical instincts — you naturally notice inconsistencies and dig until you understand why something behaves the way it does.</p></li><li><p style=\"min-height:1.5em\">You communicate findings clearly and professionally, both in writing and in conversation.</p></li><li><p style=\"min-height:1.5em\">You take ownership: once you're given a problem, you see it through rather than waiting for direction at every step.</p></li><li><p style=\"min-height:1.5em\">You're comfortable with ambiguity and change — you don't need a fixed checklist to know what to investigate next.</p></li><li><p style=\"min-height:1.5em\">You enjoy working closely with engineering, product, and operations, and you don't need a narrowly defined lane to be effective.</p></li></ul><p style=\"min-height:1.5em\"><strong>What You'll Be Doing:</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Review applications before release to ensure they meet defined requirements and identify potential compliance issues.</p></li><li><p style=\"min-height:1.5em\">Understand and investigate application behavior in order to identify technical or compliance-related problems.</p></li><li><p style=\"min-height:1.5em\">Validate application changes and document findings clearly.</p></li><li><p style=\"min-height:1.5em\">Provide actionable feedback and communicate issues effectively to engineering and operations teams.</p></li><li><p style=\"min-height:1.5em\">Work with engineers to understand and resolve identified issues.</p></li><li><p style=\"min-height:1.5em\">Verify that issues have been properly addressed before release.</p></li><li><p style=\"min-height:1.5em\">Maintain review documentation and contribute to improving internal review processes.</p></li><li><p style=\"min-height:1.5em\">Help improve and, where appropriate, automate validation and QA workflows.</p></li></ul><p style=\"min-height:1.5em\"><strong>What We Are Looking For:</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Develop a deeper understanding of our applications and systems over time and take increasing ownership of the review process.</p></li><li><p style=\"min-height:1.5em\">Help ensure a consistent, high-quality release process across products.</p></li><li><p style=\"min-height:1.5em\">Programming experience in Rust and/or Java (early-career level is fine — we care more about how quickly you learn and think through problems).</p></li><li><p style=\"min-height:1.5em\">Experience in a startup or small engineering team, particularly where you had broad responsibilities.</p></li><li><p style=\"min-height:1.5em\">Exposure to QA, testing, product validation, compliance, or security.</p></li><li><p style=\"min-height:1.5em\">Excellent written and verbal English communication skills.</p></li><li><p style=\"min-height:1.5em\">Ability to manage multiple reviews or tasks in parallel without losing track of details.</p></li></ul><p style=\"min-height:1.5em\"><strong>Why Work With Us:</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Opportunity. We are at the forefront of developing a web-scale crawler and knowledge graph that improves access to public web data and extends the value of AI to the people.</p></li><li><p style=\"min-height:1.5em\">Culture. We're a lean team with a high bar. We come to work not to be comfortable, but to find out what we're capable of and to do work that matters. We're not calling for people who keep things moving. We're calling for people who make everyone around them better. <br />We prioritize low ego and high output. This is a fully remote team.</p></li><li><p style=\"min-height:1.5em\">Compensation. You’ll receive a competitive salary, benefits and equity package.</p></li></ul>","descriptionPlain":"Who We Are:\n\nWe build infrastructure that delivers massive amounts of web data to the companies training the world’s most powerful AI models.\n\nWe're the team that helps to power and support Grass, a bandwidth-sharing network that lets us operate a massive distributed crawler, giving us unique access to high-quality public web data at global scale. On top of that, we’ve built pipelines for ingesting, segmenting, and annotating billions of videos, transcripts, and audio files, powering dataset creation for frontier labs.\n\nWe’re lean, technical, and move fast. No red tape, no slow decision-making; just a team of builders pushing to expand what’s possible for open web data and AI.\n\nThe Role:\n\nWe're looking for an early-career software engineer to help review our applications and ensure they meet compliance requirements before release. You'll be trusted to investigate application behavior independently, exercise sound judgment when identifying potential compliance issues, and take ownership of problems from discovery through resolution.\n\n\nThis role sits at the intersection of engineering, quality assurance, and compliance validation. You'll review product changes, investigate application behavior, identify potential compliance issues, document findings clearly, and work closely with engineering and operations teams to drive issues through resolution and verify that they’ve been properly addressed before deployment.\n\n\n\nPlease Note: This role requires a work schedule that overlaps sufficiently with EST business hours (online until 3:00 p.m. EST) to collaborate effectively with the team.\n\nWho You Are:\n\n - You have hands-on programming experience with Rust and/or Java — even if limited.\n\n - You're a fast learner who can pick up unfamiliar systems, applications, and requirements quickly.\n\n - You have strong analytical instincts — you naturally notice inconsistencies and dig until you understand why something behaves the way it does.\n\n - You communicate findings clearly and professionally, both in writing and in conversation.\n\n - You take ownership: once you're given a problem, you see it through rather than waiting for direction at every step.\n\n - You're comfortable with ambiguity and change — you don't need a fixed checklist to know what to investigate next.\n\n - You enjoy working closely with engineering, product, and operations, and you don't need a narrowly defined lane to be effective.\n\nWhat You'll Be Doing:\n\n - Review applications before release to ensure they meet defined requirements and identify potential compliance issues.\n\n - Understand and investigate application behavior in order to identify technical or compliance-related problems.\n\n - Validate application changes and document findings clearly.\n\n - Provide actionable feedback and communicate issues effectively to engineering and operations teams.\n\n - Work with engineers to understand and resolve identified issues.\n\n - Verify that issues have been properly addressed before release.\n\n - Maintain review documentation and contribute to improving internal review processes.\n\n - Help improve and, where appropriate, automate validation and QA workflows.\n\nWhat We Are Looking For:\n\n - Develop a deeper understanding of our applications and systems over time and take increasing ownership of the review process.\n\n - Help ensure a consistent, high-quality release process across products.\n\n - Programming experience in Rust and/or Java (early-career level is fine — we care more about how quickly you learn and think through problems).\n\n - Experience in a startup or small engineering team, particularly where you had broad responsibilities.\n\n - Exposure to QA, testing, product validation, compliance, or security.\n\n - Excellent written and verbal English communication skills.\n\n - Ability to manage multiple reviews or tasks in parallel without losing track of details.\n\nWhy Work With Us:\n\n - Opportunity. We are at the forefront of developing a web-scale crawler and knowledge graph that improves access to public web data and extends the value of AI to the people.\n\n - Culture. We're a lean team with a high bar. We come to work not to be comfortable, but to find out what we're capable of and to do work that matters. We're not calling for people who keep things moving. We're calling for people who make everyone around them better. \n   We prioritize low ego and high output. This is a fully remote team.\n\n - Compensation. You’ll receive a competitive salary, benefits and equity package.","compensation":{"compensationTierSummary":null,"scrapeableCompensationSalarySummary":null,"compensationTiers":[],"summaryComponents":[]}},{"id":"900e79ba-5668-4a43-9065-3f4bf7db6aa4","title":"Backend Engineer (TypeScript)","department":"Engineering","team":"Engineering","employmentType":"FullTime","location":"Remote","shouldDisplayCompensationOnJobPostings":false,"secondaryLocations":[],"publishedAt":"2026-09-04T21:59:08.541+00:00","isListed":true,"isRemote":null,"workplaceType":null,"address":{"postalAddress":{"addressCountry":"United States","addressLocality":""}},"jobUrl":"https://jobs.ashbyhq.com/wynd-labs/900e79ba-5668-4a43-9065-3f4bf7db6aa4","applyUrl":"https://jobs.ashbyhq.com/wynd-labs/900e79ba-5668-4a43-9065-3f4bf7db6aa4/application","descriptionHtml":"<p style=\"min-height:1.5em\"><strong>Who We Are:</strong></p><p style=\"min-height:1.5em\">We build infrastructure that delivers massive amounts of web data to the companies training the world’s most powerful AI models.</p><p style=\"min-height:1.5em\">We're the team that helps to power and support Grass, a bandwidth-sharing network that lets us operate a massive distributed crawler, giving us unique access to high-quality public web data at global scale. On top of that, we’ve built pipelines for ingesting, segmenting, and annotating billions of videos, transcripts, and audio files, powering dataset creation for frontier labs.</p><p style=\"min-height:1.5em\">We’re lean, technical, and move fast. No red tape, no slow decision-making; just a team of builders pushing to expand what’s possible for open web data and AI.</p><p style=\"min-height:1.5em\"><strong>The Role:</strong></p><p style=\"min-height:1.5em\">We are seeking a hands-on Backend Engineer to join a small, dynamic engineering team and help build, improve, and operate the backend services and production tooling that support reliable web data acquisition at scale.</p><p style=\"min-height:1.5em\">This role is ideal for someone who enjoys coding, cares about the quality of what they ship, and can work with a high degree of independence. You should be comfortable taking a high-level objective, figuring out what needs to be done, and driving it through to a reliable production result without requiring close supervision.</p><p style=\"min-height:1.5em\"><strong>Please note: </strong><em><strong>This role requires a work schedule that sufficiently overlaps with EST business hours to collaborate effectively with the team.</strong></em></p><p style=\"min-height:1.5em\"><strong>Who You Are:</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Acts with integrity and seeks out responsibility</p></li><li><p style=\"min-height:1.5em\">Demonstrates resilience, resourcefulness, and motivation for getting things done</p></li><li><p style=\"min-height:1.5em\">Organized and process-driven</p></li><li><p style=\"min-height:1.5em\">Approaches challenges as opportunities</p></li><li><p style=\"min-height:1.5em\">Curious and challenges personal assumptions regularly</p></li><li><p style=\"min-height:1.5em\">Welcomes feedback and open dialogue</p></li><li><p style=\"min-height:1.5em\">Values team success over personal recognition</p></li></ul><p style=\"min-height:1.5em\"><strong>What You'll Be Doing:</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Build and improve backend code primarily in TypeScript.</p></li><li><p style=\"min-height:1.5em\">Set up and maintain metrics dashboards for production systems.</p></li><li><p style=\"min-height:1.5em\">Update existing code to support new metrics and improve system observability.</p></li><li><p style=\"min-height:1.5em\">Set up alerts and work closely with DevOps to optimize production systems.</p></li><li><p style=\"min-height:1.5em\">Automate deployments and operational reporting to Slack.</p></li><li><p style=\"min-height:1.5em\">Test systems in production, identify and troubleshoot issues, and fix bugs.</p></li><li><p style=\"min-height:1.5em\">Help clean up and improve the existing codebase.</p></li><li><p style=\"min-height:1.5em\">Write and maintain automated tests.</p></li><li><p style=\"min-height:1.5em\">Improve and optimize deployment processes.</p></li><li><p style=\"min-height:1.5em\">Implement code enhancements as new requirements emerge.</p></li></ul><p style=\"min-height:1.5em\"><strong>Skills, Requirements and Qualifications:</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Bachelor’s degree or equivalent work experience.</p></li><li><p style=\"min-height:1.5em\">2+ years of experience in a backend software engineering role.</p></li><li><p style=\"min-height:1.5em\">Solid TypeScript development skills and strong backend software engineering fundamentals.</p></li><li><p style=\"min-height:1.5em\">Good understanding of metrics, monitoring, and alerting, with the ability to debug and troubleshoot production systems.</p></li><li><p style=\"min-height:1.5em\">Ability to write clean, maintainable, well-tested code and make reliable improvements to existing systems.</p></li></ul><p style=\"min-height:1.5em\"><strong>Nice to have:</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Experience with <strong>Go</strong>.</p></li><li><p style=\"min-height:1.5em\">Experience with Datadog, OpenObserve, and/or building Slack bots or integrations.</p></li><li><p style=\"min-height:1.5em\">Familiarity with <strong>web scraping and browser automation</strong>, including <strong>Playwright</strong>.</p></li></ul><p style=\"min-height:1.5em\"><strong>Why Work With Us:</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Opportunity. We are at the forefront of developing a web-scale crawler and knowledge graph that improves access to public web data and extends the value of AI to the people.</p></li><li><p style=\"min-height:1.5em\">Culture. We're a lean team with a high bar. We come to work not to be comfortable, but to find out what we're capable of and to do work that matters. We're not calling for people who keep things moving. We're calling for people who make everyone around them better. <br />We prioritize low ego and high output. This is a fully remote team.</p></li><li><p style=\"min-height:1.5em\">Compensation. You’ll receive a competitive salary, benefits and equity package.</p></li></ul>","descriptionPlain":"Who We Are:\n\nWe build infrastructure that delivers massive amounts of web data to the companies training the world’s most powerful AI models.\n\nWe're the team that helps to power and support Grass, a bandwidth-sharing network that lets us operate a massive distributed crawler, giving us unique access to high-quality public web data at global scale. On top of that, we’ve built pipelines for ingesting, segmenting, and annotating billions of videos, transcripts, and audio files, powering dataset creation for frontier labs.\n\nWe’re lean, technical, and move fast. No red tape, no slow decision-making; just a team of builders pushing to expand what’s possible for open web data and AI.\n\nThe Role:\n\nWe are seeking a hands-on Backend Engineer to join a small, dynamic engineering team and help build, improve, and operate the backend services and production tooling that support reliable web data acquisition at scale.\n\nThis role is ideal for someone who enjoys coding, cares about the quality of what they ship, and can work with a high degree of independence. You should be comfortable taking a high-level objective, figuring out what needs to be done, and driving it through to a reliable production result without requiring close supervision.\n\nPlease note: This role requires a work schedule that sufficiently overlaps with EST business hours to collaborate effectively with the team.\n\nWho You Are:\n\n - Acts with integrity and seeks out responsibility\n\n - Demonstrates resilience, resourcefulness, and motivation for getting things done\n\n - Organized and process-driven\n\n - Approaches challenges as opportunities\n\n - Curious and challenges personal assumptions regularly\n\n - Welcomes feedback and open dialogue\n\n - Values team success over personal recognition\n\nWhat You'll Be Doing:\n\n - Build and improve backend code primarily in TypeScript.\n\n - Set up and maintain metrics dashboards for production systems.\n\n - Update existing code to support new metrics and improve system observability.\n\n - Set up alerts and work closely with DevOps to optimize production systems.\n\n - Automate deployments and operational reporting to Slack.\n\n - Test systems in production, identify and troubleshoot issues, and fix bugs.\n\n - Help clean up and improve the existing codebase.\n\n - Write and maintain automated tests.\n\n - Improve and optimize deployment processes.\n\n - Implement code enhancements as new requirements emerge.\n\nSkills, Requirements and Qualifications:\n\n - Bachelor’s degree or equivalent work experience.\n\n - 2+ years of experience in a backend software engineering role.\n\n - Solid TypeScript development skills and strong backend software engineering fundamentals.\n\n - Good understanding of metrics, monitoring, and alerting, with the ability to debug and troubleshoot production systems.\n\n - Ability to write clean, maintainable, well-tested code and make reliable improvements to existing systems.\n\nNice to have:\n\n - Experience with Go.\n\n - Experience with Datadog, OpenObserve, and/or building Slack bots or integrations.\n\n - Familiarity with web scraping and browser automation, including Playwright.\n\nWhy Work With Us:\n\n - Opportunity. We are at the forefront of developing a web-scale crawler and knowledge graph that improves access to public web data and extends the value of AI to the people.\n\n - Culture. We're a lean team with a high bar. We come to work not to be comfortable, but to find out what we're capable of and to do work that matters. We're not calling for people who keep things moving. We're calling for people who make everyone around them better. \n   We prioritize low ego and high output. This is a fully remote team.\n\n - Compensation. You’ll receive a competitive salary, benefits and equity package.","compensation":{"compensationTierSummary":null,"scrapeableCompensationSalarySummary":null,"compensationTiers":[],"summaryComponents":[]}},{"id":"216d74ef-d925-4bf1-ae4f-9eedc3e145f8","title":"Data Engineer","department":"Analytics","team":"Analytics","employmentType":"FullTime","location":"Remote","shouldDisplayCompensationOnJobPostings":false,"secondaryLocations":[],"publishedAt":"2026-09-04T21:52:35.123+00:00","isListed":true,"isRemote":true,"workplaceType":"Remote","address":{"postalAddress":{"addressCountry":"United States","addressLocality":""}},"jobUrl":"https://jobs.ashbyhq.com/wynd-labs/216d74ef-d925-4bf1-ae4f-9eedc3e145f8","applyUrl":"https://jobs.ashbyhq.com/wynd-labs/216d74ef-d925-4bf1-ae4f-9eedc3e145f8/application","descriptionHtml":"<p style=\"min-height:1.5em\"><strong>Who We Are:</strong></p><p style=\"min-height:1.5em\">We build infrastructure that delivers massive amounts of web data to the companies training the world’s most powerful AI models.</p><p style=\"min-height:1.5em\">We're the team that helps to power and support Grass, a bandwidth-sharing network that lets us operate a massive distributed crawler, giving us unique access to high-quality public web data at global scale. On top of that, we’ve built pipelines for ingesting, segmenting, and annotating billions of videos, transcripts, and audio files, powering dataset creation for frontier labs.</p><p style=\"min-height:1.5em\">We’re lean, technical, and move fast. No red tape, no slow decision-making; just a team of builders pushing to expand what’s possible for open web data and AI.</p><p style=\"min-height:1.5em\"><strong>The Role:</strong></p><p style=\"min-height:1.5em\">We are seeking a Data Engineer to support and improve large-scale data pipelines and infrastructure. You’ll work across data collection, processing, transformation, validation, and delivery, with a focus on scalability, reliability, and performance.<br />This is a hands-on role where you’ll work with distributed systems, large datasets, web scraping infrastructure, and production data workloads.</p><p style=\"min-height:1.5em\"><strong>Please note: This role requires a work schedule that overlaps sufficiently with EST business hours to collaborate effectively with the team.</strong></p><p style=\"min-height:1.5em\"><strong>Who You Are:</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Bachelor’s degree or equivalent work experience</p></li><li><p style=\"min-height:1.5em\"><strong>Python (advanced)</strong> — strong grasp of async programming, multiprocessing, and writing production-grade code for long-running data jobs</p></li><li><p style=\"min-height:1.5em\"><strong>Web scraping at scale</strong> — hands-on experience with high-volume scraping (proxies, rate limiting, anti-bot evasion). Experience with platform APIs and large media/metadata datasets (video platforms, social media)</p></li><li><p style=\"min-height:1.5em\"><strong>Distributed data pipelines</strong> — experience designing and operating pipelines across many workers/servers using task queues (Celery, Kafka, RabbitMQ, or similar)</p></li><li><p style=\"min-height:1.5em\"><strong>Data warehousing</strong> — practical experience with columnar/analytical warehouses; Databend, ClickHouse, or BigQuery strongly preferred; comfortable with complex analytical queries, partitioning strategies, cost-aware querying on cloud warehouses</p></li><li><p style=\"min-height:1.5em\"><strong>Docker &amp; Kubernetes</strong> — containerizing workloads, writing Helm charts/manifests, managing deployments, autoscaling scraping/processing workloads</p></li><li><p style=\"min-height:1.5em\"><strong>Linux &amp; bare-metal ops</strong> — comfortable managing services on Linux servers, debugging performance issues (disk I/O, network, memory) without managed-cloud abstractions</p></li><li><p style=\"min-height:1.5em\">CI/CD for data workflows (GitHub Actions, ArgoCD)</p></li><li><p style=\"min-height:1.5em\">Writing Scalable API</p></li></ul><p style=\"min-height:1.5em\"><strong>What You'll Be Doing:</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Maintain, optimize, and troubleshoot database queries and related data systems to support efficient data access, processing, and reliability.</p></li><li><p style=\"min-height:1.5em\">Assist in creating, maintaining, and improving data pipelines used to collect, process, transform, validate, and deliver large-scale datasets.</p></li><li><p style=\"min-height:1.5em\">Support web scraping and data collection initiatives, including developing, testing, and maintaining scripts or tools used to gather publicly available data in accordance with Company requirements.</p></li><li><p style=\"min-height:1.5em\">Monitor and troubleshoot data pipeline issues, identify data quality concerns, and help implement timely fixes to maintain data accuracy and operational continuity.</p></li><li><p style=\"min-height:1.5em\">Document engineering work, including database queries, pipeline processes, scraping workflows, technical decisions, issues encountered, and resolutions implemented.</p></li><li><p style=\"min-height:1.5em\">Participate in research and development projects to improve the Company’s data products and workflows.</p></li></ul><p style=\"min-height:1.5em\"><strong>Why Work With Us:</strong></p><ul style=\"min-height:1.5em\"><li><p style=\"min-height:1.5em\">Opportunity. We are at the forefront of developing a web-scale crawler and knowledge graph that improves access to public web data and extends the value of AI to the people.</p></li><li><p style=\"min-height:1.5em\">Culture. We're a lean team with a high bar. We come to work not to be comfortable, but to find out what we're capable of and to do work that matters. We're not calling for people who keep things moving. We're calling for people who make everyone around them better. <br />We prioritize low ego and high output. This is a fully remote team.</p></li><li><p style=\"min-height:1.5em\">Compensation. You’ll receive a competitive salary, benefits and equity package.</p></li></ul>","descriptionPlain":"Who We Are:\n\nWe build infrastructure that delivers massive amounts of web data to the companies training the world’s most powerful AI models.\n\nWe're the team that helps to power and support Grass, a bandwidth-sharing network that lets us operate a massive distributed crawler, giving us unique access to high-quality public web data at global scale. On top of that, we’ve built pipelines for ingesting, segmenting, and annotating billions of videos, transcripts, and audio files, powering dataset creation for frontier labs.\n\nWe’re lean, technical, and move fast. No red tape, no slow decision-making; just a team of builders pushing to expand what’s possible for open web data and AI.\n\nThe Role:\n\nWe are seeking a Data Engineer to support and improve large-scale data pipelines and infrastructure. You’ll work across data collection, processing, transformation, validation, and delivery, with a focus on scalability, reliability, and performance.\nThis is a hands-on role where you’ll work with distributed systems, large datasets, web scraping infrastructure, and production data workloads.\n\nPlease note: This role requires a work schedule that overlaps sufficiently with EST business hours to collaborate effectively with the team.\n\nWho You Are:\n\n - Bachelor’s degree or equivalent work experience\n\n - Python (advanced) — strong grasp of async programming, multiprocessing, and writing production-grade code for long-running data jobs\n\n - Web scraping at scale — hands-on experience with high-volume scraping (proxies, rate limiting, anti-bot evasion). Experience with platform APIs and large media/metadata datasets (video platforms, social media)\n\n - Distributed data pipelines — experience designing and operating pipelines across many workers/servers using task queues (Celery, Kafka, RabbitMQ, or similar)\n\n - Data warehousing — practical experience with columnar/analytical warehouses; Databend, ClickHouse, or BigQuery strongly preferred; comfortable with complex analytical queries, partitioning strategies, cost-aware querying on cloud warehouses\n\n - Docker & Kubernetes — containerizing workloads, writing Helm charts/manifests, managing deployments, autoscaling scraping/processing workloads\n\n - Linux & bare-metal ops — comfortable managing services on Linux servers, debugging performance issues (disk I/O, network, memory) without managed-cloud abstractions\n\n - CI/CD for data workflows (GitHub Actions, ArgoCD)\n\n - Writing Scalable API\n\nWhat You'll Be Doing:\n\n - Maintain, optimize, and troubleshoot database queries and related data systems to support efficient data access, processing, and reliability.\n\n - Assist in creating, maintaining, and improving data pipelines used to collect, process, transform, validate, and deliver large-scale datasets.\n\n - Support web scraping and data collection initiatives, including developing, testing, and maintaining scripts or tools used to gather publicly available data in accordance with Company requirements.\n\n - Monitor and troubleshoot data pipeline issues, identify data quality concerns, and help implement timely fixes to maintain data accuracy and operational continuity.\n\n - Document engineering work, including database queries, pipeline processes, scraping workflows, technical decisions, issues encountered, and resolutions implemented.\n\n - Participate in research and development projects to improve the Company’s data products and workflows.\n\nWhy Work With Us:\n\n - Opportunity. We are at the forefront of developing a web-scale crawler and knowledge graph that improves access to public web data and extends the value of AI to the people.\n\n - Culture. We're a lean team with a high bar. We come to work not to be comfortable, but to find out what we're capable of and to do work that matters. We're not calling for people who keep things moving. We're calling for people who make everyone around them better. \n   We prioritize low ego and high output. This is a fully remote team.\n\n - Compensation. You’ll receive a competitive salary, benefits and equity package.","compensation":{"compensationTierSummary":null,"scrapeableCompensationSalarySummary":null,"compensationTiers":[],"summaryComponents":[]}}],"apiVersion":"1"}