Software Engineer, Data Acquisition
Nascent · Remote
This listing is no longer verified as available.
MeritLog keeps this source-backed description for reference. Availability is not verified, and there is no application link here.
Source: the employer's Ashby job board. Open the original listing for current details. Availability is not verified for this retained page.
Job details
- Work model
- Remote
- Salary
- Not listed by source
Job description
The Opportunity You'll own the systems that feed Nascent's trading and research operations with the data they need to compete - building and operating the web scraping and data acquisition infrastructure that runs 24/7 across a heterogeneous fleet of machines, providers, and operating systems. This is a roughly 50/50 split between software engineering and operations: half your time you're writing production code (primarily Rust) to build new scrapers, defeat anti-bot countermeasures, and architect resilient pipelines; the other half you're deep in logs, monitoring dashboards, and fleet health - diagnosing failures, tuning proxies, and keeping a distributed system humming across mixed infrastructure. You'll work closely with analysts and researchers who depend on the data you deliver. The problems are genuinely interesting: every major website is actively trying to stop you, infrastructure spans multiple cloud providers and bare metal, and the surface area of what breaks is enormous. This role is remote-first. Montreal proximity is preferred - the team has a growing analyst presence there and in-person collaboration matters - but it's not a hard requirement. North American or European time zones work best for team overlap. Responsibilities - Build and maintain production-grade web scraping systems - primarily in Rust - designing scrapers that are resilient to site changes, rate limiting, CAPTCHAs, and evolving anti-bot countermeasures. - Operate and monitor a heterogeneous distributed infrastructure spanning multiple cloud providers, bare metal, and mixed operating systems - you own uptime, not just deployments. - Diagnose production issues from raw logs and telemetry, building observability into every system you ship so problems surface before they cascade. - Design and implement proxy management, rotation strategies, and network-layer evasion techniques to maintain reliable data acquisition at scale. - Develop tooling for fleet health monitoring, automated alerting, and self-healing infrastructure across a diverse set of machines and hosting environments. - Collaborate with analysts and researchers to understand data requirements, prioritize new source integrations, and ensure data quality and freshness meet trading-grade standards. - Reverse-engineer web applications and APIs - inspecting network traffic, deobfuscating JavaScript, and adapting to adversarial changes in target sites. - Continuously improve system reliability, throughput, and maintainability - refactoring scraping pipelines, optimizing connection handling, and reducing operational toil through automation. - Deploy and maintain agentic workflows for data source discovery, onboarding, and troubleshooting - using LLM-based agents to automate the identification of new sources, accelerate integration, and surface and resolve failures in existing pipelines. About you - You are a builder who also runs what you build - you don't consider a project done when the PR merges; you consider it done when it's been stable in production for weeks. - You have a genuine interest in the cat-and-mouse game of web scraping at scale: anti-bot systems, browser fingerprinting, proxy rotation, and the constant adaptation it requires. - You are comfortable navigating messy, heterogeneous infrastructure - mixed OS environments, multiple hosting providers, hardware you didn't provision - and you make it better over time. - You are energized by operational puzzles: tracing a failure across distributed logs, identifying a subtle network degradation, or figuring out why a scraper that worked yesterday is now blocked. - You write clean, production-grade code and care about systems that run unattended and fail gracefully. You are proficient in at least one major programming language (Go, Python, C++, Java, or similar) - the stack is primarily Rust, and prior Rust experience is a strong plus, but if you haven't written it professionally, you're the kind of engineer who picks up new languages fast and takes ownership of the ramp. Python is useful for scripting, prototyping, and analyst-facing tooling. - You thrive in less-structured environments where you're trusted to prioritize your own work, and you take ownership of outcomes rather than waiting for tickets. - You are fluent with agentic workflows and LLM-based tooling - you know how to design and operate AI agents to discover new data sources, automate onboarding, and triage failures in production pipelines. You don't treat AI as a novelty; you reach for it when it's the right tool and build on top of it. - You have strong networking intuition - you think in terms of TCP connections, DNS resolution, HTTP headers, and proxy chains, not just API calls. Preferred experience - 2–5 years of professional software engineering experience with strong systems or backend fundamentals - production-grade code that runs in anger, not just toy projects or prototypes. Proficiency in at least one major programming language (Go, Python, C++, Java, or similar) is required; Rust experience is a strong plus but not a prerequisite. - Deep understanding of networking fundamentals: TCP/IP, DNS, HTTP internals, CDNs, proxies, load balancing, and queue handling. - Hands-on experience with web scraping or data acquisition at scale, including familiarity with anti-bot/anti-automation countermeasures and evasion techniques. - Demonstrated experience operating heterogeneous distributed infrastructure (mixed OS, hardware, hosting providers) - not just deploying to a single cloud provider. - Strong log analysis and monitoring skills: you can diagnose issues from raw logs and build observability into systems you own. - Hands-on experience building or operating agentic workflows (LLM-based agents, tool-use pipelines) for automation, data extraction, or system orchestration - this is core to how we discover and onboard new data sources. - Nice to have: - Experience with Rust - the stack is primarily Rust, and prior production experience accelerates your ramp significantly. - Proficiency in Python - useful for scripting, prototyping, and interacting with analyst-facing tooling. - Experience with cloud providers (AWS, GCP, DigitalOcean) and managing infrastructure across multiple providers simultaneously. - Background in high-throughput HTTP at scale - proxy management, VPN/routing optimization, and connection pooling. - Exposure to real-time or latency-sensitive systems (HFT-adjacent experience is a plus). About Nascent Founded in 2020, Nascent exists to build, expand, and capture opportunity, in open markets and permissionless technologies. Building from a base of permanent capital, we deploy assets across a range of both liquid and long-term strategies that ensure we are among the most active users of the open financial system we are helping to build. We’ve backed 100+ early-stage teams https://www.nascent.xyz/portfolio that we believe can change markets and expand what’s possible. At Nascent, we combine that venture + market pedigree with product-grade engineering to turn research into revenue: experiments that ship fast, are evaluated rigorously, and move P&L. We focus on models, low-latency infrastructure, and resilient systems that compound performance over time. Our Team & Culture We're an interdisciplinary team of engineers, quants, and operators who treat skill-building like a sport. The culture is competition + curiosity: push to win, share learnings, and iterate publicly through red-teams, internal games, and post-mortems. Autonomy isn't a perk - it's how we ship. Day-to-day you'll be in 2-pizza squads with end-to-end ownership, maker calendars, and a Decision SLA of <24 hours. We prefer guardrails over ceremony and measure success by repeatable edges and measurable outcomes. Principles - Compete to win - Explore, experiment, play - Always be building - Seek and speak truth - Own your shit. What We Offer - The opportunity to learn, experiment and build in an entrepreneurial environment - Remote-first, distributed team with frequent in-person sprints and retreats. - Competitive total comp with strong bonus upside tied to performance. - Hardware & home-office stipend, conference and learning budgets. - Comprehensive health benefits (medical/dental/vision), life insurance. - Open vacation policy as well as flexible work hours and location - 16 weeks fully-paid parental leave + supported return to work. - Retirement matching, open vacation policy, flexible hours. We are an equal opportunity employer and welcome diverse perspectives.
Keep exploring
Available Engineering roles
These current listings are available to explore now.
- Security Engineer - Detection & Response (Japan, X Money)SpaceXAI · On-site
- Facilities Infrastructure Engineer, Data Center Infrastructure - MemphisSpaceXAI · On-site
- Facilities Engineer, Electrical - MemphisSpaceXAI · Not provided by source
- Software Engineer - Platform Infrastructure (Rust, C++)SpaceXAI · On-site
- Software Engineer - Network Software and ServicesSpaceXAI · On-site
- Operations Engineer, Facility OperationsSpaceXAI · On-site