Backend Engineer - Data Acquisition
Indexed description
We're looking for a Backend Engineer to help build and run our web scraping platform. This isn't a ticket-taking role: you'll own the pipelines and services you work on end to end, and you'll have a real say in how they evolve.
We care a lot about proactivity. The best people on this team notice a scraper getting fragile before it breaks, ask why a target keeps failing instead of just bumping the retry count, and bring problems to the table with a proposed fix attached.
What you will be doing
- Own the extraction pipelines and backend services you work on — design, implementation, deployment, and the monitoring that keeps them healthy
- Build scrapers and data pipelines in Python, and in Go where it's the better fit
- Apply LLMs to extraction and normalisation problems within our existing evaluation and guardrail setup, and help improve accuracy and cost as you go
- Contribute to the supporting infrastructure: scheduling and orchestration, proxy and session management, queuing, storage, and observability
- Work on anti-bot systems, rate limiting, browser automation, fingerprinting, deduplication, and change detection — with support at first and growing independence
- Develop backend services and APIs that deliver structured, well-documented datasets to downstream consumers
- Surface technical risks early: brittle scrapers, silent failures, creeping costs, things that only you can see from where you're sitting
- Take part in architecture discussions and bring well-argued opinions, even when you're not the one making the final call
- Help keep the bar high on testing, code review, documentation, and monitoring, and share what you learn with the team
Requirements
- Background in Computer Science, Software Engineering, or a related technical field
- 3+ years of professional software engineering experience, with some of it in backend systems or data acquisition
- Solid production experience in Python; comfortable in Go, or motivated and able to pick it up quickly
- Experience building scrapers or data pipelines that run in production and are expected to keep running, not just one-off scripts
- Good understanding of HTTP, HTML/DOM parsing, and how browsers actually behave
- Some exposure to anti-bot countermeasures, proxy infrastructure, and rate-limit strategies, or a clear appetite to go deep on them
- Working knowledge of concurrency, distributed systems, and API design
- Experience with SQL, and ideally NoSQL, including basic schema design
- Comfortable with Docker, CI/CD, and cloud infrastructure
- Ownership mindset: you scope your own work, raise blockers early, make decisions in your area and stand behind them
- Fluent in English, both written and spoken
Perks
- Flexible Hybrid Work System
- 15th Month Bonus Salary Policy
- 4 extra vacation days
- Participation in International Projects
- Company-Paid Certifications
- Training, career progression and support
- Team Building Activities
- Innovative & Young Culture
Location
- We're based in Porto, Portugal. Remote work is possible, but we are only considering candidates located in Portugal.
Create a free Caio profile to unlock more results and save your role and location preferences.
Unlock free search