Modern Web Scraping with Erez Naveh

0
0

Today it’s estimated there are over 1 billion websites on the internet. Much of this content is optimized to be viewed by human eyes, not consumed by machines. However, creating systems to automatically parse and structure the web greatly extends its utility, and paves the way for innovative solutions and applications. The industry of web scraping has emerged to do just that. However, many websites erect obstacles to hinder web scraping. This has created a new kind of arms race between developers and anti-scraping software.




Bright Data has developed some of the most sophisticated consumer tools available to scrape public web data. Erez Naveh is an entrepreneur and former engineer at Meta. He is currently the VP of Product at Bright Data. Erez joins us in this episode to talk about Bright Data’s mission to structure the open web, and the toolkit they’ve developed to make this possible.




Paweł is the founder at flat.social the world’s first ‘flatverse’ start-up and glot.space, an AI-powered language learning app. Pawel’s background is as a full-stack software engineer with a lean and experimental approach towards product development. With a strong grounding in computing science, he spent the last decade getting early-stage products off the ground – both in startup and corporate settings. Follow Paweł on TwitterLinkedIn and his personal website – pawel.io.


Please click here to view this show’s transcript.


Sponsorship inquiries: sponsor@softwareengineeringdaily.com



The post Modern Web Scraping with Erez Naveh appeared first on Software Engineering Daily.


No comments yet...
Log in to comment
0 0 0
2025-04-22

Agentic AI at Glean with Eddie Zhou

Glean is a workplace search and knowledge discovery company that helps organizations find and access…
0 0 0
2025-04-17

Turing Award Special: A Conversation with Martin Hellman

Martin Hellman is an American cryptographer known for co-inventing public-key cryptography with Whit…
0 0 0
2025-04-15

Prometheus and Open-Source Observability with Eric Schabell

Modern cloud-native systems are highly dynamic and distributed, which makes it difficult to monitor …
0 0 0
2025-04-10

Turing Award Special: A Conversation with David Patterson

David A. Patterson is a pioneering computer scientist known for his contributions to computer archit…
0 0 0
2025-04-08

Uber’s On-Call Copilot with Paarth Chothani and Eduards Sidorovics

At Uber, there are many platform teams supporting engineers across the company, and maintaining robu…
0 0 0
2025-04-03

Turing Award Special: A Conversation with John Hennessy

John Hennessy is a computer scientist, entrepreneur, and academic known for his significant contribu…

Software Engineering Daily

Technical interviews about software topics.

Log in to Follow

More episodes from Software Engineering Daily

Top Podcasts Top rated Podcasts

Recent visits Your last viewed items

Even wachten...