I'd be more OK with this if Google had a good API for their search results. But they've deprecated it, and now there is no alternative. So I'll continue to use 3rd parties that scrape Google results, until they change their mind.
Last time I looked into this (about a month ago), there's a lot of restrictions on the use of Gemini's search grounding results. There's not even an easy or approved way to de-mangle their returned URL's to get to the real URL of the search results. Has that changed recently?
I haven't used it but they were silly about their programmatic search api in the same way. Can't use the results for anything other than showing them as-is on a results page.
EU protects a database creator if there has been a qualitative or quantitative "substantial investment" in obtaining, verifying, or presenting the content, regardless of creative expression.
In USA copyright requires a minimum degree of original creativity in the selection, coordination, or arrangement of the data.
I think it's a rather grey line to say that Google search results are just facts, but eg maps are copyrightable. There's a rather large amount of effort involved in crawling and ranking the web - the PageRank itself should be copyrightable.
It's interesting to imagine how the world would be different, if internet advertising giants were partially liable for scams/malware that they facilitate.
I recall a articale I read "somewhere" that reported the Facebook makes big profit (Billions) from scams. So they have no ( or no strong motive) motive to shut such scams down
An example is courtcase against Meta for using a Australians mining billionares likeness to promote a crypo investment scam.
I'm surprised legitimate companies don't pressure Facebook on this though. There are enough scams on Facebook that I now refuse to believe anything there, even though some of the things look useful and probably are not scams (and also are things I didn't know existed without an ad - thus filling one of the legitimate values of advertisements: informing me of things that would make my life better but I don't know exist).
Your first assertion is obviously untrue. And fake celebrity endorsements pre-date the existence of the Internet, let alone Meta. There were lawsuits back in the 1800s on this topic. This is hardly a new problem unique to Meta.
There's nothing obvious about it. Meta makes money on ads, period. Scams work and get clicks. Therefore, meta makes money on scams running rampant on their platform.
> The whole thing was just “we don’t like that this is happening, so we’re suing.”
Typical behavior from a big company with immense resources. They probably thought they would get a settlement or SerpAPI could not afford to fight. I assume they are pretty small, at least in comparison to Google (I've never heard of them).
Google has so much money that even a "loser pays" requirement on litigation probably would not disuade them.
This ruling might feel good viscerally, but it also reinforces Googles own scraping as perfectly legal. At its inception, Google probably viewed this lawsuit as win-win. Either they successfully sue a competitor into oblivion or establish a precedent that will protect themselves in the future. Google lost, but they still won.
- Google search is on the way out. I don't know any of my peers who use it anymore.
- Coding models make doing extreme depth of work possible.
- Hoards of unemployed engineers now have access to 10x coding utilities and are looking for things to do. They will start clawing away at Google products.
- Just the other day, someone cloned Google Gsuite and it looked awesome
- Drive and Search will also be fungible products
- I'm itching at the chance to build my own phone operating system, and there must be thousands of others who want to do the same.
- Chrome can probably be replaced (Firefox gained a whole percentage point last month)
I don't think Google is safe anymore.
Two caveats that I'll give them:
- YouTube still has network effects and probably can't be dislodged
I tried both and it’s not even close which interface is better when it comes to answering a question. Google gives you stuff to sift through and interpret. AI just gives the answer.
Imagine you’re in the car or hands free or disabled and you just want the question answered.
"just gives you the answer" is a problem when it is offered without context, sourcing, etc.
Such is the case for people trying to poison--ahem "influence" LLM results for "what is the best restaurant in $my_locale". You're playing a dangerous game.
Claude Sonnet 5 Medium more or less on the timestamp of the comment:
Prompt: Who won the 2026 World Cup?
Answer: Spain won the 2026 World Cup, beating Argentina 1-0 after extra time in the final at MetLife Stadium on July 19. Ferran Torres scored the only goal in the 106th minute, coming on as a substitute in the 62nd minute. It’s Spain’s second World Cup title, having also won in 2010.
Prompt: Nearby BBQ places open now?
Answer: (a geolocation permission request prompt for the browser followed by) Right in [redacted] both [redacted] (4.6 stars, open until [redacted]) and [redacted] ([redacted]) are close and currently open.
A bit further out but highly rated: [redacted]
Seems like LLM does a good job on those questions…
The LLMs I've tried don't do well with very new stuff. Like Zig for example, they tell me answers that were good for Zig 0.12 but we on 0.16 now. So I've got to feed them the latest docs, then do the AI dance.
I've been using DDG for at least 3 years. Very rarely use Google anymore, and when I do it is when DDG doesn't find much and in those cases Google usually isn't any better.
are your peers me, myself, and I? that is a wild claim. I'll start asking around but I don't think I could find one person that says they don't use google search anymore
> Hoards of unemployed engineers now have access to 10x coding utilities and are looking for things to do. They will start clawing away at Google products.
No way this is threat to google. For the same reason why the same hordes of engineers did not managed to compete with large companies up to now. And for the same reason they were not producing all that many novel small apps last 10 years.
This is winner takes all economy. Tokens or no tokes, this is not the econony of small competing companies. This is the world of big ones.
(IANAL) I think that the deeper thing from this lawsuit is that from my understanding, (inherently) Search engines are considered public indexes and the data (URL's,index) behind it is considered uncopyrighted and as such aren't protected by DMCA because DMCA only works for copyrighted contents and thus the dismissal of the lawsuit by the Judge.
Basically, search engines are publicly scrapable, though I do wonder as from a law point of view, that it must be within the murky waters as to what a search engine means in terms of seperating its search engine code/its recomendation engine and the public data much of which are intertwined with each other.
I believe that the argument that could be made is that the recommendation engine is the way it is because of all the data and its unseperable to really copyright the whole mechanism in all its glory.
Speaking of which, it seems that AI models feel really similar. Does this judge lawsuit show that AI model weights aren't copyrightable as well? If a search engine is built on public indexes then so are the AI models. I was just writing similar comment on another thread but it seems to be the case, definitely worth a blog article or thinking more about perhaps this judgement by this judge itself in general as well, I just have a vibe that this judgement has pretty far reaching consequences in its impact.
Google crawling respects robots.txt and doesn't break capcha. It is easy to tell Google to piss off. SerpAPI fully relies on end-user proxies distributed like malware (in LG tv apps for instance), it has no other way it could function because it exclusively ingests data from sources that tell it to stop. If you wanted to scrape a bot friendly site, you wouldn't need SerpApi
But only vaguely. Google uses its monopoly position in advertising to basically ensure that you allow them to scrape your site (or if not you personally, the majority of revenue driving sites). They have the benefit of being allowed by default.
They also then scrape again at the user level for users operating chrome.
They also conveniently ignore global blocks for their adsbots (you have to specifically name them to block them).
If you're not Google, you likely don't have this luxury.
My preference would be that governments force search indexes to be public. The exact mechanisms for this can be debated.
I'd be more OK with this if Google had a good API for their search results. But they've deprecated it, and now there is no alternative. So I'll continue to use 3rd parties that scrape Google results, until they change their mind.
I agree, then need to bring back Nelson Minar’s old search API, or something like it.
Google does supply search grounding with Gemini API calls, and that is handy, but not general enough.
Last time I looked into this (about a month ago), there's a lot of restrictions on the use of Gemini's search grounding results. There's not even an easy or approved way to de-mangle their returned URL's to get to the real URL of the search results. Has that changed recently?
I haven't used it but they were silly about their programmatic search api in the same way. Can't use the results for anything other than showing them as-is on a results page.
EU protects a database creator if there has been a qualitative or quantitative "substantial investment" in obtaining, verifying, or presenting the content, regardless of creative expression.
In USA copyright requires a minimum degree of original creativity in the selection, coordination, or arrangement of the data.
I think it's a rather grey line to say that Google search results are just facts, but eg maps are copyrightable. There's a rather large amount of effort involved in crawling and ranking the web - the PageRank itself should be copyrightable.
They definitely started blowing through that line when they started serving up content answers as results.
It's quite important that SERPs are scrapeable, because they keep advertising scams like ETA/ESTA sites: https://www.bbc.co.uk/news/technology-56886957
It's interesting to imagine how the world would be different, if internet advertising giants were partially liable for scams/malware that they facilitate.
I recall a articale I read "somewhere" that reported the Facebook makes big profit (Billions) from scams. So they have no ( or no strong motive) motive to shut such scams down
An example is courtcase against Meta for using a Australians mining billionares likeness to promote a crypo investment scam.
https://www.afr.com/technology/dad-it-s-a-fraud-call-that-sp...
I'm surprised legitimate companies don't pressure Facebook on this though. There are enough scams on Facebook that I now refuse to believe anything there, even though some of the things look useful and probably are not scams (and also are things I didn't know existed without an ad - thus filling one of the legitimate values of advertisements: informing me of things that would make my life better but I don't know exist).
Your first assertion is obviously untrue. And fake celebrity endorsements pre-date the existence of the Internet, let alone Meta. There were lawsuits back in the 1800s on this topic. This is hardly a new problem unique to Meta.
There's nothing obvious about it. Meta makes money on ads, period. Scams work and get clicks. Therefore, meta makes money on scams running rampant on their platform.
> The whole thing was just “we don’t like that this is happening, so we’re suing.”
Typical behavior from a big company with immense resources. They probably thought they would get a settlement or SerpAPI could not afford to fight. I assume they are pretty small, at least in comparison to Google (I've never heard of them).
Google has so much money that even a "loser pays" requirement on litigation probably would not disuade them.
DMCA needs to be reformed one day.
I am somewhat confused - does this mean we can all now legally scrape Google search results?
Always has been
Legally and ‘they won’t do everything they can to stop you’ are not the same thing of course.
Hypocrisy-rich.
If only GPT wouldn't refuse my requests to write a crawler for $site. :(
Gemini has been perfectly willing to write such things for me
Having good thoughts about Google is kind of nostalgic!
What about them taking content for their AI summaries? Have they created a system that gets content owners and creators paid in this regard yet?
These are still open: https://ec.europa.eu/commission/presscorner/detail/en/ip_25_...
https://www.epceurope.eu/post/european-publishers-council-fi...
This ruling might feel good viscerally, but it also reinforces Googles own scraping as perfectly legal. At its inception, Google probably viewed this lawsuit as win-win. Either they successfully sue a competitor into oblivion or establish a precedent that will protect themselves in the future. Google lost, but they still won.
Google has no moat anymore.
- Google search is on the way out. I don't know any of my peers who use it anymore.
- Coding models make doing extreme depth of work possible.
- Hoards of unemployed engineers now have access to 10x coding utilities and are looking for things to do. They will start clawing away at Google products.
- Just the other day, someone cloned Google Gsuite and it looked awesome
- Drive and Search will also be fungible products
- I'm itching at the chance to build my own phone operating system, and there must be thousands of others who want to do the same.
- Chrome can probably be replaced (Firefox gained a whole percentage point last month)
I don't think Google is safe anymore.
Two caveats that I'll give them:
- YouTube still has network effects and probably can't be dislodged
- Google cloud isn't going anywhere
> Google search is on the way out. I don't know any of my peers who use it anymore.
What do they use?
Of all my tech friends, colleagues, I'm the only one who uses Kagi. Another person uses searx. Everyone else uses Google.
Of my non-tech friends and colleagues, everyone uses Google.
Let me tell you about this new little-known technology that's been gaining traction in the last few years...
Sure, go ahead and ask the LLM who won the 2026 World Cup. Or nearby BBQ places open now.
Even assuming no hallucinations, an LLM can replace some use cases for search but not all.
I tried both and it’s not even close which interface is better when it comes to answering a question. Google gives you stuff to sift through and interpret. AI just gives the answer.
Imagine you’re in the car or hands free or disabled and you just want the question answered.
https://ibb.co/whYmZWxs
https://ibb.co/JFkQbmJf
"just gives you the answer" is a problem when it is offered without context, sourcing, etc.
Such is the case for people trying to poison--ahem "influence" LLM results for "what is the best restaurant in $my_locale". You're playing a dangerous game.
Claude Sonnet 5 Medium more or less on the timestamp of the comment:
Prompt: Who won the 2026 World Cup?
Answer: Spain won the 2026 World Cup, beating Argentina 1-0 after extra time in the final at MetLife Stadium on July 19. Ferran Torres scored the only goal in the 106th minute, coming on as a substitute in the 62nd minute. It’s Spain’s second World Cup title, having also won in 2010.
Prompt: Nearby BBQ places open now?
Answer: (a geolocation permission request prompt for the browser followed by) Right in [redacted] both [redacted] (4.6 stars, open until [redacted]) and [redacted] ([redacted]) are close and currently open.
A bit further out but highly rated: [redacted]
Seems like LLM does a good job on those questions…
The LLMs I've tried don't do well with very new stuff. Like Zig for example, they tell me answers that were good for Zig 0.12 but we on 0.16 now. So I've got to feed them the latest docs, then do the AI dance.
Fwiw, I've done those kinds of requests before and it did so successfully.
It did use Google to provide me with the answers though, sooo...
Whether it's a good choice or not sadly doesn't change the current reality. People now use ChatGPT as a replacement for google.
Really bad idea for a lot of search use cases.
> Of my non-tech friends and colleagues, everyone uses Google.
Me, almost, too.
I have some friends using duckduckgo - not many and only until Google is the default again...
I've been using DDG for at least 3 years. Very rarely use Google anymore, and when I do it is when DDG doesn't find much and in those cases Google usually isn't any better.
> Google search is on the way out. I don't know any of my peers who use it anymore.
bear in mind we're on Hacker News. Google's market share is still above 90% - in almost any other market this would be a ridiculous monopoly.
> - I'm itching at the chance to build my own phone operating system, and there must be thousands of others who want to do the same.
Maybe look into contributing to existing alternatives like PostmarketOS.
https://postmarketos.org/
Please don't send vibecoders to established open source projects . Your literally sabotaging them if you succeed.
Google’s moat is that websites aren’t blocking their crawler.
are your peers me, myself, and I? that is a wild claim. I'll start asking around but I don't think I could find one person that says they don't use google search anymore
> Hoards of unemployed engineers now have access to 10x coding utilities and are looking for things to do. They will start clawing away at Google products.
No way this is threat to google. For the same reason why the same hordes of engineers did not managed to compete with large companies up to now. And for the same reason they were not producing all that many novel small apps last 10 years.
This is winner takes all economy. Tokens or no tokes, this is not the econony of small competing companies. This is the world of big ones.
If only there were laws about antitrust. Oh well.
(IANAL) I think that the deeper thing from this lawsuit is that from my understanding, (inherently) Search engines are considered public indexes and the data (URL's,index) behind it is considered uncopyrighted and as such aren't protected by DMCA because DMCA only works for copyrighted contents and thus the dismissal of the lawsuit by the Judge.
Basically, search engines are publicly scrapable, though I do wonder as from a law point of view, that it must be within the murky waters as to what a search engine means in terms of seperating its search engine code/its recomendation engine and the public data much of which are intertwined with each other.
I believe that the argument that could be made is that the recommendation engine is the way it is because of all the data and its unseperable to really copyright the whole mechanism in all its glory.
Speaking of which, it seems that AI models feel really similar. Does this judge lawsuit show that AI model weights aren't copyrightable as well? If a search engine is built on public indexes then so are the AI models. I was just writing similar comment on another thread but it seems to be the case, definitely worth a blog article or thinking more about perhaps this judgement by this judge itself in general as well, I just have a vibe that this judgement has pretty far reaching consequences in its impact.
“You are trying to kidnap what I have rightfully stolen, and I think it quite ungentlemanly."
Google crawling respects robots.txt and doesn't break capcha. It is easy to tell Google to piss off. SerpAPI fully relies on end-user proxies distributed like malware (in LG tv apps for instance), it has no other way it could function because it exclusively ingests data from sources that tell it to stop. If you wanted to scrape a bot friendly site, you wouldn't need SerpApi
I'm vaguely sympathetic to this argument.
But only vaguely. Google uses its monopoly position in advertising to basically ensure that you allow them to scrape your site (or if not you personally, the majority of revenue driving sites). They have the benefit of being allowed by default.
They also then scrape again at the user level for users operating chrome.
They also conveniently ignore global blocks for their adsbots (you have to specifically name them to block them).
If you're not Google, you likely don't have this luxury.
My preference would be that governments force search indexes to be public. The exact mechanisms for this can be debated.
This sums up every single bit of AI "progress" since 2022.
Source:
Google vs. SerpApi: The Court Granted Our Motion to Dismiss
https://news.ycombinator.com/item?id=48995411