“… this Claim breaks new ground in a developing area of law related to what is copyright infringement of the copyright holder’s property within the United Kingdom on the internet and what is fair use of that copyright content. The Claimant would argue that a systemic practice by Google by the redirection of copyright content to copies of that content on aggregator websites is not remotely close to ‘fair use’ as it defrauds the original copyright holder of the property value of that media and its creative revenue generating capacity.”
“the largest image link farms are sites such as connect.in.com and zimbio.com, but there are a lot of smaller image-link farm that manage to hijack a lot of image search traffic (via google-image), a least 10 on my image traffic is being hijacked by image farms that hotlink my images (I can see that from my logs) and I estimate that the traffic hijacked by copied images (ie actually copied content with my watermarking cropped off) is much higher, given the number of copies of my images that i have located … and that image search engines lists in their results I file 50 DCMA notices every week, but it’s like a wack-a-mole game, because I have 10,000 images, and the highest ranking ones have been duplicated and are hosted by about 50 sites (for each image).”
“Hotlinking” (which is also known as inline linking) is a process whereby a website (“the first website”) displays a linked object, often in the form of an image, stored on the server hosting the content of another website (“the second website”). Hotlinking does not involve any copy being made of the image belonging to the second website. Instead, when a user visits a webpage on the first website containing, by way of example, a hotlinked image, the HTML code of that first website instructs the user’s browser to display the image directly from the server on which the content for the second website is hosted. The advantage of this approach from the first website operator’s perspective is that the transmission overhead needed to deliver the hotlinked image to the user is drawn from the second website’s resources, and not those of the first website.”
“How Google Search works 14. Google Search currently processes over 1.2 trillion search a year worldwide. This equates to 3.5 billion searches a day, and on average 40,000 searches every second … Google Search allows users to quickly access relevant information from billions of webpages on the web. Other search engines provide a similar service include Yahoo, Bing, Yandex and Baidu. 15. It would not be possible to navigate the web effectively without search engines. The internet is made up of over 130 trillion individual webpages and is constantly growing. According to www.worldwidewebsize.com as at6 November 2017 , the estimated number of indexable webpages on the internet was in excess of 45 billion. 16. In addition to searching for webpages, users of Google Search can undertake searches for images. Links to images published by third parties may be returned through Google Search results in a number of ways. In some instances, where a user carries out a search, thumbnail images may be returned as part of the results page, in addition to the usual text-based links to webpages. 17. When the user clicks on the “thumbnail” of an image on Google Search, this will open a further page in Google Search which identifies the domain name or website on which the image is published by the third party who operates the domain name or website. The user then has the option of either visiting the URL for the web page at which the image appears or navigating to the URL of the image file in its native size on the third party’s website. … Indexing Indexing 19. It is not possible for Google Search to search every web page available on the Internet in real time, and deliver results in a timeframe that would be acceptable to users. Google Search therefore compiles an index of the content of webpages, and it is this index that is examined during the search process. 20. To generate the index, software known as a “web crawler” (e.g. the “Google bot” web crawler) is used to find content that is on publicly available webpages. If the website owner does want content to be indexed by Google, the website owner can use standard techniques to ensure this … Google’s web crawler sends requests to servers hosting webpage content. If the website owners have configured these servers to respond to such requests, then the requested content will be sent to Google’s servers for indexing. Many thousands of requests are made simultaneously to populate Google’s index with information from the webpages that are being crawled. 21. The content of webpages examined by Google’s web crawler is saved and stored in a cache. The caching process of each webpage is automatic. In relation to images, a thumbnail copy of each image examined by Google’s web crawler is saved in a cache. The caching process of images is also automatic. The contents of webpages, including images, are stored in the cache only for a limited period of time so that they can be displayed rapidly when the web page is returned as part of Google Search results. The cache is rapidly updated at each exploration of the web by the web crawler to ensure that Google results reflect the evolution of webpages published online to provide users with an up-to-date mapping of the Internet. Google ensures the regular updating of the cached content by its web crawler in a timeframe which is consistent with the current state of the art and with industry practice and which is flexible depending on the importance of the page and the frequency with which it tends to be updated. The systems powering Google Search therefore only store cached content temporarily. 22. The web crawler does not cache content from webpages which are not generally accessible. The web crawler does not visit any webpages where the webmaster has instructed Google not to index its website as I described further below; it indexes webpages only where servers being configured to respond to its requests. Given that the web is made up of over 130 trillion individual webpages, it is evident that webpages indexed by Google which form Google’s index, a number in excess of 45 billion, represent only a proportion of the cached web. The caching is carried out with the implicit authorisation of the website publishers, since the act of publishing the content without restrictions on access implies that they agree that the information will be available to all including search engines such as Google Search. 23. The purpose of the cache storage and transmission is to facilitate the transmission of the content between the “recipients of the service” (as defined in Article 2 of the E-Commerce Directive and regulation 2 of the E-Commerce Regulations), ie between publishers of information on the internet and web users searching for content. The caching operations are necessary to optimise the transmission of Google Search results. Extracting and reviewing content in real time, from multiple sources and from servers located all over the world, for each individual search conducted, is not feasible from a technical standpoint. These caching operations therefore reduce the amount of data that must be transmitted over the web when undertaking searches using Google Search. The purpose of caching operations is to optimise and accelerate the flows of data over the relevant networks used by web users and Google, insofar as the cache copies will allow Google Search to display instantly the results corresponding to the queries of web users. 24. Once a search query is submitted by a user, Google Search results are then ranked in order of relative relevance to the user’s search query based upon the content of Google’s index. For a typical query, there may be thousands of webpages, or many more than that, with potentially relevant information. Accordingly, Google Search uses algorithms which rely on numerous signals to return results relevant to the query, ranked by their potential relevance. These signals include factors such as how often content on a website has been refreshed and the quality of user experience provided by a particular webpage. 25. One of the signals used by Google to rank potential relevance to the user’s query in the results returned by Google Search is known as “PageRank”
“User-agent: * Disallow” 29. “User-agent:* means that the instruction which follows applies to web crawlers visiting Mr Wheat’s website. “Disallow” has been left blank (without any text following the colon), which is an instruction to all web crawlers that all pages of Mr Wheat’s website should be indexed. 30. To prevent Google from indexing his website Mr Wheat could, at any time, configure the robots.txt file on his site to read: “User-agent: Googlebot Disallow:/”. 31. There are other methods that webmasters can use to prevent web crawlers from indexing websites, including implementing a “noindex” instruction. This instruction can either be included as a meta-tag within the HyperText Markup Language (HTML) code for a website or in an “HTTP response header” which instructs web crawlers not to index a website or parts thereof. 32. Webmasters therefore have full control over whether Google indexes their website. In general, most webmasters want to have their site indexed by web crawlers so that the website can be found by Google and other search engines. If a website is not indexed it cannot be located through search engines.”
“Intellectual Property; Brand Features”
“Other than set out expressly in the Agreement, neither party will acquire any right, title or interest in any intellectual property rights belonging to the other party or to the other party’s licensors.”