Does google scrape links from PDF files? do these links pass link juice?
-
Title is pretty much the whole question.
-
I made a test and it seems that yes, the links from pdf count for ranking.
The test is on my Romanian blog http://seogan.ro/link-building-pdf-urile-o-sursa-de-linkuri-test
You can find an English translation here: http://www.seogan.com/pdf-link-building
Hope it helps.
-
Yes it does according to Google tech spec http://code.google.com/apis/searchappliance/documentation/50/admin_crawl/Introduction.html
which specifically states if follows html links in pdf 'It follows HTML links in PDF files, Word documents, and Shockwave documents'. Google's own api docs carry more weight than a comment in a forum_._ If they are licencing this out as an application it would suggest that the same technology is available in the main engine as does Dunamis's comment about a listing in a pdf document being found in search results.
You can test for youself by publishing a pdf with a link to a info page that does not show up in any other links. Include the pdf in your sitemap but not the test page and check if it shows in googles index site:yoursite.com the next time it crawls.
This also gives some insight in an interview with Matt Cutts - http://www.stonetemple.com/articles/interview-matt-cutts-012510.shtml
Eric Enge: What about PDF files?
Matt Cutts: We absolutely do process PDF files. I am not going to talk about whether links in PDF files pass PageRank. But, a good way to think about PDFs is that they are kind of like Flash in that they aren't a file format that's inherent and native to the web, but they can be very useful. In the same way that we try to find useful content within a Flash file, we try to find the useful content within a PDF file. At the same time, users don't always like being sent to a PDF. If you can make your content in a Web-Native format, such as pure HTML, that's often a little more useful to users than just a pure PDF file.
-
This person seems to think no: http://www.google.fr/support/forum/p/Webmasters/thread?tid=14c5fe970fe84361&hl=en
but i'm not sure how much i can trust a random comment from a random source. any evidence for either argument?
EDIT: And this person seems to think they do pass link juice: http://www.whydowork.com/blog/link-building/274/
Could a mod remove the marked as answered? i don't think i am able to remove it, and the question isn't really answered.
-
yes, but do they crawl the links they find in these documents, or do they just index their contents.
-
Hmmm although i thought you had answered my question, i actually feel that you have not... Yes the links you provided state that google scrapes pdfs and even OCRs pdfs to get a better idea what is in them, but i don't see anywhere that they mention crawling the urls they find in these pdf documents.
-
Google definitely does index the contents of pdf files. I found this out the hard way as I had a real estate pdf on my site that I wanted to have listed in the index, but I didn't know that the contents would be crawled. The pdf contained some listings that I was not legally allowed to advertise on my site. (It was legal for me to give someone a report with the listings in it though).
When another realtor was searching for their own listing, my pdf came up. I got in trouble. I'm ok now though.
-
Have a look at this article http://searchenginewatch.com/article/2067225/Google-Does-PDF-Other-Changes it explains some of the doc library search for pdf files and Google's statement here http://googleblog.blogspot.com/2008/10/picture-of-thousand-words.html.
Got a burning SEO question?
Subscribe to Moz Pro to gain full access to Q&A, answer questions, and ask your own.
Browse Questions
Explore more categories
-
Moz Tools
Chat with the community about the Moz tools.
-
SEO Tactics
Discuss the SEO process with fellow marketers
-
Community
Discuss industry events, jobs, and news!
-
Digital Marketing
Chat about tactics outside of SEO
-
Research & Trends
Dive into research and trends in the search industry.
-
Support
Connect on product support and feature requests.
Related Questions
-
When we use 'link:' for who get the link, how come google show us the same domain as a link.
the search result show the domain of its own. what is is? and is it meaningful as a link?
Link Building | | onedaykorea0 -
Link Building
Hi Mozzers, I work for an IT company specialising in outsourcing Cisco Engineers worldwide. Due to being involved heavily in outsourcing, I'm finding it impossible to generate genuine and organic links to our website. My question is that is it essential to have back links back to our website to achieve high search engine rankings? I'm unable to get clients and other partners to provide links as we're the outsourced company,but I can get backlinks from unrelated c ompanies. For example we're running a CSR initiative with a National Gym relating to health & fitness, also links from our web designers etc... Do these links have any type of value? Any help and feedback would be greatly appreciated. Jason 🙂
Link Building | | 4Cornernetworks0 -
How do I search for a link within a competitors website that is linking my website?
Hello, I have just been checking my link profile and according to webmaster tools a competitor is linking to my website. Is there anyway of finding this link besides looking at every page within their website? Many thanks
Link Building | | mblsolutions0 -
When buying used domains, how do i see the links pointing to that domain? OSE not showing links
when buying used domains, how do i see the links pointing to that domain? Sometimes the open site explorer doesn't show any links to the domain, especially if the domain is parked. Obviously a domain for sale with 1000 domains linking to it has lots of SEO Value right? Thanks mozzers!
Link Building | | Ron100 -
Does Open Site Explorer show juice passing links
I'm a little confused between all the link types. Internal / External - easy Linking Root Domain - easy Followed / No followed - easy But then there's talk about "juice passing links" and I can't quite get how this is defined, and why it's something you can get from the API, but not from Open Site Explorer... or can you get that info from OSE?
Link Building | | eatyourveggies0 -
What are the new ways of Link Building which will get quality links?
Hi We are doing following kinds of Link Building: 1. Article marketing 2. Directory Submission 3. Guest Blogging 4. Press Release submission 5. Infographics (Viral Marketing) I know blog commenting and links from discussion forums are considered as greyhat / blackhat by google. So I want suggestions on some other ethical(Whitehat SEO) ways of Link Building to get quality Links. Also suggest me some Free article submission sites which give " dofollow-links" Thank You
Link Building | | Virrtuo0 -
Link values
I have an interesting question, it's more to see other peoples opinions than looking for an exact answer, as I know noone can give that. If you had a good link on harvard.edu.. how many very low quality links on low DA domains that are full of junk, would you need to match the harvard link... if all other factors were equal. I'm gonna go for 500.. I've had great results with a few high DA/trust links. I'm not looking for anyone to give me a definative answer, just curious as to the views of other SEOs 😄
Link Building | | PeterM220 -
New ways to get links (excluding on site link bait)?
Fellow SEOs, I’m looking for a new way to reach out to bloggers and site owners for links. There are quite a few link bait techniques to lure bloggers thanks to = amazing content, infographics,...but there are not so many articles on ways to contact bloggers/webmasters and ask for links... We all know that contacting site owners directly and politely asking then to place a link on their home page is as useful as a comb in Bruce willis’s hand. Link exchange is as dead as the Dodo and business partner links are usefull and easier to get but opportunities are often limited. I have been organizing contests for bloggers recently and it has been quite successful up to now, but I always like to have a spare trick up my sleeve. So if any of you have a decent method to reach out to bloggers/webmaster and get decent links please let me know, I will be forever thankful ; ) Cheers
Link Building | | ref.price0