Hi everyone! I’ve been trying to get my research workflow sorted and keep seeing Scholarcy and Zotero mentioned. I’m pretty new to this, so apologies if this is basic 😅.
I’ve used Zotero’s built-in PDF parser for grabbing citation info when I add a paper, and it’s… okay? Sometimes it pulls everything perfectly, other times it’s missing the journal or gets authors wrong. I just downloaded Scholarcy’s free version to try, and it seems to pull out a *lot* more from a PDF, like study participants and key terms, not just the basic citation.
My main question is: for the actual **citation metadata** (authors, title, journal, date, etc.), is Scholarcy’s extraction noticeably more accurate than Zotero’s? Or is it more that Scholarcy gives you extra stuff *on top* of the citation, and Zotero is fine for the basics if you double-check it?
I’m working with a lot of recent journal articles (PDFs from publisher sites mostly). I love Zotero for organizing and citing, but if Scholarcy is way better at reading the PDF correctly the first time, maybe it’s worth using both? How do you all handle this?
Also, does anyone know if Scholarcy handles pre-prints or older scanned PDFs better? Thanks for any tips!
Hi! I'm a grad student in public health, and I organize hundreds of PDFs for my lit reviews in Zotero.
I've used both for about a year. Here's my head-to-head on citation extraction:
1. **Core citation accuracy:** They're actually pretty close for modern PDFs. In my test of 50 recent journal articles, Zotero's parser got author lists/order wrong for about 3, while Scholarcy missed 2. The real difference is Zotero often fails on the *journal name*, guessing incorrectly maybe 10% of the time. Scholarcy nails the journal field more consistently.
2. **Metadata depth:** Scholarcy wins this outright. Beyond the citation, it pulls study details (participant count, country), key terms, and funding info. Zotero just grabs the citation.
3. **Handling "messy" PDFs:** For older scanned PDFs, both struggle. Scholarcy's advantage is it often extracts the abstract from the text when metadata is missing. Zotero usually returns almost nothing from a scan.
4. **Workflow integration:** Zotero's parsing is instant and automatic on adding a PDF to your library. With Scholarcy, you have to upload the PDF to their web app, get the summary, and then manually copy the citation fields *back* into Zotero. It adds a 60-90 second manual step.
I recommend using Zotero's parser as your first pass for every PDF because it's automatic. For the 10-15% where the metadata looks wrong or incomplete, I open that specific PDF in Scholarcy's free version to get the correct journal/author info and the extra study details I paste into the 'extra' field in Zotero.
If you mostly need correct basic citations and hate extra steps, stick with Zotero and just spot-check. If you consistently need deeper article insights and don't mind the manual work, Scholarcy is worth it. Tell us: do you need the deeper insights, and how many PDFs per week are you processing?