Skip to content

Comment on Papis: A CLI document and bibliography managerparent

Comments

You don't normally need to enter metadata manually. You can import it using the publication's doi, arxiv etc. Then papis gives you an opportunity to review and update the metadata. The other functionality that papis provides is a quick search through titles, authors etc.

Of course much of that can be done with doi2bib [0], arxiv2bib [1], etc., which you can combine with the wonderful bibtool [2] to keep a clean bibtex file. That's what I use and the advantage of this over papis is that you can version control it.

That said papis has its use. It's less heavy than zotero, a bit less proprietary format (the info is stored in yaml files iirc) and it provides a layer over bibtex. If it had full text indexing and search I could see myself using it more.

[0] https://github.com/bibcure/bibcure

[1] https://nathangrigg.github.io/arxiv2bib/

[2] https://www.gerd-neugebauer.de/software/TeX/BibTool/en/

Yes, many workflows are already made up of a lot of tools that are individually great but require a lot of duct tape. I guess the best path for a new effort is to work towards making any aspect of it more automatic. That's why even entering a DOI by hand seems like a regression, when Zotero can extract it automatically in most cases (as well as ISBN for books etc.).

Some areas where I would love to see innovation, and that seem entirely plausible for papis:

- Collaboratively clean metadata (i.e. share sets of papers with collaborators in a less silo-esque way than Zotero's online libraries

- Do a better job at collating related versions (i.e. multiple related talks, preprint, online first and published version) ...

- Automatically track and import new papers into staging area

- Automatically fetch PDFs from some external source (and share with collaborators)

Most of these are possible in some existing apps, but it seems that the designers always expect people to joyfully curate every nook and cranny of their libraries, when in reality I have five duplicate versions of many papers, each missing some crucial metadata, none being linked to the one PDF file that's somehow stuck on a WebDav share, and Google scholar metadata polluting BibTeX fields but omitting crucial details such as the publisher.

This field really is another case of "we can re-land and re-use rockets, but digitally organizing a theoretically well-structured information space is impossible".

I agree. Collaborative editing in this space is awful.

This field really is another case of "we can re-land and re-use rockets, but digitally organizing a theoretically well-structured information space is impossible".

Maybe because it doesn't matter that much to most people? People who live in uncommented F77 codebases are lightyears away from worrying about properly archiving information.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.