Posts

Showing posts with the label OAI-PMH

"We Need a Research Data Census" - Francine Berman

F rancine Berman's call for a research data census in the U.S. recognizes the reality that the valuable research assets produced by public (and private) research funding is uncounted, mostly unmanaged, and destined to be, or in the process of being, degraded, damaged and lost. Lost to future research, re-use, re-purposing. While a census is useful when your knowledge about a topic is effectively zero , as in this case, I don't think that it is a good ongoing solution to this particular problem. Distributed and open research data repositories, open standards like OAI-PMH , rich metadata (and the tools to create/manage them) and the will of funding agencies and research organizations can all come together to make a real-time census possible. But an initial census is clearly needed, in order properly discover the complete nature of the research data problem, in order to plan the processes, infrastructure and organizations to properly deal it. Berman, F. 2010. We Need a Research D...

Springer to acquire BioMed Central Group

I just read this happened earlier this month (more at the BioMed Central Blog ) via Peter Suber's Open Access News . I must admit I am rather surprised by this turn of events.

New Open Access Criterion: Support access by machines (m2m)

Related to my last posting ( FREE THE ARTICLES! (full-text for researchers & scientists and their machines) ) and in the light of Peter Murray-Rust's recent annoying discovery that he cannot text-mine Pubmed Central ( Can I data- and Text-mine Pubmed Central? ), I would like to suggest an additional criterion to the definition of Open Access: Open Access must include access by machines : At minimum one must allow crawls of the site/content or (to reduce the impact of badly configured crawlers) create a compressed XML file containing all metadata and either content, or direct links to content and make it available for download (and if bandwidth is still an issue put it on a P2P network like BitTorrent ). Preferable is to offer some kind of API (OTMI) or protocol (OAI-PMH) to get at content and metadata and citations. Better is to offer access to the XML of the articles in addition to the PDF and/or HTML; if the XML actually has some semantic content, then we are approaching th...