Happy Birthday Open Access News!

Open Access News is five today. OAN‘s indefatigable primary author Peter Suber has written over 10,800 OAN postings during this period. Going further back to 2001, he has written 109 issues of the SPARC Open Access Newsletter (formerly called the Free Online Scholarship Newsletter) as well as important papers on open access.

Thanks, Peter. The open access movement owes you a huge debt of gratitude for this fine work.

The REMAP Project: Record Management and Preservation in Digital Repositories

The REMAP Project at the University of Hull has been funded by JISC investigate how record management and digital preservation functions can be best supported in digital repositories. It utilizes the Fedora system.

Here’s an except from the Project Aims page (I have added the links in this excerpt):

The REMAP project has the following aims:

  • To develop Records Management and Digital Preservation (RMDP) workflow(s) in order to understand how a digital repository can support these activities
  • To embed digital repository interaction within working practices for RMDP purposes
  • To further develop the use of a WSBPEL orchestration tool to work with external Web services, including the PRONOM Web services, to provide appropriate metadata and file information for RMDP
  • To develop and test a notification layer that can interact with the orchestration tool and allow RSS
    syndication to individuals alerting them to RMDP tasks
  • To develop and test an intermediate persistence layer to underpin the notification layer and interact
    with the WSBPEL orchestration tool to allow orchestrated workflows to take place over time
  • To test and validate the use of the enhanced WSBPEL tool with institutional staff involved in RMDP activities

SWORD (Simple Web-service Offering Repository Deposit) Project

Led by UKOLN, The JISC SWORD (Simple Web-service Offering Repository Deposit) Project is developing "a prototype ‘smart deposit’ tool" to "facilitate easier and more effective population of repositories."

Here’s an excerpt from the project plan:

The effective and efficient population of repositories is a key concern for the repositories community. Deposit is a crucial step in the repository workflow; without it a repository has no content and can fulfill no further function. Currently most repositories exist in a fairly linear context, accepting deposits from a single interface and putting them into a single repository. Further deployment of repositories, encouraged by JISC and other funders, means that this situation is changing and we are beginning to see an increasingly complex and dynamic ecology of interactions between repositories and other services and systems. By and large developers are not creating repository systems and software from scratch, rather they are considering how repositories interface with other applications within institutions and the wider information landscape. A single repository, or multiple repositories, might interact with other components, such as VLEs, authoring tools, packaging tools, name authority services, classification services and research systems. In terms of content, resources may be deposited in a repository by both human and software agents, e.g. packaging tools that push content into repositories or a drag-and-drop desktop tool. The type of resource being deposited will also influence the choice of deposit mechanism. If the resources are complex packaged objects then a web service will need to support the ingest of multiple packaging standards.

There is currently no standard mechanism for accepting content into repositories, yet there already exists a stable and widely implemented service for harvesting metadata from repositories (OAI-PMH—Open Archives Initiative Protocol for Metadata Harvesting). This project will implement a similarly open protocol or specification for deposit. By taking a similar approach, the project and the resulting protocol and implementations will gain easier acceptance by a community already familiar with the OAI-PMH.

This project aims to develop a Simple Web-service Offering Repository Deposit (SWORD)—a lightweight deposit protocol that will be implemented as a simple web service within EPrints, DSpace, Fedora and IntraLibrary and tested against a prototype ‘smart deposit’ tool. The project plans to take forward the lightweight protocol originally formulated by a small group working within the Digital Repositories Programme (the ‘Deposit API’ work) . The project is aligned with the Object Reuse and Exchange (ORE) Mellon-funded two-year project by the Open Archives Initiative, which commenced in October 2006. Members of the SWORD project team are represented on its Technical and Liaison Committees. . . . . The SWORD project is not attempting to duplicate work being done being done by ORE, but seeks to build on existing work to support UK-specific requirements whilst feeding into the ongoing ORE project.

Position Papers from the NSF/JISC Repositories Workshop

Position papers from the NSF/JISC Repositories Workshop are now available.

Here’s an excerpt from the Workshop’s Welcome and Themes page:

Here is some background information. A series of recent studies and reports have highlighted the ever-growing importance for all academic fields of data and information in digital formats. Studies have looked at digital information in science and in the humanities; at the role of data in Cyberinfrastructure; at repositories for large-scale digital libraries; and at the challenges of archiving and preservation of digital information. The goal of this workshop is to unite these separate studies. The NSF and JISC share two principal objectives: to develop a road map for research over the next ten years and what to support in the near term.

Here are the position papers:

Friday’s OAI5 Presentations

Presentations from Friday’s sessions of the 5th Workshop on Innovations in Scholarly Communication in Geneva are now available.

Here are a few highlights from this major conference:

  • Doctoral e-Theses; Experiences in Harvesting on a National and European Level (PowerPoint): "In the presentation we will show some lessons learned and the first results of the Demonstrator, an interoperable portal of European doctoral e-theses in five countries: Denmark, Germany, the Netherlands, Sweden and the UK."
  • Exploring Overlay Journals: The RIOJA project (PowerPoint): "This presentation introduces the RIOJA (Repository Interface to Overlaid Journal Archives) project, on which a group of cosmology researchers from the UK is working with UCL Library Services and Cornell University. The project is creating a tool to support the overlay of journals onto repositories, and will demonstrate a cosmology journal overlaid on top of arXiv."
  • Dissemination or Publication? Some Consequences from Smudging the Boundaries between Research Data and Research Papers (PDF): "Project StORe’s repository middleware will enable researchers to move seamlessly between the research data environment and its outputs, passing directly from an electronic article to the data from which it was developed, or linking instantly to all the publications that have resulted from a particular research dataset."
  • Open Archives, The Expectations of the Scientific Communities (RealVideo): "This analysis led the French CNRS to start the Hal project, a pluridisciplinary open archive strongly inspired by ArXiv, and directly connected to it. Hal actually automatically transfers data and documents to ArXiv for the relevant disciplins; similarly, it is connected to Pum Med and Pub Med Central for life sciences. Hal is customizable so that institutions can build their own portal within Hal, which then plays the role of an institutional archive (examples are INRIA, INSERM, ENS Lyon, and others)."

(You may want to download PowerPoint Viewer 2007 if you don’t have PowerPoint 2007).

Thursday’s OAI5 Presentations

Presentations from Thursday’s sessions of the 5th Workshop on Innovations in Scholarly Communication in Geneva are now available.

Here are a few highlights from this major conference:

  • Business Models for Digital Repositories (PowerPoint): "Those setting up, or planning to set up, a digital repository may be interested to know more about what has gone before them. What is involved, what is the cost, how many people are needed, how have others made the case to their institution, and how do you get anything into it once it is built? I have recently undertaken a study of European repository business models for the DRIVER project and will present an overview of the findings."
  • DRIVER: Building a Sustainable Infrastructure of European Scientific Repositories (PowerPoint): "Ten partners from eight countries have entered into an international partnership, to connect and network as a first step more than 50 physically distributed institutional repositories to one, large-scale, virtual Knowledge Base of European research."
  • On the Golden Road : Open Access Publishing in Particle Physics (RealVideo): "A working party works now to bring together funding agencies, laboratories and libraries into a single consortium, called SCOAP3 (Sponsoring Consortium for Open access Publishing in Particle Physics). This consortium will engage with publishers towards building a sustainable model for open access publishing. In this model, subscription fees from multiple institutions are replaced with contracts with publishers of open access journals where the SCOAP3 consortium is a single financial partner."
  • Open Access Forever—Or Five Years, Whichever Comes First: Progress on Preserving the Digital Scholarly Record (RealVideo): "The current state of the curation and preservation of digital scholarship over its entire lifecycle will be reviewed, and progress on problems of specific interest to scholarly communication will be examined. The difficulty of curating the digital scholarly record and preserving it for future generations has important implications for the movement to make that record more open and accessible to the world, so this a timely topic for those who are interested in the future of scholarly communication."

(You may want to download PowerPoint Viewer 2007 if you don’t have PowerPoint 2007).

OpenDOAR API

The OpenDOAR project has announced the availability of an API for accessing digital repository data in their database.

Here’s an excerpt from the press release:

OpenDOAR, as a SHERPA project, is pleased to announce the release of an API that lets developers use OpenDOAR data in their applications. It is a machine-to-machine interface that can run a wide variety of queries against the OpenDOAR Database and get back XML data. Developers can choose to receive just repository titles & URLs, all the available OpenDOAR data, or intermediate levels of detail. They can then incorporate the output into their own applications and ‘mash-ups’, or use it to control processes such as OAI-PMH harvesting. . . .

OpenDOAR is a continuing project hosted at the University of Nottingham under the SHERPA Partnership. OpenDOAR maintains and builds on a quality-assured list of the world’s Open Access Repositories. OpenDOAR acts as a bridge between repository administrators and the service providers who make use of information held in repositories to offer search and other services to researchers and scholars worldwide.

A key feature of OpenDOAR is that all of the repositories we list have been visited by project staff, tested and assessed by hand. We currently decline about a quarter of candidate sites as being broken, empty, out of scope, etc. This gives a far higher quality assurance to the listings we hold than results gathered by just automatic harvesting. OpenDOAR has now surveyed over 1,100 repositories, producing a classified Directory of over 800 freely available archives of academic information.

Wednesday’s OAI5 Presentations

Presentations from Wednesday’s sessions of the 5th Workshop on Innovations in Scholarly Communication in Geneva are now available.

Here are a few highlights from this major conference:

  • MESUR: Metrics from Scholarly Usage of Resources (PowerPoint): "The two-year MESUR project, funded by the Andrew W. Mellon Foundation, aims to define and validate a range of usage-based impact metrics, and issue guidelines with regards to their characteristics and proper application. The MESUR project is constructing a large-scale semantic model of the scholarly community that seamlessly integrates a wide range of bibliographic, citation and usage data."
  • OAI Object Re-Use and Exchange (PowerPoint): "In this presentation, we will give an overview of the current activities, including: defining the problem of compound documents within the web architecture, enumerating and exploring several use cases, and identifying likely adopters of OAI-ORE."
  • OpenDOAR Policy Tools and Applications (RealVideo): "OpenDOAR has developed a set of policy generator tools for repository administrators and is contacting administrators to advocate policy development."
  • State of OAI-PMH (PowerPoint): "The OAI-PMH was released in 2001 and stabilized at v2.0 in 2002. Since then there has been steady growth in adoption of the protocol. Support for the OAI-PMH is assumed for base-level interoperability between institutional repositories, and is also provided for many other collections of scholarly material. I will review the current landscape and reflect on some milestones and issues."

(You may want to download PowerPoint Viewer 2007 if you don’t have PowerPoint 2007).

The Depot: A UK Digital Repository

The JISC Repositories and Preservation program has established the Depot, so that researchers who do not have an institutional repository can deposit digital postprints and other digital objects.

Here’s an excerpt from the press release:

The general strategy being adopted in the UK is that every university should develop and establish its own institutional repository (IR), as part of a comprehensive ‘JISC RepositoryNet’. Many researchers can already make use of the IRs set up in their institution, but that is not (yet) the case for all. A key purpose for The Depot is to bridge that gap during the period before all have such provision, and to provide a deposit facility that will enable all UK researchers to expose their publications to readers under terms of Open Access.

The Depot will also have a re-direct function to link researchers to the appropriate home pages of their own institutional repositories. The end result should be more content in repositories, making it easier for researchers and policy makers to have peer-reviewed research results exposed to wider readership under Open Access. . . .

The principal focus for The Depot is the deposit of post-prints, digital versions of published journal articles and similar items. There are plans to include links to places for depositing other digital materials, such as research datasets and learning materials. As indicated, The Depot helps provide a level-playing field for all UK researchers and their institutions, especially when deposit under Open Access is required by grant funding bodies. It may also become a useful facility for institutions as they implement and manage their own repositories, helping to promote the habit of deposit among staff, with the simple message, ‘put it in the depot’.

The Depot is based on E-Prints software and is compliant with the Open Archive Initiative (OAI), which promotes standards for repository interoperability. Its contents will be harvested and searched through the Intute Repository Search project. It offers a redirect service, UK Repository Junction, to ensure that content that comes within the remit of an extant repository is correctly placed there instead of in The Depot.

Additionally, as IRs are created, The Depot will offer a transfer service for content deposited by authors based at those universities, to help populate the new IRs. The Depot will therefore act as a ‘keepsafe’ until a repository of choice becomes available for deposited scholarly content. In this way, The Depot will avoid competing with extant and emerging IRs while bridging gaps in the overall repository landscape and encouraging more open access deposits.

A Depot FAQ is available.

OpenLOCKSS Project

Led by the University of Glasgow Library, the new JISC-funded OpenLOCKSS project will preserve selected UK open access publications.

Here’s an excerpt from the project proposal:

Although LOCKSS has initially concentrated on negotiations with society and commercial publishers, there has always been an interest in smaller open-access journals, as evidenced by the LOCKSS Humanities Project1, where twelve major US libraries have collaborated to contact more than fifty predominantly North American open access journal titles, enabling them to be preserved within the LOCKSS system. . . .

At present, much open access content is under threat, and is difficult to preserve for posterity under standard arrangements, at least until the British Library, and the other UK national libraries, are able to take a more proactive and comprehensive stance in preserving websites comprising UK output. Many open access journals are small operations, often dependent on one or two enthusiastic editors, often based in university departments and/or small societies, concerned with producing the next issue, and often with very little interest in or knowledge of preservation considerations. Their long term survival beyond the first few issues can often be in doubt, but their content, where appropriate quality controls have been applied, is worthy of preservation.

LOCKSS is an ideal low-cost mechanism for ensuring preservation, provided that appropriate contacts can be made and plug-in developments completed, and sufficient libraries agree to host content, on the Humanities Project model. . . .

Earlier in 2006, a survey was carried out by the LOCKSS Pilot Project, to discover preferences for commercial/society publishers to approach with a view to participating in LOCKSS, and Content Complete Ltd have been undertaking this work, as well as negotiating with the NESLi2 publishers on their LOCKSS participation. . . .

We propose to consider initially the titles with at least six votes (it may not be appropriate to approach all these titles, for example we shall check that all are currently publishing and confirm that they appear to be of appropriate quality), followed by those with five or four votes. We propose that agreements for LOCKSS participation are concluded with at least twelve titles, with fifteen as a likely upper limit.

Repository 66: OA Digital Repository Map Mashup

Stuart Lewis of the University of Wales Aberystwyth has created a Google Map mashup called Repository 66 that shows worldwide open access digital repositories using data from ROAR and OpenDOAR. (Route 66 was a famous highway in the US.)

Dr. John Hoey Joins the Scholarly Exchange Board

Julian Fisher, Managing Director of the Scholarly Exchange, has announced that Dr. John Hoey has joined the Scholarly Exchange Board.

Here’s an excerpt from the SPARC-OAForum announcement:

Dr. Hoey is the former editor-in-chief of the Canadian Medical Association Journal and long an advocate of open access publishing. A specialist in community medicine and internal medicine, he is Professor of Medicine (adjunct) at Queen’s University and a Special Advisor to the Principal on Public Health.

Scholarly Exchange, Inc. has eliminated a major obstacle in starting open access journals by providing a free and fully supported e-publishing platform. Combining Open Journal Systems public-domain software with complete hosting and support, this service offers scholars unrivaled freedom and flexibility to produce academic journals at a price that fosters the open access model. It also develops tools and methods to promote and support open access journals.

Report About Users’ Digital Repository Needs at the University of Hull

The RepoMMan Project at the University of Hull has published The RepoMMan User Needs Analysis report.

Here’s an excerpt from the JISC-REPOSITORIES announcement:

The document covers the repository needs of users in the research, learning & teaching, and administration areas. Whilst based primarily on needs expressed in interviews at the University of Hull the document is potentially of wider applicability, drawing from an on-line survey of researchers elsewhere and a survey of the L&T community undertaken by the CD-LOR Project.

DRAMA Project’s Fedora Authentication Code Alpha Release

The DRAMA (Digital Repository Authorization Middleware Architecture) project has released an alpha version of its Fedora authentication code. DRAMA is part of the RAMP (Research Activityflow and Middleware Priorities Project) project.

Here’s an excerpt from the fedora-commons-users announcement about the release’s features:

  • Federated authentication (using Shibboleth) for Fedora.
  • Extended XACML engine support via the introduction of an XML database for storing and querying policies and XACML requests over web services.
  • Re-factoring of Fedora XACML authorization into an interceptor layer which is separate from Fedora.
  • A new web GUI for Fedora nicknamed "mura" (Note: that we will be changing the GUI name to a new one soon).

Polimetrica Publisher: An Open Access Book Publisher

Polimetrica Publisher is a scientific open access book publisher. It has published a number of books in the areas of applied, pure, and human sciences.

This excerpt from its "Our Open Access Manifesto" describes its philosophy and business model:

Polimetrica Publisher works from a simple premise: that for a better future of the people it’s possible to disseminate the knowledge by publishing innovative books freely accessible to anyone in the world who might be interested.

Informed by that premise, we’re trying to build a new model of scientific publishing that embraces economic self-subsistence, openness, and fairness; the model is based on the following elements:

  1. each scientific book is published in two editions: a printed edition, available in the market, and an electronic edition, freely available through the web; both editions are identified by a different ISBN code.
  2. each scientific book is edited in collaboration with universities or with authoritative professors or specialists.
  3. the printed edition is distributed on the international market.
  4. the electronic edition is free access through the Polimetrica web site.
  5. Polimetrica pays to the author or to the academic institution on all sales of the printed edition a 10% royalty of the net receipts.
  6. each scientific publication is funded by a contribution of 1.500 Euros about.
  7. anyone interested in our activities is encouraged to buy a membership; the members will have access to special conditions. Additional information are at the page
    http://www.polimetrica.com/main/membership.php

Polimetrica Publisher currently has three membership options that provide a specified number of books on CD-ROM/DVD, discount prices, and newsletters.

To download free digital book, the user fills out a form providing name, country, and e-mail address. A download link is sent to the provided e-mail address.

A book that may be of particular interest to DigitalKoans readers is Open Access: Open Problems.

Report on Sharing and Re-Use of Geospatial Data in Repositories

The GRADE project has released a report titled Designing a Licensing Strategy for Sharing and Re-Use of Geospatial Data in the Academic Sector.

The JISC-REPOSITORIES announcement indicates that the report presents "a licensing strategy for the sharing and re-use of geospatial data within the UK research and education sector," and that it "puts forward a conceptual framework for resolving those described rights management issues raised in relation to repositories."

Here is an excerpt from the report that describes it further:

Geospatial material created in the education sector can be highly complex, incorporating data created elsewhere either as found, or customised to fit the particular need of the academic or lecturer. The downstream rights can become very complex, as it is necessary to ensure that permissions have been gained to reuse or repurpose the data, and it is usually essential that correct attribution is made. There are currently concerns and confusion over the assertion of IPR and copyright of created geospatial data particularly where third party data are included.

This report considers a licensing strategy for the sharing and re-use of geospatial data within the UK research and education sector.

Petition for Public Access to Publicly Funded Research in the United States

AALL, ALA, ACRL, the Alliance for Taxpayer Access, Public Knowledge, SPARC, and other organizations have initiated the Petition for Public Access to Publicly Funded Research in the United States.

The petition states:

We, the undersigned, believe that broad dissemination of research results is fundamental to the advancement of knowledge. For America’s taxpayers to obtain an optimal return on their investment in science, publicly funded research must be shared as broadly as possible. Yet too often, research results are not available to researchers, scientists, or the members of the public. Today, the Internet and digital technologies give us a powerful means of addressing this problem by removing access barriers and enabling new, expanded, and accelerated uses of research findings.

We believe the US Government can and must act to ensure that all potential users have free and timely access on the Internet to peer-reviewed federal research findings. This will not only benefit the higher education community, but will ultimately magnify the public benefits of research and education by promoting progress, enhancing economic growth, and improving the public welfare.

We support the re-introduction and passage of the Federal Research Public Access Act, which calls for open public access to federally funded research findings within six months of publication in a peer-reviewed journal.

The petition follows a similar effort in the European Union, the Petition for Guaranteed Public Access to Publicly-Funded Research Results, which was signed by over 23,000 individuals and organizations.

UK Council of Research Repositories Established

SHERPA Plus has announced the launch of the UK Council of Research Repositories.

It is described as follows: "UKCoRR will be an independent professional body to allow repository managers to share experiences and discuss issues of common concern. It will give repository managers a group voice in national discussions and policy development independent of projects or temporary initiatives."

PRESERV Project Report on Digital Preservation in Institutional Repositories

The JISC PRESERV (Preservation Eprint Services) project has issued a report titled Laying the Foundations for Repository Preservation Services: Final Report from the PRESERV Project.

Here’s an excerpt from the Executive Summary:

The PRESERV project (2005-2007) investigated long-term preservation for institutional repositories (IRs), by identifying preservation services in conjunction with specialists, such as national libraries and archives, and building support for services into popular repository software, in this case EPrints. . . .

PRESERV was able to work with The National Archives, which has produced PRONOMDROID, the pre-eminent tool for file format identification. Instead of linking PRONOM to individual repositories, we linked it to the widely used Registry of Open Access Repositories (ROAR), through an OAI harvesting service. As a result format profiles can be found for over 200 repositories listed in ROAR, what we call the PRONOM-ROAR service. . . .

The lubricant to ease the movement of data between the components of the services model is metadata, notably preservation metadata, which informs, describes and records a range of activities concerned with preserving specific digital objects. PRESERV identified a rich set of preservation metadata, based on the current standard in this area, PREMIS, and where this metadata could be generated in our model. . . .

The most important changes to EPrints software as a result of the project were the addition of a history module to record changes to an object and actions performed on an object, and application programs to package and disseminate data for delivery to an external service using either the Metadata Encoding and Transmission Standard (METS) or the MPEG-21 Part 2: Digital Item Declaration Language (DIDL). One change to the EPrints deposit interface is the option for authors to select a licence indicating rights for allowable use by service providers or users, and others. . . .

PRESERV has identified a powerful and flexible framework in which a wide range of preservation services from many providers can potentially be intermediated to many repositories by other types of repository services. It is proposed to develop and test this framework in the next phase of the project.

UK EThOSnet ETD Project Funded

A UK-wide ETD project called EThOSnet has been funded for a two-year period by JISC and CURL (Consortium of Research Libraries). When the project concludes, the British Library will establish the EThOS service based on the work done by EThOSnet.

An excerpt from the press release is below:

The project builds on earlier exploratory work, also funded by JISC and CURL, which between 2004 and 2006 developed a prototype for the service. Independent evaluation has since given the prototype strong backing and suggested further developments, while a recent consultation resulted in expressions of interest from over 70 HE institutions to participate in the emerging e-theses service.

EThOSnet builds on these firm foundations and through collaboration with the British Library and the HE community will transform access to theses in the UK by providing the full text of theses through a single point of entry. In addition, in tandem with the emerging network of institutional repositories in the UK, it promises to become a central element of the national infrastructure for research.

Fez 1.3 Released

Christiaan Kortekaas has announced on the fedora-commons-users list that Fez 1.3 is now available from SourceForge.

Here’s a summary of key changes from his message:

  • Primary XSDs for objects based on MODS instead of DC (can still handle your existing DC objects though)
  • Download statistics using apache logs and GeoIP
  • Object history logging (premis events)
  • Shibboleth support
  • Fulltext indexing (pdf only)
  • Import and Export of workflows and XSDs
  • Sanity checking to help make sure required external dependencies are working
  • OAI provider that respects FezACML authorisation rules

For further information on Fez, see the prior post "Fez+Fedora Repository Software Gains Traction in US."