Sunday, 4 May 2014

Interview with Kathleen Shearer, Executive Director of the Confederation of Open Access Repositories

In October 1999 a group of people met in New Mexico to discuss ways in which the growing number of “eprint archives” could co-operate.
 
Kathleen Shearer
Dubbed the Santa Fe Convention, the meeting was a response to a new trend: researchers had begun to create subject-based electronic archives so that they could share their research papers with one another over the Internet. Early examples were arXiv, CogPrints and RePEc.

The thinking behind the meeting was that if these distributed archives were made interoperable they would not only be more useful to the communities that created them, but they could “contribute to the creation of a more effective scholarly communication mechanism.”

With this end in mind it was decided to launch the Open Archives Initiative (OAI) and to develop a new machine-based protocol for sharing metadata. This would enable third party providers to harvest the metadata in scholarly archives and build new services on top of them. Critically, by aggregating the metadata these services would be able to provide a single search interface to enable scholars interrogate the complete universe of eprint archives as if a single archive. Thus was born the Open Archives Initiative Protocol for Metadata Harvesting (OAI-PMH). An early example of a metadata harvester was OAIster.

Explaining the logic of what they were doing in D-Lib Magazine in 2000, Santa Fe meeting organisers Herbert Van de Sompel and Carl Lagoze wrote, “The reason for launching the Open Archives initiative is the belief that interoperability among archives is key to increasing their impact and establishing them as viable alternatives to the existing scholarly communication model.”

As an example of the kind of alternative model they had in mind Van de Sompel and Lagoze cited a recent proposal that had been made by three Caltech researchers.

Today eprint archives are more commonly known as open access repositories, and while OAI-PMH remains the standard for exposing repository metadata, the nature, scope and function of scholarly archives has broadened somewhat. As well as subject repositories like arXiv and PubMed Central, for instance, there are now thousands of institutional repositories. Importantly, these repositories have become the primary mechanism for providing green open access — i.e. making publicly-funded research papers freely available on the Internet. Currently OpenDOAR lists over 3,600 OA repositories.

Work in progress


Fifteen years later, however, the task embarked upon at Santa Fe still remains a work in progress. Not only has it proved hugely difficult to persuade many researchers to make use of repositories, but the full potential of networking them has yet to be realised, not least because many repositories do not attach complete and consistent metadata to the items posted in them, or they only provide the metadata for a document, not the document itself. As a consequence, locating and accessing content in OA repositories remains a hit and miss affair, and while many researchers now turn to Google and Google Scholar when looking for research papers, Google Scholar has not been as receptive to indexing repository collections as OA advocates had hoped.

For scholars, the difficulties associated with accessing papers in repositories is a continuing source of frustration. Meanwhile, critics of green OA argue that the severe shortage of content in them means that any hope of building an effective network of OA repositories is a lost cause anyway.

For their part, conscious that green OA poses a potential threat to their profits, publishers have responded to the growing calls for open access by offering pay-to-publish gold OA journals as an alternative.


It was against this background that in 2012 the Finch Committee concluded that in order for the UK to make an effective transition to OA “a clear policy direction should be set towards support for publication in open access or hybrid journals, funded by APCs, as the main vehicle for the publication of research.”

Explaining the decision to prioritise gold OA, Finch argued that repositories had failed to deliver on their promise. “Despite the best efforts of repository managers and librarians … rates of deposit and usage of published materials remain fairly low; and a number of issues will need to be addressed if institutional repositories are to fulfil a bigger and more effective role in the research communications landscape.”

For that reason, Finch added, repositories should in future be viewed as being merely “complementary to formal publishing, particularly in providing access to research data and to grey literature, and in digital preservation”

The Finch Report proved highly controversial, particularly when Research Councils UK (RCUK) responded by introducing a new gold-preferred OA Policy conforming to its recommendations. Many OA advocates in particular felt betrayed.

But we need to ask: did Finch have a point?

We should not doubt that huge challenges remain in getting content into repositories. However, the whys and wherefores of this have been well rehearsed elsewhere, so we won’t dwell on them here.

Instead, let’s consider the current state of the repository infrastructure, particularly with regard to interoperability and discoverability. Why, for instance, do many repositories not expose adequate metadata?  Why do they sometimes provide just the metadata and not the full text? When will the sophisticated search functionality that researchers need become standard in repositories? Will it? And what new developments might help here? More generally, what does the future hold for the OA repository?

Investing for the long term


Who better to put these questions to than Kathleen Shearer, Executive Director of the Confederation of Open Access Repositories (COAR)? Launched in October 2009, COAR’s mission is to “enhance the visibility and application of research outputs through a global network of open access digital repositories” and its membership currently includes over 100 institutions from around the world.

Reading Shearer’s replies below one has to conclude that there is much still to be done. Scholars and scientists will therefore clearly need to be patient. And while new repositories are constantly being created, and existing ones improved (as are cross-repository search services like BASE), the truth is that if the vision articulated in New Mexico fifteen years ago is to be fully realised the research community is going to have to invest a great deal more time, effort and money to developing its repositories.

But should it? Now that most if not all scholarly publishers offer gold OA is further investment in repositories justified?

Shearer believes it is — for two reasons. First, she says, wide-scale take up of green OA would contain publishers’ prices; second, the time has in any case come for the research community to take back control of the scholarly communication system, and repositories will be vital in doing that.

As Shearer puts it, “[T]he Green Road is key. We must collectively build and maintain a global system of repositories. It introduces competition into the system and will act as an important deterrent to arbitrary price increases by publishers.”

She adds, “It will also demonstrate the important role that institutions play in the stewardship of research outputs. To that end, institutions should devote more resources to their repository operations in order to improve repository services and increase the size of their collections.”

As I read it, the promise is that any investment made in OA repositories today will more than pay for itself in the long term.

The interview begins


RP:  Can you say who you are, where you are based and what role you play within COAR?

KS: I am the Executive Director of COAR and I am based in Montreal, Canada, although the COAR office is located in Göttingen, Germany. I have been working in the area of open access and digital repositories for about a dozen years now, mainly in the Canadian context as a consultant and a research associate with the Canadian Association of Research Libraries. In June 2013, I became the Executive Director of COAR.

RP: Briefly, what is COAR, how is it funded, and what is its purpose?

KS: COAR, the Confederation of Open Access Repositories, is an association of repository initiatives with an international membership.

We have over 100 members in 35 countries around the world. Our members come from a variety of communities including universities/libraries, research institutions, funding agencies, intergovernmental organizations and government departments — any organization that may have an interest in repository development and wants to be connected with the international community.

COAR’s mission is to raise the visibility of research outputs through a global network of repositories. We are active on two levels: (1) At the practical level, we support communities of practice around areas of importance for our members mainly in terms of best practices, interoperability and monitoring trends in the repository landscape and (2) At the strategic level, we aim to facilitate greater alignment of regional and national repository networks around the globe.

COAR is funded mainly through membership fees, although we receive in-kind support for our office space from the University of Göttingen and some partnership funding as well.

We are quite a light-weight organization with about 1.5 full time positions in total and an Executive Board chaired by Norbert Lossau, Vice-President of the University of Göttingen. Most of our activities are undertaken by the active participation of our members. 

RP: The mission of COAR, you said, is to “raise the visibility of research outputs through a global network of repositories”. I think it might help if we tried to clarify what this means in practice. In other words, what do we mean by repository here, and what role exactly do we expect that repository to play? Are we talking about a global network of institutional repositories, or does repository here encompass more than that (i.e. central subject-based repositories like PubMed Central and arXiv too, and perhaps other content management systems and databases?)

Likewise, should we assume the role of the repository remains as it was originally conceived — a tool to support green OA by providing a place where papers published in subscription journals can be self-archived in order to ensure that free copies are always available outside the subscription paywall?

Or do we assume that the repository can now also act as a publishing platform on which institutions can publish their own journals — as currently planned, for instance, by University College London?

Alternatively, perhaps the assumption is that today the repository should be viewed as little more than what the Finch Report assumed it to be: something “complementary to formal publishing, particularly in providing access to research data and to grey literature, and in digital preservation” (A model that assumes open access is provided by means of gold rather than green OA)?

KS: Repositories are evolving and play a number of roles. At their core, a ‘repository’ could be theoretically defined as a set of services that provide open access to research outputs (along the lines of Cliff Lynch’s original definition in 2003). However, in practice, repository services and infrastructures are diverse and there is a lot of overlap with other systems. Perhaps most significantly, practices and technologies are changing quickly, making it a challenge to concretely define their services. My feeling is that we need to be flexible in the way we conceptualize repositories.

In terms of COAR, we are a community brought together by a set of shared principles and common practices rather than by a narrowly delineated concept of repository. So yes, we would include disciplinary repositories and content management systems (if they provide open access to full text) in our global network.

In terms of a complement to formal publishing, I expect that traditional publishing will soon be going through some pretty big transitions, likely some very disruptive changes. I agree with Dominique Babini, Jean-Claude Guédon and others that we should aim for a basic, open, and interoperable system that is free to both access and contribute to. Value-added services by publishers and others can be built on top of this content.

One way of thinking about repositories is that they represent an institutional commitment to the stewardship of research outputs. In this sense, they address two important problems in the current system: sustainability and stewardship.

I believe institutions should assume greater responsibility for managing, providing access and preserving the content created through research. It will alleviate some of the inflationary aspects of scholarly publishing and enable us to have more influence on future directions. This was the traditional mission of libraries in the print world, which has been somewhat lost in the transition to digital content. How this plays out in terms of models will likely vary according to content type, discipline, and region.

Interoperability


RP: I would like to focus on the issue of interoperability. I am aware of a number of current initiatives devoted to getting institutional repositories to interact/interoperate, including DRIVER, DRIVER II, euroCRIS, OpenAIRE and no doubt there are others too. How do these various initiatives fit together (do they?), and why are there so many initiatives that — to the layperson at least — might seem to be duplicating effort?

KS: There are several initiatives that have evolved from different requirements, regions, and with differing aims.

DRIVER and DRIVER II were European Commission-funded projects to support the implementation of repositories in EU countries. The aim was to have repositories adopt common guidelines for organizing their content so they could be harvested and searched through the DRIVER search service. 

OpenAIRE has built upon work of DRIVER to implement further standards that enable the European Commission to track the open access research output they fund. Each of these three projects required some level of interoperability between participating repositories.

There are similar initiatives in other regions, such as La Referencia in Latin America and SHARE in the US that will also require some level of interoperability across those repository networks.

COAR is a forum whereby all of these regional initiatives can work together to identify issues in common and, where appropriate, agree on standardized practices. COAR will be intensifying efforts in this area and has just launched an initiative to address some of the differences between repository networks that are evolving.

EuroCRIS is a European association that is looking at interoperability between research administrative systems. The objective of these systems is to manage and report on research activities. Unlike repositories, CRIS systems do not usually manage full text content.

We have seen in the last few years some merging between CRIS systems and repositories, with some repositories being integrated with CRIS's, or at least interoperability between repositories and CRIS. 

COAR has also been working with EuroCRIS to identify strategies for greater interoperability between research administration systems and repositories.

RP: The concept of networking repositories dates back at least to 1999, and the Santa Fe Convention. I believe it was in the wake of the Santa Fe meeting that the OAI-PMHprotocol was developed. However, I assume that both the thinking and the technology have developed somewhat since then.

As I understand it, for instance, OAI-PMH was based on the principle that services would be developed to harvest metadata from repositories in order to aggregate their holdings and provide a centralised discovery service. I guess this assumed that records in repositories would consist of metadata but not the full text (so the goal presumably was to signal where papers were held, not to provide direct access to them).

I would think that the emphasis today is more on providing direct access to full-text documents not just their metadata. Briefly, therefore, can you say how thinking has developed since 1999, and how the technologies and protocols have changed to reflect this?

KS: OAI-PMH was developed on the principle that a service would harvest the metadata record that would then point the user back to the full text content in the repository. So in that sense it does facilitate access to the full text, but without having to aggregate the content into a central archive.

OAI-PMH is still the common denominator for metadata exposure in repositories and it remains standard practice for cross-repository search services to harvest metadata and then point the user back to the repository to access the full text. Full text harvesting is much more demanding, requiring large storage space to house the content in a central location and there are other technical challenges attached to full text harvesting.

The disadvantage of metadata harvesting is that the search services are based on the metadata supplied by the repositories, which isn't always comprehensive, complete or consistent. COAR aims to improve the current situation by identifying and encouraging the adoption of common standards and metadata globally. However, for better discoverability, and especially for other services such as text mining, using full text search is highly desirable.

In terms of discovery, repository managers have found that most users find the content in repositories through search engines such as Google and Google Scholar, not from metadata harvesting services or by directly searching the repository. Therefore, the repository community has put significant efforts into exposing their content to commercial search engines through various optimization techniques. 

Beyond discoverability, there are other areas of repository networking and interoperability, like content transfer, usage data, etc. where new technologies and standards/protocols have been created. COAR is a forum whereby interoperable practices can be agreed upon globally.

Full text


RP: You say that it remains standard practice for cross-repository search services to harvest metadata and then point back to the full text in the repository, and you said that COAR assumes OA repositories will “provide open access to full text”. This would seem to imply that an OA repository always now includes the full-text as well as the metadata (and indeed most people would presumably expect that of an OA repository).

However, not all records in OA repositories do provide access to the full-text, and many seem to offer little more than the bibliographic details. Even a poster child of the OA movement — Harvard’s DASH repository — has been criticised for not providing the full text (e.g. here). These criticisms were made a few years ago, but DASH does still today contain records without any full-text attached. Moreover, some do not even provide a link to the full-text (and DASH does not seem to have a RequestCopy Button). When I looked in DASH the other day, for instance, I found (at random) five examples of this (one, two, three, four, five).

I think this cannot be a consequence of publisher embargoes since the articles concerned date back as far as 1993, with the two most recent published five years ago (and in any case the Harvard OA Policies claim to moot publisher embargoes). Moreover, where in a couple of cases the DASH records do pointto the full-text this is a link to the publisher’s version, where the user is asked to pay for access ($35 in one case). This cannot be described as OA.

You may not want to comment specifically on DASH, but do you think it problematic when records in OA repositories do not always provide access to the full-text, and maybe don’t even link to a free copy of it? If so, what can/is COAR do/doing to address the situation, in concrete terms?

KS: Ideally, all records in the repository will have the full text attached. However, as you point out, this isn’t always the case. I’m not sure about the specific case of DASH, but this really speaks to the collection policy of the individual repository.

As I said earlier, more and more repositories are now being used to track research output. In that case the objective may be to collect information about all of the publications at the institution, regardless of whether they are open access or not. Still other repositories may be inputting metadata records without the full text as a strategy to encourage authors to upload their documents.

If we look at the OpenAIRE portal as an example, they are currently harvesting 8.4 million records from over 400 sources (mostly repositories, but also open access journal articles). Over 8.2 million of those records are open access. So, I believe that the vast majority of content in repositories is open access, with a small percentage of metadata-only records. The portion of open access, of course, will vary depending on the repository.

In my opinion, the most effective way to improve the proportion of full text in repositories is to continue to advocate for open access policies at funding agencies and institutions around the world. These are the levers that will have a real influence on the policies and practices of the individual repositories. More staffing and resources directed towards repository operations would also help.

RP: You said that rather than searching directly in repositories, or exploiting metadata harvesting services (like OAIster perhaps?), researchers tend to rely on search services like Google and Google Scholar for the discovery of scholarly content in repositories.

Does this mean that the repository community tends today to assume that the research community should rely on mainstream search services, rather than trying to build sophisticated repository search services itself?

If so, I am conscious that OA advocates frequently complain that Google is not supportive enough of their needs, and not as keen to index repository collections as they would like. Would you agree? What is the current situation with regard to mainstream search services like Google, Bing and Yahoo in terms of indexing repositories, and what future developments do you envisage that might improve the situation so far as searching repositories is concerned?

KS: It’s not really about what the repository community believes is the best solution, but rather a practical response to user behaviour.

It would be erroneous to assume all information seekers are the same. However, we do know that even for well-developed disciplinary services, such as PubMed Central and Medline, the majority of users access articles directly from commercial search engines like Google and Google Scholar.

According to my COAR colleague Eloy Rodrigues, Director of the University of Minho Documentation Services, most well developed institutional repositories have about 3/4 of their traffic coming from Google and other generic search engines. Repository managers take that as very positive sign of the visibility and accessibility of the content in the repository. 

In terms of mainstream search engines and Google Scholar there has been ongoing discussion about their efficacy in retrieving scholarly content. It really depends on if you are looking for something you know exists (i.e. you search the title or author’s name) or you are searching using key words.

As reported in an article published in the Online Journal of Public Health Information (Giustini and Boulos, 2013), “Google Scholar’s constantly-changing content, algorithms and database structure make it a poor choice for systematic reviews.”

If you are looking for a specific document in a repository and you know the title, the search engine will likely point to it. However, searching by key words, content in repositories are not always high in the rankings.

The problem of visibility is likely even more acute for repositories with non-English content as there does seem to be a bias towards English language content in these search engines. 

This will remain an ongoing challenge for repositories as technology continues to change rapidly.

Inherent tension


RP: Certainly there seems to be some disappointment amongst researchers that 15 years after the Santa Fe meeting they still find it extremely difficult, if not impossible, to search effectively in and across OA repositories. I saw this view expressed most recently by Cambridge University chemist Peter Murray-Rust who tweeted, “IF libraries provide modern search I'd change my mind; but articles in repos are difficult to discover”. His conversation can be viewed here.

Does Murray-Rust have a point? What can you say to convince him that his needs will be met soon? Can you? If so, how will they be met?

KS: There is an inherent tension that exists in the repository community. On the one hand, we aim to make the deposit process as easy as possible so that creators will contribute (or repository staff costs are manageable); on the other hand, we want to assign good quality metadata (which takes time and effort) because we know it will enable greater interoperability and improve discoverability of content. So far, the former has been a greater priority.

There is some truth to Peter Murray-Rust’s comments in that complex search services, such as those developed for some discipline-based repositories, require quite a high level of curation, especially for non-textual material. Datasets, for example, need to be accompanied by fairly comprehensive metadata describing them and those metadata elements need to be standardized across each item.

It is a far greater challenge to develop complex searching across numerous repositories containing different disciplines, languages and formats. To facilitate advanced searching in this context, there needs to be interoperability across repositories. COAR has been working on this and this is one of our top priorities; but it takes time to realize this across a very diverse repository landscape.

That being said, there are already a number of cross-repository search services, for example BASE, CORE, and OpenAIRE, which are working to improve the retrieval of content in repositories. They have advanced search options that allow you, for example, to limit your search to publication type, geographic location, publication year and so on. You can’t do all of these things in Google Scholar.

OpenAIRE enables users to identify publications related to the projects for which they are funded. These services (and others) will continue to develop and will incorporate more sophisticated tools to improve discovery in the future.

Personally, I can envision a time not too far in the future when more complex search services are built on top of repository networks. What individual repositories should focus on, in my opinion, is ensuring that their content is open, can be indexed, and is attached with the necessary metadata in order to facilitate the development of these services.

RP: From what you have said would it be accurate for me to conclude the following: Users tend to prefer using commercial search engines and Google Scholar for discovering research papers in repositories. However, this is not always the best approach.

We don’t yet know exactly what the role of the OA repository will be, nor what form it might eventually take (indeed, repositories will likely take a number of different forms, and play a variety of different roles).

For these reasons it is important that repository managers ensure their content is open, that it has appropriate metadata attached, and that it can be indexed. Doing this will provide sufficient flexibility for future developments.

Finally, we are still some years out from the point where researchers with sophisticated search needs can expect the level of discoverability that they want/need?

Have I understood correctly?

KS: Yes, you are for the most part correct in summarizing my opinion.

A couple of small clarifications: We know from repository managers that the majority of users are coming to repositories from commercial search engines and not through harvesting services or the search facility built into the repository; and we know from user studies that the starting point to find information for many researchers is through Google or Google Scholar.

Currently, as things stand, the content in repositories is not highly ranked in Google Scholar, and in terms of Google, repositories are indexed alongside billions of other pages. So, no, this is not ideal for the discoverability of repository content, particularly for key word or topic-based searching.

I note that in the early days of Google Scholar, the open access community advocated for the search results to be tagged as open access (or not). Obviously we were not successful, but this would have enabled users to limit results to open access content and certainly been a boost for the visibility of repository content in this context.

I do believe the discoverability of repository content will improve greatly in the coming years. Refining the cross-repository search services, those that are based on harvested metadata, will depend on improving the standardization and comprehensiveness of metadata records. Technology will help with this. There are new, automated methods for assigning metadata and repository software platforms can build-in standard vocabularies and metadata elements.

The greater challenge is coming to an agreement about common terminologies and approaches across the entire repository community. COAR will play an important role by acting as a forum whereby the repository community can make these kind of collective decisions. 

There will also likely be a number of services developed in the coming years to facilitate full-text searching through harvesting the content. According to Petr Knoth (Knowledge Media Institute, The Open University, UK) who has been doing research in this area through the CORE initiative referenced earlier, there still are a number of technical and legal barriers to full text harvesting from repositories.

However, in the coming years, I expect that the repository community will begin to address these barriers, especially the technical ones.

Again, I hope that COAR can play a role in developing solutions and disseminating best practices.

SHARE or CHORUS?


RP: You said (or at least implied) that repositories should be viewed as tools to enable the research community to “assume greater responsibility for managing, providing access and preserving the content created through research”. And you cited SHARE as an example of an initiative focussed on providing interoperability between repositories.

It is worth noting that SHARE is a response by librarians to the OSTP Memorandum, which directs US Federal agencies to develop plans to ensure that the published results of research they have funded is made OA. As such, SHARE could be viewed as a good example of how research institutions can try to take greater responsibility for scholarly communication, since it would put librarians in charge of managing access to papers released as a result of the OSTP Memorandum.

However, you will know that publishers have proposed an alternative model based on CHORUS. The aim of CHORUS is to ensure that it is publishers rather than librarians who manage access to these papers, and it demonstrates their wish to remain firmly in control of scholarly communication, even after research papers have been made OA.

How would you respond to someone who argued the following: Since the research community is finding it difficult to fill repositories (a point frequently made, not least by the Finch Report), and both difficult and time-consuming to create the necessary infrastructure to ensure repository content is optimally discoverable, might it not make more sense to outsource the task to publishers via initiatives like CHORUS? After all, CHORUS will deliver OA, and since publishers have greater resources they might be expected to undertake the task more effectively, and more quickly. Moreover, since it is they who publish the papers in the first place, they already have all the content in place.

KS: My major concern about CHORUS is that the publishing community would have too much control of the scholarly communication system. A number of large publishers have already demonstrated that they don’t support the principle of open access (remember PRISM).

Frankly, the interests of publishers often lie elsewhere and they may be motivated by things such as profit margin not the public good.

On the other hand, at the core of the mission of the university and the library is the advancement and dissemination of knowledge. It seems to me that the world’s collective knowledge created through research should rest in the hands of long-term actors whose raison d’etreis to ensure that it is preserved and remains accessible to all.

CHORUS may seem like an appealing option for the US agencies at the moment, but the long-term implications are that the research community will have little control or ability to influence the future directions of scholarly communication if we take that route.

I’m also very concerned about the costs of such a system. Article processing fees are already way too high for many researchers, especially in developing countries. The recent study of APCs undertaken by the Wellcome Trust and others found that the average per article APC is $1,418 USD for open access publishers. I don’t believe this can scale globally and will ultimately result in disadvantaging a large number of researchers who can’t afford to pay.

RP: You are right that speed and effectiveness is one thing, cost and ownership something else. And as you suggested earlier, if the research community were to take greater responsibility for managing access to research it could hope to “alleviate some of the inflationary aspects of scholarly publishing and enable us to have more influence on future directions.”

This reminds me of what your colleague Eloy Rodrigues said to me last year. The future of scholarly communication, and its cost to the research community, he suggested, will depend on whether there is a “research-driven”’ transition to open access or a “publishing-driven” transition (in order words, whether the transition prioritises the needs of the research community or the needs of publishers). I would think that the competing SHARE and CHORUS initiatives are representative of these two approaches, and this suggests to me that in the coming years we will see publishers and librarians jostling for control of the scholarly communication system. And if that is right, the institutional repository will surely become a key battleground in the struggle.

Would you agree? And if it wants to ensure a “research-driven” transition to OA what should the wider research community be doing in your view?

KS: The choices that institutions make now about how they are going to invest in scholarly communications are absolutely critical.

First of all, I think the Green Road is key. We must collectively build and maintain a global system of repositories. It introduces competition into the system and will act as an important deterrent to arbitrary price increases by publishers.

It will also demonstrate the important role that institutions play in the stewardship of research outputs. To that end, institutions should devote more resources to their repository operations in order to improve repository services and increase the size of their collections.

Secondly, we should encourage and sponsor the development of new publishing models and value-added services that conform to our vision.

In terms of repositories, this would include better cross-repository discovery services, text mining capabilities, disciplinary views, and the development of overlay journals. Leslie Chan, for example, makes the case that the distinctions between “journal” and “repository” are increasingly blurred and that “mega-journals” are essentially repositories with overlay services.

We should be participating in projects that demonstrate the added value of repositories and repository networks across the research life cycle. Of course, this will require that we take some risks, which is a difficult case to make in hard economic times to (often) risk adverse organizations.

Global discussion


RP: You said that the way in which scholarly communication develops will vary “according to content type, discipline, and region.” Certainly, as OA develops we do appear to be seeing distinctive regional differences emerging. For instance, where the pay-to-publish gold OA model is being pushed heavily by the UK and The Netherlands there is still more of a focus on green OA in North America. Meanwhile, in Africa and Latin America a repository-based publishing model currently appears to dominate.

As things stand I would expect to see the Global North increasingly move to a pay-to-publish gold OA model and the Global South to a free-to-publish/free-to-read repository-based publishing model similar to that pioneered by SciELO and AJOL. If that proves the case, however, will it be the best outcome in a global research environment?

When I spoke to Dominique Babini last year she said “[W]e owe ourselves a global discussion about the future of scholarly communication”. And she added, “Now that OA is here to stay we really need to sit down and think carefully about what kind of international system we want to create for communicating research, and what kind of evaluation systems we need, and we need to establish how we are going to share the costs of building these systems.”

This would seem to imply a more global approach than we are currently seeing develop. Would you agree with Babini? If so, who should organise the global discussion she has called for, and who should take part in it?

KS: Yes, I agree, and I would add that we should consider carefully the unintended consequences of adopting the various models.

“What kind of system do we want to create for communicating researcher?” I would propose that we want one in which all researchers can access and contribute to, regardless of geographic location or discipline; and where the knowledge created is assessed on its real value, rather than on the region from which it emerges or the so called “impact” of the journal in which it is being published.

A dual system as you describe above is not ideal and I believe it will create inherent inequalities across the regions. Especially if we continue to rely on impact measures that do not reflect the quality of the research, but rather serve to prop up the traditional publishing system.

I believe there is a general lack of awareness in the “north” about the “southern” perspective and that we do need to ensure that the voices from the south are heard.

In terms of the global discussion, we already have a number of international forums for exchange: the funding agencies have the Global Research Council; libraries have organizations such as the SPARCs and IFLA; the repository community has COAR; and, publishers have their own venues.

UNESCO, and the governments represented there, has also become interested in open access. We could begin the global discussion by facilitating greater dialogue across these different stakeholder organizations.

One missing but very important link is the research community. It’s clear that many researchers have not been sufficiently engaged with the issues of open access to understand the nuances. For example many researchers still equate open access with open access journals. So we need a mechanism for bringing those communities into the discussion as well.

It is illuminating to note that a parallel global discussion is currently occurring in the area of research data through the Research Data Alliance (RDA). It has been comparatively easy in the context of research data to bring together the key stakeholders — researchers, data repositories, institutions, and funding agencies — to adopt a common vision and agree on practical strategies for moving forward.

Why haven’t we been able to do that for publications? The essential difference is that for publications, there are some parties that have a significant financial interest in maintaining control of the system. This makes the global discussion far more challenging.

RP:  Thank you very much for taking the time to speak with me.



Saturday, 5 April 2014

Interview with Jean-Gabriel Bankier, President & CEO of bepress

Founded in 1999 by three Berkeley professors, bepress (formerly Berkeley Electronic Press) spent the first decade of its existence building up a portfolio of peer-reviewed journals — much like any scholarly publisher. In 2011, however, it  took what might seem like a surprising decision: it decided to sell all its journals to De Gruyter and reinvent itself as a technology company.
Jean-Gabriel Bankier

Instead of publishing journals, bepress is now focussed on developing and licensing the publishing technology it created for its earlier publishing activities, and its flagship product is a cloud-based institutional repository/publishing platform called Digital Commons.

Digital Commons is currently licensed to more than 320 academic institutions, who use the software to publish over 700 journals, 94% of which are open access. This publishing activity is invariably managed by the institution’s library, and often includes the publishing of books, conference proceedings, data sets, audio-visual collections, and other digital content types too.

Is this a sign of things to come: Publishers becoming technology companies and librarians becoming publishers? President and CEO of bepress Jean-Gabriel Bankier believes it is. As he puts it in the Q&A below, “Library-led publishing is an integral strategy in the university taking back ownership of scholarly communication.” As such, he adds, the future of scholarly publishing now “lies in the hands of libraries and scholars.”

To support his argument Bankier cites a US study in which 55% of the universities and colleges surveyed said that they are offering or considering offering library publishing services.

Moreover, bepress is not the only game in town for libraries looking for a publishing platform. In 2001 the Public Knowledge Project released the first version of the open-source publishing software Open Journals Systems (OJS), and today OJS estimates that over 6,000 journals are being published using its software. Many of these journals are undoubtedly being published (or soon will be published) by university libraries — e.g. the library at University College London and Stellenbosch University library

We could also note that in 2012 US-based Amherst College announced that it was launching its own press. This will publish peer-reviewed books in the liberal arts, and will be managed by the library. 

What all this means, says Bankier, is that if publishers “want to continue to play a significant role in supporting the changing needs of the research community” they will need to consider following the example of bepress, and morph from content provider to technology company.

Doubtless other publishers would challenge this assertion. But whatever the future holds, I think anyone interested in open access, or scholarly communication more generally, will find what Bankier has to say below of great interest.

The Q&A begins


RP: As I understand it, bepress was founded as a scholarly publisher in 1999. Can you say briefly who founded it and what the initial goal was? Is it for profit or non-profit?

J-G B: In 1999, UC Berkeley professors Robert Cooter, Aaron Edlin, and Ben Hermalin banded together to launch Berkeley Electronic Press, now simply called bepress.

The heart of bepress has always been about listening to faculty and responding with simple technology-based solutions that support scholars in the rapidly changing world of scholarly communications.

Initially, for us, that meant exploring alternatives to commercial scholarly journal publishing which were plagued by slow turnaround times, limited access, and unreasonable prices. Later, that meant providing authors and universities themselves with the means to publish their research openly and widely. We are a for-profit company.

RP: You say bepress is a for-profit company. I assume the shareholders are the three founders? Can you tell me what the company’s revenues and profits were for the last financial year?

J-G B: Yes, the founders are shareholders. The company was born with just a little seed money from the founders, parents, and incredibly supportive friends and neighbours, most of whom continue to own part of the company. Bepress has never had venture capital or private equity. Berkeley isn't far from Silicon Valley, but we weren't that kind of start-up. Our first office, after we moved out of one of the founder's kitchen, had no natural lighting and ceilings so low that it necessitated skidding around mismatching, three-wheeled chairs to avoid banging one's head on the ceiling.

I'm happy to report that around 15 years later, we've now got offices with actual windows and chairs that don't wobble. Our business doesn't wobble any more either. Our revenues are around $10 million a year with an unbelievably low cancelation rate for subscribers (below 1% in 2013). We are very stable and run at a modest profit. It is a great feeling to finally be able to send small dividend checks to those friends and family members who put their faith in us back in the beginning.

RP: How many journals did bepress come to publish, in what fields, and were they subscription journals or OA journals?

J-G B: We started with a few titles in economics, and eventually grew to publish roughly 65 journals altogether. In addition to economics, our emphasis was in education, health and medicine, law, policy, science and technology, and statistics.

We simultaneously sold subscriptions to libraries and made the articles freely available online. Many people suggested that we were crazy to use such a model, but it was really the best of all worlds. The authors and editors got to maximize readership and share their ideas, readers got access to high quality scholarship, and libraries got a sustainable journal subscription model. We raised prices on our journals by a grand total of only three percent over 10 years.

RP: In 2011 bepress sold all its journals to the German publisher De Gruyter in order to focus on developing software tools to assist the research community publish its own journals. Why?

J-G B: The library community had little appetite for new subscription journals, but a big appetite to support publishing on their campuses. According to one study produced by the libraries of Purdue University, Georgia Institute of Technology, and University of Utah in 2012, “55% of universities and colleges are offering or considering library publishing services.”

We saw this change in appetite first-hand. In the one-year period predating the sale, we launched only four new journals under the bepress moniker. Over the same period, we helped libraries launch over 100 new open-access journal titles using our software. Starting new commercial journals had become a challenge. By our count Springer, Taylor & Francis, and Sage and Wiley together introduced fewer new journal titles to market in 2012 than the publishing programs of the Digital Commons’ community of libraries. 

Given this change, it became clear that the best way for us to advance scholarly communications was to focus our energy on supporting library-led publishing, so we exited the commercial subscription-based journal business.

It is worth noting that some of the editors who started their journals in Digital Commons had originally approached commercial publishers. One acquisitions editor from a major publisher looked at the list of journals that publish on Digital Commons and wistfully listed those that had approached her but were impossible to take on because of the economics of subscription sales.

Behind the scenes something very interesting was going on in scholarly communications. Libraries had made it more risky for publishers to start new journals. In turn, some publishers tried passing that risk onto editors, societies, and universities in the form of minimum number of guaranteed subscriptions or minimum annual guaranteed payments. Luckily, many of those journals had another option: to come to their libraries for publishing support.

Today Digital Commons is used by over 700 journals. We believe the future of scholarly publishing lies in the hands of libraries and scholars.

RP: I am particularly interested in the products Digital Commons and SelectedWorks. Can you say something about these and give me some indication of pricing?

J-G B: Digital Commons is a hosted platform designed for the full spectrum of scholarly publishing. We provide libraries the tools, training, and expertise to build essential publishing programs on their campuses.

Digital Commons supports journals as well as books, conference proceedings, data sets, audio-visual collections, and all other digital content types that showcase the whole of an institution’s publications and scholarly output. The content is all the institution’s own; we provide the platform, the support and the expertise.   

University of Wollongong Research Online and Purdue University’s ePubs are two great examples of the Digital Commons publishing platform in action.

SelectedWorks is the author equivalent to Digital Commons. It allows faculty to organize and share the breadth of their research, works, and achievements. A SelectedWorks site looks like this. These faculty profiles can be integrated with Digital Commons to allow authors to have their works and profile discoverable alongside the institution's broader scholarly corpus. Integration with the Digital Commons Network takes that connection one step further by tying together the works of authors with the works of others in their discipline.

As for cost, Digital Commons and SelectedWorks together cost a fraction of an FTE; far less than the cost of having local staff to support users and open source publishing systems in a comparable manner.

Inspiring to see


RP: You say that Digital Commons costs a fraction of an FTE. I wonder if you could expand on that. Since it is a hosted service I assume Digital Commons is billed by means of an annual subscription, rather than a one-off payment. Could you give me some specific examples of what different sorts of libraries could expect to pay per annum to use the service?

J-G B: You are correct. Both Digital Commons and SelectedWorks are sold as annual subscriptions.  Size, as determined by FTE, plays an important role in pricing. As you might guess, larger institutions pay more than smaller ones. The average annual subscription fee for Digital Commons and SelectedWorks in 2013 for an academic library was around $20,000 and $8,000 a year respectively. 

RP: As I understand it, Digital Commons is designed to allow universities to publish their own journals on top of their institutional repositories. Presumably it does not matter which repository software they use? Can you say how many universities are currently publishing journals in this way, and what has motivated them to do so? Are they all US-based institutions?

J-G B: Digital Commons is an integrated platform to support the full spectrum of scholarly publishing and institutional repository needs. Imagine ContentDM, DSpace, Dryad, Figshare, ePrints, Open Journal Systems, Open Monograph Press, and Open Conference Systems all rolled into one solution, and you have the Digital Commons platform.

For some institutions, Digital Commons on its own meets all of their publishing and repository needs. Others use Digital Commons in combination with a myriad of local digital preservation and digital asset management solutions. There is, nonetheless, a clear trend toward gradual consolidation as institutions migrate collections stored on DSpace, ContentDM, Omeka, and OJS onto Digital Commons.

There are more than 320 academic institutions that currently use Digital Commons for the full spectrum of scholarly publishing. Three hundred of those are U.S. based institutions and the remaining 20 institutions are in Canada, Australia, Asia, and Europe. These publishing programs are being led by campus libraries that see the opportunity to meet the publishing needs of their faculty and students in a scalable way. It’s been inspiring to see how much faculty and student editors appreciate the library in this new role, and how it has improved their perception of the library’s role on campus.

RP: How does Digital Commons differ from OJS? For instance, Digital Commons is not open source I think. If not, presumably Digital Commons is a more expensive option? And if so, what extra value does bepress offer?

J-G B: Digital Commons is not open source like OJS. I’m a fan of open source software, but I’m an even bigger fan of software in the cloud. Cloud-based ejournal publishing software is the future; it is easier to use, scale, build upon, and support. Digital Commons is delivered via the cloud, while OJS is mostly locally installed.  

As I mentioned earlier, Digital Commons costs a fraction of an FTE, which makes it cheaper than installing and maintaining local software whether open source or otherwise. That is all the more true when one considers the many systems people wish to integrate with that continuously change and therefore require code changes. On top of that, as soon as you take advantage of the open source software by modifying the code, you will typically find yourself stranded in a given version or investing serious money to port your modifications to a newer version.

Let me give you a concrete example: I’m sure you won’t be surprised to hear that readership on tablets and mobile devices is booming. Faculty want the fruits of their labor — their journal — to look great on their iPads and iPhones. In response to this growing demand, last year we rolled out responsive designs for all of the journals, repositories, and other collections on Digital Commons and SelectedWorks.

This means that the web pages respond to the visitor's screen size and change the layout accordingly, so it means the journals look great on an iPhone. Check out this Digital Commons journal example on your mobile phone and then go look at your favourite OJS journal. It was a significant amount of work for us to produce a beautiful mobile interface for journals, but the beauty of a service in the cloud is that once we got it right it was a snap to roll it out for everyone in the community.

But, while we are low cost, we compete mainly by being high value. Moving beyond where the code is stored, Digital Commons also differs from OJS because Digital Commons is a service, not just software. By that I mean it includes people with publishing and technology knowledge who support editors with site design, workflow design, support, training, and publishing expertise. And, for those journals ready for it, Digital Commons includes additional professional publishing services such as abstracting and indexing, DOIs, and preservation in CLOCKSS and Portico.

Finally, we provide service without limits: unlimited storage, unlimited users, unlimited training, unlimited support, unlimited journals, unlimited bandwidth, unlimited upgrades...you get the point.

Our approach frees libraries and journal editors from the technical hassles of running software so they can focus on running a journal. We all know it is a lot of hard work to launch a new journal—e.g., recruiting an editorial board, recruiting authors, and preparing a first issue.

We see our role as helping with everything else. We’ve seen that when given such support that library-publishing programs flourish. Brigham Young University and Kansas State University’s New Prairie Press are both recent examples of converts from OJS to Digital Commons.

Philosophical and practical questions


RP: You said that institutions are migrating to Digital Commons from platforms like DSpace, Omeka and OJS. The latter are open source platforms, whereas Digital Commons is a proprietary platform. I suspect that some OA advocates might argue a) that open source solutions are more appropriate for library-led publishing (since, for instance, open source could be said to better fit with a university’s mission) and, b) proprietary platforms introduce the danger of vendor lock in. I wonder if you could respond to these two points.

J-G B: I know your question reflects the viewpoint of some in the community. I believe it’s something felt stronger on your side of the Atlantic than on my side these days, but I also know that the libraries that use our services sometimes face these types of questions as well.

I think the question of “are you using the ‘right’ kind of software for your open access initiative?” has the unfortunate effect of dividing the community. The bepress community of scholarly communications libraries are no less committed to open access than libraries using open source software, but taken to the extreme, this question devalues the accomplishments of the librarians who build successful OA publishing programs on their campuses using options like Digital Commons.

You were spot on in your Q & A about the state of open access when you identified the lack of cohesion in the OA movement as its biggest failure to date. For those who missed the piece, you wrote: “I wonder if the movement has not been its own worst enemy. After all, it is far from unified, and it has spent a great deal of valuable time arguing with itself rather than winning hearts and minds or taking the practical action needed to achieve its objectives.” I would argue that Digital Commons is one such practical option for libraries to pursue open access library publishing.  

Now I’ll hop down from my soapbox and respond to the two parts of your question. The first part, (a), is philosophical and principled while the second part, (b), is practical. I’m going to tackle the practical argument about vendor lock-in first. I would argue that established open standards are pretty good protections against vendor lock-in. It is really not complicated to migrate from one platform to another platform because all platforms in the space, including Digital Commons, support an open standard, called Open Archives Initiative (OAI). We’ve supported the move of more than 50 collections from different locally-hosted, open source platforms to Digital Commons (and a handful the other way as well).  At the end of the day, the content and metadata are easily portable.

To your first point (a), the connection between open source and open access is strong in the library community, no doubt about it. But here’s the rub: library-led publishing is not a service intended for librarians. It is a service for faculty and students provided to them by their library.

Faculty and students are a demanding bunch; they are increasingly comfortable evaluating all sorts of digital tools, and they expect the best in the tools they choose for themselves. They have a certain goal — to publish a successful journal, for example — and they will be grateful when the library provides them with an excellent publishing experience. If they are poorly served, they will either abandon their library as a publishing partner, or perhaps worse yet, accept the service and complain about it. To meet its mission, the library should ask, “How can I provide the best publishing services to my campus, given the available resources?”

Librarian Karen G. Schneider summed it up better than I ever could when she wrote: “I hate the idea that for some librarians if a particular software is open source, hands down, it’s the right choice. The right choice is the software that meets the mission. While the principles behind open source are admirable, when an open-source product doesn’t meet your library’s needs, your first obligation is to your users.” Karen G. Schneider on July 26, 2006.

RP: Realistically, how simple is it to start publishing a journal using Digital Commons? Is much training required?

J-G B: From our perspective, it is possible to launch a journal and train the editors in a day. If your question is "how easy is it for editors to learn to use the platform?" then the answer is: very easy.

Editors tell us that ours is the easiest to use of all the manuscript submission systems out there. We do all the work of customizing workflows for them. They learn the system in an hour (or an hour and a half if there are lots of editors learning simultaneously), but we train and re-train them as often as necessary.

This model of unlimited support is especially valued by libraries that would like to support student publishing efforts. Students graduate, and the editorship changes hands; it’s a big relief to the library to always have bepress available to train students and answer all of their questions, and to serve as institutional memory when the editorship changes hands.

RP: Are universities (presumably the librarians) able to do all the work themselves, or do they need to buy support services? If so, what kind of support services, and how much are these likely to cost?

J-G B: Librarians certainly could do all of this work themselves. They have to do all the work with OJS already, after all. With Digital Commons, however, they get teams of publishing experts and platform experts at the ready to help them support their faculty and students. There are no extra costs, so they don’t have to second-guess themselves if they ever want more help or services. They can say “yes” to all the requests and questions from their editors and authors. Unlimited support and service is incredibly important to us for this very reason.

Trend is towards openness


RP: I am told that most of the journals using Digital Commons are OA journals. Are you able to give me any stats on this, and current trends?

J-G B: 94% of the journals using Digital Commons are open access, and of the 6% that are subscription-controlled, the majority make their back issues openly available.

The current trend is towards openness as more journal editors realize that the more open they make their journals, the greater their success. It’s been several years since we’ve seen editors of an open access journal using Digital Commons jump ship to sign on with a publisher and move their content behind a pay wall.

In fact, we’ve seen just the opposite — editors of subscription journals making their back issues open access, in turn increasing readership and subscriptions revenues.

We’ve also seen editors of subscription journals with open back archives on Digital Commons take the next step and abandon the subscription model altogether for open access.

We’ve even had a few journals leave publishers for the library in order to make their journals open access. If you look at the Law Review Commons alone, you will find over 170 journals that are open access, almost all of which used to be exclusively subscription access.

RP:  How do universities fund the OA journals they publish using Digital Commons, and can you say something about the average total annual cost of running an OA journal on Digital Commons? Do many of the journals operate article-processing charges?

J-G B: With one exception, none of the journals operate an article-processing charge. Authors who publish in these journals can be sure that their scholarship has been accepted by the journal on its merits alone. The one exception is the “Open Journal of Occupational Therapy” at Western Michigan University. It charges an article processing fee of $200, and $100 for student authors.

Because the Digital Commons publishing model is unlimited, once the university makes the initial investment, they can support as many journals and other publications as they would like.

Most libraries consider supporting a publishing program on campus to fall into their core service offerings, and the fact that publishing capabilities are simply part of our integrated platform makes it easy for them to support the full spectrum of needs, such as conferences, data sets, multimedia publications, textbooks, and so on. There are so many publishing needs on campus when you understand it as a full spectrum, and librarians are eager to meet those needs.

RP: You say that the publishing program is usually treated as part of a library’s core service offering. Do I understand from this that we cannot know what it costs to run a journal on Digital Commons as libraries don’t break out the costs?

J-G B: I see that I didn’t answer your last question very well. Thanks for giving me a second shot. This is actually a really tricky question because the incremental cost of running a journal on Digital Commons is zero. There is no charge per journal. Digital Commons is kind of like those passes you used to get to travel all around Europe by train. I think they were called Eurail in the US and InterRail in Europe. Your question is like asking me “how much did the trip from London to Paris cost?” Since I paid one flat fee for the pass, dividing it up to figure out exactly how much one specific trip cost would depend on how many other trips I took.  

One thing is for sure; when I travelled by train in my youth I was certainly incentivized to take full advantage of my pass. It allowed me the opportunity to explore and experiment. We find that libraries that use Digital Commons are equally incentivized to get the maximum out of their investment. We think it is clearly a model that works.  

However, with Digital Commons the benefits of the unlimited service are actually much better than an unlimited train pass because the benefits are cumulative — so journals launched in any given year are always covered in all future years, with no related change in cost.

RP: Can you share with me any other stats that would give readers a sense of what is happening in the scholarly communication space currently, especially with regard to OA?

J-G B: We recently presented a poster at the Library Publishing Forum in Kansas City that offered a number of interesting statistics. I think one of the most interesting ones is what we’re seeing in terms of the growth and success of library-led, open access publishing.

We’re seeing an increasing number of institutions supporting not just two or three journals, but ten, fifteen, or even twenty journals.  These journals are often serving the needs of underserved disciplines, such as arts and humanities and education. 

RP: There is a lot of talk today about the need for the research community to “take back ownership” of scholarly communication. What are you views on this, and where do you see Digital Commons fitting into the debate? Are the journals being published using Digital Commons merely a fringe activity in your view, or are they the first buds of a revolution that will see today’s scholarly publishers disintermediated?

J-G B: Library-led publishing is an integral strategy in the university taking back ownership of scholarly communication. For the library to offer publishing as a campus-based service means that faculty and students have access to professional publishing tools from their own university, without needing to deal with commercial platforms or predatory publishers.

I’ve been attending the SPARC conference (the major OA event in the US) since 2008 and this year was the first time I had an overwhelming feeling that there is simply no slowing down the open access movement. It is raining green and gold. The benefits to authors, editors, readers and academic institutions are too great. Our momentum is unstoppable.

And let me add one trend that I think speaks volumes about the fact that this is far beyond the first buds of a revolution. 143 of the journals hosted on Digital Commons are edited not by faculty, but by graduate students and undergraduate students.

In an article about the award-winning Journal of Critical Thought and Praxis, founder Cameron Beatty explained the value of their new journal succinctly when he said “We were doing social justice work and some of the journals we wanted to get published in weren’t necessarily interested in the work that we were doing. So we were like, ‘What does this mean for us as graduate students? We need to get a job; we need to get published. Where is there a space for us that want to do critical social justice work? Why don’t we create our own?’”

Let’s turn now to the undergraduate journals. As a research apprenticeship, students are learning far more than the principles of peer review or how to deal with an editorial board (though that’s all really useful). They’re learning that they can have direct access to scholarly publishing, and that they can take a journal from idea to fully realized publication just by talking to their library. It’s going to be awfully hard to explain to them why they can’t just keep doing that when they become professional academics themselves.

If you are interested to learn why faculty and students are excited about this trend in their own words, I highly recommend this Webinar. I also would suggest following Scholarly Communications Librarian and Associate Professor Stephanie Davis-Kahl, from Illinois Wesleyan University. In March she was named the 2014 ACRL/EBSS Distinguished Librarian.  

Transition from publisher to technology company


RP: There is also a view that scholarly publishers will need to morph from being content providers to technology companies. Would you say that this is the path that bepress has taken? If so, do you expect all or most scholarly publishers to have to take the same path in the coming years?

J-G B: For us, the move from publisher to technology company was organic. We only became a technology company when the tools we created and honed to support our own journals became popular. Editors started asking us if they could use our tools when working with other publishers. We began licensing the software as a service and realized that there was a great need for both the software and the support bepress provided, including technical updates, publishing expertise, optimized web discovery, and customer support. We saw that, in line with our mission, we could serve the unmet needs of scholars by offering this software service more widely.

All in all, our complete transition from publisher to technology company took ten years. We were early, but I do think other scholarly publishers are going to have to follow a similar path from content provider to technology company if they want to continue to play a significant role in supporting the changing needs of the research community.

RP: What other changes do you see for scholarly communication going forward: the phasing out of pre-publication peer review? The death of the traditional journal and paper? Other likely developments?

J-G B: The consideration of new forms of review is extremely valuable, but I do not ultimately see the phasing out of pre-publication peer-review. If anything, I see a phasing in of pre-publication peer review, especially among undergraduates — we’ve seen it become a valuable teaching tool that improves the quality of the student work at many institutions. It is a just a matter of time before peer-review tools also find their way into primary education. My oldest son is nine, by the time he is in high-school, peer-review publishing will be part of his curriculum. There are already a few such journals.

Though print is slowly becoming less prevalent, I do not see the passing of the traditional journal model. These will remain as the cornerstones of scholarly communications for some time. This doesn’t mean that there will be no change — there already has been.

The biggest that we’ve seen is an increasing acceptance that there is a continuum of scholarship so that other works, in addition to the peer-reviewed articles, will count toward tenure and promotion. So in addition to providing tools to publish traditional journals, we are constantly seeking ways to support the publication of these other types of works.

Ultimately, the biggest change I see happening is increased awareness on the part of academic libraries that library-led publishing should be a core service that they offer to their faculty and to their students, and that this is a natural role for them that fits into their existing skills, their mission, and the future of their profession.

RP: Thank you for taking the time to answer my questions.