First objective of the JISC-supported Sonex initiative was to identify and analyse deposit opportunities (use cases) for ingest of research papers (and potentially other scholarly work) into repositories. Later on, the project scope widened to include identification and dissemination of various projects being developed at institutions in relation to the deposit usecases previously analyzed. Finally, Sonex was recently asked to extend its analysis of deposit opportunities to research data.






Showing posts with label Events. Show all posts
Showing posts with label Events. Show all posts

Saturday, 24 March 2012

Northwest England DCC roadshow at the University of Salford: a report



  Although not directly related to current Sonex work on analysing requirements for dataset transfer via Sword, the international workgroup was interested in attending a DCC roadshow for gathering a view on -and providing its own input to- RDM-related training initiatives that complement direct institutional experience in RDM acquired through JISC MRD projects. So when a chance showed up to attend the Northwest England DCC roadshow at the University of Salford, we were happy to engage with the UK Digital Curation Centre for being there at the University Library on March 20th and 21th.

Training initiatives regarding research data management were also thoroughly discussed at the 'Research Data Management: Activities and Challenges' workshop organised by the Knowledge Exchange Primary Data Workgroup in Bonn last November, which Sonex also attended (and provided a report for), so there were opportunities in Salford for identifying synergies between national and international RDM initiatives in this regard. This was also the right occasion for highlighting an example of best practice in RDM-related training activities based on local network building and promoting extensive debate on where to start and how to carry on with the work, while disseminating the appropriate tools to do it along the way.

The DCC roadshow proved to be a very effective complement indeed to JISC MRD programme and other RDM-related initiatives for reaching the 'common university' - i.e. those ones where preliminary efforts -be it at researcher survey level- are taking place to build some kind of RDM infrastructure but with no particular 'official' support outside -sometimes even inside- their institutions. The event in Salford gathered representatives from many NW universities -Salford, Manchester, Liverpool, Sheffield, Leeds- and debates along the roadshow were very much enriched by the mixture of institutional profiles attending it, from librarians to research office managers to researchers to ethics committee members. The experienced DCC roadshow team -Martin Donnelly, Andrew McHugh and Patrick McCann- were also very efficient in promoting dialogue and passing on guidelines and expertise along the event.

A shorter schedule was applied to this NW England roadshow, that took just two days instead of three: a first day devoted to presentations on data management initiatives taking place in the region and a second day for group discussions on research data management needs and how to use tools provided by DCC to identify them (such as DAF, Cardio or DMPOnline). The DCC Data management roadshow is also an evolving creature and there are slight variations in content among different editions thereof, this meaning that the event focus can be adapted to different levels of regional RDM implementation: emphasis can be for instance made on advanced RDM tools such as DMPOnline where RDM initiatives are well under way while mainly focusing on the Data Assessment Framework initiative in regions where RDM lies yet at a preliminary implementation stage.

Highlights from the first day included an estimulating 'Towards Open Worlds' keynote speech by Professor Martin Hall, a Vice Chancellor showing an unusually high commitment to Open Access and keen to debate related issues with the audience. An inspiring presentation of the two-stage MaDAM/MiSS JISC MRD project at U of Manchester was also delivered by Meik Poschen, providing some guiding light for preliminary initiatives in RDM currently being carried out at other institutions. Finally, Day I sessions were closed with a four expert panel discussion, in which presenters at the event were asked to stress a specific issue in RDM they considered worth deeper examination. The answers were: Cost model (Meik Poschen, UoM), Limits to researcher time availability (Rachel Kane, U Sheffield), Who shoud lead RDM tasks - is the Library able to? (Julie Berry, Salford U) and Research motivation as a decisive argument (Graham Pryor, DCC).

Along subsequent discussions Sonex became aware of three relevant points:

- Benefits can arise regarding these issues from a deeper analysis of international RDM initiatives -including ongoing and forthcoming European projects- connected to the institutional activity in the area,

- Besides disseminating specific funder mandates, finding a way for estimating the institutional costs derived from universities not managing their research data could be a potentially very effective argument for engaging universities with RDM activity.

- There is much emphasis in discussions on how to train researchers, but not so much on how to set up and train a team of dedicated data librarians - this strongly depending on Library staff figures and on whether or not librarians see themselves as fit for the task.


On the roadshow Day II the EPSRC policy framework on RDM and its implications for RDM strategy implemention at universities and research centres were discussed, and several joint RDM planning activities were carried out by different groups using DCC tools for examining aspects such as benefits to be obtained from RDM, where each institution stands in terms of RDM implementation strategy or how to deeper engage research groups and university management into RDM.

From a Sonex point of view, attending the roadshow proved very useful for identifying successful models of RDM training and dissemination, and we would humbly recommend to provide this RDM training initiative an international profile once complete so that similar efforts may be applied to a broader context. It would also be useful that participants in the DCC roadshows could provide feedback on the impact of their taking part in the initiative on their institution's work on RDM implementation a few months afterwards. Any future reporting from the DCC roadshow team on their initiative will be a very interesting read indeed and we shall be following dissemination initiatives outside the UK -such as the talk on DMPOnline at the Future Perfect 2012 Conference in Wellington this week- and hoping they'll soon arrive to continental Europe, where their work on Data Management Plans may be particularly valuable in the near future.




Tuesday, 13 December 2011

Thematic parallel session on metadata - actions to be taken


  On Day II of the JISC MRD Programme 2011-13 launch event in Nottingham, last Dec 2nd, specific subject-based discussion sessions were held among the different JISCMRD02 Projects for research data management in order to promote synergies and joint work on common issues. This is a brief report on the outcomes of such discussions at the parallel session on metadata - some other were simultaneously held for Institutional, Life Sciences, Engineering or Archaeology MRD projects, whose discussions have been reported elsewhere (and there are also other posts summarizing talks for this one too).

It was really hard for some of us to pick a single of those groups, since many projects actually belonged to several strands (some lucky ones had also two representatives at the event, it should be noted). The session on metadata was attended, among others, by:

- Anna Clements (U St Andrews)
- Simon Kerridge (U Sunderland)
- Kevin Ginty (U Sunderland)
- Charlotte Pascoe (British Atmospheric Data Centre)
- Pablo de Castro (SONEX Workgroup)
- Simon Hodson (JISC MRD Programme manager)
- David Shotton (U Oxford)
- Louise Corti (UK Data Archive)
- Marco Fabiani (Queen Mary U London)
...


Discussion

Metadata standards were repeatedly discussed along the session - there was a joint (and unsuccessful) attempt to recall whether anyone knew about a metadata standard registry available for different disciplines. Representatives from CERIF4Datasets Project, University of Sunderland, mentioned they were using the MEDIN metadata standard for their work in marine sciences data management. The Core Scientific Metadata Model (CSMD) standard, developed at STFC for the I2S2 Project was also mentioned as an interesting approach to multi-disciplinary metadata standard for structural sciences such as Chemistry, Materials Sciences, Earth Sciences or Biochemistry. Finally, the PIMMS Project (BADC/U Reading), mentioned Metafor as a Climate Science metadata standard and their goal of using PIMMS software tool to generate CIM-based content.

At some point the idea catched up that metadata standards should perhaps be mandated by publishers in order to harmonise discipline-specific data description procedures. Publishers are actually involved in several very successful international RDM projects, such as Dryad, but -save for REWARD- are significantly missing in JISCMRD02 projects.

Having previously developed the Semantic Publishing and Referencing (SPAR) Ontologies, David Shotton said he was now working on their extension to CERIF-based metadata description of datasets, which is closely linked to dataset CERIFication work being carried out at the CERIF4Datasets Project.


Actions

The following actions were proposed for improving the chances of metadata standard harmonisation - hence enhancing dataset discoverability:

  • Trying to locate (or otherwise collect) an already existing registry of metadata standards for different disciplines, in order to offer researchers from a given discipline an already tested metadata schema they can re-use,

  • Mapping metadata standards to each other aiming to produce a minimum-sufficient-information metadata set that may be widely applicable accross disciplines,

  • Taking steps towards organising a workshop in order to have metadata issues discussed among relevant stakeholders. ANDS Metadata Workshop in 2010 might be a potential source of inspiration for this with all those discipline-based approaches to metadata standards. Proposed dates for this Metadata WS were spring-summer 2012.


Finally, there was a wrap-up by different subject-based project groups which showed strong possibilites for a more stable cooperation among them (Biomedical/Healthcare projects even discussed the possibiity of building a common wiki). Some cooperation frameworks (googlegroups, mailing lists) might be set for promoting this disciplinar trans-project collaboration. Regarding the metadata strand, it should be noted it was also an issue in discussions held at most subject-specific workgroups, so it would potentially allow contributions from all of them.

Saturday, 17 September 2011

Progress on Researcher ID initiatives: IRISC 2011 Helsinki


  
The problem with names...

Prof. Carlos Martínez-Alonso is a renowned Spanish senior biochemist. He was actually President of the Spanish National Research Council (CSIC) when the Berlin Declaration was signed by the institution in January 2006. Prof. Martínez-Alonso has published hundreds of papers in high impact factor journals. However, when retrieving a complete list of his publications from PubMed database, you find out it is not possible unless several parallel author queries are carried out: there is a Martinez-A C entry under which most of his publications get listed [222]. But then there's also Martinez-Alonso C [21] and even Alonso CM [1].

It might be argued it's all about funny Spanish names with two surnames in them. That's a problem alright. Not just for Spanish names though: it's quite the same for Portuguese/Brazilian authors as well. Not to mention transliteration of Asian author names (see "Which Wei Wang?" Phys Rev 2007 editorial). PubMed is presently running its Author ID project in order to tackle this problem, which is by no means exclusive of theirs: around 2/3 of the over 6 million authors in MEDLINE share a last name and first initial with at least one other author, and an ambiguous name refers to 8 persons on average (Torvik and Smalheiser, "Author name disambiguation in MEDLINE").

Name disambiguation and proper attribution is a well-known problem in the scholarly publishing ecosystem. There have been and there are lots of initiatives trying to tackle this complex issue at subject, institutional or even national level - with remarkable success in the case of the Dutch Digital Author Identifier (DAI).

However, this is not an issue to be tackled at national nor subject level, but globally. Commercial stakeholders such as ThomsonReuters or Elsevier-Scopus are then in a privileged position to implement some international author unique identification schema. From a knowledge discovery viewpoint there are however some problems in this commercial-stakeholder approach: the ResearcherID, ThomsonReuter's author identifier, will provide seamless integration with ISI Web of Knowledge and show all author publications registered in that database, but will otherwise leave out most of the research output.

Some joint effort between public institutions and private stakeholders (remarkably publishers) must therefore be attempted to unify the multiple author identification standards and devise a single, comprehensive one at a global level. And that's where ORCID comes in.

  
... and strategies to tackle it: IRISC 2011 workshop

The Open Researcher & Contributor ID (ORCID) initiative started in Dec 2009 as a non-profit organisation. Currently over 240 participants have joined the project for developing the one research identifier which is not limited to discipline, institution or geographical area. Many other projects are working in this issue at the same time (such as abovementioned discipline-based PubMed Author ID and Cornell University initially institutional then grown to national VIVO initiative).

ORCID and VIVO were two of the main topics of the IRISC 2011 Workshop on Identity in Research Infrastructure and Scientific Communication held this week (Sep 12-13) in Helsinki - see the event programme with attached presentations. Gudmundur "Mummi" Thorisson, Research Associate at University of Leicester and member of ORCID Technical Working Group, was IRISC 2011 main organizer.

There were two major IRISC 2011 strands: identity regarding knowledge discovery and identity for security & access control (focusing mainly on identity federation). A third big cross-issue along the Helsinki event was research data management, from three different perspectives:

i) dealing with a rapidly increasing amount of biomedical research data (Andrew Lyall, EMBL, ELIXIR Project)

ii) dealing with clinical research sensitive data (see Tony Brookes GEN2PHEN Project presentation)

iii) benefits the ORCID implementation might bring to research data attribution and management (mentioned in most ORCID-related presentations and discussions along the workshop)


There were several presentations dealing both with ORCID and closely resembling VIVO initiatives. Martin Fenner, Hannover Medical School and member of ORCID Board of Directors announced the ORCID registration service will start operating in spring 2012. ORCID will be open: researchers will be able to manage & maintain their profiles, filed data will be openly available, ORCID-related software will be released as open source, and researchers will control their privacy settings (with a chance too to share with particular members). Finally, for ORCID identity definition purposes, self-claim as well as external claiming sources will be used.

Brian Lowe, University of Cornell, presented the already running NIH-funded, institutionally-managed VIVO initiative. VIVO is aiming for an extensible semantic model-based more comprehensive approach than ORCID. However, links have already been established between both initiatives and ORCID is hoping to build upon VIVO success in the US.


Breakout sessions were held on IRISC Day 2 on the workshop's two main strands: "Unique identifiers and the Digital Scholar" (lead by Cameron Neylon and Jason Priem) and "What do researchers need from the authentication and authorisation infrastructure (AAI)?" (chaired by Michael Linden, CSC). Breakout session #1 was devoted to discussing potential tools and services to researchers ORCID could provide in the short term (6 months from adoption). Several groups were set up for the purpose and proposed ideas were later voted and discussed for selecting three main future worklines for ORCID to deal with. The proposed and selected use cases were the following:

-> data submission to repositories (multiple task attribution)

service to enable attribution or comment

pre-populate ORCID data

-> manuscript/grant tracking system

ORCID app gallery

-> automatic CV maintenance (potentially including data citations in CVs)

connecting different author research & social network profiles

Selected ORCID use cases were later introduced by Cameron Naylon along his talk 'ORCID and researchers' at the second annual ORCID Outreach Meeting held at CERN on Sep 16th, 2011.

Sunday, 10 July 2011

Gettin' on...


  After quite a long, not totally intended silence - schedules get so hectic every now and then- it is the purpose of the Sonex workgroup to update the project blog by briefly reporting on recently held workshops we have attended since last post. These have been, inter alia, the 2nd euroCRIS/CNR-IRPPS workshop on CRIS and OAR (Rome, May 23-24), euroCRIS membership meeting 2011 (Bologna, May 26-27), CERN Workshop on Innovations in Scholarly Communication (OAI7, Geneva, June 22-24) and LIBER 40th Annual Conference 2011 (Barcelona, Jun 29-Jul 2).

Wednesday, 26 May 2010

Recently held and upcoming events on CERIF-CRIS/IR integration

An euroCRIS-organised event related to CERIF-CRIS/IR integration was recently held at CNR Rome, Italy, and forthcoming CRIS2010 will be taking place next June 2nd to 5th in Aalborg, Denmark:

  • Workshop on CRIS, CERIF and Institutional Repositories: Maximising the Benefit of Research Information for Researchers, Research Managers, Entrepreneurs and the Public (Istituto di ricerche sulla Popolazione e le Politiche Sociali, IRPPS, Consiglio Nazionale delle Ricerche, CNR, Rome, Italy, May 10-11, 2010).

  • CRIS2010: Connecting Science with Society: The Role of Research Information in a Knowledge-Based Society (10th International Conference on Current Research Information Systems, Aalborg, Denmark, June 2-5, 2010).

Tuesday, 11 May 2010

'Learning how to play nicely: Repositories and CRIS: a report', by Richard Jones

Friday 7th was the joint JISC and ARMA event "Learning how to play nicely: Repositories and CRIS", aimed at stirring up some discussion around the relationship and integration between these two kinds of system. Such integration has been talked about for some time, and I find myself recalling the Knowledge Exchange workshop in Utrecht where JISC, in partnership with SURF and DEFF and DFG initiated similar discussions in 2007. It is good to see that this discussion has moved from the domain of Repository, CRIS and CERIF developers into the mainstream of Research and Repository Managers, where requirements can more appropriately be sourced. For this technical observer the event was somewhat too non-technical, but I think this was the intention and for the best.

Andy McGregor from JISC set the scene for the event, giving us a little background on JISC involvement, and talking about different approaches that could be taken to integration, such as the use of CERIF or of Linked Data for the sharing of information. He then passed us over to Simon Kerridge from ARMA, who discussed in a bit more detail what a CRIS is; he also gave us some better terminology that we might prefer to use: RMAS (Research Management and Administration System) and ERA (Electronic Research Administration). The briefing paper that accompanies the event tells us that "by communicating research information more effectively ... the process of sharing data becomes more efficient, duplication of effort is reduced and information becomes more accurate", and this clearly drives the purpose of the day. Particularly, there is no intention here to merge CRIS and Repositories - the two communities have sufficiently different use cases that this is unlikely to happen - but simply to enhance communication between them in the correct way.

Anna Clements then introduced the CRIS that they use at St Andrews, while William Nixon and Valorie McCutcheon from the University of Glasgow presented Enlighten. Particularly, Enlighten is an interesting case as it is based on the EPrints software, and started life as an institutional repository in around 2003, but has now grown into a fully fledged publications management system. The presentations were then wrapped up by Jackie Knowles, from the Welsh Repository Network (the event organisers), who gave us an insight into things that went well and things that didn't during development of CRIS and Repository systems at institutions around the country. The ones that stuck for me were:

  • Don't overcomplicate your requirements

  • Don't develop DIY solutions which turn into single points of failure (i.e. ensure they are robust against staff changes)

  • Ensure that your requirements are well specified and met; she cites an unfortunate and extreme tale of a team who lost their jobs after failing to successfully implement a system which had no formal requirements in the first place!


The afternoon of the event was given over to discussion among delegates, and this observer did not attend due to his position as representing a supplier - the event coordinators felt that without the suppliers present the conversation would be more candid. The results of those discussions should be made available soon, and we'll link them when they are. Meanwhile, I therefore represented Symplectic in the exhibition stall, alongside Avedas, EPrints, Atira, ARMA, ThomsonReuters, IDEATE and DuraSpace; it was busy for much of the afternoon, which I think shows a clear interest in this space at this time.

Sunday, 9 May 2010

Learning how to play nicely: Repositories and CRIS event

Last Friday May 7th a joint JISC and ARMA one-day event on repositories and Current Research Information Systems (CRISes) was held at Leeds Metropolitan University. Organised by the Welsh Repository Network (WRN), this event brought together representatives from both research administration and repository management functions within institutions to explore the synergies, overlaps and opportunities in our role of curating institutional research and publication management information.

Following issues -among others- were discussed at the meeting (see event programme for contributions):

  • Why a CRIS? The perspective from the repository and research management communities
  • The ideal CRIS: a view from euroCRIS
  • DIY Success: Case study from the University Glasgow - How repository and research management systems have been successfully integrated
  • Where did it all go wrong?: Case study on how repository and research management systems have not been so successfully integrated

Tweets about the event were saved, and presentations are already available online as well. Finally, Richard Jones from Sonex workteam was attending the seminar at Leeds Met and will also be delivering a brief report on the main issues dealt with at the event.

Wednesday, 14 April 2010

Upcoming Sonex-related events

• OR10: The 5th International Conference on Open Repositories (Madrid, Spain, Jul 6-9, 2010)

• CRIS2010: Connecting Science with Society (Aalborg, Denmark, Jun 2-5, 2010)

• Learning how to play nicely: Repositories and CRIS (Leeds Metropolitan University, Leeds, UK, May 7, 2010)

• Repository Multiple Deposit meeting (London, UK, Apr 8, 2010)

• Readiness for REF (R4R) Workshop (King's College London, Mar 23, 2010)

• OpenAIRE Inaugural Conference (Athens, Greece, Jan 13-14, 2010)

• JISC Deposit Show-and-Tell Barcamp (University College London, Oct 12, 2009)