semantic web (16 out of 71)
Review of "Utilising Semantic Web Ontologies to publish Experimental Workflows"
In reply to:
I can see how the ultimate goal described by the article - to publish semantic representations of experimental workflows - contributes towards the vision of decentralised scholarly communication. Unfortunately I'm missing how the work described in the article contributes towards this ultimate goal. The abstract says you want to see how semantic workflow creation can "be combined with traditional forms of documentation and publication" but I don't see a result pertaining to this.
To begin with, the purpose of the tool is quite narrow. How do you know for sure that "advancing knowledge about the use of vocabularies in facilitating sharing and repeatability of experiments and replication of results" ultimately contributes towards reproducability of results, which is the high level goal of this work? How does knowledge of OPMW tie in with the authoring or experimentation process, for example? Are researchers expected to document their workflows as they go along, or retrospectively after the fact (I imagine this depends on the task at hand). For which stage is your tool intended? Or is it simply a teaching tool rather than designed for actually documenting workflows? This isn't clear.
I thought I might understand better by running the tool, but it seems to be broken.
I'm skeptical about the open world assumption and the "nature of Linked Data" being used as a reason to not provide any instruction for using the tool... I would have thought that whether instruction is needed is a UI concern.
There is no results or analysis section, and the tense makes it sound like the described experiment hasn't actually been carried out. Are you asking for feedback about the design of the experiment? If so, you should state this clearly in the introduction and abstract. if not then I'd like to read this again when your findings are ready.
The background work sections appear to be fairly comprehensive, but this is not my area of expertise, so if there is related work or other background information missing I am unable to point it out. It's not entirely clear how all of the related work describe relates to the problem at hand though so I'd like to see this be made explicit.
The discussion about nuances of licensing is interesting, for example licensing different parts of a workflow separately, and conveying this to people who want to replicate experiments and use the data produced. I think maybe the licensing topic deserves an article and experiment all of its own. Colour coding or different levels of alerts for different licensing to help people understand is an interesting UI challenge. I think ongoing work on Data Terms of Use might be interesting to you (this is about personal data rather than experimental data).
In summary, the stated goals make this worth further discussion, but it's not clear how the work you've done so far meets these goals. I'd like to know what are your next steps forward for this work, and technically how this could integrate with other projects related to exposing more semantically enriched academic research to the world.
Review of Instrumenting Continuous Knowledge Extraction, Sharing, and Benchmarking
In reply to:
I like how this article seeks to accommodate a broad view of viable data sources for research, and particularly encourages data sharing and reuse between researchers.
The authors provide three examples of existing tools which could do (or be adapted to do) parts of the suggested pipeline. I hope that publishing this encourages others who are developing tools along these lines to come forward and let the authors and others know so that the community can start a comprehensive directory, as I'm sure there are plenty more.
It would be helpful to also have a characterisation of what is definitely missing as far as the authors know, and what the authors think are good directions to priortise for near term research and development.
Obviously I agree with the authors' call to open source such tooling for community benefit. I'd be particularly interested to hear your thoughts on the "agreed-upon integration platform" and what you think the best forum for discussing such a platform would be. Hopefully we can come up with ideas for that during the EDSC workshop discussion sessions!
+ Recogito annotation platform
Amy added http://recogito.pelagios.org/rhiaro to https://rhiaro.co.uk/bookmarks/
🔁 https://twitter.com/philarcher1/status/778515246109646848
Amy shared https://twitter.com/philarcher1/status/778515246109646848
Vocabulary development and maintenance, W3C namespace control. 15:30 today, room 1.04 at TPAC2016- Phil
In reply to:
+ http://kidehen.blogspot.cz/2015/09/what-happened-to-semantic-web.html
Amy added http://kidehen.blogspot.cz/2015/09/what-happened-to-semantic-web.html to https://rhiaro.co.uk/bookmarks/
+ http://semprivacy.com/papers/privon2013.pdf
Amy added http://semprivacy.com/papers/privon2013.pdf to https://rhiaro.co.uk/bookmarks/
+ https://hackpad.com/PrivOn-2015-1w1AVtigY92
Amy added https://hackpad.com/PrivOn-2015-1w1AVtigY92 to https://rhiaro.co.uk/bookmarks/
+ http://dig.csail.mit.edu/2009/presbrey/UAP.pdf
Amy added http://dig.csail.mit.edu/2009/presbrey/UAP.pdf to https://rhiaro.co.uk/bookmarks/
+ http://www.lsrn.org/semweb/rdfpost.html
Amy added http://www.lsrn.org/semweb/rdfpost.html to https://rhiaro.co.uk/bookmarks/
+ http://video.dataversity.net/video/semantic-web-10-years-of-achievement/
Amy added http://video.dataversity.net/video/semantic-web-10-years-of-achievement/ to https://rhiaro.co.uk/bookmarks/
OH: "The thing Semantic Web people are best at is jumping on bandwagons
ESWC2015
This is a summary of a few bits and pieces that stood out to me from ESWC2015. I haven't covered every session I attended or paper I saw, just the ones that remained with me (other people will do full summaries of all the paper sessions I'm sure, or you can refer to the programme or Fabien Gandon's closing slides which have an excellent summary). For a more 'live' overview of my view on the conference you can see everything I posted during it.
Overall, I had a great experience, met some fantastic people and absorbed lots of interesting ideas. I feel more positive about work in linked data; I'd been slacking off following the community for a while, but I've been reassured that there are plenty of practical-minded researchers out there who are doing great things, and I'll be paying more attention again henceforth. Daily swims in the sea probably didn't hurt.
SemDev2015
The developers workshop was great, full of people positive about building tools and applications, and finding ways to make the power of linked data accessible to actual end users. The focus was on building for web developers rather than on end-user applications, with projects being great libraries and tooling for working with linked data, as a way to bridge the gap. There was an air of frankness, with attendees keen to address problems openly, without handwaving or glossing over things that weren't working out. There was even live debugging during presentations.
Here's the program, with links to projects and repos.
I missed the final discussion session, but this was recorded and I hear it was good.
Philoweb
For a philosophy of the web workshop, the talks and discussions during this workshop were around pretty pragmatic issues. In particular how we can obtain true decentralisation, problems with centralised DNS and internet infrastructure, the lack of attention paid in this community to security issues, and the importance of understanding social processes and current practice for ensuring the web continues to function and that we don't "break it by accident" (Henry Thompson). These aren't things that tend to get much of a forum at conferences like ESWC, but semantic web academics being at the forefront of a truly linked information space should definitely be encouraged to think about the effects of our work on society, particularly underprivileged and minorities.
USEWOD
This workshop - usage analysis and the web of data - had a general focus on understanding and getting the most out of the web as we know it today, in order to shape the web we want in the future. As well as traditional paper submissions, they were also accepting submissions via blog posts, and will continue to accept articles on an ongoing basis, which is a great way to keep the discussion alive. I was gutted to miss Max van Kleek's keynote "Not in my Castle" because I got the timing wrong, but I hear it was awesome.
In Use & Industry
Harry Halpin and Francesca Bria worked on an interesting project to map social innovation projects (like hacklabs, open data initiatives, community enterprises) across Europe. It wouldn't have been strictly necessary to use linked data for this, and doing so might have actually caused the site to be pretty slow. However, it allowed them to do a bunch of interesting network analysis on the hundreds of different projects and organisations mapped and gain some insights into how to strengthen such initiatives (for example, by increasing collaboration opportunities). Also, I suspect technologies for building sites on the back of linked data have probably improved quite a bit since this work was started, so the speed issue might easy to overcome. I paid attention cos I'm generally interested in putting stuff on maps but I'd really like to see more decentralised mapping things; projects/organisations publishing their information independently as linked data, such that they're in control over what's available, rather than having to submit their info to a centralised service.
Crowdsourcing and web science
Seyi Feyisetan from Southampton discussed different factors that affected the performance of crowd workers, by looking at features of the tasks themselves rather than the platform or rewards, when asking workers to classify entities in tweets. It was suggested that their results could be used to work out if NER on your microposts dataset would be better performed by machine, human experts or crowdworkers, depending on the contents of the dataset.
Revanthy Krishnamurthy presented about using general background knowledge and the contents of tweets to detect the location of twitter users, as most twitter users don't have geolocation enabled when posting. They did smart stuff like correlating mentions of events, landmarks and slang terms with physical places, but I don't remember them saying much about respecting the privacy of people who actively don't use geolocation..
Demos
The demos and poster session I thought was particularly lively. I liked that it didn't overlap with any other sessions, and breakfast cakes and fruit were distributed through the demos area. It was pretty cramped though, and possibly would have been better off earlier in the week too (it was in the morning of the last day).
I appreciate the principles and technologies behind Sarven Capadisli's Linked Research project, to the point that I implemented it for one of my own papers immediately. Encouraging web scientists to publish their research using the native web stack, using RDF to make research queryable and discoverable on a more granular level - in other words, to practice what we preach - is a worthy goal. And it was really easy to set up. Everyone should do it.
Entity annotation isn't something I know much about, but I think I understood a bit more after talking to Ricardo Usbeck about GERBIL, a tool for evaluating entity annotators. This easy to use online tool lets you compare some of the different annotators available against different types of datasets, to see which would perform the best for your particular use case, without you needing to access any of the test datasets yourself (as you often have to pay for licenses). You just plug your annotator in and leave it running, and it returns results. This also allowed them to check reported results of popular annotators and compare them according to different standards, for a more well rounded view of their capabilities. I dunno if the preceding paragraph made much sense, but that's what I got.
Other
Had some great discussions about microformats, RDFa, schema.org, federated/decentralised social web stuff and whatnot. I was also approached by a bunch of people who knew who I was from reading my WWW2015 post o.O Which was weird, but most people seemed to like it..
Finally, this was one of the better catered conferences I've been to, so props for that :)
+ http://csarven.ca/statistical-linked-dataspaces#linked-data-pages
Amy added http://csarven.ca/statistical-linked-dataspaces#linked-data-pages to https://rhiaro.co.uk/bookmarks/
"I wrote this code and it doesn't work really." Realtalk from @KKjernsmo xD