Guest

This post is by a guest poster. If you would like to write something for the Open Knowledge Foundation blog, please see the submissions page.

More Reading

Post navigation

3 Comments

  • a commercial enterprise would not be able to release a product incorporating share-alike data or resources derived from it under the same conditions.

    That’s not true. A commercial enterprise could release a product incorporating copyleft/ShareAlike works, including their modifications under the same terms. Happens all the time. The enterprise might not want to, and it’s understandable why if you desire maximum immediate use you’d want to avoid copyleft, but “would not be able to” is misleading. Furthermore, copyleft requirements are typically only triggered when a derivative work is created. It’s perfectly possible for a proprietary software program to ship with copyleft content/data, and vice versa.

    We would really like to see something like Open Data Commons Attribution License (ODC-BY) become the license that authors automatically reach for when they publish language data on the web,

    Could you give an example of language data ODC-BY would be ideal for? Note that the license covers the database, not the database contents. I imagine language data is varied, but the first thing I think of regarding IP problems would be copyright on text one would one to include in a corpus — as you say later in the paragraph (re fair use (go fair use!)) “whole texts”.

    in the way the CC-BY-SA-NC license is now.

    I was not aware CC-BY-NC-SA is now widely used for language data. Please share examples.

    ODC-BY was developed primarily for databases, but it would not take much to apply it to language data, if it has not been done already (see, e.g., the Definition of Free Cultural Works)

    I’m scratching my head over what the above parenthetical means.

    Finally, the link in the last paragraph should be to http://www.anc.org/contribute.html (lowercase ‘C’).

  • Great post Nancy — I can hardly thing of a better exemplar for open data that the king of annotated linguistic corpora you are creating.

    In fact, this example strikes as having interesting similarities with what goes on bioinformatics: there people are ‘annotating’ resources like the the genome and have a similar requirements for openness to be able to share easily.

Leave a Reply

Your email address will not be published. Required fields are marked *

back to top