Skip to content
Merged
Show file tree
Hide file tree
Changes from 9 commits
Commits
Show all changes
23 commits
Select commit Hold shift + click to select a range
8de77c2
closes #159 Move information about old GSoC projects to CC Open Source
Cronus1007 Dec 15, 2020
6392f98
open source blog posts 2012 completed
Cronus1007 Dec 16, 2020
f3c94f7
open source blog posts 2010 completed
Cronus1007 Dec 16, 2020
8d364ad
open source blog posts 2009 completed
Cronus1007 Dec 16, 2020
f62624d
open source blog posts updated and PR ready for review
Cronus1007 Dec 16, 2020
32b1182
Merge branch 'master' into move
TimidRobot Dec 18, 2020
2e52d5f
made all the required changes as requested
Cronus1007 Dec 21, 2020
9344fbc
Merge branch 'move' of https://github.com/Cronus1007/creativecommons.…
Cronus1007 Dec 21, 2020
f3f266b
Merge branch 'master' into move
TimidRobot Jan 6, 2021
55265da
Merge branch 'master' into move
Cronus1007 Jan 6, 2021
1df9fb6
made the required changes as requested by timid robot
Cronus1007 Jan 6, 2021
1e371f4
fixed deploying failing errors
Cronus1007 Jan 6, 2021
d23a254
Merge branch 'master' into move
TimidRobot Jan 14, 2021
920299b
Merge branch 'master' into move
Cronus1007 Jan 24, 2021
861022e
correct names
TimidRobot Feb 2, 2021
4277558
Merge branch 'master' into move
Cronus1007 Feb 3, 2021
3fecf5b
correct names
Cronus1007 Feb 3, 2021
4692f61
Merge branch 'master' into move
Cronus1007 Feb 10, 2021
4877fc6
Merge branch 'master' into move
Cronus1007 Feb 17, 2021
285aa2c
Merge branch 'master' into move
Cronus1007 Feb 20, 2021
5bfaaa7
added link text for additional context
TimidRobot Feb 20, 2021
42bddc4
minor whitespace updates
TimidRobot Feb 20, 2021
9d74fe9
fixed title and links
TimidRobot Feb 20, 2021
File filter

Filter by extension

Filter by extension

Conversations
Failed to load comments.
Loading
Jump to
Jump to file
Failed to load files.
Loading
Diff view
Diff view
10 changes: 10 additions & 0 deletions content/blog/authors/Ethan-Lim/contents.lr
Original file line number Diff line number Diff line change
@@ -0,0 +1,10 @@
username:
---
name: Ethan Lim
---
md5_hashed_email:
---
about:


Ethan Lim([`@ethanlim`](https://github.com/ethanlim) on GitHub) was a Google Summer of Code 2013 intern with Creative Commons. They developed [CC Medua Fingerprinting Library](<#>) as a part of their internship.
9 changes: 9 additions & 0 deletions content/blog/authors/Ishan-Thilina/contents.lr
Original file line number Diff line number Diff line change
@@ -0,0 +1,9 @@
username:
---
name: Ishan Thilina
---
md5_hashed_email:
---
about:

<Ishan Thilina Somasiri> ([`@ishanthilina`](<Ishan Thilina Somasiri>) on GitHub) was a Google Summer of Code 2012 intern with Creative Commons. They developed [<Libre Office>](<https://github.com/cc-archive/cc.libreoffice>) as a part of their internship.
9 changes: 9 additions & 0 deletions content/blog/authors/THanish/contents.lr
Original file line number Diff line number Diff line change
@@ -0,0 +1,9 @@
username:
---
name: THanish
---
md5_hashed_email:
---
about:

Thanish([`@mnmtanish`](https://github.com/mnmtanish) on GitHub) was a Google Summer of Code 2013 intern with Creative Commons. They developed [Pasteboard](<https://github.com/cc-archive/pasteboard>) as a part of their internship.
9 changes: 9 additions & 0 deletions content/blog/authors/akila/contents.lr
Original file line number Diff line number Diff line change
@@ -0,0 +1,9 @@
username:
---
name: Akila Wajirasena
---
md5_hashed_email:
---
about:

<Akila Wajirasena> ([`@<akilaw>`](<https://github.com/akilaw>) on GitHub) was a Google Summer of Code 2010 intern with Creative Commons. They developed [<Open Office Plugin Updates>](<https://github.com/cc-archive/cc.libreoffice>) as a part of their internship.
9 changes: 9 additions & 0 deletions content/blog/authors/blaise/contents.lr
Original file line number Diff line number Diff line change
@@ -0,0 +1,9 @@
username:
---
name: Blaise Alleyne
---
md5_hashed_email:
---
about:

<Blaise Alleyne> ([`@<balleyne>`](<https://github.com/balleyne>) on GitHub) was a Google Summer of Code 2009 intern with Creative Commons. They developed [<UPDATE AND EXPAND DRUPAL CREATIVE COMMONS MODULE>](<#>) as a part of their internship.
9 changes: 9 additions & 0 deletions content/blog/authors/dinishi/contents.lr
Original file line number Diff line number Diff line change
@@ -0,0 +1,9 @@
username:
---
name: Dinishika Weerarathna
---
md5_hashed_email:
---
about:

<Dinishika Weerarathna> ([`@<dinishi>`](<#>) on GitHub) was a Google Summer of Code 2009 intern with Creative Commons. They developed [<RDFa Plugin for WordPress>](<https://github.com/cc-archive/ExternalData_RDFa>) as a part of their internship.
9 changes: 9 additions & 0 deletions content/blog/authors/erlehmann/contents.lr
Original file line number Diff line number Diff line change
@@ -0,0 +1,9 @@
username:
---
name: Nils Dagsson Moskopp
---
md5_hashed_email:
---
about:

<Nils Dagsson Moskopp> ([`@<erlehmann>`](<https://github.com/erlehmann>) on GitHub) was a Google Summer of Code 2010 intern with Creative Commons. They developed [<Wordpress Plugin>](<https://github.com/cc-archive/wordpress-cc-plugin>) as a part of their internship.
1 change: 1 addition & 0 deletions content/blog/categories/gsoc-2009/contents.lr
Original file line number Diff line number Diff line change
@@ -0,0 +1 @@
name:gsoc-2009
Comment thread
TimidRobot marked this conversation as resolved.
Outdated
1 change: 1 addition & 0 deletions content/blog/categories/gsoc-2010/contents.lr
Original file line number Diff line number Diff line change
@@ -0,0 +1 @@
name: gsoc-2010
1 change: 1 addition & 0 deletions content/blog/categories/gsoc-2012/contents.lr
Original file line number Diff line number Diff line change
@@ -0,0 +1 @@
name: gsoc-2012
1 change: 1 addition & 0 deletions content/blog/categories/gsoc-2013/contents.lr
Original file line number Diff line number Diff line change
@@ -0,0 +1 @@
name: gsoc-2013
89 changes: 89 additions & 0 deletions content/blog/entries/cc-Drupal/contents.lr
Original file line number Diff line number Diff line change
@@ -0,0 +1,89 @@
title: Update and Expand Drupal Creative Commons Module
---
categories:
open-source
gsoc-2009
---
author: blaise
---
pub_date: 2009-07-15
---
body:

Creative Commons provides vast number of public copyright licenses for people who want to enable free distribution of their work. Creative Commons licenses currently covers over **1.6 billion resources**. These license files are then translated to multiple different languages and ported for [different jurisdictions](https://wiki.creativecommons.org/wiki/CC_Ports_by_Jurisdiction) for international usage. People link to the respective licenses along with their licensed works. These license files are in the form of `html` files, stored in [creativecommons/creativecommons.org](https://github.com/creativecommons/creativecommons.org/tree/master/docroot/legalcode) repo.

### Problem Statement
There are currently two Creative Commons (CC) modules for Drupal. CC Lite only offers limited functionality, but the CC module has not been updated for a long time and doesn't support current versions of Drupal.

### Planning
* update the Creative Commons module, referencing and possibly building on Creative Commons lite, to support Drupal 6\
* expand its functionality to embed and detect license information for some file uploads.

### Summary
The Creative Commons module allows users to assign a Creative Commons license to
the content of a node, or to specify a site-wide license. It uses to Creatve
Commons API to retrieve up-to-date license information. Licenses are diplayed
using a Creative Commons Node License block and the Creative Commons Site
License block. The module also supports some license metadata fields. License
information is output using ccREL RDFa inside the blocks, and can optionally be
output as RDF/XML in the body of a node.

Creative Commons search is available at /search/creativecommons/, and (if the
Views module is installed and enabled) a Creative Commons view is available at
/creativecommons. Creative Commons license information and metadata are
available to the Views module.

For a full description of the module, visit the project page:
http://drupal.org/project/creativecommons

To submit bug reports and feature suggestions, or to track changes:
http://drupal.org/project/issues/creativecommons

### Requirements
None

### Installation
* Install as usual, see http://drupal.org/node/70151 for further information.

### Configuration
* Configure user permissions in Administer >> User management >> Permissions >>
creativecommons module:

- administer creative commons

Users can customize the module settings in Administer >> Settings >>
Creative Commons

- attach creative commons

Users will be able to attach license information to the content of a node.

- use creative commons user defaults

Users will be able to set their own defaults, independent of site defaults
(but still subject to site license availability settings).

* Set available license types, required/available metadata and display settings
Administer >> Settings >> Creative Commons. To make it mandatory to specify a
license, simply make the 'None' type unavailable.

* Set default license type and jurisdiction in Administer >> Settings >>
Creative Commons >> site defaults. Here, you can set the default license to be
used as a site-wide license if you wish, and you can include any relevant
metadata.

* Enable Creative Commons licensing for desired content types in Administer >>
Settings >> Creative Commons >> content types. For example, you might wish to
allow Creative Commons licensing for blog posts, but not forum posts.

* In your Drupal user account settings, you can set a jurisdiction or default
license to override the site defaults.


CC Drupal is only possible due to the support and guidance of my mentors [Kevin Reynen](http://drupal.org/user/48877) and `CC Tech Staff Member`, who have been very supportive on every step of the project. Also I would like to thank engineering director [Kriti Godey](https://creativecommons.org/author/kgodey) for her continuous support.


The project is approaching its completion. Can't wait to see it in production.

*Signing off
Blaise Alleyne*
94 changes: 94 additions & 0 deletions content/blog/entries/cc-RDFA-plugin-wordpress/contents.lr
Original file line number Diff line number Diff line change
@@ -0,0 +1,94 @@
title: RDFa Plugin for WordPress
---
categories:
open-source
gsoc-2009
---
author: dinishi
---
pub_date: 2009-07-15
---
body:

Creative Commons provides vast number of public copyright licenses for people who want to enable free distribution of their work. Creative Commons licenses currently covers over **1.6 billion resources**. These license files are then translated to multiple different languages and ported for [different jurisdictions](https://wiki.creativecommons.org/wiki/CC_Ports_by_Jurisdiction) for international usage. People link to the respective licenses along with their licensed works. These license files are in the form of `html` files, stored in [creativecommons/creativecommons.org](https://github.com/creativecommons/creativecommons.org/tree/master/docroot/legalcode) repo.

### Problem Statement
Conversion of information to a format compatible with RDF is a cumbersome endeavor. Though a complete automation of the conversion process would be a goal which is hard to achieve it can be made lot simpler with a wordpress plug-in. The plug-in will enable the users to convert a generic HTML in to a RDF compatible format with ease.

### Overview
External Data is an extension to MediaWiki that allows for retrieving data
from various sources: external URLs and local wiki pages (in CSV, JSON and
XML formats), database tables, and LDAP servers

The extension defines five parser functions - #get_external_data,
get_db_data, #get_ldap_data, #external_value and #for_external_table:

`get_external_data retrieves the data from a URL that holds XML, CSV or
JSON, and assigns it to local variables or arrays.`

`get_db_data retrieves data from a database, using a SQL query, and assigns
it to local variables or arrays.`

`get_ldap_data retrieves data from an LDAP server and assigns it to
local variables.`

`external_value displays the value of any retrieved variable, or the
first value if it's an array.`

`for_external_table applies processing onto multiple rows retrieved by
get_external_data.`

In addition, the extension defines a new special page, 'GetData', that
exports selected rows from a wiki page that holds CSV data, in a format that
is readable by #get_external_data.

For more information, see the extension homepage at:
[RDFa](http://www.mediawiki.org/wiki/Extension:External_Data)


### Requirements
This version of the External Data extension requires MediaWiki 1.8 or higher.

### Installation

To install the extension, place the entire 'ExternalData' directory
within your MediaWiki 'extensions' directory, then add the following
line to your 'LocalSettings.php' file:

`require_once( "$IP/extensions/ExternalData/ExternalData.php" );`

To cache the data from the URLs being accessed, you can call the contents
of ExternalData.sql in your database, then add the following to
LocalSettings.php:

`$edgCacheTable = 'ed_url_cache';`

You should also add a line like the following, to set the expiration time
of the cache, in seconds; this line will cache data for a week:

`$edgCacheExpireTime = 7 * 24 * 60 * 60;`

You can also set for string replacements to be done on the URLs you call,
for instance to hide API keys:

`$edgStringReplacements['MY_API_KEY'] = 'abcd1324';`

You can create a "whitelist" to allow retrieval of data only from trusted
sites, in the manner of MediaWiki's $wgAllowExternalImagesFrom - if you
are hiding API keys, it is very much recommended to create such a
whitelist, to prevent users from being able to discover theire values:


Finally, to use the database or LDAP retrieval capabilities, you need to
set connection settings as well - see the online documentation for more
information.


CC RDFa Plugin for WordPress is only possible due to the support and guidance of my mentors [Nathan Yergler](https://github.com/nyergler) and `CC Tech Staff Member`, who have been very supportive on every step of the project. Also I would like to thank engineering director [Kriti Godey](https://creativecommons.org/author/kgodey) for her continuous support.

You can follow the project on Github: [cc-archive/ExternalData_RDFa](https://github.com/cc-archive/ExternalData_RDFa).

The project is approaching its completion. Can't wait to see it in production.

*Signing off
Dinishika Weerarathna*
47 changes: 47 additions & 0 deletions content/blog/entries/cc-libreoffice(Open Office)/contents.lr
Original file line number Diff line number Diff line change
@@ -0,0 +1,47 @@
title: Creative Commons LibreOffice (OpenOffice) plugin updates
---
categories:
gsoc-2012
open-source
---
author: Ishan-Thilina
---
pub_date: 2012-07-15
---
body:

This year Creative Commons has a much more limited focus on what we would like to see out of Google Summer of Code than in previous years.
### Installation
The latest version of the plugin is 0.7.0, available for download from the [OpenOffice.org Extension site](https://extensions.openoffice.org/en/project/ccooo)
To install:

* Download the plugin
* Open OpenOffice.org
* Go to the `Tools` menu and select `Extension Manager`. Click `Add` and select the file you downloaded.
* After the installation process completes, restart OpenOffice.org. The Creative Commons menu options will be located at the bottom of the `Insert` menu.

Debian/Ubuntu (& other variants) users:
* Debian/Ubuntu (& other variants) users may have problems if the plugin installed using the Extension Manager.
* They may need to install the plugin using the terminal.
* To install use /usr/lib/openoffice/program/unopkg gui -f ccooo.oxt

### Problem Statement
The current Creative Commons LibreOffice (OpenOffice) plugin is written in Java. This project is aimed to port the current plugin to Python while rectifying the problems in the current plugin.
### Solution
The Python port of the Creative Commons Add-in for LibreOffice and OpenOffice.org which allows license information to be embedded in OpenOffice.org and LibreOffice Writer, Impress, Draw and Calc documents.

For more information visit [the official addin page](https://wiki.creativecommons.org/wiki/OpenOfficeOrg_Addin) of Creative Commons

###Testing
Download the out.oxt file. Install it via the LibreOffice Extension Manager (Tools --> Extension Manager).

Then go to `Insert` `--> ` `Creative Commons`. By using the first option you can insert the license information to the document. The second option is for changing those data.
### Knowledge Prerequisite
`Python` ,`Shell`

CC LibreOffice is only possible due to the support and guidance of `CC tech staff member`, who have been very supportive on every step of the project. Also I would like to thank engineering director [Kriti Godey](https://creativecommons.org/author/kgodey) for her continuous support.

The project is approaching its completion. Can't wait to see it in production.

*Signing off
Ishan Thilina Somasiri*
49 changes: 49 additions & 0 deletions content/blog/entries/cc-media-fingerprinting-library/contents.lr
Original file line number Diff line number Diff line change
@@ -0,0 +1,49 @@
title: Creative Commons Media Fingerprinting Library
---
categories:
gsoc-2013
open-source
---
author: Ethan-Lim
---
pub_date: 2013-07-15
---
body:

CC would prefer that all content on the Web include correct licensing metadata. Alas, that is not the case. So we're interested in code that will allow us to identify a given item across the Web, even if there's no metadata alongside (or within) it. The tricky part is: people often crop or resize images, clip videos, re-encode content, or quote only pieces of text. So a simple hash is not sufficient: we need more intelligent fuzzy matching. That's what this project is about.

### Expected Results
A library that provides two methods:
* Given a media file, output a fingerprint, and
* Given a file and a fingerprint, return the likelihood of the file matching the original file.

You can focus your efforts on only one or two media types, or you can do more if it's possible.
The library can be in a low-level language (C/C++) or you can use a higher-level language (JavaScript) if it's feasible. Speed is not a major concern at this point.
Bonus: An additional API/method to detect content inside other files (e.g., a PowerPoint file that includes a CC licensed image, or a still image inside a video).

### Notes/Resources
The first task is to decide on a strategy to compare two items and decide how similar they are. Some choices are:
* Hamming distance (bitwise AKA Manhattan distance)
* Euclidean distance (plane distance, also good in higher dimensions)
* Set similarity (Jaccard index; MinHash)

For this project, set similarity seems like the best choice. It would potentially allow us to detect works remixed into other works, if some portion of them has remained intact in some way. The technique involves distilling a document into a set of things, and comparing two documents is simply the ratio of things they have in common to things they do not.

A good way to start is with text, and involves a technique called shingling. For something like images, we'll need more work to determine which "interesting" features of the image to consider (to generate the set of things). This is called "keypoint extraction" and involves using standard algorithms to find vectors of floats that describe each keypoint. Since for images two keypoint vectors might be very similar but not identical, some additional work in clustering and mapping to example keypoints is required for images.
Some reading:
* Chapters 1 and 3 of [Mining Massive Datasets](http://infolab.stanford.edu/~ullman/mmds.html)
* [building shingles in text](https://lingpipe-blog.com/2011/01/12/scaling-jaccard-distance-deduplication-shingling-minhash-locality-sensitive-hashi/)
* [Introduction to Information Retrieval](https://nlp.stanford.edu/IR-book/)
* [OpenCV](https://opencv.org/) for extracting things (features) of images
* BRISK / FREAK: algorithms for "keypoint extraction", for images
* [pHash.org](http://www.phash.org/) might be something we can use.

### Knowledge Prerequisite
`Media formats/encodings` ,`JavaScript` ,`C/C++.`

CC MEDIA FINGERPRINTING LIBRARY is only possible due to the support and guidance of my mentors [Dan Mills](#) or ` other CC tech staff member`, who have been very supportive on every step of the project. Also I would like to thank engineering director [Kriti Godey](https://creativecommons.org/author/kgodey) for her continuous support.

The project is approaching its completion. Can't wait to see it in production.

*Signing off
Ethan Lim*
Loading