Wednesday, October 04, 2006

Poll-itics

The following appeared in the Melbourne newspaper The Age:

THE Australian Democrats are refusing to publish an online survey about God and government after a campaign by Christian groups "skewed" the results.

Democrats leader Lyn Allison said 40 times the usual number responded to the survey, and overwhelmingly took the position advocated by some Christian leaders. Normally the party would be happy with 1000 responses, but the church and state survey got 40,000.

Senator Allison said it was ironic that a survey on the influence of churches should attract such an intense effort by churches to apply influence.

Christian groups have urged the Democrats to release the results, saying it was dishonest that a party that was founded on a claim "to keep the bastards honest" should keep the results secret because they were not what the Democrats wanted.

[The Age - National]

When this appeared in the Sydney Morning Herald's blog column "Stay In Touch", there was vigorous (even rancorous) debate about whether the Democrats should publish the results, and the significance thereof.

But what intrigued me more was the fact that the Democrats typically expect only about 1000 people to respond to their online surveys.

Online polls are actually fairly common, but as a few comments on the SMH blog pointed out, they hardly qualify as good scientific or statistical method, and no one should make too much out of the results. One comment noted that the Democrats website itself states "online surveys are useful because they are fast, easy and inexpensive but they do not typically gather in-depth, rigorous scientifically valid information" [here] and then asked the obvious - "then why do it in the first place"?

This then got me thinking - obviously the Democrats have done these surveys before, and have a fair idea of the typical response pattern. Also fairly obvious is that they have published the results on previous occasions, when the number of responses has been around the 1000 mark. Given that 1000 responses is about one-hundredth of 1% of the voting population of this country, do they believe that the results of such polls are in any way indicative of the overall views of the voting public? (As distinct from "rigorous scientifically valid information", I might add.)

It seems a fairly safe deduction that online polls will usually measure predominantly, perhaps almost exclusively, the opinions of those who frequent your website. Who else is going to go to your website? SBS Sport frequently polls its viewers on the SBS website, but if you don't watch SBS Sport or frequent the SBS website, you won't even know the poll is there, let alone participate in it. Ergo, the results reflect the opinions of the watchers of SBS Sport, not the opinions of sports enthusiasts in general.

In the same way, the Democrats' online surveys, which are not advertised broadly in the media, will usually measure only the opinions of those people who are aware of the poll, i.e. those who frequent the website or who keep track of Democrats-related news and events. In other words, those who are interested in the Democrats and their policies, which will primarily be supporters of the party.

It's therefore quite interesting to see that they will not release the results of the latest online survey because these results have been "skewed" - the results of the Democrats' online surveys are almost certainly skewed in any case, normally in the direction of those who agree with the Democrats.

In a nutshell, online polls measure the responses of interest groups of one kind or another. In the case of this latest poll, the subject of the poll intersected another, larger, interest group - Christians. It's a shame that the Democrats won't release the results.

And now I find myself being a touch cynical. I find it hard to credit that the Democrats would not be aware that the usual results of their surveys are dominated by Democrats supporters, so why publish results which you know do not in all probability reflect the opinions of the wider public? Sadly, there is an obvious answer.

The upshot of the Democrats withholding the survey results is they have made themselves look either incompetent, churlish or duplicitous, depending on who you talk to. If political parties are going to be taken seriously on the web, they need to avoid silly games like this, and think very carefully before conducting online surveys.

A final observation in light of all this, useful for both online surveys and elections:
Don't vote, it only encourages them. (Author Unknown)

Sunday, October 01, 2006

Copywrong

Miguel Guhlin linked to this on his blog:

There has been a major shift in how (some) Scientific Publishers see the purpose and practice of scholarly communication. Listening to the words used, “database” has replaced “journal” and “users” has replaced “readers”. I suspect the latter word conflates “purchasing officers” with “readers” into an unhappy anonymous entity. Moreover there is a tension between the publisher and the users - significant content is illegally downloaded and an important role of the publisher is acting as “policeman” making sure that content is not stolen. ...[snip]... Now, I have never advocated breaking or abolishing copyright, but it is clear that this is creating a tension in the publisher/reader community. I’ve been involved in setting or being on the board of scientific journals and I see their major purpose as enhancing scholarly communication. I’m worried that we are losing sight of this, where journals in non-profit organisations are seen as a way of subsiding other activities of the society. If the publishers see “users” as a group who have a major motive to steal content, I suspect things will get worse. At some stage we seem to have flipped from a community where publishers interpreted the wishes of the community and served them - for a reasonable fee - to a world where publishers make the rules and police their non-compliance. Did anyone in the reader community:
* actually ask for journals to be transformed to databases?
* actually ask for content to be limited in time to the duration of a subscription (we used to have physical journals we could take home and even hand down to our descendants or give to needy institutions)
It worries me that this has happened almost silently. I remember in ca. 1970 (when I was too inexperienced to notice) that authors were asked to transfer copyright to publishers. These requests came from trusted societies - national societies and international unions (At that stage there were essentially no commercial publishers - Pergamon was a few years later). I didn’t think twice about it - but it was one of the biggest mistakes of my scientific life. Are we sleepwalking into something just as serious? Objectively I have some sympathy with publishers whose content is illegally downloaded - I do believe in copyright. But pragmatically is the way forward to be increasingly draconian with readers (sorry, users)?

Unilever Centre for Molecular Informatics, Cambridge - petermr’s blog » Blog Archive » Do you read journals, or “use a database”?

Reading this reminded me of something else I'd come across recently...

British Academy Says Copyright Hindering Scholarship
A report from the British Academy, to be launched on 18 September, expresses fears that the copyright system may in important respects be impeding, rather than stimulating, the production of new ideas and new scholarship in the humanities and social sciences. ...[snip]... Existing UK law provides exemption from copyright for fair dealing with material for purposes of private study and non-commercial research, and for criticism and review. "There is, however, little clarity about the precise scope of these exemptions, and an absence of case law" said John Kay, who is Chair of the Working Group which oversaw the Review. "Publishers are risk-averse, and themselves defensive of existing copyrights. "The situation is aggravated by the increasingly aggressive defence of copyright by commercial rights holders, and the growing role - most of all in music - of media businesses with no interest in or understanding of the needs of scholarship. It is also aggravated by the unsatisfactory EU Database Directive, which is at once vague and wide-ranging, and by the development of digital rights management systems, which may enable publishers to use technology to circumvent the exceptions to copyright which are contained in current legislation. ...[snip]... This report parallels a report from the Royal Society, 'Keeping science open: the effects of intellectual property on the conduct of science (2003),' which expresses related worries about the ways in which intellectual property, its interpretation and its use, impact on the progress of science.

Managing Information News (found via Stephen Downes' blog)

Together, these blog entries paint a disturbing picture of publishers as an hindrance to research. I'm sure there's plenty more examples out there.

The comment at the end of the Cambridge blog is one I can empathise with. I do believe in copyright, but in proportion -- at what point will it become impossible to even quote a single sentence from a published work without first acquiring the express permission of the publisher (presumably for a fee)?

Stephen Downes has argued that locking down the use of other people's words is in itself a type of theft; his article in the April 2003 edition of the Journal of the United States Distance Learning Association (the link is on Stephen's blog, under the title Copyright, Ethics and Theft) eloquently outlines his thoughts on the subject.

Regardless of whether or not you believe in the concept of copyright, it should be becoming apparent to everyone that the Internet and associated technologies have created a situation that current laws about copyright and ownership do not properly address, and in fact cannot address while they remain tied to ideas about ownership that fail to recognise how the Internet has changed the landscape of expression and distribution of information.

The only people who could possibly be happy about the situation are lawyers.

Thursday, September 28, 2006

Gettin' Wikijiggered

Wikijiggered? Trust me, it's a politer word than what I'm thinking at the moment!

I've been wrestling with DokuWiki, a wiki that the TechGnome (aka Nayth) mentioned to me. It's neat - a wiki that relies only on PHP and the standard HTML/Javascript/CSS blend, no backend database (unless you want one) and no java classes (I'll come back to those in a minute). One upshot of this is that the pages are saved as text files -- yes, I said text files, vanilla flavoured, human readable, completely portable text files.

So why am ready to rip a few limbs off something? ACLs.

Access Control Lists. Or should that be Administrators Cursing Loudly? Without ACL features turned on, DokuWiki performs as expected, all is well. Turn the ACL features on, and register a new user, get a password emailed back, and login... except the username/password combo doesn't work. Aargh! How can be so close and yet so far?!

I want to set up this wiki in a school, but I don't want anyone in the world being able to hack up students' work, I'd like a certain level of security around this - if student X vandalised student Y's work, I'll be able to track X down and kick his sorry ar... erm... reprimand him. But I can't do that with someone on the other side of the planet. Hence security. Hence ACLs. And hence "Aargh".

The security issue also came up the other day with blogs. KP wanted a blog that only she and her students would access. I was looking at setting up a blogging system anyway, but had assumed that the TechGnome would know how to lock up subdirectories of our webserver - what do I tell my students until they're tired of hearing it? Never assume.

Nayth, to his credit, had not been simply resting on his laurels - the issue was that Apple had changed things significantly when they introduced OS X, Nayth had simply not had the time to dig through the manuals/forums/online support, and no one had pushed him to find out the answer.

At this point, God said "Come, let us go down and confuse them" (actually, that was the people of the Shinar plain [Genesis 11], but it works here too), and reminded Nayth of something he had seen. Mac OS X Server has a built-in weblog setup.

A quick look at this revealed that it could be locked down to specific users who would be authenticated by LDAP against the users on our file server. Hooray! Problem solved. (Never assume.)

Further inspection of this wondrous blogging system revealed what I should have suspected. It's not something Apple cooked up from scratch - in fact, it's a pruned version of Blojsom, a java-based implementation of Blosxom. Pruned? Yes - Apple have removed several aspects of the standard Blojsom implementation, added their own OS X branded skins, tied it to LDAP, and so on. The point is obviously to make it work neatly with OS X Server, which it does, but at the sacrifice of features and the user interface.

Some of this can be worked around (e.g. themes), but others present a problem. For example, WYSIWYG editing is a standard feature of most blogging systems, but Apple have removed that bit. You can probably install TinyMCE, but if you do, forget about using Safari. There's no easy way to upload images - you can put them into a folder in the web server, but then you have to add the image tags manually. (Unless I can get one of the Blojsom plugins to do this for me. And here I'm starting to get out of my depth - Blojsom is java-based, and its plugins are java classes. If it's just a case of editing some text files to enable the appropriate classes, fine, but if something goes belly-up, well... oh look, a decoy!)

And if that wasn't enough? A user can only have one blog, based on his/her username. Of course, it's possible to create a blog for a 'group', and we could create as many 'groups' as we like, but this means creating 'groups' that aren't actually groups, and cluttering up user lists, and I hate kludges like that at the best of times.

In short, Apple's Mac OS X Server weblog system is wikijiggered blogshlock. (For that matter, so is Safari. Safari could be the best browser on the planet, except that when it comes to some things, like using rich text editors in blogging systems, it's hopeless. Hmm, I can feel another rant coming on.)

Software that takes you nine-tenths of the way, and then leaves you flat on your laurels, just short of your goal - don't you hate being wikijiggered?

Update:
Oh, the shame. The ACLs weren't the problem. The problem? My bad.

It turned out that in adjusting one of the files to set up the 'superuser', Dreamweaver was saving the file with 'Mac-style' CR line endings instead of 'Linux-style' LF line endings, and this was screwing up Dokuwiki's attempts to read the user names and password hashes.

I suppose I could blame the documentation which directed me to make the changes to that file but never mentioned the necessity of ensuring that it is saved with LFs rather than CRs, but then again, that documentation is in a wiki, so I suppose the responsible thing is to go back and add this myself.

So I guess I wikijiggered my own blog post. Ain't irony a beautiful thing?

Tuesday, September 26, 2006

Onions Make Me Cry

From Miguel Guhlin's blog for Sep 24:

TorPark, an anonymizing browser (available for Windows) that doesn't have to be installed on your computer, is now out, shares Ben Horst at SolidOffice.org. ...[snip]... As far as I know, while it's possible to block, many school districts don't have a clue. . .the best they can do is block the download site that students might use. Now, all students have to do is load TorPark onto a USB drive (some schools are providing USB Flash drives to students as alternatives to floppies, so all they have to do is load it on there...convenient, huh?). ...[snip]... It's also available in different languages!

Plug it into any internet terminal whether at home, school, work, or in public. Torpark will launch a Tor circuit connection, which creates an encrypted tunnel from your computer indirectly to a Tor exit computer, allowing you to surf the internet anonymously. How much does Torpark cost? IT'S FREE.


What are the implications for something this easy to use? Well, don't worry, info-tech people will be getting all excited about blocking this! Quick, Quick...get started blocking and filtering!! What is fascinating is how K-12 schools can continue to try and block sites like this when there are communities of developers figuring out ways to bypass the blocks and filtering...it's a free speech, protect your privacy kind of use that many Americans see as fundamental.

To achieve strong anonymity, intermediate services may be employed to thwart attempts at identification, even by governments. These attempt to use cryptography, passage through multiple legal jurisdictions, and various methods to thwart traffic analysis to achieve this. A more recent approach in internet anonymity involves the use of an onion router such as Tor. Onion routers send information over encrypted protocols to several intermediate computers around the world in order to make identification more difficult. This has been countered with advances in text analysis, in which the identity of a writer is determined by comparing the writing style of a piece to styles of pieces in which the author is known.
[Source]


This last point--text analysis--almost reminds me of TurnItIn, and the ongoing controversy of using it.

At McLean High School, in Virginia, students collected more than 1,100 signatures on a petition opposing mandatory use of the service, according The Washington Post. The anti-Turnitin faction argues that the database violates students’ intellectual-property rights. And the high school’s use of Turnitin creates the sense that students are guilty until proved innocent, says Ben Donovan, a senior at McLean. "It’s like if you searched every car in the parking lot or drug-tested every student," he says. Source: The Chronicle Campus Blog [Source]


Around the Corner - MGuhlin.net - Courage can't see around corners, but goes around them anyway. - Mignon McLaughlin


I hadn't heard of onion routers until I read this blog entry and then did a little research, starting with the TorPark site itself. It's not hard to see why some school administrators would be getting nervy - this sort of technology can easily make a mockery of a school's filtering efforts.

And the more I think about that, the more convinced I am that trying to use technology to thwart people from using technology is about as sensible as washing grease stains with olive oil. (I was going to say, 'painting over wallpaper', but then I remembered my father... /sigh/.)

That's not to say that we don't use technology at all in managing technology in our school. Obviously Internet content filtering in a K-12 school is good practice, if only to protect little ones from the "nasty stuff". At the same time, relying almost exclusively on technology to handle what students do with technology is just crazy -- if we imagine that filtering systems will stop older/smarter students who are determined to bypass it, we're deluding ourselves.

The TurnItIn issue surrounds the use of software to detect plagiarism -- one gripe with it is that every student is "assumed guilty until proved innocent". As one commenter to the Chronicle Campus Blog pointed out, the software "will only catch the laziest students that simply buy a paper or copy a website off of the Internet, but it does nothing to stop a student from using a thesaurus to change enough words to fool the software" [comment 9 on the aforementioned Blog]. As almost every word processor now comes with a thesaurus, casting through a paper and 'adjusting' enough words to beat the software takes relatively little effort. (How many of these teachers have actually considered changing their assessment methods to address the problem?)

In the same vein, but on a different scale (but not that different), the music publishing companies who are currently pushing legislators in several countries to treat everyone as music pirates (and remember, you're guilty until proven innocent) are following the same flawed logic -- digital rights management is all about using technology to stop you from using technology. But can it really? I read somewhere recently that the new DRM technology in the latest incarnation of iTunes was broken in three days. If you can think of a way to lock it, someone else out there can conceive a way to unlock it.

I honestly think it's naive to believe that you can defeat technology with technology. But there's plenty of people out there who are going to try. And there will be lots of tears before they're through. Mostly their own, I suspect.

Tuesday, September 19, 2006

Authority? What's That?

Tim Bray's latest blog entry raises a question I expect as a teacher and IT Coordinator I will hear more and more: how can we be sure that the info on a website is reliable and authoritative?

First, let me quote from Tim's blog:

Let’s ask an interesting real-world question that real-world people might ask: for each of the ten provinces of Canada, what is its population? Let’s suppose you’re not a Canadian insider who knows that the Source Of All Numbers is Statistics Canada. So, you could go to Wikipedia, which would be easy and quick. From East to West you’d look at http://en.wikipedia.org/wiki/Newfoundland, http://en.wikipedia.org/wiki/Prince_Edward_Island, and, well, I’ll stop there, because the pattern is obvious. On each of those pages you’ll find the population, along with a lot of other basic facts, presented crisply and legibly, no further steps required. ¶But you know, that’s just the Wikipedia; some joker might have gone in and changed the number by couple hundred thousand up or down, just for fun. Wouldn’t you be better off going to a source with some real authority?


ongoing · Wikipedia: Resistance is Absent


Tim goes on to explain how he then searched on the government websites of the various provinces, and was confronted by lousy web design, URIs that only a machine or an über-geek could conceive (for example, www.gov.on.ca/ont/portal/!ut/p/.cmd/cs/.ce/7_0_A/.s/7_0_252/_s.7_0_A/7_0_252/_l/en?docid=EC001035 -- wha...?) and he concludes that "Wikipedia is going to win". Given his original premise -- if you want to have authority on the Web, you have to show up on the Web... And those who ought to enjoy more authority than Wikipedia aren’t [emphasis mine] -- it seems safe to conclude that Tim isn't entirely happy with this situation.

So let me make a few observations.

Tim is probably right in thinking that far too many people would read the Wikipedia articles and be satisfied with that, not bothering to check further. Students often tend to do this. Mine would if I let them... but I don't.

I've taken to telling my students that websites are not "nuggets of information" waiting for them to come and pick them up, but signs on a trail leading to the "real answer". The trail metaphor is a handy one, since it suggests that they have to continue on, following the links, occasionally doubling back from dead-ends to re-find the trail, and so on.

Some of my students would have certainly found the stats Tim wanted much faster than Tim -- they would have scanned through the Wikipedia article's links and found at the bottom this link: StatCan 2001 Census which is sort of where Tim ended up, but via a longer route.

Then there's the issue of people changing entries in Wikipedia. There's no question, people do weird and stupid things, and changing entries in Wikipedia is one of them. But an even stranger thing also happens -- people fix the mistakes! They get very defensive about it. And it's why Wikipedia works.

But there's one phrase in Tim's blog that really stands out for me: Wouldn’t you be better off going to a source with some real authority? [Emphasis mine]

Define real authority. (Actually, I might pose this as a question to a senior Computing class.) Government websites? (Is that laughter in the background?) Newspaper columns? University sites? Books? (Remember Margaret Mead vs Derek Freeman? I won't even mention Derrida.)

I know that for many, the suggestion that "real authority" is largely ephemeral will ring of heresy. But I'm part of a generation that has grown up realising that the "authorities" all too often spoke (and still speak) a lot of BS.

I would like to think that students who are now growing up with the Web, like my own daughters, will turn out to be fairly savvy when it comes to evaluating info from the Web, from the media, from wherever. They'll know that there's a need to check and cross-check and evaluate and never take any of it for granted. And they may not need to know how to spell authoritative.

Monday, September 18, 2006

Web to point.. oh?

Last Thursday and Friday, I was at an IT Integrators Conference in North Sydney. I got to see an interesting cross-section of things being done in Australian schools, as well as hearing from some compelling speakers, particularly Jim Mullaney and Stephen Downes. The common theme was "Web 2.0" (though there was some discussion that maybe that term has now been copyrighted? Oh, please - next someone will try to take out a trademark on "ugg boot"... oh... right).
So now the places where IT and education are coming together are blogs and wikis and newsfeeds and learning management systems. That's fine for me, I'm familiar with all these things, but in a short time I will become responsible for staff who are almost totally unfamiliar with these things and who are still trying to get their heads around how to integrate web browsers and email and word processors and Powerpoint into their classroom practice. "Web 2.0? I'm still at Web 0.2, thanks!"
The challenge is how to bring these staff up to speed on what they can do in the classroom with IT, but on the positive side, Web 2.0 presents far more opportunity for students to be involved in the technology. I always worried (still do) about how Powerpoint is used in classrooms - I've seen too many people (from Principals down) using several thousand dollars worth of equipment to do what could be done with an old-fashioned overhead projector. (The term is "powerpointlessness" - thank you, Jamie MacKenzie - check out From Now On.) Web 2.0 tools - blogging, wikis - allow students to put their own thoughts and ideas online and participate in a dialog that can be larger than the classroom and longer than the lesson.
That's not to say that Powerpoint presentations, email and Excel spreadsheets don't have their place - obviously they still do. But Web 2.0 tools have the potential to redefine pedagogy in a way that "office" software and older web-base software didn't.
I think the key idea is dialog, with students being participants in the processes of uncovering and connecting disparate components of knowledge. If that's a little hard to follow, I suppose it's because I'm not exactly a constructionist, nor a connectivist in how I view knowledge.
How well I get this across to my colleagues remains to be seen.

Blogging with Flock

So I'm using Flock to manage blogs and newsfeeds, but in checking to see how it handled the feed from my own blog, it's not been too happy.

The problem would appear to lie in the way Flock is formatting the HTML it sends to Blogger. Or perhaps in what Blogger is doing with it.

As an experiment, I'm sending this post to Blogger via Flock, but I'm keeping a copy of Flock's html so I can compare it to what ends up in Blogger.

Should be interesting.


Update 1: So far, there's no problem, so I'll try editing in Blogger and see if it takes. (This is where things went awry before.) I'll add styling in this post, maybe that's where the problem lies.

The other issue is that the new post shows up in Vienna, but not in Flock itself. Very strange.

Update 2: adding styling hasn't caused problems.

I'll try blockquotes and links. Here's Ongoing.

What else can I try?

Update 3: the HTML gremlins have gone away, but Desultoration still won't show as a new post in Flock. $#@!

Update 4: it's the next morning, and suddenly the new posts in Desultoration have appeared in Flock. ?!?!? Okay, it's only beta software, and I should know better than to expect it to work perfectly. Still, if it continues to happen, I'll be giving the News part of Flock a wide berth.


Blogged with Flock