Monday, 22 March 2010

Il milione

Tra Pdl e questura, guerra di numeri sulla piazza [...] gli organizzatori [...] parlano di un milione di presenze, e la questura [...] stima 150mila presenti. [Il Giornale]
In queste manifestazioni tutti danno numeri che fa comodo dare [Berlusconi]
Verrà fatto un decreto d'urgenza stasera stessa che garantirà che i partecipanti sono stati comunque 1 milione. Sarà così possibile conteggiare anche quelli che avrebbero voluto andarci ma non potevano. Oppure oggi erano impegnati ma se fosse stata domani ci sarebbero andati di sicuro. Oppure erano a mangiarsi un panino. [anonimo]
Siete dei comunisti in malafede, "un milione di partecipanti" va interpretato come "un milione *in* partecipanti": sono i partecipanti che ci siamo potuti comprare con un milione. [anonimo]
Effettivamente se non ci fosse da piangere ci sarebbe da ridere...
S.

Thursday, 11 March 2010

Citazione

“La tecnologia si sta affrancando dalla scienza e il mare che da sempre divide il dire dal fare oggi si naviga all’incontrario, nel senso che è molto più semplice fare che dire, cioè capire e spiegare” 
(Giuseppe O. Longo, Homo Technologicus, Meltemi, 2001)
S.

Wednesday, 10 March 2010

Frase

"Di imparare non si finisce mai, e quel che non si sa è sempre più importante di quel che si sa già" (Gianni Rodari).

S.

Thursday, 11 February 2010

Eh...

"Sono Guido, buongiorno... Sono atterrato in quest'istante dagli Stati Uniti, se oggi pomeriggio, se Francesca potesse... io verrei volentieri... una ripassata"
Se non ci fosse da piangere ci sarebbe da ridere...

S.

Friday, 13 November 2009

IIR 2010

I though that I would never have organized something again, but I was wrong. Yes: errare humanum est, sed perseverare diabolicum. But this sounds interesting.


I'm organizing the Italian Information Retrieval workshop. Feel free to contact me for further infos. And feel free to submit your papers, of course!!

S.

Wednesday, 7 October 2009

E

Tra un po' non avremo più tesisti e subito dopo nemmeno studenti, figurarsi i dottorandi.

E

Tuesday, 6 October 2009

Research evaluation and bibliometrics - again

Another REF report. Quotations:
bibliometric indicators alone cannot provide a sufficiently robust measure of quality to drive funding allocations in any discipline at present. (p. 27, emphasis added)
Through those initial consultations – and the more recent bibliometrics pilot exercise – a widespread consensus has emerged that while metrics should inform expert review they are not sufficiently robust to replace expert review. While we remain concerned to reduce the burden of assessment, we believe we have exhausted the main options for any radically different alternative approach. The REF will be driven by a process of expert review, informed by metrics. (p. 27, emphasis added)
As I wrote time ago: it is stupid not to rely on bibliometric indicators. And it is stupid to rely on bibliometric indicators only.

S.

Monday, 28 September 2009

Promemoria

Non discutere con uno stupido: la gente potrebbe non notare la differenza.
Non discutere con uno stupido: devi scendere al suo livello,  e poi lui ti batte con l'esperienza.

S.

Friday, 25 September 2009

Almalaurea, statistiche, e dintorni.

Uno degli argomenti più frequenti nei consigli della mia facoltà è la differenza fra gli studenti di Informatica e TWM (lauree triennali) e Informatica e Tecnologie dell'informazione (lauree specialistiche). Quando si solleva la questione, di solito è sempre per dire che:
  • gli studenti di TWM sono di qualità inferiore, e me ne accorgo durante il mio insegnamento di XXX;
  • il corso di TWM è più facile.
Io faccio sempre notare tre cose. Primo, che forse è vero, e forse no. Secondo, che non è da stupirsi che gli studenti di TWM abbiano maggiori difficoltà dato che il corso è più giovane, ancora in assestamento, e che TWM viene sempre visto come il fratello minore di Informatica, e che quindi se c'è da scegliere fra i due (ad esempio mutuare un corso da uno all'altro, o decidere a chi di due docenti assegnare il corso a Informatica e chi a TWM, o ecc.) la scelta privilegia sempre Informatica (mutuando si mantiene il programma del corso di Informatica; il docente più "bravo" va a Informatica; ecc.). Terzo, che forse l'approccio scientifico sarebbe diverso: non basarsi sulle opinioni e sensazioni, ma guardare i dati. Finalmente i dati ci sono, grazie ad Almalaurea che, recentemente, ha introdotto la possibilità di vedere i dati per i singoli corsi di laurea. E i dati sono interessanti: Posted from Diigo.

S.

Wednesday, 23 September 2009

PhD in Computer Science @ Udine

We have a PhD degree in Computer Science at the University of Udine, Dept. of Mathematics and Computer Science. If you're interested, you have only two days left to submit an application (and you can do that online).

S.

Friday, 31 July 2009

Come una foto qui nella mia testa

Era da un po' che F. non era uno spettacolo...

- F., ti ricordi di quando abbiamo fatto quel giro in bici su al lago e abbiamo fatto il picnic?
- No...
- Sicuro? Era qualche anno fa, avrai avuto tre anni, stavi sul seggiolino della bici e io pedalavo, poi siamo arrivati davanti alla galleria, faceva fresco, ci siamo seduti lì e abbiamo mangiato i panini, pane, salame, marmellata, ...
- Ahh, sii, è vero... non me lo ricordo bene, ma un po' sì... è come una fotografia, qui nella mia testa...
- :-) E com'è fatta questa fotografia?
- Siamo io e te sulla bici, io sul seggiolino rosso, e io mi giro e vedo il lago là sotto, dietro di noi...
- Magari un giorno mi fai un disegno...
- Papà, ma perché non me lo ricordo tutto?

S.

Thursday, 23 July 2009

Invited talks

I'll be giving two invited talks in the next months:
  • Readersourcing: Scholarly publishing, peer review, and barefoot cobbler's children @ FDIA 2009 @ ESSIR 2009, 01/09/2009.

    Abstract:

    I will start from an introduction to the field of scholarly publishing, the main knowledge dissemination mechanism adopted by science, and I will pay particular attention to one of its most important aspects, peer review. I will present scholarly publishing and peer review aims and motivations, and discuss some of their limits: Nobel Prize winners experiencing rejected papers, fraudulent behavior, sometimes long publishing time, etc. I will then briefly mention Science 2.0, namely the use of Web 2.0 tools to do science in a hopefully more effective way.

    I will then move to the main aspect of the talk. My thesis is composed of three parts.

    (i) Peer review is a scarce resource, i.e., there are not enough good referees today. I will try to support this statement by something more solid than the usual anecdotal experience of being reject because of bad review(er)s --- that I'm sure almost any researcher has experienced.

    (ii) An alternative mechanism to peer review is available right out there, it is already widely used in the Web 2.0, it is quite a hot topic, and it probably is much studied and discussed by researchers: crowdsourcing. According to Web 2.0 enthusiasts, crowdsourcing allows to outsource to a large crowd tasks that are usually performed by a small group of experts. I think that peer review might be replaced --- or complemented --- by what we can name Readersourcing: a large crowd of readers that judge the papers that they read. Since most scholarly papers have many more readers than reviewers, this would allow to harness a large evaluation workforce. Today, readers's opinions usually are discussed very informally, have an impact on bibliographic citations and bibliometric indexes, or stay inside their own mind. In my opinion, it is quite curious that such an important resource, which is free, already available, used and studied by the research community in the Web 2.0 field, is not used at all in nowadays scholarly publishing, where the very same researchers publish their results.

    (iii) Of course, to get a wisdom of the crowd, some readers have to be more equal than others: expert readers should be more influential than naive readers. There are probably several possible choices to this aim; I suggest to use a mechanism that I proposed some years ago, and that allows to evaluate papers, authors, and readers in an objective way. I will close the talk by showing some preliminary experimental results that support this readersourcing proposal.

    Disclaimer: This talk might harm your career; don't blame me for that.

    P.S. Yes, this is somehow related to a previous post...

  • Two Tales on Relevance Crowdsourcing: Criteria and Assessment @ GIScience Colloquium, University Zurich-Irchel, 13/10/2009.

    Abstract (DRAFT):

    In Information Retrieval (IR) and Web search, relevance is a central notion. I will discuss how to outsource to the crowd two relevance-related tasks. The first task concerns the elicitation of relevance criteria. After some results obtained in the 90es, relevance criteria (i.e., the features of the retrieved items that determine their relevance) seem well known and stable. We conjectured that for e-Commerce / product search the criteria might be different, and we used Amazon Mechanical Turk, a crowdsourcing platform, to find a confirmation of our hypothesis.

    The second task concerns effectiveness evaluation. A common evaluation methodology for search engines and IR systems is to rely on a benchmark (a.k.a. test collection); benchmarks need relevance assessment, i.e., to assess the relevance of documents to information needs; usually this task is done by experts, either paid for their work or participating in the evaluation exercise themselves. Again, we used Mechanical Turk, this time to re-assess some TREC topic/document pairs and thus see if we can "get rid of" relevance assessors by replacing them with a crowd working remotely on the Web. I'll discuss the preliminary results on the reliability of the crowd of assessors.

    (this is joint work with Omar Alonso, A9.com; thanks also to Dan Rose)
I'll publish the slides here, once ready. Meaning: after the talks :-)
S.

Wednesday, 22 July 2009

Wednesday, 8 July 2009

Perché la democrazia non funziona

Guidavo. Era sera. Pioveva. Ascoltavo Travaglio. Lo ascolto spesso ultimamente. E a un certo punto ho improvvisamente capito. Perché la democrazia non funziona più.

Due motivi:
1. Corsa alle armi informazionale: la complessità aumenta sempre e capire cosa è giusto e cosa è sbagliato è sempre più difficile. Ognuno di noi cresce linearmente con le proprie conoscenze, mentre la conoscenza totale globale cresce esponenzialmente?
2. Ignoranza. Se ognuno di noi capisce meno, siamo tutti più ignoranti. Non in termini assoluti, non rispetto ai nostri avi. Ma rispetto a quello che servirebbe. E se non c'è cultura, educazione, la democrazia non funziona. Come ha detto quel comunista di Franklin D. Roosevelt, "Democracy cannot succeed unless those who express their choice are prepared to choose wisely. The real safeguard of democracy, therefore, is education. ". Perché vince (=prende voti) non chi fa le cose giuste ma chi parla agli istinti, anche ai più bassi istinti, delle persone. E diventa tutto un MagnaMagna.

Ho capito.

Uhm.

D'altro canto, oggi le tecnologie consentono di accedere alle informazioni in modo molto più veloce ed efficace di quanto accadenva in passato. Quindi è vero che ci sono più informazioni e più complesse ma anche più possibilità di accedervi. Mah. Ho improvvisamente capito. O almeno credo. O forse no. Ho perso le parole eppure ce le avevo là un attimo fa.

S.