<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"
	xmlns:content="http://purl.org/rss/1.0/modules/content/"
	xmlns:wfw="http://wellformedweb.org/CommentAPI/"
	xmlns:dc="http://purl.org/dc/elements/1.1/"
	xmlns:atom="http://www.w3.org/2005/Atom"
	xmlns:sy="http://purl.org/rss/1.0/modules/syndication/"
	xmlns:slash="http://purl.org/rss/1.0/modules/slash/"
	>

<channel>
	<title>R. Stuart Geiger</title>
	<atom:link href="http://stuartgeiger.com/wordpress/feed/" rel="self" type="application/rss+xml" />
	<link>http://stuartgeiger.com/wordpress</link>
	<description>computational ethnographer &#124; ethnographer of computation </description>
	<lastBuildDate>Sun, 21 Aug 2016 17:49:45 +0000</lastBuildDate>
	<language>en-US</language>
	<sy:updatePeriod>hourly</sy:updatePeriod>
	<sy:updateFrequency>1</sy:updateFrequency>
	<generator>https://wordpress.org/?v=4.5.2</generator>
	<item>
		<title>Trace Ethnography: A Retrospective</title>
		<link>http://stuartgeiger.com/wordpress/2016/03/trace-ethnography-a-retrospective/</link>
		<comments>http://stuartgeiger.com/wordpress/2016/03/trace-ethnography-a-retrospective/#respond</comments>
		<pubDate>Mon, 28 Mar 2016 18:55:01 +0000</pubDate>
		<dc:creator><![CDATA[stuart]]></dc:creator>
				<category><![CDATA[Blog Posts]]></category>

		<guid isPermaLink="false">http://stuartgeiger.com/wordpress/?p=788</guid>
		<description><![CDATA[This is a cross-post of a post I wrote for Ethnography Matters, in their &#8220;The Person in the (Big) Data&#8221; series When I was an M.A. student back in 2009, I was trying to explain various things about how Wikipedia worked to my then-advisor David Ribes. I had been ethnographically studying the cultures of collaboration in the encyclopedia project, and I had gotten to the point where I could look through the metadata documenting changes to Wikipedia and know quite a bit about the context of whatever activity was taking place. I was able to do this becauseWikipedians do this: they leave publicly accessible trace data in particular ways, in order to make their actions and intentions visible to other Wikipedians. However, this was practically illegible to David, who had not done this kind of participant-observation in Wikipedia and had therefore not gained this kind of socio-technical competency. For example, if I added “{{db-a7}}” to the top an article, a big red notice would be automatically added to the page, saying that the page has been nominated for “speedy deletion.” Tagging the article in this way would also put it into various information flows where Wikipedia administrators would review it. If any of Wikipedia’s administrators agreed that the article met speedy deletion criteria A7, then they would be empowered to unilaterally delete it without further discussion. If I was not the article’s creator, I could remove the {{db-a7}} trace from the article to take it out of the speedy deletion process, which means the person who nominated it for deletion would have to go through the standard deletion process. However, if I was the article’s creator, it would not be proper for me to remove that tag — and if I did, others would find out and put it back. If someone added the “{{db-a7}}” trace to an article I created, I could add “{{hangon}}” below it in order to inhibit this process a bit — although a hangon is a just a request, it does not prevent an administrator from deleting the article. Wikipedians at an in-person edit-a-thon (the Women’s History Month edit-a-thon in 2012). However, most of the time, Wikipedians don’t get to do their work sitting right next to each other, which is why they rely extensively on trace data to coordinate render their activities accountable to each other. Photo by Matthew Roth, CC-BY-SA 3.0 I knew all of&#8230; ]]></description>
				<content:encoded><![CDATA[<p><em>This is a cross-post of <a href="http://ethnographymatters.net/blog/2016/03/23/trace-ethnography-a-retrospective/">a post I wrote</a> for Ethnography Matters, in their <a href="http://ethnographymatters.net/editions/the-person-in-the-big-data/">&#8220;The Person in the (Big) Data&#8221; series</a></em></p>
<p>When I was an M.A. student back in 2009, I was trying to explain various things about how Wikipedia worked to my then-advisor David Ribes. I had been ethnographically studying the cultures of collaboration in the encyclopedia project, and I had gotten to the point where I could look through the metadata documenting changes to Wikipedia and know quite a bit about the context of whatever activity was taking place. <em>I</em> was able to do this because<em>Wikipedians</em> do this: they leave publicly accessible trace data in particular ways, in order to make their actions and intentions visible to other Wikipedians. However, this was practically illegible to David, who had not done this kind of participant-observation in Wikipedia and had therefore not gained this kind of socio-technical competency.</p>
<p>For example, if I added “{{db-a7}}” to the top an article, <a href="https://en.wikipedia.org/wiki/Template:Db-a7">a big red notice</a> would be automatically added to the page, saying that the page has been nominated for “<a href="http://enwp.org/WP:CSD">speedy deletion</a>.” Tagging the article in this way would also put it into various information flows where Wikipedia administrators would review it. If any of Wikipedia’s administrators agreed that the article met speedy deletion criteria A7, then they would be empowered to unilaterally delete it without further discussion. If I was not the article’s creator, I could remove the {{db-a7}} trace from the article to take it out of the speedy deletion process, which means the person who nominated it for deletion would have to go through the standard deletion process. However, if I was the article’s creator, it would not be proper for me to remove that tag — and if I did, others would find out and put it back. If someone added the “{{db-a7}}” trace to an article I created, I could add “{{hangon}}” below it in order to inhibit this process a bit — although a hangon is a just a request, it does not prevent an administrator from deleting the article.</p>
<div class="wp-caption aligncenter" style="width: 755px; border: 0;">
<p><img class="aligncenter" src="http://i2.wp.com/upload.wikimedia.org/wikipedia/commons/thumb/5/5b/Wiki_Women%27s_Edit-a-thon-1.jpg/800px-Wiki_Women%27s_Edit-a-thon-1.jpg?resize=680%2C453&amp;ssl=1" alt="File:Wiki Women's Edit-a-thon-1.jpg" width="745" height="496" /></p>
<p class="wp-caption-text"><em>Wikipedians at an in-person edit-a-thon (the Women’s History Month edit-a-thon in 2012). However, most of the time, Wikipedians don’t get to do their work sitting right next to each other, which is why they rely extensively on trace data to coordinate render their activities accountable to each other. Photo by <a href="https://en.wikipedia.org/wiki/File:Wiki_Women%27s_Edit-a-thon-1.jpg">Matthew Roth, CC-BY-SA 3.0</a></em></p>
</div>
<p>I knew all of this both because Wikipedians told me and because this was something I experienced again and again as a participant observer. Wikipedians had documented this documentary practice in many different places on Wikipedia’s meta pages. I had first-hand experience with these trace data, first on the receiving end with one of my own articles. Then later, I became someone who nominated others’ articles for deletion. When I was learning how to participate in the project as a Wikipedian (which I now consider myself to be), I started to use these kinds of trace data practices and conventions to signify my own actions and intentions to others. This made things far easier for me as a Wikipedian, in the same way that learning my university’s arcane budgeting and human resource codes helps me navigate that bureaucracy far easier.<span id="more-10287"></span></p>
<p>This “trace ethnography” emerged out of a realization that people in mediated communities and organizations increasingly rely on these kinds of techniques to render their own activities and intentions legible to each other. I should note that this was not my and David’s original insight — it is one that can can be found across the fields of history, communication studies, micro-sociology, ethnomethodology, organizational studies, science and technology studies, computer-supported cooperative work, and more. As we say in the paper, we merely “assemble their various solutions” to the problem of how to qualitatively study interaction at scale and at a distance. There are jargons, conventions, and grammars learned as a condition of membership in any group, and people learn how to interact with others by learning these techniques.</p>
<p>The affordances of mediated platforms are increasingly being used by participants themselves to manage collaboration and context at massive scales and asynchronous latencies. Part of the trace ethnography approach involves coming to understand why these kinds of systems were developed in the way that they were. For me and Wikipedia’s deletion process, it went from being strange and obtuse to something that I expected and anticipated. I got frustrated when newcomers didn’t have the proper literacy to communicate their intentions in a way that I and other Wikipedians would understand. I am now at the point where I can even morally defend this trace-based process as Wikipedians do. I can list reason after reason why this particular process ought to unfold in the way that it does, independent of my own views on this process. I understand the values that are embedded in and assumed by this process, and they cohere with other values I have found among Wikipedians. And I’ve also met Wikipedians who are massive critics of this process and think that we should be using a far different way to deal with inappropriate articles. I’ve even helped redesign it a bit.</p>
<p>Trace ethnography is based in the realization that these practices around metadata are learned literacies and constitute a crucial part of what it means to participate in many communities and organizations. It turns our attention to an ethnographic understanding of these practices as they make sense for the people who rely on them. In this approach, reading through log data can be seen as a form of participation, not just observation — if and only if this is how members themselves spend their time. However, it is crucial that this approach is distinguished from more passive forms of ethnography (such as “lurker ethnography”), as trace ethnography involves an ethnographer’s socialization into a group prior to the ability to decode and interpret trace data. If trace data is simply being automatically generated without it being integrated into people’s practices of participation, if people in a community don’t regularly rely on following traces in their everyday practices, then the “ethnography” label is likely not appropriate.</p>
<p>Looking at all kinds of online communities and mediated organizations, Wikipedia’s deletion process might appear to be the most arcane and out-of-the-ordinary. However, modes of participation are increasingly linked to the encoding and decoding of trace data, whether that is a global scientific collaboration, an open source software project, a guild of role playing gamers, an activist network, a news organization, a governmental agency, and so on. Computer programmers frequently rely on GitHub to collaborate, and they have their own ways of using things like issues, commit comments, and pull requests to interact with each other. Without being on GitHub, it’s hard for an ethnographer who studies software development to be a fully-immersed participant-observer, because they would be missing a substantial amount of activity — even if they are constantly in the same room as the programmers.</p>
<p><strong>More about trace ethnography</strong></p>
<p>If you want to read more about “trace ethnography,” we first used this term in “<a href="http://www.stuartgeiger.com/papers/cscw-sustaining-order-wikipedia.pdf">The Work of Sustaining Order in Wikipedia: The Banning of a Vandal</a>,” which I co-authored with my then-advisor David Ribes in the proceedings of the CSCW 2010 conference. We then wrote <a href="http://www.stuartgeiger.com/trace-ethnography-hicss-geiger-ribes.pdf">a followup paper in the proceedings of HICSS 2011</a> to give a more general introduction to this method, in which we ‘inverted’ the CSCW 2011 paper, explaining more of the methods we used. We also held a workshop at the 2015 iConference with Amelia Acker and Matt Burton — the details of that workshop (and the collaborative notes) can be found at<a href="http://trace-ethnography.github.io/">http://trace-ethnography.github.io</a>.</p>
<p><b>Some examples of projects employing this method:</b></p>
<p>Ford, H. and Geiger, R.S. “Writing up rather than writing down: Becoming Wikipedia literate.” <i>Proceedings of the Eighth Annual International Symposium on Wikis and Open Collaboration</i>. ACM, 2012. <a href="http://www.stuartgeiger.com/writing-up-wikisym.pdf">http://www.stuartgeiger.com/writing-up-wikisym.pdf</a></p>
<p>Ribes, D., Jackson, S., Geiger, R.S., Burton, M., &amp; Finholt, T. (2013). Artifacts that organize: Delegation in the distributed organization. <i>Information and Organization</i>, <i>23</i>(1), 1-14. <a href="http://www.stuartgeiger.com/artifacts-that-organize.pdf">http://www.stuartgeiger.com/artifacts-that-organize.pdf</a></p>
<p>Mugar, G., Østerlund, C., Hassman, K. D., Crowston, K., &amp; Jackson, C. B. (2014). Planet hunters and seafloor explorers: legitimate peripheral participation through practice proxies in online citizen science. In<i>Proceedings of the 17th ACM conference on Computer supported cooperative work &amp; social computing</i> (pp. 109-119). ACM. <a href="http://dl.acm.org/citation.cfm?id=2531721">http://dl.acm.org/citation.cfm?id=2531721</a></p>
<p>Howison, J., &amp; Crowston, K. (2014). Collaboration Through Open Superposition: A Theory of the Open Source Way. <i>Mis Quarterly</i>, <i>38</i>(1), 29-50. <a href="http://aisel.aisnet.org/cgi/viewcontent.cgi?article=3156&amp;context=misq">http://aisel.aisnet.org/cgi/viewcontent.cgi?article=3156&amp;context=misq</a></p>
<p>Burton, M. (2015). Blogs as Infrastructure for Scholarly Communication. Doctoral Dissertation, University of Michigan.<a href="http://deepblue.lib.umich.edu/bitstream/handle/2027.42/111592/mcburton_1.pdf">http://deepblue.lib.umich.edu/bitstream/handle/2027.42/111592/mcburton_1.pdf</a></p>
]]></content:encoded>
			<wfw:commentRss>http://stuartgeiger.com/wordpress/2016/03/trace-ethnography-a-retrospective/feed/</wfw:commentRss>
		<slash:comments>0</slash:comments>
		</item>
		<item>
		<title>Come to the Trace Ethnography workshop at the 2015 iConference!</title>
		<link>http://stuartgeiger.com/wordpress/2015/01/register-to-the-trace-ethnography-workshop-at-the-2015-iconference/</link>
		<comments>http://stuartgeiger.com/wordpress/2015/01/register-to-the-trace-ethnography-workshop-at-the-2015-iconference/#respond</comments>
		<pubDate>Thu, 08 Jan 2015 16:23:09 +0000</pubDate>
		<dc:creator><![CDATA[stuart]]></dc:creator>
				<category><![CDATA[Uncategorized]]></category>
		<category><![CDATA[community]]></category>
		<category><![CDATA[conference]]></category>
		<category><![CDATA[ethnography]]></category>
		<category><![CDATA[qualitative]]></category>
		<category><![CDATA[research]]></category>
		<category><![CDATA[trace data]]></category>
		<category><![CDATA[virtual]]></category>

		<guid isPermaLink="false">http://stuartgeiger.com/wordpress/?p=726</guid>
		<description><![CDATA[We&#8217;re organizing a workshop on trace ethnography at the 2015 iConference, led by Amelia Acker, Matt Burton, David Ribes, and myself. See more information about it on the workshop&#8217;s website, or feel free to contact me for more information. Date: March 24th 2015, 9:00-a.m.-5:00 p.m. Location: iConference venue, Newport Beach Marriott Hotel &#38; Spa, Newport Beach, CA Deadline to register through this form: Feb 1st, 2015. Note: you will also have to register through the official iConference website as well. Notification: Feb 15th, 2015 Description: This workshop introduces participants to trace ethnography, building a network of scholars interested in the collection and interpretation of trace data and distributed documentary practices. The intended audience is broad, and participants need not have any existing experience working with trace data from either qualitative or quantitative approaches. The workshop provides an interactive introduction to the background, theories, methods, and applications–present and future–of trace ethnography. Participants with more experience in this area will demonstrate how they apply these techniques in their own research, discussing various issues as they arise. The workshop is intended to help researchers identify documentary traces, plan for their collection and analysis, and further formulate trace ethnography as it is currently conceived. In all, this workshop will support the advancement of boundaries, theories, concepts, and applications in trace ethnography, identifying the diversity of approaches that can be assembled around the idea of ‘trace ethnography’ within the iSchool community.]]></description>
				<content:encoded><![CDATA[<p>We&#8217;re organizing a workshop on trace ethnography at the 2015 iConference, led by Amelia Acker, Matt Burton, David Ribes, and myself. See more information about it on<a href="http://trace-ethnography.github.io/"> the workshop&#8217;s website</a>, or feel free to contact me for more information.</p>
<p>Date: March 24th 2015, 9:00-a.m.-5:00 p.m.</p>
<p>Location: iConference venue, Newport Beach Marriott Hotel &amp; Spa, Newport Beach, CA</p>
<p>Deadline to register <a style="font-weight: inherit; font-style: inherit; color: #007edf;" href="http://goo.gl/forms/Cc4G1ULyXv">through this form</a>: Feb 1st, 2015. Note: you will also have to register through the <a style="font-weight: inherit; font-style: inherit; color: #007edf;" href="http://ischools.org/the-iconference/">official iConference website</a> as well.</p>
<p>Notification: Feb 15th, 2015</p>
<p>Description: This workshop introduces participants to trace ethnography, building a network of scholars interested in the collection and interpretation of trace data and distributed documentary practices. The intended audience is broad, and participants need not have any existing experience working with trace data from either qualitative or quantitative approaches. The workshop provides an interactive introduction to the background, theories, methods, and applications–present and future–of trace ethnography. Participants with more experience in this area will demonstrate how they apply these techniques in their own research, discussing various issues as they arise. The workshop is intended to help researchers identify documentary traces, plan for their collection and analysis, and further formulate trace ethnography as it is currently conceived. In all, this workshop will support the advancement of boundaries, theories, concepts, and applications in trace ethnography, identifying the diversity of approaches that can be assembled around the idea of ‘trace ethnography’ within the iSchool community.</p>
]]></content:encoded>
			<wfw:commentRss>http://stuartgeiger.com/wordpress/2015/01/register-to-the-trace-ethnography-workshop-at-the-2015-iconference/feed/</wfw:commentRss>
		<slash:comments>0</slash:comments>
		</item>
		<item>
		<title>A dynamically-generated robots.txt: will search engine bots recognize themselves?</title>
		<link>http://stuartgeiger.com/wordpress/2014/05/robots-txt/</link>
		<comments>http://stuartgeiger.com/wordpress/2014/05/robots-txt/#comments</comments>
		<pubDate>Wed, 14 May 2014 01:37:02 +0000</pubDate>
		<dc:creator><![CDATA[stuart]]></dc:creator>
				<category><![CDATA[Blog Posts]]></category>
		<category><![CDATA[algorithms]]></category>
		<category><![CDATA[bots]]></category>
		<category><![CDATA[communication]]></category>
		<category><![CDATA[discourse]]></category>
		<category><![CDATA[infrastructure]]></category>
		<category><![CDATA[internet]]></category>
		<category><![CDATA[philosophy]]></category>
		<category><![CDATA[technology]]></category>

		<guid isPermaLink="false">http://stuartgeiger.com/wordpress/?p=653</guid>
		<description><![CDATA[In short, I built a script that dynamically generates a robots.txt file for search engine bots, who download the file when they seek direction on what parts of a website they are allowed to index. By default, it directs all bots to stay away from the entire site, but then presents an exception: only the bot that requests the robots.txt file is allowed full reign over the site. If Google&#8217;s bot downloads the robots.txt file, it will see that only Google&#8217;s bot gets to index the entire site. If Yahoo&#8217;s bot downloads the robots.txt file, it will see that only Yahoo&#8217;s bot gets to index the entire site. Of course, this is assuming that bots identify themselves to my server in a way that they recognize when it is reflected back to them. What is a robots.txt file? Most websites have one of these very simple file called &#8220;robots.txt&#8221; on the main directory of their server. The robots.txt file has been around for almost two decades, and it is now a standardized way of communicating what pages search engine bots (or crawlers) should and should not visit. Crawlers are supposed to request and download a robots.txt file from any website they visit, and then obey the directives mentioned in such a file. Of course, there is nothing which prevents a crawler from still crawling pages which are forbidden in a robots.txt file, but most major search engine bots behave themselves.  In many ways, robots.txt files stand out as a legacy from a much earlier time. When was the last time you wrote something for public distribution in a .txt file, anyway? In an age of server-side scripting and content management systems, robots.txt is also one the few public-facing files a systems administrator will actually edit and maintain by hand, manually adding and removing entries in a text editor. A robots.txt file has no changelog in it, but its revision history would be a partial chronicle of a systems administrator&#8217;s interactions with how their website is represented by various search engines.You can specify different directives for different bots by specifying a user agent, and well-behaved bots are supposed to look for their own user agents in a robots.txt file and follow the instructions left for them. As for my own, I&#8217;m sad to report that I simply let all bots through wherever they roam, as I use a sitemap.tar.gz file which a WordPress plugin generates for me on a regular basis and submits to&#8230; ]]></description>
				<content:encoded><![CDATA[<p>In short, I built a script that dynamically generates a robots.txt file for search engine bots, who download the file when they seek direction on what parts of a website they are allowed to index. By default, it directs all bots to stay away from the entire site, but then presents an exception: only the bot that requests the robots.txt file is allowed full reign over the site. If Google&#8217;s bot downloads the robots.txt file, it will see that only Google&#8217;s bot gets to index the entire site. If Yahoo&#8217;s bot downloads the robots.txt file, it will see that only Yahoo&#8217;s bot gets to index the entire site. Of course, this is assuming that bots identify themselves to my server in a way that they recognize when it is reflected back to them.</p>
<p><span id="more-653"></span></p>
<p><span style="color: #292f33;">What is a robots.txt file? Most websites have one of these very simple file called &#8220;robots.txt&#8221; on the main directory of their server. The robots.txt file has been around for almost two decades, and it is now a standardized way of communicating what pages search engine bots (or crawlers) should and should not visit. Crawlers are supposed to request and download a robots.txt file from any website they visit, and then obey the directives mentioned in such a file. Of course, there is nothing which prevents a crawler from still crawling pages which are forbidden in a robots.txt file, but most major search engine bots behave themselves. </span></p>
<p><span style="color: #292f33;">In many ways, robots.txt files stand out as a legacy from a much earlier time. When was the last time you wrote something for public distribution in a .txt file, anyway? In an age of server-side scripting and content management systems, robots.txt is also one the few public-facing files a systems administrator will actually edit and maintain by hand, manually adding and removing entries in a text editor. A robots.txt file has no changelog in it, but its revision history would be a partial chronicle of a systems administrator&#8217;s interactions with how their website is represented by various search engines.</span><span style="color: #292f33;">You can specify different directives for different bots by specifying a user agent, and well-behaved bots are supposed to look for their own user agents in a robots.txt file and follow the instructions left for them. </span>As for my own, I&#8217;m sad to report that I simply let all bots through wherever they roam, as I use a sitemap.tar.gz file which a WordPress plugin generates for me on a regular basis and submits to the major search engines. So my robots.txt file just looks like this:</p>
<pre><span style="color: #000000;">User-agent: *
</span>Allow: /</pre>
<p><span style="color: #292f33;">An interesting thing about contemporary web servers is that file formats no longer really matter as much as they used to. In fact, files don&#8217;t even have to exist as we they are typically represented in URLs. When your browser requests the page http://stuartgeiger.com/wordpress/2014/05/robots-txt, there is a directory called &#8220;wordpress&#8221; on my server, but everything after that is a fiction. There is no directory called 2014, no a subdirectory called 05, and no file called robots-txt that existed on the server before or after you downloaded it. Rather, when WordPress receives a request to download this non-existent file, it intercepts it and interprets it as a request to dynamically generate a new HTML page on the fly. WordPress queries a database for the content of the post, inserts that into a theme, and then has the server send you that HTML page &#8212; with linked images, stylesheets, and Javascript files, which often do actually exist as files on a server. The server probably stores the dynamically-generated HTML page in its memory, and sometimes there is caching to pre-generate these pages to make things faster, but other than that, the only time an HTML file of this page ever exists in any persistent form is if you save it to your hard drive. </span></p>
<p><span style="color: #292f33;">Yet robots.txt lives on, doing its job well. It doesn&#8217;t need any fancy server-side scripting; it does just fine on its own. Still, I kept thinking about what it would be like to have a script dynamically generate a robots.txt file on the fly whenever it is requested. Given that the only time a robots.txt file is usually downloaded is when an automated software agent requests it, there is something strangely poetic about an algorithmically-generated robots.txt file. It is something that would, for the most part, only ever really exist in the fleeting interaction between two automated routines. So of course I had to build one.</span></p>
<p>The code required to implement this is trivial. First, I needed to modify how my web server interprets requests, so that whenever a request was made to robots.txt, the server would execute a script called robots.php and send the client the output as robots.txt. Modify the .htaccess file to add:</p>
<pre><span style="color: #000000;">RewriteEngine On
RewriteBase /
RewriteRule ^robots.txt$ /robots.php</span></pre>
<p>Next, the PHP script itself:</p>
<pre>&lt;?php
header('Content-Type:text/plain');
echo "User-agent: *" . "\r\n";
echo "Allow: /" . "\r\n";
?&gt;
</pre>
<p>Then I realized that this was all a little impersonal, and I could do better since I&#8217;m scripting. With PHP, I can easily query the user-agent of the client which is requesting the file, the identifier it sends to the web server. Normally, user agents define the browser that is requesting the page, but bots are supposed to have an identifiable user-agent like &#8220;Googlebot&#8221; or &#8220;Twitterbot&#8221; so that you can know them when they come to visit. Instead of granting access to every user agent with the asterisk, I made it so that the user agent of the requesting client is the only one that is directed to have full access.</p>
<pre>&lt;?php
header('Content-Type:text/plain');
echo "User-agent:" . $_SERVER['HTTP_USER_AGENT'] . "\r\n";
echo "Allow: /" . "\r\n";
?&gt;</pre>
<p>After making sure this worked, I realized that I needed to go out there a little more. If the bots didn&#8217;t recognize themselves, then by default, they would still be allowed to crawl the site anyway. robots.txt works on a principle of allow by default. So I needed to add a few more lines which made it so that the robots.txt file the bot downloaded would direct all <strong>other</strong> bots to <strong>not</strong> crawl the site, but give full reign to bots with the user agent it sent the server.</p>
<pre>&lt;?php
 header('Content-Type:text/plain');
 echo "User-agent: *" . "\r\n";
 echo "Disallow: /" . "\r\n";
 echo "User-agent:" . $_SERVER['HTTP_USER_AGENT'] . "\r\n";
 echo "Allow: /" . "\r\n";
 ?&gt;</pre>
<p>This is what you get if you download it in Chrome:</p>
<pre>User-agent: *
Disallow: /
User-agent: Mozilla/5.0 (Windows NT 6.3; WOW64) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/34.0.1847.131 Safari/537.36 
Allow: /</pre>
<p>The restrictive version is now live, up at <a href="http://www.stuartgeiger.com/robots.txt">http://www.stuartgeiger.com/robots.txt</a>. I&#8217;ve also put it up <a href="https://github.com/staeiou/robots.txt.php">on github</a>, because apparently that&#8217;s what cool kids do. I&#8217;m looking forward to seeing what will happen. Google&#8217;s webmaster tools will notify me if its crawlers can&#8217;t index my site, for whatever reason, and I&#8217;m curious if Google&#8217;s bots will identify themselves to my servers in a way that they will recognize.</p>
]]></content:encoded>
			<wfw:commentRss>http://stuartgeiger.com/wordpress/2014/05/robots-txt/feed/</wfw:commentRss>
		<slash:comments>1</slash:comments>
		</item>
		<item>
		<title>Bots, bespoke code, and the materiality of software platforms</title>
		<link>http://stuartgeiger.com/wordpress/2014/01/bots-bespoke-code-and-the-materiality-of-software-platforms/</link>
		<comments>http://stuartgeiger.com/wordpress/2014/01/bots-bespoke-code-and-the-materiality-of-software-platforms/#respond</comments>
		<pubDate>Mon, 06 Jan 2014 21:22:54 +0000</pubDate>
		<dc:creator><![CDATA[stuart]]></dc:creator>
				<category><![CDATA[Academic Works]]></category>
		<category><![CDATA[bots]]></category>
		<category><![CDATA[discourse]]></category>
		<category><![CDATA[governance]]></category>
		<category><![CDATA[infrastructure]]></category>
		<category><![CDATA[internet]]></category>
		<category><![CDATA[power]]></category>
		<category><![CDATA[software]]></category>
		<category><![CDATA[technology]]></category>
		<category><![CDATA[wikipedia]]></category>

		<guid isPermaLink="false">http://stuartgeiger.com/wordpress/?p=564</guid>
		<description><![CDATA[This is a new article published in Information, Communication, and Society as part of their annual special issue for the Association of Internet Researchers (AoIR) conference. This year&#8217;s special issue was edited by Lee Humphreys and Tarleton Gillespie, who did a great job throughout the whole process. Abstract: This article introduces and discusses the role of bespoke code in Wikipedia, which is code that runs alongside a platform or system, rather than being integrated into server-side codebases by individuals with privileged access to the server. Bespoke code complicates the common metaphors of platforms and sovereignty that we typically use to discuss the governance and regulation of software systems through code. Specifically, the work of automated software agents (bots) in the operation and administration of Wikipedia is examined, with a focus on the materiality of code. As bots extend and modify the functionality of sites like Wikipedia, but must be continuously operated on computers that are independent from the servers hosting the site, they involve alternative relations of power and code. Instead of taking for granted the pre-existing stability of Wikipedia as a platform, bots and other bespoke code require that we examine not only the software code itself, but also the concrete, historically contingent material conditions under which this code is run. To this end, this article weaves a series of autobiographical vignettes about the author&#8217;s experiences as a bot developer alongside more traditional academic discourse. Official version at Information, Communication, and Society Author&#8217;s post-print, free download [PDF, 382kb] &#160; &#160;]]></description>
				<content:encoded><![CDATA[<p>This is a new article published in <em>Information, Communication, and Society</em> as part of their annual special issue for the Association of Internet Researchers (AoIR) conference. This year&#8217;s special issue was edited by Lee Humphreys and Tarleton Gillespie, who did a great job throughout the whole process.</p>
<p>Abstract: This article introduces and discusses the role of <i>bespoke code</i> in Wikipedia, which is code that runs alongside a platform or system, rather than being integrated into server-side codebases by individuals with privileged access to the server. Bespoke code complicates the common metaphors of platforms and sovereignty that we typically use to discuss the governance and regulation of software systems through code. Specifically, the work of automated software agents (bots) in the operation and administration of Wikipedia is examined, with a focus on the materiality of code. As bots extend and modify the functionality of sites like Wikipedia, but must be continuously operated on computers that are independent from the servers hosting the site, they involve alternative relations of power and code. Instead of taking for granted the pre-existing stability of Wikipedia as a platform, bots and other bespoke code require that we examine not only the software code itself, but also the concrete, historically contingent material conditions under which this code is run. To this end, this article weaves a series of autobiographical vignettes about the author&#8217;s experiences as a bot developer alongside more traditional academic discourse.</p>
<p><a href="http://www.tandfonline.com/doi/full/10.1080/1369118X.2013.873069" target="_blank">Official version at Information, Communication, and Society</a></p>
<p><a href="http://stuartgeiger.com/bespoke-code-ics.pdf" target="_blank">Author&#8217;s post-print, free download [PDF, 382kb]</a></p>
<p>&nbsp;</p>
<p>&nbsp;</p>
]]></content:encoded>
			<wfw:commentRss>http://stuartgeiger.com/wordpress/2014/01/bots-bespoke-code-and-the-materiality-of-software-platforms/feed/</wfw:commentRss>
		<slash:comments>0</slash:comments>
		</item>
		<item>
		<title>When the Levee Breaks: Without Bots, What Happens to Wikipedia’s Quality Control Processes?</title>
		<link>http://stuartgeiger.com/wordpress/2013/08/when-the-levee-breaks-without-bots-what-happens-to-wikipedias-quality-control-processes/</link>
		<comments>http://stuartgeiger.com/wordpress/2013/08/when-the-levee-breaks-without-bots-what-happens-to-wikipedias-quality-control-processes/#respond</comments>
		<pubDate>Sat, 17 Aug 2013 03:49:42 +0000</pubDate>
		<dc:creator><![CDATA[stuart]]></dc:creator>
				<category><![CDATA[Academic Works]]></category>
		<category><![CDATA[bots]]></category>
		<category><![CDATA[governance]]></category>
		<category><![CDATA[infrastructure]]></category>
		<category><![CDATA[quantitative]]></category>
		<category><![CDATA[software]]></category>
		<category><![CDATA[technology]]></category>
		<category><![CDATA[wikipedia]]></category>

		<guid isPermaLink="false">http://stuartgeiger.com/wordpress/?p=534</guid>
		<description><![CDATA[I&#8217;ve written a number of papers about the role that automated software agents (or bots) play in Wikipedia, claiming that they are critical to the continued operation of Wikipedia. This paper tests this hypothesis and introduces a metric visualizing the speed at which fully-automated bots, tool-assisted cyborgs, and unassisted humans review edits in Wikipedia. In the first half of 2011, ClueBot NG – one of the most prolific counter-vandalism bots in the English-language Wikipedia – went down for four distinct periods, each period of downtime lasting from days to weeks. Aaron Halfaker and I use these periods of breakdown as naturalistic experiments to study Wikipedia’s quality control network. Our analysis showed that the overall time-to-revert damaging edits was almost doubled when this software agent was down. Yet while a significantly fewer proportion of edits made during the bot’s downtime were reverted, we found that those edits were later eventually reverted. This suggests that human agents in Wikipedia took over this quality control work, but performed it at a far slower rate.]]></description>
				<content:encoded><![CDATA[<p>I&#8217;ve written a number of papers about the role that automated software agents (or bots) play in Wikipedia, claiming that they are critical to the continued operation of Wikipedia. <a href="http://stuartgeiger.com/wikisym13-cluebot.pdf">This paper</a> tests this hypothesis and introduces a metric visualizing the speed at which fully-automated bots, tool-assisted cyborgs, and unassisted humans review edits in Wikipedia. In the first half of 2011, ClueBot NG – one of the most prolific counter-vandalism bots in the English-language Wikipedia – went down for four distinct periods, each period of downtime lasting from days to weeks. Aaron Halfaker and I use these periods of breakdown as naturalistic experiments to study Wikipedia’s quality control network. Our analysis showed that the overall time-to-revert damaging edits was almost doubled when this software agent was down. Yet while a significantly fewer proportion of edits made during the bot’s downtime were reverted, we found that those edits were later eventually reverted. This suggests that human agents in Wikipedia took over this quality control work, but performed it at a far slower rate.</p>
<p><a href="http://stuartgeiger.com/wordpress/2013/09/when-the-levee-breaks-without-bots-what-happens-to-wikipedias-quality-control-processes/revert/" rel="attachment wp-att-535"><img class="size-full wp-image-535 aligncenter" src="http://stuartgeiger.com/wordpress/wp-content/uploads/2013/09/revert.png" alt="revert" width="480" height="524" srcset="http://stuartgeiger.com/wordpress/wp-content/uploads/2013/09/revert.png 480w, http://stuartgeiger.com/wordpress/wp-content/uploads/2013/09/revert-274x300.png 274w" sizes="(max-width: 480px) 100vw, 480px" /></a></p>
]]></content:encoded>
			<wfw:commentRss>http://stuartgeiger.com/wordpress/2013/08/when-the-levee-breaks-without-bots-what-happens-to-wikipedias-quality-control-processes/feed/</wfw:commentRss>
		<slash:comments>0</slash:comments>
		</item>
		<item>
		<title>About a bot: reflections on building software agents</title>
		<link>http://stuartgeiger.com/wordpress/2013/08/about-a-bot-reflections-on-building-software-agents/</link>
		<comments>http://stuartgeiger.com/wordpress/2013/08/about-a-bot-reflections-on-building-software-agents/#respond</comments>
		<pubDate>Tue, 13 Aug 2013 04:22:22 +0000</pubDate>
		<dc:creator><![CDATA[stuart]]></dc:creator>
				<category><![CDATA[Uncategorized]]></category>
		<category><![CDATA[bots]]></category>
		<category><![CDATA[ethnography]]></category>
		<category><![CDATA[infrastructure]]></category>
		<category><![CDATA[materiality]]></category>
		<category><![CDATA[philosophy]]></category>
		<category><![CDATA[software]]></category>
		<category><![CDATA[technology]]></category>
		<category><![CDATA[wikipedia]]></category>

		<guid isPermaLink="false">http://stuartgeiger.com/wordpress/?p=550</guid>
		<description><![CDATA[This post for Ethnography Matters is a very personal, reflective musing about the first bot I ever developed for Wikipedia. It makes the argument that while it is certainly important to think about software code and algorithms behind bots and other AI agents, they are not immaterial. In fact, the physical locations and social contexts in which they are run are often critical to understanding how they both &#8216;live&#8217; and &#8216;die&#8217;.]]></description>
				<content:encoded><![CDATA[<p><a href="ttp://ethnographymatters.net/2013/08/13/about-a-bot/">This post for Ethnography Matters</a> is a very personal, reflective musing about the first bot I ever developed for Wikipedia. It makes the argument that while it is certainly important to think about software code and algorithms behind bots and other AI agents, they are not immaterial. In fact, the physical locations and social contexts in which they are run are often critical to understanding how they both &#8216;live&#8217; and &#8216;die&#8217;.</p>
]]></content:encoded>
			<wfw:commentRss>http://stuartgeiger.com/wordpress/2013/08/about-a-bot-reflections-on-building-software-agents/feed/</wfw:commentRss>
		<slash:comments>0</slash:comments>
		</item>
		<item>
		<title>Bots and Cyborgs: Wikipedia&#8217;s Immune System</title>
		<link>http://stuartgeiger.com/wordpress/2012/10/bots-and-cyborgs-wikipedias-immune-system/</link>
		<comments>http://stuartgeiger.com/wordpress/2012/10/bots-and-cyborgs-wikipedias-immune-system/#respond</comments>
		<pubDate>Wed, 17 Oct 2012 04:13:31 +0000</pubDate>
		<dc:creator><![CDATA[stuart]]></dc:creator>
				<category><![CDATA[Uncategorized]]></category>
		<category><![CDATA[bots]]></category>
		<category><![CDATA[governance]]></category>
		<category><![CDATA[information]]></category>
		<category><![CDATA[infrastructure]]></category>
		<category><![CDATA[power]]></category>
		<category><![CDATA[wikipedia]]></category>

		<guid isPermaLink="false">http://stuartgeiger.com/wordpress/?p=544</guid>
		<description><![CDATA[My frequent collaborator Aaron Halfaker has written up a fantastic article with John Riedl in Computer reviewing a lot of the work we&#8217;ve done on algorithmic agents in Wikipedia, casting them as Wikipedia&#8217;s immune system. Choice quote:  &#8220;These bots and cyborgs are more than tools to better manage content quality on Wikipedia—through their interaction with humans, they’re fundamentally changing its culture.&#8221;]]></description>
				<content:encoded><![CDATA[<p>My frequent collaborator Aaron Halfaker has written up a <a href="http://stuartgeiger.com/bots-cyborgs-halfaker.pdf" target="_blank">fantastic article</a> with John Riedl in <em>Computer</em> reviewing a lot of the work we&#8217;ve done on algorithmic agents in Wikipedia, casting them as Wikipedia&#8217;s immune system. Choice quote:  &#8220;These bots and cyborgs are more than tools to better manage content quality on Wikipedia—through their interaction with humans, they’re fundamentally changing its culture.&#8221;</p>
]]></content:encoded>
			<wfw:commentRss>http://stuartgeiger.com/wordpress/2012/10/bots-and-cyborgs-wikipedias-immune-system/feed/</wfw:commentRss>
		<slash:comments>0</slash:comments>
		</item>
		<item>
		<title>An apologia for instagram photos of pumpkin spice lattes and other serious things</title>
		<link>http://stuartgeiger.com/wordpress/2012/09/on-instagram-photos-of-pumpkin-spice-lattes-and-other-serious-things/</link>
		<comments>http://stuartgeiger.com/wordpress/2012/09/on-instagram-photos-of-pumpkin-spice-lattes-and-other-serious-things/#respond</comments>
		<pubDate>Sun, 09 Sep 2012 21:10:09 +0000</pubDate>
		<dc:creator><![CDATA[stuart]]></dc:creator>
				<category><![CDATA[Blog Posts]]></category>
		<category><![CDATA[communication]]></category>
		<category><![CDATA[community]]></category>
		<category><![CDATA[context collapse]]></category>
		<category><![CDATA[discourse]]></category>
		<category><![CDATA[facebook]]></category>
		<category><![CDATA[instagram]]></category>
		<category><![CDATA[internet]]></category>
		<category><![CDATA[photography]]></category>
		<category><![CDATA[pumpkin spice lattes]]></category>
		<category><![CDATA[social media]]></category>
		<category><![CDATA[timeline]]></category>

		<guid isPermaLink="false">http://stuartgeiger.com/wordpress/?p=480</guid>
		<description><![CDATA[I don&#8217;t normally pick on people whose work I really admire, but I recently saw a tweet from Mark Sample that struck a nerve: &#8220;Look, if you don&#8217;t instagram your first pumpkin spice latte of the season, humanity&#8217;s historical record will be dangerously impoverished.&#8221;  While it got quite a number of retweets and equally snarky responses, he is far from the first to make such a flippant critique of the vapid nature of social media.  It also seriously upset me for reasons that I&#8217;ve been trying to work out, which is why I found myself doing one of those shifts that researchers of knowledge production tend to do far too often with critics: don&#8217;t get mad, get reflexive.  What is it that makes such a sentiment resonate with us, particularly when it is issued over Twitter, a platform that is the target of this kind of critique?  The reasons have to do with a fundamental disagreement over what it means to interact in a mediated space: do we understand our posts, status updates, and shared photos as representations of how we exist in the world which collectively constitute a certain persistent performance of the self, or do we understand them a form of communication in which we subjectively and interactionally relate our experience of the world to others? This comment also got me thinking because it reminded me of an interaction I had at Media in Transition 6, the first major conference I attended as a presenter.  The year&#8217;s theme was &#8220;storage and transmission,&#8221; and there were a lot of well-established scholars from a variety of fields talking about social media in terms of archives and memory practices.  I remember one discussion where people were talking about how exciting it was see the widespread emergence of Facebook photo albums, arguing that youth who share photos on Facebook were engaging in the 21st century equivalent of scrapbooking – a once-common cultural practice which had been in serious decline.  I raised my hand and made a comment I&#8217;m not sure was fully grasped: that as one of the youngest people in the room, my friends and I understood photo sharing not a form of archiving but a mode of communication.  In other words, I take a photo of the MIT Media Lab and share it on Facebook primarily to tell my friends that I&#8217;m in Boston at a conference.  Sure, there is archival value to this kind&#8230; ]]></description>
				<content:encoded><![CDATA[<p>I don&#8217;t normally pick on people whose work I really admire, but I recently saw <a href="https://twitter.com/samplereality/status/244151842974609408">a tweet</a> from Mark Sample that struck a nerve: &#8220;Look, if you don&#8217;t instagram your first pumpkin spice latte of the season, humanity&#8217;s historical record will be dangerously impoverished.&#8221;  While it got quite a number of retweets and equally snarky responses, he is far from the first to make such a flippant critique of the vapid nature of social media.  It also seriously upset me for reasons that I&#8217;ve been trying to work out, which is why I found myself doing one of those shifts that researchers of knowledge production tend to do far too often with critics: don&#8217;t get mad, get reflexive.  What is it that makes such a sentiment resonate with us, particularly when it is issued over Twitter, a platform that is the target of this kind of critique?  The reasons have to do with a fundamental disagreement over what it means to interact in a mediated space: do we understand our posts, status updates, and shared photos as representations of how we exist in the world which collectively constitute a certain persistent performance of the self, or do we understand them a form of communication in which we subjectively and interactionally relate our experience of the world to others?</p>
<p><span id="more-480"></span></p>
<p>This comment also got me thinking because it reminded me of an interaction I had at Media in Transition 6, the first major conference I attended as a presenter.  The year&#8217;s theme was &#8220;storage and transmission,&#8221; and there were a lot of well-established scholars from a variety of fields talking about social media in terms of archives and memory practices.  I remember one discussion where people were talking about how exciting it was see the widespread emergence of Facebook photo albums, arguing that youth who share photos on Facebook were engaging in the 21st century equivalent of scrapbooking – a once-common cultural practice which had been in serious decline.  I raised my hand and made a comment I&#8217;m not sure was fully grasped: that as one of the youngest people in the room, my friends and I understood photo sharing not a form of archiving but a mode of communication.  In other words, I take a photo of the MIT Media Lab and share it on Facebook primarily to tell my friends that I&#8217;m in Boston at a conference.  Sure, there is archival value to this kind of activity, but that is an added benefit which we occasionally utilize – and always after the fact.  We don&#8217;t take a picture to remember an event and then later remember that event.  We take a picture to communicate an event, later remembering a strange hybrid of the event itself and all the interactions we had about the event.  This is especially the case with something like Facebook&#8217;s timeline: instead of carefully assembling scrapbooks ourselves, we have delegated these memory practices to Facebook&#8217;s algorithms.</p>
<p>Returning to instagram photos of pumpkin spice lattes, I admit that as a twentysomething techie-hipster in the Bay Area, I use not just Twitter, but instagram, Tumblr, and a variety of other social media platforms.  I also enjoy pumpkin spice lattes, perhaps because they are delicious, but also because I really do take in all those little things that tell me that summer is ending and autumn will soon begin.  We don&#8217;t have that much seasonal variation in the Bay Area, and coffee is a big deal here as it is everywhere – it is the world&#8217;s most popular drug.  All this to say that the first advertisement for pumpkin spice lattes plastered on the side of a Starbucks is something I notice.  And so I take photos of them, which I share with my friends and strangers.  Some of them are in the Bay Area and have the same seasonal cues I do, while some are in completely different parts of the world, where frozen water falls from the sky and other crazy things like that.  Together, we engage not so much in an act of collective sensemaking, but the sharing of a common experience: thanks to this and a hundred other little reminders, we know that winter is coming.</p>
<p>I don&#8217;t do it because I think I&#8217;m contributing to some grand archive of humanity&#8217;s historical record.  Not even close.  In fact, if that is how I thought about most of my social media practices, I would be so anxious about choosing what to post and when that I wouldn&#8217;t make use of it at all.  I know this because there was a time when I did think of my social media usage in such a way, and that is exactly what happened.  Today, I am self-conscious enough to realize that there are people who would harshly judge me for the fact that I do come to know and understand the changing of the seasons – such a timeless and universal force of &#8216;nature&#8217; that humanity is always subjected to – in part through a multi-national corporation&#8217;s advertising campaign.  So, fearing context collapse, I don&#8217;t publish those same kinds of photos and have those same kinds of interactions in the same place as I publish my academic musings.</p>
<p>Yet the important thing to realize is that in posting these instagram photos of pumpkin spice lattes, I am likely contributing to some grand archive of humanity&#8217;s historical record – or at least there are people who think I am, which is probably even more important for this argument.  In fact, there are uncountably many digital artifacts on the Internet documenting the excitement leading up to everything from the McRib coming back to a new season of Mad Men premiering.  These are the kinds of interactions which are being recorded and increasingly preserved at a startling rate, compared to what kinds of materials we have typically chosen to preserve.  If we as a society preserve them not like members of previous generations individually preserved letters and memorabilia, but instead stored these interactions in massively-indexed digital archives, they will likely be an irresistible resource for future generations of historically-minded humanists and social scientists.  Perhaps this is where the tension lies: it could be that many people don&#8217;t want the records we leave for posterity to be filled with what is certainly not a representative sample of our collective cultural experience.  I somewhat agree with this sentiment, because I know that the people who post the most on these sites are probably some of the least representative of humanity.</p>
<p>However, I must argue that if a future historian (or a contemporary social scientist or humanist) wants to seriously delve into what it is like for a certain segment of the population to be human and experience the world in 2012, they have to understand that they <em>ought</em> to be looking a lot of nearly-identical photos of Starbucks products.  Not because pumpkin spice lattes themselves are such a culturally important phenomenon which reveal so much about the human condition – that&#8217;s completely the wrong way of looking at this.  Rather, the activity of sharing nostalgia-filtered instagram photos of the first pumpkin spice latte of the year is one way in which some members of a globalized, corporate consumer culture collectively experience the changing of the seasons.  If you&#8217;re not a part of a social group that engages in these kinds of practices, then you probably see the stray instagram photo that someone publishes to their Twitter stream as, well, something to be ridiculed.  You also may think that someone who has let a multi-million dollar corporate advertising campaign overcode their experience of nature is also independently deserving of ridicule, which I also disagree with, but that&#8217;s another issue entirely.</p>
<p>On a side note, this &#8216;photography as documentation versus experience&#8217; issue may also be why instagram, with all its filters and frames, gets so much hate. If you&#8217;re a photographic realist and understand photo sharing as a way of documenting the present world for an other who is not present in time and/or space, then those silly filters and frames seriously invalidate a core assumption behind such a practice.  However, if you instead understand photo sharing as a mode of communication in which we seek to not so much objectively document the external world<em> for others</em> as subjectively express our experience <em>with others, </em>then filters and frames are probably one of the most innovative &#8216;features&#8217; added to the social activity that is photography since the caption.</p>
<p>This is also where I disagree with the critiques of photography from theorists like Barthes and Sontag, or more accurately, I think their critiques are only specific to the kinds of photo sharing practices which were prevalent in their time.  A photojournalist who waits for days to take an unrepresentative snapshot of a war zone is doing a completely different kind of &#8216;manipulation&#8217; than someone who adds a washed-out filter to a smartphone photo of an empty street so that it more accurately conveys the dreariness they feel.  Sure, I&#8217;ll be the first to admit that instagram filters are also so prevalent because they enforce an aesthetic field in which almost any photo – even those blurry, overexposed shots quickly taken in poor lighting with crappy smartphone cameras – can be made to look &#8220;good.&#8221;  But that only strengthens my point: &#8220;serious&#8221; photographers who see instagram as a platform for collectively engaging in a centuries-old craft in which the world is captured onto a fixed medium don&#8217;t get that it is actually a platform for collectively engaging in a much older craft: conversation and storytelling.  In fact, I see these critiques as essentially the same ones Plato had of writing and rhetoric: How dare you make it easier for people to competently relate their experiences in a way that has meaning to themselves and the people around them!?!</p>
<p>Those who study youth and social networking practices should already know that this entire issue is one of context collapse, but it is a more expanded case than the standard media narrative about college students posting wild photos that their parents or potential employers can see.  The issue is usually framed as stemming from the need to use the same platform to interact with multiple, overlapping, simultaneously-existing social worlds that hold different values about what is acceptable behavior and what is not.  However, I think that both of these cases also arise from a much less-discussed disagreement regarding the way in which participation in social networking sites is understood: When I share a photo of a party I attend, am I objectively documenting an event that happened to me, recording what took place so that my social network – including those people who I later friend on Facebook – can go through my profile and see how I&#8217;ve always been a cool party-goer?  Or am I sharing the photo to the people I currently interact with on Facebook, communicating that I was just at a fun party not as something that will stand on its own for all time, but instead something to serve as the basis for a conversation?  Either way, I will have to deal with the standard context collapse issues about how I should act in a social space where people from different social worlds are watching me, but this distinction is something more than that.</p>
<p>This issue about the profile as a performance of the self versus the profile as a by-product of interactions seems to be my main frustration with something like Facebook&#8217;s now-mandatory Timeline feature.  Tensions over the rollout of Timeline, which aggregates your entire past on Facebook in an easy-to-read summary of your life, seem to be part of a larger trend that seeks to conflate these two understandings of what it means to engage in social activity online.  And as a side note, it is interesting that Timeline conflates this distinction with code, as opposed to cultural critics who conflate this with discourse.</p>
<p>Anyways, after a terrible context collapse incident as a freshman in college, I like to think that I&#8217;ve always been a savvy Facebook user, self-censoring when I&#8217;m interacting in a space that could in any way be public.  Still, I just spent quite a long time trying to remove as much as I could from my Timeline, not because it contains anything that I would be seriously embarrassed about, but because it doesn&#8217;t represent who I am now in any way.  The people who I was friends with in 2005 aren&#8217;t the same as the people who I&#8217;m friends with in 2012, the things that mattered to me aren&#8217;t the same, the photos of me look nothing like I do now, and so on.</p>
<p>Especially because there weren&#8217;t that many ways of interacting as there are now, since 2005 I have understood and used my Facebook profile as a carefully-curated representation of myself, working hard to remove those little interactions about how awesome last night was after they served their immediate communicative purposes.  However, that is getting harder and harder to do, which is why I move to other platforms – partially because their code is written in such a way that does not essentialize my interactions to form my profile, but also because the people who I communicate with share my same understanding of what it means to interact in such a space.</p>
]]></content:encoded>
			<wfw:commentRss>http://stuartgeiger.com/wordpress/2012/09/on-instagram-photos-of-pumpkin-spice-lattes-and-other-serious-things/feed/</wfw:commentRss>
		<slash:comments>0</slash:comments>
		</item>
		<item>
		<title>The ethnography of robots: interview at Ethnography Matters</title>
		<link>http://stuartgeiger.com/wordpress/2012/08/the-ethnography-of-robots-interview-at-ethnography-matters/</link>
		<comments>http://stuartgeiger.com/wordpress/2012/08/the-ethnography-of-robots-interview-at-ethnography-matters/#respond</comments>
		<pubDate>Tue, 14 Aug 2012 17:55:17 +0000</pubDate>
		<dc:creator><![CDATA[stuart]]></dc:creator>
				<category><![CDATA[Uncategorized]]></category>
		<category><![CDATA[actor-network theory]]></category>
		<category><![CDATA[ANT]]></category>
		<category><![CDATA[bots]]></category>
		<category><![CDATA[ethnography]]></category>
		<category><![CDATA[latour]]></category>
		<category><![CDATA[network]]></category>
		<category><![CDATA[technology]]></category>

		<guid isPermaLink="false">http://stuartgeiger.com/wordpress/?p=455</guid>
		<description><![CDATA[This was an interview I did with the wonderful Heather Ford, originally posted at Ethnography Matters (a really cool group blog) way back in January. No idea why I didn&#8217;t post a copy of this here back then, but now that I&#8217;m moving towards my dissertation I&#8217;m thinking about this kind of stuff more and more.  In short, I argue for a non-anthropocentric yet still phenomenological ethnography of technology, studying not the culture of the people who build and program robots, but the culture of those the robots themselves. Heather Ford spoke with Stuart Geiger, PhD student at the UC Berkeley School of Information, about his emerging ideas about the ethnography of robots. “Not the ethnography of robotics (e.g. examining the humans who design, build, program, and otherwise interact with robots, which I and others have been doing),” wrote Geiger, “but the ways in which bots themselves relate to the world”. Geiger believes that constructing and relating an emic account of the non-human should be the ultimate challenge for ethnography but that he’s getting an absurd amount of pushback from it.” He explains why in this fascinating account of what it means to study the culture of robots. HF: So, what’s new, almost-Professor Geiger? SG: I just got back from the 4S conference — the annual meeting of the Society for the Social Study of Science — which is pretty much the longstanding home for not just science studies but also Science and Technology Studies. I was in this really interesting session featuring some really cool qualitative studies of robots, including two ethnographies of robotics. One of the presenters, Zara Mirmalek, was looking at the interactions between humans and robots within a modified framework from intercultural communication and workplace studies. I really enjoyed how she was examining robots as co-workers from different cultures, but it seems like most people in the room didn’t fully get it, thinking it was some kind of stretched metaphor. People kept giving her the same feedback that I’ve been given — isn’t there an easier way you can study the phenomena that interest you without attributing culture to robots themselves? But I saw where she was going and asked her about doing ethnographic studies of robot culture itself, instead of the culture of people who interact with robots — and it seemed like half the room gave a polite chuckle. Zara, however, told me that she loved the&#8230; ]]></description>
				<content:encoded><![CDATA[<p>This was an interview I did with the wonderful <a href="http://hblog.org">Heather Ford</a>, <a href="http://ethnographymatters.net/2012/01/15/the-ethnography-of-robots/">originally posted</a> at <a href="http://www.ethnographymatters.com">Ethnography Matters</a> (a really cool group blog) way back in January. No idea why I didn&#8217;t post a copy of this here back then, but now that I&#8217;m moving towards my dissertation I&#8217;m thinking about this kind of stuff more and more.  In short, I argue for a non-anthropocentric yet still phenomenological ethnography of technology, studying not the culture of the people who build and program robots, but the culture of those the robots themselves.</p>
<p><span id="more-455"></span></p>
<p><em> Heather Ford spoke with Stuart Geiger, PhD student at the UC Berkeley School of Information, about his emerging ideas about the ethnography of robots. “Not the ethnography of robotics (e.g. examining the humans who design, build, program, and otherwise interact with robots, which I and others have been doing),” wrote Geiger, “but the ways in which bots themselves relate to the world”. Geiger believes that constructing and relating an emic account of the non-human should be the ultimate challenge for ethnography but that he’s getting an absurd amount of pushback from it.” He explains why in this fascinating account of what it means to study the culture of robots.</em></p>
<p>HF: So, what’s new, almost-Professor Geiger?</p>
<p>SG: I just got back from the 4S conference — the annual meeting of the Society for the Social Study of Science — which is pretty much the longstanding home for not just science studies but also Science and Technology Studies. I was in this really interesting session featuring some really cool qualitative studies of robots, including two ethnographies of robotics. One of the presenters, Zara Mirmalek, was looking at the interactions between humans and robots within a modified framework from intercultural communication and workplace studies.</p>
<p>I really enjoyed how she was examining robots as co-workers from different cultures, but it seems like most people in the room didn’t fully get it, thinking it was some kind of stretched metaphor. People kept giving her the same feedback that I’ve been given — isn’t there an easier way you can study the phenomena that interest you without attributing culture to robots themselves? But I saw where she was going and asked her about doing ethnographic studies of robot culture itself, instead of the culture of people who interact with robots — and it seemed like half the room gave a polite chuckle. Zara, however, told me that she loved the idea and we had a great chat afterwards about this.</p>
<p>HF: What do you think people are upset about?</p>
<p>SG: The more middle-of-the-road stances come from people who don’t personally have a strong reaction either way, but tell me that I’ll have to fight an uphill battle from angry humanists who I’ll talk about later. These people aren’t really against the idea, but they don’t really see the value added in ascribing culture to the non-humans. They tell me that there are better and more non-controversial ways of analyzing, say, distributed cognition in a heterogeneous network of humans and robots. It’s a response that I appreciate, because it would be futile to have to go through all of this work on an ethnography of robots if my analysis is otherwise identical to an ethnography of robotics. And then the most polite responses I get are people who tell me it is interesting, and then when I prod them further to ask them if they actually buy it, they tell me that they don’t *yet* think it can be done, but would like to see what I end up with.</p>
<p>Some of the really negative responses I get involve a visceral reaction against attributing ‘culture’ to the realm of the non-human. I understand this — anthropology is, by definition, anthropocentric: it is concerned with the human condition, as it is constituted in various localities and peoples. This is the same fight we Latourians have with sociologists about the term “agency”: there is a very deeply-rooted assumption that humans have some innate, unique qualities that distinguish us from not only mere matter but other animals as well. When someone comes along and makes a very nuanced point about how objects have agency, the most immediate and natural response is first of all anthropomorphism, which is easy to rebut.</p>
<p>But then comes a much more worthy ontological argument from people who really know their stuff: that when Latour ascribes agency to objects, he actually manages to do so by keeping the agency of humans and the agency of non-humans symmetrical. Against the standard, boring objection that he ascribes too many human characteristics to non-humans, what is really going on is that he accomplishes so much by taking away so many of those ‘uniquely human’ qualities from human agents. This is why Latour never goes inside of anyone’s head, why he rarely tries to give a psychological or cognitive account in the actor-networks he studies. (Read Latour’s review of “Cognition in the Wild” by Ed Hutchins for more on this, and you can see that he loves the idea that these seemingly human abilities like cognition are not pre-given but themselves an effect of a heterogeneous network of humans and non-humans.)</p>
<p>Anyways, far from being an anthropomorphism, Latour’s ontology is flat, in which all entities have the same capacities. That is, they have the same a priori capabilities, but they are definitely not equal after socio-technical relations emerge and start operating. This all means that against the vulgar interpretations of ANT, objects don’t have intentionality or consciousness, because — and this is the really important point — neither do humans. Or, in another interpretation, perhaps humans do have intentionality or consciousness, but it makes no difference one way or another. A good actor-network theorist is able to take some existing system in which there are far too many explanations based on those uniquely human qualities and give an alternative account that relies instead on materials, technologies, infrastructures, documentation, and other modes of externalized practices. It is not to make the more futile argument that norms and consciousness and all those warm fuzzy humanisms don’t exist, but that they’re not necessary.</p>
<p>Anyways, the same thing happens with me in my ethnography of robots, as I’m effectively taking life out of culture. You can see why both sociologists and anthropologists object to this, albeit for slightly different reasons. Sociologists will allow, for example, some analysis of the sociality of bees, while anthropologists will reject out of hand an ethnography of bees (which like robots/robotics, is different from an ethnography of bees-with-humans). But both seem opposed to attributing sociality or culture to a fundamentally non-living set of individuals. Or even calling non-living entities ‘individuals’ in the first place. And I won’t fall into the trap of saying that robots are living and then mapping human categories onto robot phenomena (e.g. consciousness = statefulness, cognition = code), even though it might seem to make things easier in the short-term. More on that later, but for now be content that all of these things are possible without robots having some advanced AI.</p>
<p>Any ethnography of a non-human society would have to fight this same kind of battle that Latour fought over agency, if it didn’t wish to succumb to the very tempting but misguided prospect of simply importing and mapping existing ontological categories from sociology: e.g. norms in a robot society are found in protocols. This, by the way, would then be using ‘culture’ and indeed the entire ethnographic framework as one massively-stretched analogy, which isn’t the point. The argument is not so much that a robot society is very much ‘alive’ in the same way that human societies have, say, deviant individuals, fluid norms, fascinating rituals, internal contradictions, complicated power relations, and many more weirdly beautiful and complex aspects hidden just below the surface.</p>
<p>Rather, the point of anthropology is typically to locate a people who are typically strange and foreign to us, and then relate the way in which those people live, showing not only how they are different from us but also how they are the same. In doing so, we learn not only about others, but also ourselves. So in that framework, I tend to agree with the critics who say that only way to give a vitalistic account of a robot society is by projecting too many human qualities onto the non-human. What is then left is a non-vitalistic ethnography: an account of a culture devoid of life. Like with Latour and agency, once we show that life is not a necessary criterion for this thing called culture, then the fun really begins — and you can see why lots of people would oppose this.</p>
<p>HF: No friends for robot anthropology, then?</p>
<p>SG: I do have some allies and kindred spirits, and I keep returning to this quote from Deleuze and Guattari’s A Thousand Plateaus on music: “Of course, as Messiaen says, music is not the privilege of human beings: the universe, the cosmos, is made of refrains … The question is more what is not musical in human beings, and what is already musical in nature. Moreover, what Messiaen discovered in music is the same thing ethnologists discovered in animals: human beings are hardly at an advantage, except in the means of overcoding, of making punctual systems.” Music is but one of many domains that is typically seen as inherently social and therefore uniquely human, and the anthropocentric perspective tends to reduce everything to how it functions in the human experiential frame. And on a side note, this is why I’m so excited by Ian Bogost’s upcoming book “Alien Phenomenology: Or What It’s Like To Be A Thing” — the title just says it all, doesn’t it?</p>
<p>And before you start to think that I’m envisioning some sort of AI-based fantasy of the singularity in which robots start to replace all of us social humans — therefore locating the sociality of robot culture in its ability to stand in for humans — that’s definitely the exact opposite of where I’m going. Robots can be said to have their own culture precisely because they don’t need to copy our sociologisms in order to be social, although what they do in their own social realm may not easily map on to things we do in our social realm. This is probably what fascinates me most about this project. And it is precisely for this reason that we must absolutely resist the temptation to make cheap analogies between things that happen in robot culture and human culture, such as saying that protocols are just robot norms.</p>
]]></content:encoded>
			<wfw:commentRss>http://stuartgeiger.com/wordpress/2012/08/the-ethnography-of-robots-interview-at-ethnography-matters/feed/</wfw:commentRss>
		<slash:comments>0</slash:comments>
		</item>
		<item>
		<title>Closed-source papers on open source communities: a problem and a partial solution</title>
		<link>http://stuartgeiger.com/wordpress/2011/06/closed-source-papers-on-open-source-communities-a-problem-and-a-partial-solution/</link>
		<comments>http://stuartgeiger.com/wordpress/2011/06/closed-source-papers-on-open-source-communities-a-problem-and-a-partial-solution/#comments</comments>
		<pubDate>Sun, 12 Jun 2011 18:24:27 +0000</pubDate>
		<dc:creator><![CDATA[stuart]]></dc:creator>
				<category><![CDATA[Blog Posts]]></category>
		<category><![CDATA[Wikis]]></category>
		<category><![CDATA[academia]]></category>
		<category><![CDATA[copyright]]></category>
		<category><![CDATA[creative commons]]></category>
		<category><![CDATA[education]]></category>
		<category><![CDATA[intellectual property]]></category>
		<category><![CDATA[open access]]></category>
		<category><![CDATA[research]]></category>
		<category><![CDATA[wikimedia foundation]]></category>
		<category><![CDATA[wikipedia]]></category>

		<guid isPermaLink="false">http://www.stuartgeiger.com/wordpress/?p=447</guid>
		<description><![CDATA[In the Wikipedia research community &#8212; that is, the group of academics and Wikipedians who are interested in studying Wikipedia &#8212; there has been a pretty substantial and longstanding problem with how research is published. Academics, from graduate students to tenured faculty, are deeply invested and entrenched in an system that rewards the publication of research. Publish or perish, as we&#8217;ve all heard.   The problem is that the overwhelming majority of publications which are recognized as &#8216;academic&#8217; require us to assign copyright to the publication, so that the publisher can then charge for access to the article.  This is in direct contradiction with the goals of Wikipedia, as well as many other open source and open content creation communities &#8212; communities which are the subject of a substantial amount of academic research. Freely-accessible or freely-licensed? There are actually two issues here, the first being that members of these communities want access to research about themselves without having to pay the average $20-$30 an article.  While important, this also overshadows a more fundamental concern: communities like Wikipedia, Apache, Creative Commons, and OLPC were founded on the idea of providing free and open software, hardware, or educational content to the world.   The Wikimedia Foundation&#8217;s mission statement is &#8220;to empower and engage people around the world to collect and develop educational content under a free license or in the public domain.&#8221;  That is pretty clear-cut, and those of us with obligations to both our own academic community and the Wikipedia community are having more and more problems with negotiating those competing tensions. In a sense, this is related to how the major ethical dilemma with 19th and early 20th century anthropologists wasn&#8217;t about giving &#8216;their natives&#8217; a copy of their manuscripts. Rather, it was that most anthropologists were participating in systems of colonialism, which were in direct opposition to the interests of the people they studied.  Now, I am in no way arguing that the same kind of power relation exists between academics who study Wikipedians and the Wikipedian community, or that the issue open educational sources is on the same ethical level as colonialism.   As an aside, contemporary anthropologists have documented this shift from &#8216;studying down&#8217; to &#8216;studying up&#8217;, although I would say that most academics who research open communities like Wikipedia are now &#8216;studying across&#8217; &#8212; but that interesting subject is for another blog post.   But I bring it up&#8230; ]]></description>
				<content:encoded><![CDATA[<p>In the Wikipedia research community &#8212; that is, the group of academics <em>and Wikipedians</em> who are interested in studying Wikipedia &#8212; there has been a pretty substantial and longstanding problem with how research is published.  Academics, from graduate students to tenured faculty, are deeply invested and entrenched in an system that rewards the publication of research.  Publish or perish, as we&#8217;ve all heard.   The problem is that the overwhelming majority of publications which are recognized as &#8216;academic&#8217; require us to assign copyright to the publication, so that the publisher can then charge for access to the article.  This is in direct contradiction with the goals of Wikipedia, as well as many other open source and open content creation communities &#8212; communities which are the subject of a substantial amount of academic research.</p>
<p><span id="more-447"></span><strong>Freely-accessible or freely-licensed?</strong></p>
<p>There are actually two issues here, the first being that members of these communities want access to research about themselves without having to pay the average $20-$30 an article.  While important, this also overshadows a more fundamental concern: communities like Wikipedia, Apache, Creative Commons, and OLPC were founded on the idea of providing free and open software, hardware, or educational content to the world.   The <a href="http://wikimediafoundation.org/wiki/Mission_statement">Wikimedia Foundation&#8217;s mission statement</a> is &#8220;to empower and engage people around the world to collect and develop educational content under a <a title="w:en:free content" href="http://en.wikipedia.org/wiki/en:free_content">free license</a> or in the public domain.&#8221;  That is pretty clear-cut, and those of us with obligations to both our own academic community and the Wikipedia community are having more and more problems with negotiating those competing tensions.</p>
<p>In a sense, this is related to how the major ethical dilemma with 19th and early 20th century anthropologists wasn&#8217;t about giving &#8216;their natives&#8217; a copy of their manuscripts. Rather, it was that most anthropologists were participating in systems of colonialism, which were in direct opposition to the interests of the people they studied.  Now, I am in no way arguing that the same kind of power relation exists between academics who study Wikipedians and the Wikipedian community, or that the issue open educational sources is on the same ethical level as colonialism.   As an aside, contemporary anthropologists have documented this shift from &#8216;studying down&#8217; to &#8216;studying up&#8217;, although I would say that most academics who research open communities like Wikipedia are now &#8216;studying across&#8217; &#8212; but that interesting subject is for another blog post.   But I bring it up because unlike with the Trobriand Islanders, the communities that we study are now beginning to articulate their concerns with how we perform and publish our research, and it is something that we need to listen to.</p>
<p>So to return to the core issue at hand: why is the Wikipedian community (and the Wikimedia Foundation) supporting research that will be copyrighted and bound up in publications which further support an intellectual property regime they clearly stand against?   And what does it mean for us as academic researchers to give back to the communities we study?   It obviously goes beyond being willing to send a copy of a PDF to an interested Wikipedian over e-mail, or even hosting a freely-accessible copy of our copyrighted PDFs on our websites (which many of us do, even when we&#8217;re not supposed to).  For those of us studying Wikipedia, Creative Commons, Scratch, or a number of open content creation communities, it means releasing our research under <a href="http://creativecommons.org/licenses/">a Creative Commons license</a>, as this has become the standard for releasing everything other than code.</p>
<p>Now, the moment I say this, all the academics breathe a heavy sigh, knowing that such a request is impossible, given the current academic system in which we are entrenched.  Even the <a href="http://onlinelibrary.wiley.com/journal/10.1111/(ISSN)1083-6101">Journal of Computer Mediated Communication</a>, one of the few top-tier open access journals in the social sciences, is copyrighted by the publisher.  Some academic superstars like Lawrence Lessig have been able to get their books published from a university press while still being released under a CC license, but not all of us are Lawrence Lessig.  Especially for graduate students and junior faculty, who are desperately trying to get their research published anywhere, when the paper finally gets accepted and that copyright assignment form comes in your inbox, the last thing you want to do is start a losing battle over CC-BY-SAing your paper. However, I do have to give a shoutout to Joseph Reagle, who spent a massive amount of effort getting MIT Press to let him publish <a href="http://reagle.org/joseph/2010/gfc/">his book on Wikipedia</a> under a CC license (although with a number of restrictions), but it is unclear the extent to which this will continue in the future.</p>
<p><strong>A partial solution: freely-licensed figures, &#8216;used with permission&#8217; in copyrighted research papers</strong></p>
<p>So now I finally get to the solution that this blog post was supposed to be entirely about.  We academics who study open content communities have an obligation to release our research under free licenses.  This does not mean that we have to release our <em>research papers</em> under <a href="http://creativecommons.org/licenses/by-sa/3.0/">CC-BY-SA</a>, which is all but impossible for most of us.  What it means is that we must release our findings, results, and conclusions under such licenses, and thanks to how copyright works, we can do this through the existing system.  Conclusions and abstracts are easy: we just re-write them.  We should actually be in the habit of re-writing our densely-worded abstracts and conclusions under a more succinct and human-readable for the communities we study anyway.</p>
<p>However, there is also a way to do this with figures, charts, and graphs.  This idea came to me when I saw a copyrighted article in the ACM library (from the Association for Computing Machinery, where a significant amount of Wikipedia research is published) which used a photo someone else took &#8220;with permission.&#8221;  This kind of thing happens regularly enough for the ACM to have <a href="http://www.acm.org/publications/policies/copyright_policy">a rather sane policy</a> on it: &#8220;The author&#8217;s copyright transfer applies only to the work as a whole, and not to any embedded objects owned by third parties. An author who embeds an object, such as an art image that is copyrighted by a third party, must obtain that party&#8217;s permission to include the object, with the understanding that the entire work may be distributed as a unit in any medium.&#8221;  I haven&#8217;t checked any other publication houses, but I&#8217;ve seen this kind of situation happen in so many different books and papers that it could provide a nice loophole in for most of academia.</p>
<p>For most research on Wikipedia, the figures, charts, and graphs are the most interesting aspects of the research, and these can be released under a <a href="http://creativecommons.org/licenses/by/3.0/">CC-BY</a> or <a href="http://creativecommons.org/licenses/by-sa/3.0/">CC-BY-SA</a> license, and then used with permission in an ACM article.  The ACM&#8217;s main concern is that they need authors to assign copyright to them in order to make sure publication goes smoothly, and as long as the &#8216;original author&#8217; of the image is completely fine with having the image in the work and published by the ACM, everyone is happy.  I&#8217;m no lawyer, but I think this would work with releasing figures, charts, and graphs, even though the copyright policy only qualifies the legal phrase with an example of art images copyrighted by third parties.  This doesn&#8217;t work as well with many forms of qualitative research, such as historical or interview-based research in which the goal is to elaborate on specific case studies.  Still, figures and conceptual diagrams are also useful in those kinds of papers, and can be added to an alternative documentation of a research project, which is possibly co-extensive with <em>but not identical to</em> the research paper.</p>
<p>I&#8217;ve actually been putting my charts and graphs up on <a href="http://commons.wikimedia.org">Wikimedia Commons</a> for quite some time (you can check them all out on <a href="http://commons.wikimedia.org/wiki/Special:ListFiles/Staeiou">my user gallery</a>), even before I realized that copyright was even an issue.   These figures are present in my published papers, many of which are copyrighted by the ACM.  Thankfully, it turns out that this is actually compatible copyright-wise, but this is only solid because I uploaded them to Commons before assigning copyright to the ACM.  It is less clear if someone can retroactively release such images.</p>
<p>But that issue aside, my graphs and charts can live in both worlds, serving members of both communities.  For my quantitative research, these graphs contain my core findings about the rise of bots and assisted editing tools, for example. I have yet to document my previous research projects in a way that would be helpful to others.  More on that in the section below, but I think that even just uploading figures to Commons is a good start.  And it is incredibly painless, especially given that uploading to Commons is a lot easier now than it has been in the past.</p>
<p><strong>Research documentation on Meta-Wiki</strong></p>
<p><strong></strong>Documentation of research projects could take place quite nicely in <a href="http://meta.wikimedia.org/wiki/Research:Projects">a new Research: namespace</a> that some great people at the Wikimedia Foundation have provided to document planed, current, and past research projects on Meta-Wiki, the wiki that is used to coordinate many tasks which are common to all language versions of Wikipedia, as well as projects like Wikisource or Wiktonary.  You can see a very rough example of one of these that I am working on with as part of my summer research  fellowship with the Wikimedia Foundation: <a href="http://meta.wikimedia.org/wiki/Research:Alternative_lifecycles_of_new_users">an incomplete but still interesting study of new users</a> that fellow-Fellow Jonathan Morgan and I are doing.</p>
<p>The documentation page is not written like an academic article, although it does give Wikipedians and researchers alike something that is arguably more important.  It gives information necessary to replicate the study, for example, how we sampled for new users and what coding schema we used to track new user participation in community spaces.  It also contains a few sentences about the motivation of the study, and a few sentences about each of the results. And critically, it contains the graphs which clearly indicate that since 2004, fewer users are participating in community spaces in their first thirty days of joining the project.  If I wanted to write this up into an academic article (which I do plan to), I can do so in such a way that is both suitable for the ACM or another academic publisher, while keeping all the existing content on the documentation page freely-licensed.</p>
<p>Now, to be on the safe side, it may be wise to release these graphs under a <a href="http://creativecommons.org/licenses/by/3.0/">CC-BY</a> license instead of a <a href="http://creativecommons.org/licenses/by-sa/3.0/">CC-BY-SA</a> one, because the <a href="http://en.wikipedia.org/wiki/Share-alike">Share Alike</a> requirement might require some other researcher to release an entire academic paper under a CC-BY-SA license if they use one of my CC-BY-SA figures.   However, I do not think this is the case, because as I am the original copyright holder, I can choose to give permission to using images in my own academic papers.   This is a common misconception with Share Alike and CC licenses in general &#8212; while I can never revoke my license once I make it, I am not bound by those terms in my own work, and can release the image under as many free and non-free licenses as I choose.   For example, if it is entirely my own image that I license with CC-BY-SA, I do not have to release every work that builds on it under CC-BY-SA, just as I can license the work for commercial use even if I choose a CC license that prohibits commercial use.</p>
<p><strong>Research isn&#8217;t a paper</strong></p>
<p>In all, I think that many of the seemingly-intractable problems stem from the false assumption that research projects are entirely encapsulated in a series of papers, and so the demand to &#8216;freely license your research&#8217; is heard as &#8216;freely license your papers&#8217;.  However, academics already think of research projects as these long processes which spawn multiple papers, and so there is no reason why a research project could not also spawn a freely-licensed documentation space which does not prohibit the publishing of research papers.  Certainly there are many aspects of research papers which would not be included, and there is a risk that these documentation spaces would be second-class reports which are always incomplete compared to the research paper.  Though it is a bit patronizing to universally assume that community members don&#8217;t want that dense theoretical analysis of how distributed cognition flows in the actor-network, I think that a facts, figures, and abstracts version would suffice for most.</p>
<p>Given the current academic systems in which we are currently entrenched, I think that this is a good short-term solution, especially for graduate students and other junior scholars who do not have the political capital to change the way in which existing publication regimes operate.  And who knows, perhaps by creating alternative, freely-licensed spaces for documenting research, these publications will recognize the need to make research, though not necessarily research papers, freely accessible and open to all.</p>
]]></content:encoded>
			<wfw:commentRss>http://stuartgeiger.com/wordpress/2011/06/closed-source-papers-on-open-source-communities-a-problem-and-a-partial-solution/feed/</wfw:commentRss>
		<slash:comments>5</slash:comments>
		</item>
	</channel>
</rss>
