Thursday, November 13, 2008

Week 10 Comments

Comment #1:
I commented on Jen’s “Intro to Information Technology” page:
https://www.blogger.com/comment.g?blogID=1475137707322366107&postID=7838913334025505790&page=1

Comment #2:
I commented on Megan’s “The Alley View” page:
https://www.blogger.com/comment.g?blogID=1139180432200060758&postID=5904996233988888078

Reading Response #10: Harvest Time

Of the readings assigned this week, I found “Current Developments and Future Trends for the OAI Protocol for Metadata Harvesting” to be the most interesting as it peripherally touched upon the relationship between metadata harvesting and “The Deep Web.”

Typically, harvesting culls information from Web pages using the metadata tags embedded in the HTML code. Initially, there was some dispute as to whether the function should be performed digitally or manually, the concern being that people would create too many disparate terms while a purely software-based procedure would exclude important semantic relationships. With Dublin Core rapidly becoming a standard for metadata schema, Web pages are increasingly adhering to a similar standard and format. But does this improvement in search methods extend to the “Deep Web”?

The “Deep Web” includes information that is available to the public but resides outside the scope of traditional search engines because data is stored in proprietary databases, only accessible by direct inquiries. However, the OIA Protocol allows search engines without normal access to this information to index pages hosted on the “Deep Web” through OAI repositories. As a significant portion of digital information resides on the “Deeb Web” this new component of accessibility is important as it helps to promote open-source and transparent policies in regards to public information.

Muddiest Point #10

I was curious to why uploading my website to the Pitt's page didn't work the first time but it did the second. In the first instance, I followed Dr. He's instructions explicitly but I kept getting a 403 error message. The second time, I used instructions from the Technology Help desk. The only difference between the Help Desk's instructions and Dr. He's was that I typed in a telnet address in my browser. This address led to a request for a source of software to open it and I selected FileZilla. The FileZilla screen popped open and when I sent the file, it was visible on Pitt's page.

Saturday, November 8, 2008

Assignment #6

To view my completed web page, please click here.

Thursday, November 6, 2008

Week 9 Comments

Comment #1:
I commented on Lauren’s blog “LIS 2600 Land”
https://www.blogger.com/comment.g?blogID=4181925387762663697&postID=3072857614832667163

Comment #2:
I commented on Theresa’s blog “Intro to Information Technology”
https://www.blogger.com/comment.g?blogID=5586031599791302355&postID=9132805377535596301

Reading Response #9: Gone Fishing

Michael Bergman’s article “The Deep Web: Surfacing Hidden Value” is an important white paper because it addresses both the limitations of current search engine formats and the structure of information on the Web. Google, today’s most popular search engine, relies on an aggregate formula to create a list of query results intended to minimize duplicates and increase the number of relevant resources. However this format is inherently imperfect because it relies on a system of popularity through web citation (similar to the way the most prominently published scientific journal articles cite each other). As a result, a web page with relevant information might end up further down on a list of query results because it has not been cited adequately by other web pages.

A bigger problem with this search method is that it skims over the larger repository of information available in the Deep Web. Most of the information here is digitally available but instead of being hosted on a “surface page”, it is embedded in proprietary databases that are linked to but function off of the Internet.

The Deep Web should be a primary concern for several reasons. Currently a great deal of development is being done on more semantic and comprehensive search capabilities. For this work to be functional and current, it has to be able to adapt to the exponential increase in digital information as well as its location, both on surface web pages and the Deep Web.

Also, the availability of information is one of the most important components of digital network systems because without it, the democratic intention of the web is meaningless. Bergman gives the example of several federal organizations that post their information online but not in a format accessible by commercial web engines; the majority of the information is hidden in the “Deep Web.” Though not intentionally deceptive, this unexplored territory of information could inadvertently become an intentional iron curtain. As the format of information transitions from analog to digital, it is important that the same amount of information be readily available.

Muddiest Point #9

This week's lecture went over my head. I thought I had a basic understanding of HTML but realized I didn't when I wasn't able to discern the difference between HTML and XML.

Friday, October 31, 2008

Week 8 Comments

Comment #1:

I posted a comment on Jacqui Taylor’s “Qui Quandaries” blog:
https://www.blogger.com/comment.g?blogID=2005895256228614061&postID=1597573534668681094

Comment #2:

I posted a comment on Sean Kilcoyne’s “spk” blog:
https://www.blogger.com/comment.g?blogID=1129785935180596689&postID=7878297980523559430

Reading Response #8

I reread the literature on XML and I am failing to see the dramatic difference between XML and HTML except that the former provides a more guided experience for users although it doesn't utilize a standardized coding format. Also, XML does have more specific identifying parameters within the coding set but creators are able to create their own DTDs (Document Type Definition). Will this affect how the document is searched on the web and is this new format more compatible with Web 2.0?

Another concern the reading raised was about uniformity. Currently, there is the struggle to create uniform metadata tags to generate more effective web searches. Similarly, semantic web research is trying to find away to incorporate the diversity of contexts but through a uniform metadata scheme so that information networks can create more efficient, streamlined queries. But if XML allows the creator to provide their own tag systems (although I am not familiar enough with XML to know if this affects search dynamics) this format could foster more user control but compound the problem of cataloguing information to make it more readily accessible.

Muddiest Point #8

We covered HTML in class but I was wondering what is the difference between regular HTML and semantic HTML? Is Web 2.0 based on regular HTML or semantic HTML?

Saturday, October 18, 2008

Assignment #5

I created a virtual shelf with references on the art and films of Peter Greenaway.

Tuesday, October 14, 2008

Week 7 Comments

Comment #1:
I posted a comment on Sean’s blog, spk blog.
https://www.blogger.com/comment.g?blogID=1129785935180596689&postID=6165812584986651423

Comment #2
I posted a comment on Tamoul’s blog
https://www.blogger.com/comment.g?blogID=7114620464717775258&postID=3250005245059352985

Reading Response #7: Fair Isn't Always Equal

“Beyond HTML: Developing and Re-imagining Web Guides in a Content Management System” is a case study that delineates a university library undergoing the transition from independent web postings to a format streamlined by a CMS. One of the important lessons it demonstrates is that the digital divide doesn’t just affect users. Due to varying levels of expertise, different liaison web interfaces had radically different information and accessibility levels as well as duplicated information.

Ironically, one of the primary functions of libraries is to provide a readily accessible information format but technology and generational divides have created inconsistencies. CMS can alter that. Uniform templates can be created allowing for a modicum of flexibility to accommodate librarians from different disciplines. Most importantly, it creates in a single database with identical vocabulary which not only reduces storage capacity (by eliminating duplicates) but also creates a more familiar interface for users.

One thing that did strike me in the article was the question of using open-source software. GSU didn’t use it because it was deemed incompatible with their Windows systems. I think that it is important for public libraries to consider moving away from commercial products and adopting open-source software. Yes, there are constant upgrades but this is true with any type of software. Open-source reduces budgetary demands and can be specifically modified to adapt to individual libraries’ needs. This is not unfeasible as exemplified by the study. GSU had the money and resources to create an in-house database system. Time and money could have been saved using an open-source software.

Finally, CMS are important because libraries don’t have uniform technical training and, ultimately, interfaces most accommodate the user and provide the most efficient access to information.

Muddiest Point #7

I am understand that wireless internet is available through the distribution and receipt of radio waves but I don't understand how computers are able to block access to something as intangible as radio waves and how WEPs physically work.

Wednesday, October 1, 2008

Muddiest Point #6

I was wondering if there is a difference between Library Thing and Good Reads and if both sites work on the aggregate function the way that Google does. Does an aggregate dynamic changes how much information is available in a purely referential information system as opposed to a search engine?

Week 6 Comments

Week 6 Comments

Comment #1 (on Lauren’s blog)
https://www.blogger.com/comment.g?blogID=4181925387762663697&postID=481855389631007759

Comment #2 (on Sean’s blog)
https://www.blogger.com/comment.g?blogID=1129785935180596689&postID=1090369698999357745

Tuesday, September 30, 2008

Reading Response # 6: It's Not Free and It's Not Fair

This week’s two assigned articles were an interesting combination. Jeff Tyson’s piece outlines the history and physical infrastructure of the Internet while Andrew Pace’s “Dismantling Integrated Library Systems” chronicles libraries’ struggles with adopting and paying for ILS. Despite the loss of physical paraphernalia and brick and mortar institutions, there is nothing lightweight and accessible about the price of information.

Tyson states that because of its design as interconnected networks within networks, the Internet isn’t owned; it’s a shared information resource. This creates a contradictory dynamic because it costs money to access it, whether through individual ISPs or companies and libraries paying flat fees to provide free access to their users.

I understand that I focus a great deal on the economics of technology but that is only because when heralding the benefits of the digital age, democratic and open source are ubiquitous descriptions. Whatever the original motivation behind its creation, the Internet is only as democratic as the society it functions in. For example, more prohibitive societies monitor websites, censor public information, and restrict access. We have a more democratic approach to the exchange of information but because we are a capitalist democracy, our Internet functions like one. It’s a shared network but privately owned companies make a profit from it by reformatting it into a paid service. It’s not as if there is a free point of access and ISPs are faster, more dynamic alternatives; they are the only alternatives. The same holds true with effective access to the glut of available information. Libraries access to networked information is only as good as the ILS they are able to afford. Not only does this widen the "digital divide" but quality become a privilege available only to those with enough money to buy it.

I’d really like to learn the economic history of the Internet, to understand how shared resources become a utility cost. I think learning about the dynamics of this transformation is important because the Internet, while no longer in its nascent stages, is still open to paradigm changes and could still become a democratic resource. Otherwise, we are only fooling ourselves if we believe that true democracy is a hand out waiting to be paid.

Saturday, September 27, 2008

Assignment #3: Zotero/CiteULike

http://www.citeulike.org/user/rag55

The resources found through citeulike have the tag "from-citeulike"
The imported resources found through Zotero/Google Scholar have the tag "from-zotero"

Friday, September 26, 2008

Week 5 Comments

Comment #1:
https://www.blogger.com/comment.g?blogID=5586031599791302355&postID=8265205753100140876

Comment#2:
https://www.blogger.com/comment.g?blogID=1129785935180596689&postID=7650461811986294684