Off the Top: Information Aggregation Entries

20052010201520202025

Showing posts: 106-120 of 142 total posts


3 October 2004

Feed On This

The "My" portal hype died for all but a few central "MyX" portals, like my.yahoo. Two to three years ago "My" was hot and everybody and their brother spent a ton of money building a personal portal to their site. Many newspapers had their own news portals, such as the my.washingtonpost.com and others. Building this personalization was expensive and there were very few takers. Companies fell down this same rabbit hole offering a personalized view to their sites and so some degree this made sense and to a for a few companies this works well for their paying customers. Many large organizations have moved in this direction with their corporate intranets, which does work rather well.

Where Do Personalization Portals Work Well

The places where personalization works points where information aggregation makes sense. The my.yahoo's work because it is the one place for a person to do their one-stop information aggregation. People that use personalized portals often have one for work and one for Personal life. People using personalized portals are used because they provide one place to look for information they need.

The corporate Intranet one place having one centralized portal works well. These interfaces to a centralized resource that has information each of the people wants according to their needs and desires can be found to be very helpful. Having more than one portal often leads to quick failure as their is no centralized point that is easy to work from to get to what is desired. The user uses these tools as part of their Personal InfoCloud, which has information aggregated as they need it and it is categorized and labeled in a manner that is easiest for them to understand (some organizations use portals as a means of enculturation the users to the common vocabulary that is desired for use in the organization - this top-down approach can work over time, but also leads to users not finding what they need). People in organizations often want information about the organization's changes, employee information, calendars, discussion areas, etc. to be easily found.

Think of personalized portals as very large umbrellas. If you can think of logical umbrellas above your organization then you probably are in the wrong place to build a personalized portal and your time and effort will be far better spent providing information in a format that can be easily used in a portal or information aggregator. Sites like the Washington Post's personalized portal did not last because of the cost's to keep the software running and the relatively small group of users that wanted or used that service. Was the Post wrong to move in this direction? No, not at the time, but now that there is an abundance of lesson's learned in this area it would be extremely foolish to move in this direction.

You ask about Amazon? Amazon does an incredible job at providing personalization, but like your local stores that is part of their customer service. In San Francisco I used to frequent a video store near my house on Arguello. I loved that neighborhood video store because the owner knew me and my preferences and off the top of his head he remembered what I had rented and what would be a great suggestion for me. The store was still set up for me to use just like it was for those that were not regulars, but he provided a wonderful service for me, which kept me from going to the large chains that recorded everything about me, but offered no service that helped me enjoy their offerings. Amazon does a similar thing and it does it behind the scenes as part of what it does.

How does Amazon differ from a personalized portal? Aggregation of the information. A personalized portal aggregates what you want and that is its main purpose. Amazon allows its information to be aggregated using its API. Amazon's goal is to help you buy from them. A personalized portal has as its goal to provide one-stop information access. Yes, my.yahoo does have advertising, but its goal is to aggregate information in an interface helps the users find out the information they want easily.

Should government agencies provide personalized portals? It makes the most sense to provide this at the government-wide level. Similar to First.gov a portal that allows tracking of government info would be very helpful. Why not the agency level? Cost and effort! If you believe in government running efficiently it makes sense to centralize a service such as a personalized portal. The U.S. Federal Government has very strong restriction on privacy, which greatly limits the login for a personalized service. The U.S. Government's e-gov initiatives could be other places to provide these services as their is information aggregation at these points also. The downside is having many login names and password to remember to get to the various aggregation points, which is one of the large downfalls of the MyX players of the past few years.

What Should We Provide

The best solution for many is to provide information that can be aggregated. The centralized personalized portals have been moving toward allowing the inclusion of any syndicated information feed. Yahoo has been moving in this direction for some time and in its new beta version of my.yahoo that was released in the past week it allows the users to select the feeds they would like in their portal, even from non-Yahoo resources. In the new my.yahoo any information that has a feed can be pulled into that information aggregator. Many of us have been doing this for some time with RSS Feeds and it has greatly changed the way we consume information, but making information consumption fore efficient.

There are at least three layers in this syndication model. The first is the information syndication layer, where information (or its abstraction and related metadata) are put into a feed. These feeds can then be aggregated with other feeds (similar to what del.icio.us provides (del.icio.us also provides a social software and sharing tool that can be helpful to share out personal tagged information and aggregations based on this bottom-up categorization (folksonomy). The next layer is the information aggregator or personalized portals, which is where people consume the information and choose whether they want to follow the links in the syndication to get more information.

There is little need to provide another personalized portal, but there is great need for information syndication. Just as people have learned with internet search, the information has to be structured properly. The model of information consumption relies on the information being found. Today information is often found through search and information aggregators and these trends seem to be the foundation of information use of tomorrow.



16 September 2004

43folders for Refining Your Personal InfoCloud

I have been completely enjoying Merlin Mann's 43folders the past couple weeks. It has been one of my guilty pleasures and great finds. Merlin provides insights to geeks (some bits are Mac oriented) on how to better organize the digital information around them (or you - if the shoe fits). This is a great tutorial on refining your Personal InfoCloud, if I ever saw one.

Everytime I read this I do keep thinking about how Ben Hammersley has hit it on the head with the Two Emerging Classes. The volume of information available, along with the junk, and the skills needed to best find and manage the information are not for the technically meek.



20 August 2004

Fixing Permalink to Mean Something

This has been a very busy week and this weekend it continues with the same. But, I took two minutes to see if I could solve a tiny problem bugging me. I get links to the main blog, Off the Top, from outside search engines and aggregators (Technorati, etc.) that are referencing content in specific entries, but not all of those entries live on the ever-changing blog home page. All of the entries had the same link to their permanant location. The dumb thing was every link to their permanant home was named the same damn thing, "permalink". Google and other search engines use the information in the link name to give value to the page being linked to. Did I help the cause? No.

So now every permanent link states "permalink for: incert entry title". I am hoping this will help solve the problem. I will modify the other pages most likely next week sometime (it is only a two minute fix) as I am toast.



11 August 2004

Back and Digging Out

Coming back from six plus days of being untethered from the net I found I had 1117 unread RSS feeds. This is worse than my personal e-mail stack, which was just over 550 (I get to my work e-mail stack tomorrow, which averages about 80 e-mails per day). The RSS feeds really threw me as I was not expecting it to have snowballed like that.

There were a few things I was wanting to follow that I knew may pop their heads up while I was away so I followed these on my Treo 600 on Google News and del.icio.us aggregator. I was able to find most of what I was looking for and do a quick read and then e-mail an annotated link to one of my personal e-mail accounts. I did find some things on del.icio.us that I just copied into my del.icio.us bookmarks so I could come back to them later.

I got far less done on the writing front as my son was along for the vacation, which made it a real family vacation and not the usual working vacation with the laptop on my lap on the front porch when I am not playing in the waves. No, I would not say I am rested, but I do have more wonderful memories of a great summer get away. Our time schedules shifted to a 10 month-old's eating and sleeping schedule. When we drifted to our normal shore vacation schedule we had a cranky kid, which only took two days to convert to a vacation fully focused on the kid. We met many wonderful new people, stayed in a different B&B, and found a new restaurant to add to our favorites.

I am now ready for the last two days of the week and to start responding to e-mail tomorrow. I am also ready to tackle my writing assignments that are well over due. My laptop is also fully updated with OS and software updates that make it really sing, too bad Windows updates never make the machine perceivably faster.



17 July 2004

Now Delicious

Time has been very thin of late. In the past six months or so started noticing an increasing number of links from del.icio.us and started pulling the feeds of some folks I like to follow their reading list into my site feed aggregator. I had about four or five del.icio.us feeds in my aggregator (meta aggregation of other's meta aggregations - MetaAg MetaAg). This past week I was taking medicine that tweaked by sleep patterns so I had some free awake time after midnight and I finally set up my own vanderwal del.icio.us feed.

I like having the ability to pull [meta] tags aggregations that others have used, like security, which is a great help during the day at work. I can also track some topics I keep finding myself at the periphery and ever more interested in as they tie to some personal projects.

I did consider something similar with Feedster, but it was down for updating recently when I had the tiny bit of time to fiddle with setting something up. By the way, Feedster is now Standards-based (not fully valid, but rather close) and it loads very quickly (most of the time).



9 July 2004

Tantek Mulls Contact Info Updating

Tantek mulls a means to keep contact info upto date. This should be much easier than Tantek has made out. This could be as easy as publishing one's own vcard that is pointed to with RSS. When the vcard changes the RSS feed notifies the contact info repositories and they grab the vcard and update the repository's content. This is essentially pulling content information into the user's Personal InfoCloud. (Contact info updating and applications are a favorite subject of mine to mull over.)

Why vcard? It is a standard sharing structure that all contact information applications (repositories understand). Most of us have more than one contact repository: Outlook at work; Lotus Organizer on the workstation at home; Apple Address Book and Entourage on the laptop; Palm on the Cellphone PDA; and Addresses in iPod. All of these applications should synch and perfectly update each other (deleting and updating when needed), but they do not. Keeping vcard field names and order constant should permit the info to have corrective properties. The vCard RDF W3C specifications seem to layout existing standards that should be adopted for a centralized endeavor.

What not Plaxo? Plaxo is limited to applications I do not run everywhere (for their download version) and its Web version is impractical as when I need contact information I am most often not in front of a terminal, I am using a Treo or pulling the information out of my iPod.

While Tantek's solution is good and somewhat usable it is not universal as a vCard RDF would be with an application that pinged the XML file to check for an update daily or every few days.



30 June 2004

Future of Local Search on Mac

One of the best things I found to come out of the Apple WWDC keynote preview of the next update of the OS X line, Tiger, Spotlight. Spotlight is the OS file search application. Not only does Spotlight search the file name, file contents (in applications where applicable), but in the metadata. This really is going to be wonderful for me. I, as a user, can set a project name in the metadata and then I can group files from that point. I can also set a term, like "synch" and use AppleScript and Search to batch the files together for synching with mobile devices, easily. Another nice feature is the searches can be saved and stored as a dynamic folder. This provides better control of my Personal InfoCloud.

Steven Johnson provides the history of search in Apple, which has nearly the same technology in Cosmo slated for release in 1996.



17 June 2004

Malcolm McCullough Lays a Great Foundation with Digital Ground

Today I finished reading the Malcolm McCullough book, Digital Ground. This was one of the most readable books on interaction design by way of examining the impact of pervasive computing on people and places. McCullough is an architect by training and does an excellent job using the architecture role in design and development of the end product.

The following quote in the preface frames the remainder of the book very well:

My claims about architecture are indirect because the design challenge of pervasive computing is more directly a question of interaction design. This growing field studies how people deal with technology - and how people deal with each other, through technology. As a consequence of pervasive computing, interaction design is poised to become one of the main liberal arts of the twenty-first century. I wrote this book because I ran into many people who believe that. If you share this belief, or if you just wonder what interaction design is in the first place, you may find some substance here in this book.

This book was not only interesting to me it was one of the best interaction books I have read. I personally found it better than the Cooper books, only for the reason McCullough gets into mobile and pervasive computing and how that changes interaction design. Including these current interaction modes the role of interaction design changes quite a bit from preparing an interface that is a transaction done solely on a desktop or laptop, to one that must encompass portability and remote usage and the various social implications. I have a lot of frustration with flash-based sites that are only designed for the desktop and are completely worthless on a handheld, which is often where the information is more helpful to me.

McCullough brings in "place" to help frame the differing uses for information and the interaction design that is needed. McCullough includes home and work as the usual first and second places, as well as the third place, which is the social environment. McCullough then brings in a fourth place, "Travel and Transit", which is where many Americans find themselves for an hour or so each day. How do people interact with news, advertisements, directions, entertainment, etc. in this place? How does interaction design change for this fourth place, as many digital information resources seem to think about this mode when designing their sites or applications.

Not only was the main content of Digital Ground informative and well though out, but the end notes are fantastic. The notes and annotations could be a stand alone work of their own, albeit slightly incongruous.



11 April 2004

Stitching our Lives Together

Not long ago Jeffrey Veen posted about Will you be my friend, which brought up some needs to better stitch together our own disperse information. An excellent example is:

For example, when I plan a trip, I try to find out who else will be around so I have people to hang out with. So my calendar should ask Upcoming.org, "Hey, Jeff says he's friends with Tim. Will he be in New York for GEL?"

This example would allow up to interact with our shared information in a manner that keeps it within our extended Personal InfoCloud (the Personal InfoCloud is the information we keep with us, is self-organized, and we have easy access to). Too many of the Web's resources where we store our information and that information's correlation to ourselves (Upcoming.org, LinkedIn, etc.) do not allow interactivity between online services. Some, like Upcoming and Hilton Hotels do provide standard calendaring downloads of the events and reservations you would like to track.

Some of this could be done with Web Services, were their standards for the interaction. Others require a common API, like a weblogging interface such as Flickr seems to use. The advent of wide usage of RSS feeds and RSS aggregators is really putting the user back in control of the information they would like to track. Too many sites have moved toward the portal model and failed (there are large volumes of accounts of failed portal attempts, where the sites should provide a feed of their information as it is a limited quantity). When users get asked about their lack of interest in a company's new portal they nearly always state, "I already have a portal where I aggregate my information". Most often these portals are ones like My Yahoo, MSN, or AOL. Many users state they have tried keeping more than one portal, but find they loose information very quickly and they can not remember, which portal holds what information.

It seems the companies that sell portal tools should rather focus on integration with existing portals. Currently Yahoo offers the an RSS feed aggregator. Yahoo is moving toward a one stop shopping for information for individuals. Yahoo also synchs with PDA, which is how many people keep their needed information close to themselves.

There are also those of us that prefer to be our own aggregators to information. We choose to structure our large volumes of information and the means to access that information. The down side of the person controlling the information is the lack of common APIs and accessible Web Services to permit the connecting of Upcoming to our calendar (it can already do this), with lists of known or stated friends and their interests.

This has been the dream of many of us for many years, but it always seems just around the corner. Now seems to be a good time to just make it happen. Now is good because there is growing adoption of standards and information that can be personally aggregated. Now is good because there are more and more services allowing us to categorize various bits of information about our lives. Now is good because we have the technology. Now is good because we are smart enough to make it happen.



23 January 2004

Keeping the Found Things Found

This weeks New York Times Circuits article: Now Where Was I? New Ways to Revisit Web Sites, which covers the Keep the Found Things Found research project at University of Washington. The program is summarized:

The classic problem of information retrieval, simply put, is to help people find the relatively small number of things they are looking for (books, articles, web pages, CDs, etc.) from a very large set of possibilities. This classic problem has been studied in many variations and has been addressed through a rich diversity of information retrieval tools and techniques.

This topic is at the heart of the Personal Information Cloud. How does a person keep the information they found attracted to themselves once they found that information. Keeping the found information at hand to use when the case to use the information arises is a regular struggle. The Personal Information Cloud is the rough cloud of information that follows the user. Users have spent much time and effort to draw information they desire close to themselves (Model of Attraction). Once they have the information, is the information in a format that is easy for the user or consumer of the information to use or even reuse.



18 January 2004

Portable Personal Information Repository

MIT's Technology Review discusses Randolph Wang's wireless PDA for personal information storage (registration for TR may be required). This brief description (I could find no longer nor explicit description at Wang's Princeton pages nor searching CiteSeer) is very interesting to me.

One PC at work, another at home, a laptop on the plane, and a personal digital assistant in the taxicab: keeping all that data current and accessible can be a major headache. Randolph Wang, a Princeton University computer scientist, hopes to relieve the pain with one mobile device. Designed to provide anytime, anywhere access to all your files, the device stores some data, but its main job is to wirelessly retrieve files from Internet-connected computers and deliver them to any computer you have access to. Wangís prototype is a PDA with both cellular and Wi-Fi connections, but the key is his software, which grabs and displays the most current data stored on multiple computers. Wang has tested his prototype with more than 40 university and home computers on and around the Princeton campus. He eventually wants to shrink the device down to the size of a wristwatch to make carrying it a snap.

This is really getting to a personal information cloud that follows the user. This really is getting to the ideal. Imagine having everything of interest always with you and always available to use. Wang's solution seems to solve one of the ultimate problems, synching. The synching portion of this seems to stem from PersonalRAID: Mobile Storage for Distributed and Disconnected Computers, which was presented at a USENIX conference. I really look forward to finding out more about this product.



11 January 2004

Blogs highlighted on Meet the Press

While I am not a huge blog-for-blog-sake person, Meet the Press has a relatively long roundtable discussion on blog and the Democratic presidential campaigns. The talk about how Joe Trippi not only uses the blog to communicate with potential Dean supporters, but how he and others cull ideas from the blogs.

This highlight how new innovative ideas can quickly get posted by individuals, culled, and directly or with modification get implemented to better an endeavor. One clever idea that was culled from a weblog was the ability to have individuals share their unused weekend cell phone minutes and have the campaign use these minutes to call voters in Iowa. A rather clever idea and even smarter use of culling the Internet's vast cacophony of voices to find good helpful ideas to run a business/organization better and smarter.



1 December 2003

Catalog of the Cool

Wired's article on Kevin Kelley's Catalog of the Cool had me aching to get home and take a peek at the Cool Tools. I was a fan of the Whole Earth Catalog (now an online magazine) when I was growing up. I think we only had one copy that was bought when I was in fourth grade, but I always flipped through it. I was fascinated with all the different stuff that filled its pages. I was a particular fan of the log cabin and some of the other outdoor items.

I was very impressed with the Catalog of the Cool as it kept the same spirit, but with things I can find more useful. My only WEC did not have items that were of great use to me, but it inspired my imagination thinking of the adventures one would have with the items. It is even better than a J. Peterman catalog.



17 November 2003

More on Urban Tapestries

More on Urban Tapestries:

Urban Tapestries is a framework for understanding the social, cultural, economic and political implications of pervasive location-based mobile and wireless systems. To investigate these issues, we are building an experimental location-based wireless platform to allow users to access and author location-specific content (text, audio, pictures and movies). It is a forum for exploring and sharing experience and knowledge, for leaving and annotating ephemeral traces of peoplesí presence in the geography of the city.


14 November 2003

NBA does Moneyball

Fans of Moneyball will like the Washington Post story on the NBA wiz kid executive. The focus of the article is the San Antonio Spur's Sam Presti, age 27, who is applying MBA tactics to the NBA. Yes, quantitative analysis to mitigate risk and control cost is behind the NBA version of Moneyball, just as it is in Major League Baseball.



This work is licensed under the Creative Commons Attribution-NonCommercial-ShareAlike License.