Tuesday, June 28, 2005

"Inbound" vs. "Outbound" definitions: Vertical search is horizontal!

One of the questions today at the SDForum Vertical LEAP conference has been "what is vertical search?

There are two simple answers to this. Firstly, there's what I'd call "Inbound" vertical search. This means that anyone in the world, at some given point in time, wants to do a certain thing:
Shop
Travel
Get a job
etc.

So the "Inbound" is people (temporarily) participating in a certain type of activity.

Visual:



Second, there's the "Outbound" concept of vertical. This is the sort of vertical that business-to-business publishers like Primedia and CMP and Reed Elsevier have been working on forever. Oh, and there was this cool "bricks and clicks" company called eBusinessMedia...

Here, the vertical is based on "who I am" rather than "what I do". The music business is a vertical. Agriculture is a vertical. Medicine, with several sub-verticals, is a vertical. Music, agriculture, and medicine all need employment/jobs services -- all need secondary markets for used capital equipment -- all need dedicated industry news sources -- but the "vertical" here is the fact that medical jobs, markets, and news are entirely different from agriculture jobs, markets, and news.

Visual:



Oddly enough, from the standpoint of this traditional view of "vertical," the distinctions that are driving this conference are actually horizontal.

Visual:



Hmmm. Leave it a bunch of technologists in Silicon Valley to mix up their X and Y axes ;-)

Local Media, Technorati, and Vertical Search

I am sitting here at SDForum's "Vertical LEAP" vertical search conference, where I've sat through a local search presentation from Google, Yahoo, Ask Jeeves, and MSN which pretty much missed the point entirely. I asked a question about a) events and b) local search as local media, since neither (enormous, gigantic, huge) topic had been mentioned by any of the assembled gigantic players. Credit to Brady Forrest, PM at MSN for their local search, for at least mentioning mobile -- yet another (enormous, gigantic, huge) topic they'd entirely slid past for a whole hour. If these guys get it, then they're holding their cards very close to their chests.

I'm working on an events search company called Zvents, which heretofore is unannounced. So there. I announced it. Secret closed demos on the web in a month, live beta by September, or faster if Tyler starts adding Red Bull to his coffee.

So I am spending a lot of time thinking about local media. And none of these local guys get it at all, and at the same time, Technorati (hats off to Dave Sifry!) gets it in a (enormous, gigantic, huge) way. I'll be the Nth guy to link to the Live8 page, and the second or third guy to link to Peter Caputa's discussion of that page (right on Peter, and we should talk about events) and I'll even throw in a bonus link to Chew Shop on Caputa on Technorati.

So what do I think?

Well here are a few relevant facts.
1) Local media is enormously fragmented. There are 70 newspapers in the Bay Area, plus tens of TV and radio stations, and they are *all* supported by ad dollars.

2) Classified ads were a nice chunk taken out of local media, but those dollars are largely gone; high-value categories like jobs, housing, and for-sale went to Monster, eBay, and the like long ago.

3) The three legs of local media are 1) reporting 2) editorial and 3) advertising. Let's take these in order.

Reporting: When we look at a project like Dan Gillmor's open media / bayosphere / We the People, you see the foundations of local reporting. Remember that this is not just going to be people doing this in their spare time -- we are going to get latter-day Ben Franklins running around researching stuff and writing it up (or podcasting or vblogging it) and publishing it online. Reportage = conceptually sorted, 20 years of implementation to go.

Editorial: I have a whole separate riff about how political blogging and tech blogging represent the explosion of the vertically integrated media model -- guys like Kevin Drum or Joshua Marshall provide me with more editorial insights than does the New York Times; but for the moment, they are reliant on the "MSM" mainstream media to provide the basic grist for their editorial mills. Most interesting to me is the extension of editorial blog voices beyond the narrow reaches of politics and emerging technology. I firmly believe that the reason these "went first" is that the foundation of bloggable news was firmly in place in textual form, and the blogging tools that sprang up made manipulation of that text very simple for bloggers. Myspace is actually an example of music blogging, I think -- 18 million people go there because, fundamentally, they gave indie bands the (different) tools they needed to post MP3s, talk about their shows, and post their band/brand image online. I think that with the emergence of "structured blogging" standards for other categories -- shopping, events, finance, and so forth -- we will see and explosion of new blogging editorial categories as, for instance, a financial blogger can quickly and easily point to financial data and charting information to make his (her) points.

Advertising: This is the part of local that has gone the furthest, and the least; local classified (in certain categories) went 10 years ago, but local newspaper display, local TV and radio, and lots of classified -- plus flyers, posters, billboards, etc. -- has yet to go. And we are talking tens of billions of dollars. It will follow the eyeballs, which will follow the tools, editorial voices, and reportage that is already moving to new media.

Technorati gets the media thing, as demonstrated by the Live8 page -- but it's not clear that they get the local media thing, which fundamentally has shaken out to be an oligopoloy the first time around between Knight-Ridder, Hearst, Tribune, Gannet, etc. with very little local competition. Will it happen that way again? I'm not sure - and this Technorati thing may be the emergence of a new NBC or a new USA Today, not a new Knight-Ridder -- but what I'm sure of is that none of the local papers get this, none of the big portals get this, and there is money laying all over the table to be scooped up.

Game on.

P.S. How much has the world changed? This places is full of VCs and entrepreneurs and the nicest thing that's happened to me today is that Marc Canter linked to me in his blog. Thanks, Marc!

Monday, June 27, 2005

Response to Kevin Drum - Three Thoughts on Future Media

In light of the recent Supreme Court decision on file-sharing, political blogger Kevin Drum asks:

"The year is 2015 and Columbia has just released Spiderman 7. The next day, 10 million people with no technical savvy at all go to their computers, stick a Blu-ray disc into their DVD drive, log on to Movies4Free... Three minutes later they have a 100% perfect DVD...As bandwidth increases, DVD technology improves, and software becomes as easy to use as a toaster, every piece of digital content on the planet will be available within minutes...Is this OK? Or do you have a different vision for the future? If profit-based movie/music distribution becomes essentially impossible, do you think the content industry will somehow adapt and get its revenue elsewhere? Or will content creators continue creating but just make a lot less money at it?"

Three thoughts:

1) Right now, radio broadcasters make more money from music than music labels do. U.S. CD sales are $17 billion, radio advertising is $20 billion. Broadcasters' ability to take the labels to the cleaners for the past, oh, ninety years is due to a court decision in the early days of the 20th century that set the recompense rate exceedingly low. The NAB has had enough political clout to keep it there ever since. Cool, huh? Note that they make all this moolah by... giving songs away for free.

2) Before Caruso (correct me if I'm wrong on this) there were no recorded-media stars, and instead there were many more local peformers of moderate skill who played directly to the people (y'know, live) and made a decent living doing so. If the hyper-promotion of mediocrity machine that is the labels goes away, perhaps we can move somewhat back in the direction of this former model. Quality will still rise to the top; with free distribution of songs, it would do so as a matter of course. But if the primary payment mechanism for musicians was for *playing music* instead of recorded album sales, then, it would be terrible! No one would get more rich than Jerry Garcia. Have you ever seen Jerry Garcia's house? That sounds OK to me.

3) Movies are a very social medium with extremely high costs of production. Right now, they emphasize cost-of-production over, say, quality. I get better acting at my local theatre than I do in watching a Jerry Bruckheimer $200m extravaganza -- but the "production values" of the local theatre are a bit lower. If production technologies get cheaper (and they will!) then the local theatre can start to compete -- meaning that we'll get a much wider base of content producers, similar to today's documentary situation. Unlike music, it's not unreasonable to hyper-encrypt movies during their initial "theatrical" phase and only show them there -- people seem to like watching movies together, and after you've made back your (much smaller) costs of production plus some profit, you throw the thing into the maw of the collective and start working on your next piece. Again, surprisingingly similar to the local-theater model, which seems to work just fine.

Trackback

Tuesday, June 21, 2005

So what about tag spam?

Tyler asks me another question:
"I'm a little concerned about tag spam. Is this a problem on Flickr and other sites that allow communal tagging?"

So I went and did a search.

I haven't found any discussion of it as an issue. Comment spam on blogs and comment trolls, on the other hand are well-known and much-discussed problems; none of the solutions are perfect, but include:

- Registration (know who posts something)
- Moderation (increase or decrease visibility by editorial intervention)
- Deletion (the ultimate in editorial intervention)
- Complication (making spam hard to do quickly and automatically)

Tag spam might be different than blog comments, because:
a) You can't include a URL (which spam needs to do in most cases to be effective)
b) You can't express yourself to others in a personalized way (which most trolls seem to want to do).

So it may be anodyne enough to not really be useful to vandal communities. Just a guess.

Some related links:
A loooong but fairly interesting post about tags. No mention of spam. John Dvorak pans tagging in PC Magazine. Some brief thoughts on vandalism in reply to Dvorak -- apparently del.icio.us has had some spam Another reply to Dvorak. A flat-out awesome post on the meaning of RSS

Are Tag Clouds Useful? And how to make them more so

Tyler asked:
"Is there a better way to display tag utilization than the tag cloud?"

I would say yes. The "tag cloud" was a cool concept but I find it of limited usefulness. Its axes of meaning include:
1) inclusion (vs. exclusion - the cloud is finite)
2) font size
3) order

I have seen clouds ranked both alphabetically (another one) and ranked by use; when listed alphabetically that means that there is only one axis of meaning (aside from the set limitation of inclusion) which is font size.

It would be possible to include font color as another axis, and even bold vs. italic, but what does that tell the user? It will not be clear.

Absolute popularity is a moderately interesting feature of tagclouds that, frankly, is best represented by a list; I assume that the reason that tagclouds were first done is that the resultant visual mass fits more neatly on to a web page than a 100-item vertical or horizontal list.

Here is an idea that might make tag clouds more useful: Create an actual honest-to-god 2D map representation of them. Sort of like a network diagram, but all node and no connection between node. User types in a tag -- "chocolate" and you create a custom 2D map based on "events that include tag chocolate also include these tags" and then let the user scroll around on that map or click on that map, in Google Maps fashion (can you tell that I have been profoundly influenced by Google Maps? ;-) and get to "food" "wine" "laborador" "ghirardelli" "roses" or whatever, and gradually scroll away to the outer bounds of the related-tags-surface.

Now what you have is a tagcloud that shows not overall popularity (a "who cares" feature -- when was the last time you checked out Google Zeitgeist, much less derived some useful information from it?) but relatedness to something the user cares about, presented in a navigable visual form.

Might be interesting.

Tuesday, February 22, 2005

How Google, A9, and Flickr create those awesome new web interfaces - "Ajax"

I recently wrote a short post about the fast and fabulous new web interfaces we're starting to see, which are changing how web apps work, and making them ever so much more rich and interactive -- for the first time, a real threat and alternative to the desktop.

Now we are getting some great insight and writing about how this is done.

Quoting Jesse James Garrett:
Take a look at Google Suggest. Watch the way the suggested terms update as you type, almost instantly. Now look at Google Maps. Zoom in. Use your cursor to grab the map and scroll around a bit. Again, everything happens almost instantly, with no waiting for pages to reload. Google Suggest and Google Maps are two examples of a new approach to web applications that we at Adaptive Path have been calling Ajax.
The name is shorthand for Asynchronous JavaScript + XML, and it represents a
fundamental shift in what’s possible on the Web. Ajax isn’t a technology. It’s really several technologies, each flourishing in its own right, coming together in powerful new ways.
Again, all credit to the link which led me there:

Preoccupations: "Ajax, who he?"

Marc Canter is a genius -- the "Digital Lifestyle Aggregator"

This is one of those off-the-top-of-the-head posts that would be oh so much better if I polished it and edited it and moderated the tone a bit. But I am really, really geeked out at finding Marc Canter's 10-month-old post on "Digital Lifestyle Aggregators" which is near enough to a spittin' image of what I have been working on for the last year, knowing nothing of this essay. We call our version the Extended Conversation, and since we're coming in from a telco tack, we are spending a lot of time thinking about voice and phones as a crucial component of this -- something that I think the largely web-based thinking of the U.S. folks may have missed. Or, I could just be a prisoner of the installed base of my client -- as Marc says, "...Based upon what assets, IP, market position, resources and existing technology base – each DLA customer will have their own unique set of requirements and implementation details..."

So true.

Marc says he works for Tribe.net, though he also says he hopes they fire him. Is this what Tribe is doing? Not that I was aware of... will find out.

Updated: Mark runs Broadband Mechanics, which is a technology provider to Tribe.net among [apparently] many other firms.

All credit due to the links that led me there:

"Thoughts on the Digital Lifestyle Aggregator"

"From blog to DLA" (tb)

Thursday, February 17, 2005

Three quick links - Web2 calendaring and mapping

Want to capture these though I don't have time to comment right now.

1) JWZ aka Jamie Zawinski of Netscape (past) and DNA Lounge (present) on "Groupware, Bad!" and the perils of not getting your customers laid... I found this via the peripatetic Tim Bray.

2) A guy I never heard of before, JGWebber (Joel), who I found via JWZ's site. He has a very detailed analysis of the nutz n boltz driving Google Maps, which incidentally was created by a friend of a friend. Under the radar, Google is doing a lot of hiring-by-acquiring in the Bay Area. About a year ago, I told this guy that he was totally wrong to be working on this project. So much for my omnipotent authority (re: another forthcoming post on that subject...)

3) A company I never heard of before, SandCodex, which has some interesting interface and mapping technology. Well, the demo is cool, anyway. I got this from JGWebber's site.

That is an example of the extended conversation... Maybe I should revisit del.icio.us for this kind of thing.

Thursday, January 27, 2005

So what does A9 Yellow Pages Mean?

I try to avoid me-too "look what just happened" blog posts, and despite a ream of content on that last amazon post, a lot of said content fell into the "wow how cool" hand-waving bucket.

Before I natter on a bit more, John Battelle (as usual) has nailed it in pithy form in his Business 2.0 story:

It's pushing Amazon's virtual-commerce business model -- where you can buy
anything you want online -- into the bricks-and-mortar local retail space.
Amazon wants to be part of any kind of commerce.

Update: And finally, before I natter, check out this very funny note from John Aboud, via Seth Godin.

Commence my nattering. Here's what this announcement actually means:

1) Amazon used to sell stuff. Call that the "database model" where they shipped first information, and then goods, from a centralized location.

2) Amazon created the marketplace. Call that the "network model" where they acted as an info-mediary who brought together buyers and sellers, with goods shipped around the periphery of the infospace, peer-to-peer, without Amazon actually ever touching them.

Note that quarter to quarter, Amazon derives either a majority, or all, of its profits from the marketplace.

3) Now Amazon has created what I would call the "embedded network model" which extends the network model into the fabric of everday life. Contrast to the marketplace -- you don't care who the seller is, or where the seller is, you just want the goods you found on this mystical Amazon site to show up at your doorstep. In the "embedded network model" Amazon is enabling you -- with your real location, real needs, and real-world context-- to seek out merchants who are also local, physical, and with whom you may interact directly. Today, it's not really embedded -- it's just a first-step extension of the marketplace model. But look where it can go:

1) Amazon links to the catalogs and inventory of the merchants, and you can search that. Amazon provides "merchant tools" (which eBay and PayPal are also working on) to help merchants do this.

2) Amazon offers its Visa card/shopping cart as a pre-payment system for its users. Want a pizza? Find it online, order it online, pay online, and go pick it up.

3) Amazon, working with a delivery partner, enables local delivery from local merchants for same-day, or perishable, goods. Speciality grocery + Amazon = WebVan 2.0??

And these are just the commerce implications. The non-commercial (social) implications are even more interesting -- see my previous comment on the Viper Room. Very, very interesting.

Google Voice - The Extended Conversation

The problem with the extended conversation it that it happens in so many places that sometimes, it's hard to consolidate, even in your own mind.

January 17:
First, there were rumors of Google getting into the voice business.
Then, John Battelle posted a piece on the rumor.
Where I wrote a comment.

For a week, it was all quiet on the Western Front.

January 24:
And then the whole world went absolutely apeshit.

And, incidentally (or not?) some guy from Bulgaria named Dimitar Vesselinov posted a comment on Battelle, nailing the issue:
It makes sense. Why? Do you want to search for your voice data? What about
transcripts? People want this kind of information and I suppose they will get
it. It’s about the data, stupid!

Throughout this all, I somehow had an image in my mind that I'd written a pithy post about all this and put it up here, but somehow, I never did so. This is now that (non-pithy) post.

Ahem. Let me discuss the Extended Conversation.

Oh, hell, let me just demonstrate it by linking to a photo on Flickr of a slide that I've made which introduces the "why" for both the Extended Conversation, justifies Google's move, and agrees with Dimitar, all in one go.

See? Wasn't that cool? It comprised multiple forms of media, hyperlinking, tagging, user-created content, diary --> conversation --> publication fuzziness, and about twelve other principles of What's Going On here in this new new world.

We used to have "conversations" on the telephone. Some of us had "correspondence" via letters, but not since the days of Voltaire could letters be considered "conversations". (Incidentally, did you know that in London, the Post Office delivered mail four times a day before the advent of telephony?) Then we added email, and IM, and suddenly the web went broadband and storage got cheap, and somewhere in there search engines actually started to work, and then a bunch of crazies went and invented the whole blogging paradigm, and there were digital cameras everywhere, and social networks had their moment of glory but left a lasting impression, and then (finally, in the US!) cellphones started getting smart instead of dumb (if only mobile operators would follow, the dummies!) and we are starting to see the emergence of this big gnarly thing called the extended conversation that is going to encompass everything in exactly the same way that the phone system used to touch on everything, except this time we aren't just stuck in voice, and we can tie it all together in a pretty package with a bright red ribbon.

Some people call this Web 2.0. I call it the Extended Conversation.

So let's look at Google. I think they get all, or most of this, already. Firstly, they recognize that search changes how you use information, most especially when you've got a lot of info to manage. Do you even look at the call record on your phone bill? Probably too much information. Do you search back through your email inbox to find stuff you wrote a month ago? Of course you do, because you can, especially with Lookout or Google Desktop or one of the really useful personal search engines.

They get user-created content, too. They have blogger, they have Picasa, they have GMail. They just launched Desktop. They get that user-created content equals communication... I think. When you look at the SIMS study, it's clear that the vast majority of "the world's information" which Google is determined to search is user-created content (see the Hal Varian "How Much Information" SIMS Berkeley study) or read the fantastic article "What's Next for Google" that Charles Ferguson wrote for MIT's Technology Review a few months ago, you can't help but think (if you are a Googler) OK, what are we gonna search/own/monetize next? And the big flaming answer is voice, voice, voice.

So yeah, they're gonna do Google Voice. But maybe not yet.

Why are they gonna do it?

Because all the content in the world comes from people, and most of it is voice.

How are they gonna do it?

Well, if I were a Googler told to do this tommow, I'd turn some of those 250K Linux servers they have into POTS NAPs around the United States and Canada, and I'd run Asterix on them (or a highly robust Googlized version of Asterix) and I'd throw the UI to it in with GMail (or keep it separate, but hold hands with GMail behind the scenes) and make a great inbox and outbox and CDR function, and record every call (allowing the user to delete, etc., just like email) and allow calls to be annotated and hyperlinked and blogged and so forth... and maybe even think about real-time linking of points within the call and what I am doing in the browser, for instance while I am talking (via Google Voice) I do a Google search, and that search is tagged and linked to that moment of the conversation. Whammo, you have a really compelling application-level extension to voice, and another knee in the groin of the dumb transport companies.

Oh, my. This is gonna be a fun decade.

Amazon takes a whack at telcos -- application trumping transport

A9 just annouced a way-cool extension to the Yellow Pages concept. From the Washington Post:

For the past several months, a fleet of nondescript vans equipped with digital
cameras and proprietary computer software has been traversing the streets of
several major U.S. cities, continuously photographing businesses on every block.
Operating secretly under the code name "Project Mercury," the vans have
transmitted more than 20 million images to a database compiled by A9.com Inc., a
wholly owned subsidiary of Amazon.com Inc. A9 set up shop last fall in Silicon
Valley, far from the Internet retailer's Seattle headquarters, with the goal of
creating new search technology for computer users to hunt for information and
products on the Internet.

Today, millions of the photos, coupled with their corresponding electronic yellow page listings, are scheduled to become available on the A9.com Web site. Only businesses that have paid to be in the A9 yellow pages will be featured. A9 will invite those retailers to take digital photos of their own that can be added to the Web site, potentially taking computer users inside stores to view merchandise.

A9's offering includes a way for computer users to take notes in an online "diary," and save and search the comments. It also offers a feature that automatically connects users to businesses by phone. Users input a phone number on the Web page, and, with a click of the mouse, can effectively be connected to any of the businesses listed free of charge. Special software connects the call by ringing a user's phone and the line at the business.


Here is an example search result from Boston.

There are two important points here. The first point is that geography is really, really important. The fact that all these businesses have a location matters to their customers; the fact that Amazon has GPS-tagged all the information around the pictures and listings is what makes this service work.

The second, even more important point, is that communications is becoming an extended conversation, and applications are trumping transport. When the world was about plain vanilla POTS, being a transport company (hello, AT&T) was a very valuable thing to be. But on the day it is announced that SBC might buy an AT&T that is a shadow of its former self, Amazon announces a feature that -- almost incidentally -- offers a free phone call between the user of its website and the merchant they are trying to reach. Given that the original intent of the yellow pages was to stimulate phone calls to businesses (and charge for the phone calls) that's an incredible turnabout.

Coupled with the potential for Google to get into the VoIP market -- again, as an extension to its current application-centric business of enabling conversations, not communications -- this now qualifies as a trend.

Incidentally, this also demonstrates that Time Warner was exactly right to buy AOL. They just bought at the wrong time, and may have bought the wrong company.

Also incidentally, the ability of users to add their own photos is really important. If Amazon is smart, they will allow a lot of clustering and linking to happen that has *nothing* to do with the core commerce focus of their yellow pages -- for instance, people linking their party photos from the Viper Room in LA to the geo/location of the Viper Room. There are all sorts of eVite possibilities here, too... it just goes on and on. Welcome to the future.

Finally, last incidental point -- Flickr had better get all the way to platform status soon, or sell themselves to the highest bidder, or they are going to be overwhelmed by the rollout of smart photo-inclusive features like this one.

Voice will be taken over by application providers if the telcos don't do something smart, or something sneaky (AKA regulatory relief). Get hot, boys...

Monday, January 24, 2005

Bellster, Technorati, Slashdot?

When I do a search on Technorati (5:40pm PST, 1/24/2005) for "Bellster" Slashdot is not any of the 41 results, even though Slashdot has a 128-comment post/article on Bellster that was created at 1/24/2005 at 5:08pm EST (3.5 hours ago). Why is that? Too recent? Slashdot hasn't implemented Dave Sifry's server-side code? Will try again tomorrow.

Update: It showed up on 1/25 at about noon.

Jeff Pulver at play again: Bellster launches

As noted by Kevin Werbach and others, Jeff Pulver, the originator of FreeWorldDialup, the Pulver Innovations WiSIP phone, and a host of other telco-tweaking apps, is at it again. He's launched Bellster, which allows its participants more-or-less no-extra-charge phone calls worldwide. "No extra charge" is significant because it's not free -- in order for the service to work, each participant has to have some sort of reasonably bulk usage plan in their local calling area/country.

My take? RIght now, the short answer is, it's a stunt. A cool stunt that is interesting and will cause big telcos more worries than actual revenue problems. Like BitTorrent, its gnarly technical requirements will keep it from touching more than 1% of the userbase. Basically, you have to set up a dedicated Linux Asterix PBX server in your house to share your phone line with others. Not something that the average user is going to do.

Now... when someone rewrites Asterix into a PC application that runs on Windows, with a slick user interface a la Skype, and adds Bellster functionality, and says to every dummy (cough) user in the world, "Plug your phone line into your PC, download our app, and you get free phone calls worldwide" ... then, the telcos have a problem.

Friday, December 17, 2004

The three-axis auto-match and Match.com

I met my girlfriend, Andrea, through Match.com, and thus I view the capabilities of such matching services in a fairly positive light. However, onece one sets out on the quest of finding across distance that "perfect someone" via multilple search categories, one must commit to the notion that an ever-more-perfect search mechanism will yield ever-more-perfect results.

The previous post on Matt Jones' bookshelf and my forthcoming post on location-based blogging beg the question -- what are fast and easy ways to identify whether someone is likely your type? I would argue that you can analyze three axes and get very good results on the interesting-conversation front, and probably the I-want-to-date-you front as well:

1) Personal: Height/weight/gender/gender preference/availability. Just the basics.
2) Bookshelf: What do you have on yours?
3) Geo-log -- where have you been, and how much time have you spent there?

Voila. A whole new way of doing dating.

More on Amazon - "My Work Bookshelf"


My Work Bookshelf
Originally uploaded by blackbeltjones.

Someone with the screen-name 'blackbeltjones' has uploaded his work bookshelf onto Flickr and annotated the jpg with the actual titles of the books. Can I just say that a) Flickr rocks b) this is a cool idea and c) this is an obvious extension of the content of my previous post on this subject. Prabal, are you taking note?

Click through on this photo to the right, and the full coolness will be revealed.

Tuesday, December 14, 2004

Real search engines and great search engine humor (no, really!)

If you think that Google is the end-all and be-all, read John Battelle on IBM's WebFountain. Wow. That is how a search engine is supposed to work, but the reason you can't access it is that it costs too much per query to give to you for free, and you're too cheap to pay for it.

I can't remember where it was that someone pointed out to me that the true genius of Google wasn't that they created such a great search result, but that they created a pretty decent search result on a less-computational-cost-per-search than could be repaid by advertising. Per might remember, I need to ask him. (task placeholder)

I am going to cross-post on WebFountain over at Onohoku (link placeholder) because this is the scary shit that is destroying the middle class as we know it, in both corporate and governmental applications. Technology can make you free, and it can also make you a robot (which means 'slave labor' in the original Czech)...

The Battelle piece also led me to a truly wonderful self-parody hidden within the Google corporate site. I won't spoil it other than to say fly over there now and check it out!

Monday, December 13, 2004

More on Amazon and the extended conversation

On a day when Google announces that they're digitizing a number of major libraries, Prabal and I happened to be having a conversation about Amazon and where all this virtual/physical stuff is going.

Here are a couple of datapoints.

Firstly, Amazon has this handy "search inside the book" feature which allows full-text searching of books.

Secondly, Amazon now does this very handy "this book references... this book is referenced by..." relational search for references to and from the book.

Now. Most people's libraries are made up of books published in the last 20 years. The vast majority of these books have bar codes on them, which either contain or link directly to the ISBN number of the book in some database online somewhere. Here's a scenario:

January 1, 2005: Amazon places a special offer on its website. By clicking in a special box on their order form, for just $5.99, Amazon will include in your order a USB bar-code reader and a CD with special book-organizing software -- kind of like iTunes for your books. When you receive the software and plug in the reader, you can scan the bar code on every book you own; and for the ones that don't have bar codes, you can type in the ISBN or Library of Congress catalog number. That combination will allow your average book owner to get 90% of their library online at a rate of about five books a minute, or a few hours for a reasonable library of 1000 books. Press a button, and your entire library is uploaded to Amazon.

Some trust and fair-use issues would need to be resolved; Amazon might negotiate with the holders of the copyrights that anyone who physically scanned a bar code would get unrestricted search and access to the text of the books online. Perhaps this would be limited to books you'd bought through Amazon, and they would effectively get into the DRM business. Books that you'd typed in the LC or ISBN number would need to be managed a bit more closely; perhaps such books would drive a query from the site to look up and enter a quotation from a chapter in the book, or some such method of ensuring that the customer actually held a physical copy.

This would also lead to issues with people walking into bookstores and scanning books. However, this battle will be fought soon over cell phones with good cameras in them. Already in Japan, it is reported as a serious issue that kids will enter bookstores and snap photos of pages in magazines that are of particular interest to them.

What would the result of such an upload be? Suddenly, one could manage and search one's personal library online. Search from a text perspective, anyway. Soon I'm going to get around to writing about the "fruit theory" which causes me to believe that books on bookshelves -- even lots of them, such as my library of about 3000 books -- are a highly efficient mental model by which to keep track of information. What's a great example of this? Prabal takes pictures of his bookshelves. He has shelves and shelves of books in Ohio, whereas he is in California -- and he knows where on the shelf a certain book is (efficient mental model). With the picture, he can a) remind/search to a greater degree of precision and b) tell his wife Wendy using a coordinate system that he might not remember, but can now communicate, e.g. "it's a little to the left of the big red book on the 2nd shelf from the bottom," and FedEx does the rest.

For manifold reasons, people are always? or at least for a very long time, going to want real physical books. This combined system would be a beautiful way to combine the benefits of the virtual and physical, and Amazon would be dumb or nuts (and they are neither) not to pursue it.

And then... Steve Wozniak is doing virtual/physical integration via his "Wheels of Zeus" startup -- which is aimed at the "where's my stuff" personal RFID market. For a deluxe package of $59.95, Amazon could mail out a combined RFID/bar code scanner and a big sticker sheet of WoZ tags... and not only you, but your entire family , could suddenly find and access all those great books on your shelves.

The combination of web-based management, personal possessions which come from a limited dataset, and scanning/tagging technologies is going to be flat-out amazing. Imagine managing your wine collection this way. Your tools. Your shoes. Anything (unlike CDs) where the virtual must remain no matter how far digitization goes.

This is gonna be cool.

And the best thing is, besides this great conversation, we had a billion-dollar idea tonight! Now we just need a billion-dollar implementation... ;-)

Friday, December 03, 2004

The Extended Conversation

I stole this citation from John Battelle's Searchblog. But it's cool enough that I want to write about it myself.

"Rageboy" has discovered, and briefly discussed, that Amazon is now providing hypertextual citation in both directions -- showing both the books that are cited in Book A, and all the books that cite Book A. This means that you can now not only find out all about Book A, you can quickly ascertain where Book A sits in the entire pantheon of scholarship in its topical field. The similarity of this to Google's PageRank algorithm will not be lost on any of the search cognoscenti, and the potential of that alone is enormous -- in addition to reading review, finding out "people who bought also bought," and looking at best-seller lists, I can now choose my books on the basis of how often they are cited, or how broadly they cite -- powerful applications of the "hubs and authorities" model which underlies PageRank. When this is suitably combined with a nice visualization and discovery tool like TouchGraph, we might really be on to something. Three cheers for social networks of data!!

Thursday, November 04, 2004

Final Results: Givers for Kerry, Takers for Bush

A few weeks ago, I posted a comparison between the Presidential polls and the Tax Foundation's tax tables on state-by-state net federal tax receipts (taxes collected vs. federal spending). Now that the people have spoken, I have updated this table to reflect the final results. It's still a very strong correlation:

Givers for Kerry, Takers for Bush.

No wonder we have an enormous deficit...


tax to vote final 2004
Originally uploaded by onohoku.


Tuesday, October 19, 2004

Types of Social Gatherings

Brainstorming with Andy about different types of social gatherings and their purposes.
This is neither complete nor fully sorted. (DRAFT)

Sports --> Teams get together to compete, fans get together to be entertained and vicariously compete.
House party --> people get together to socially network, which is fun.
Conference --> Information exchange and social networking
Bazaar --> Buying and selling; commerce.
"Activities" --> pleasurable activity -- like dinner (tasty) skiing (fun) hiking (fun and exercise)
"Activities" (b) --> Practice - increase skill for performance or competition
Performance (e.g. symphony) --> expertise/skill (of players) pleasure (of audience)
Cause --> result (fundraiser, politics, etc.)
Collaboration --> work/task
Education --> information exchange only (vs. +social NW of conference)
Interfaces (Waltzbot/DDR; karate-bot; musical instruments that are net-enabled and data-capture-laden)

Types of outputs of work/collaboration:
--> document
--> object a car, a chair, a house)
--> idea (this brainstorm... if we didn't write it down - docs are frozen ideas)
--> other action (e.g. Karl Rove gets you to vote Republican; his action begets your action)
--> performance? How does practicing for a play fit in? What is created there?

Thursday, October 14, 2004

Divded States: Givers for Kerry, takers for Bush?

Timothy Noah wrote an interesting piece on Slate today, which defends Massachusetts against the implicit attacks which George W. Bush levelled against it last night. The core of Noah's argument is that the much-maligned liberal hotbed of "Taxachussets" pays a heck of a lot more into federal coffers than it takes out -- in other words, it is a net giver to the country as a whole.

He links to Tax Foundation table of federal expenditures and receipts state-by-state.
I found this table absolutely fascinating.

I dropped this data into Excel, re-ordered the states by their FY2002 rank, and created a new column which represents the current (as of today) state poll standings for Kerry vs. Bush, taken from electoral-vote.com. Using Excel's handy conditional formatting, I colored this column blue if Kerry is +4 or greater, red if Bush is +4 or greater, and grey if the state is from -3 to +3.
The results are quite interesting, to say the least.
I am happy to send this original Excel document to anyone who wants it, but I have attached, for easy viewing purposes, a .jpg of the table. The main impact is visual, and the message it conveys is simple:

States which are net givers to the federal pie predominantly vote for Kerry, while states that are net takers from the federal pie predominantly vote for Bush.

This Bush hegemony of the dependent includes a large number of Western and Great Plains states where the "culture of self-sufficiency" and derision of poor urban welfare freeloaders are wildly predominant points of view.

A couple other observations jump out:

1) Just 16 states (about 1/3) are net givers, while 32 (about 2/3) are net takers. Two (FL, OR) are even at 1.00. This means that, grossly speaking, the givers have large economies, while the takers have small economies. Since the leading 16 are dramatically more liberal than the trailing 32, we can also conclude that as one's economy grows and one's state becomes more prosperous, it also becomes more liberal.

2) This is most aptly demonstrated by the case of Colorado, which has leaped 17 places on the list between 1992 and 2002, from significant net taker (1.06) to significant net giver (.79). In the same 10 years, the modern economy of Colorado has exploded; its population has vastly increased; and the state has become significantly more liberal, to the point where it remains very much in play for the Presidental election, despite its as-of-today +6 for Bush poll standing.

I could go off in any number of directions with this data set, but I'll leave that to the members of the list. If I have the time, I intend to drop in some state GNP data, and also electoral-vote data and possibly population data, so we can see the typical profile of liberal vs. conservative states.

Two points that I will add are:

a) I am just now finishing Neal Stephenson's 3-volume, 3000 page novel on the Baroque era, when both the Scientific Revolution and the commercial revolution transformed first England, and then the world. Stephenson has done a superb job of laying out that the competing factions of Whigs & Tories in England fought, quite explicitly, over whether the wealth of England would be based on commerce (the Whigs) or on land (the Tories, allies of the aristocracy). Fortunately for all of us, the Whigs won - I'm much happier as I am that I would have been as a gentleman's man-servant, I suspect. This recent exposure to these ideas, when added to the visual impact of this attached data, causes me to wonder if similar arguments can be made -- on the basis of this hard data -- about the different political parties of the United States.

b) I am going to go way out on a limb, and say that not only will Kerry win the election, but that he will win the election by capturing swing states Colorado, Ohio, Wisconsin, Pennsylvania, Michigan, Nevada, Florida, and New Jersey. A glance at the list will demonstrate why I chose these.

Tuesday, August 03, 2004

Watch the election unfold... poll by poll

The sort of information aggregation that Electoral Vote is doing is really quite impressive. With a little bit of elbow grease and a good web implementation, they are presenting a very interesting and valuable piece of present knowlege out of a flurry of confusing numbers and statistics. This sort of interest-driven site is a classic example of user-created content, and in the same way that the PTA makes your kid's school run better, these people are subtly but definitely improving the workings of our world. Well done.


Wednesday, May 05, 2004

Social Network Theory applied to Political Books

I continue to find social networks of information more interesting and informative than social networks of people. Perhaps this is because the richness of data available for mining on the Web of social information is so much greater. In five years, I can imagine that the existence of long data exhausts from Match.com, Friendster, political donation lists, Yahoo groups, etc., will create equally viable information sources for human social networks.

In the meantime, what we have to play with are social networks of information. Google, of course, is the exemplar of this field, with its PageRank algorithm based heavily on "what other sites think" of a given website. Amazon's 'people who bought X, also bought Y' is another fertile field; and we have a fascinating new analysis of political bestsellers from Valdis Krebs.

Here is the network map of the top 100 political books on Amazon, arranged according to their degree of relatedness. It demonstrates what Krebs rightly calls 'echo chambers' of thought on both the left and the right, with "debate replaced by hate" at the extreme margins, where partisans of both the left and right consistenly buy a tightly grouped set of books that most strongly reinforce each other.

This is a demonstration of the increasing segmentation of America, where sophisticated marketing, geographic mobility, and the twin philosophies of voter and consumer choice have enabled seemingly similar American citizens to exist in completely separate realities.

As an aside, I highly, highly recommend Charlie Wilson's War, which is one of the few books in this diagram read by both left and right. It's a superb story about rea-world people, and real-world politics, accomplishing literally unbelievable feats during the the closing years of the Cold War.

Wednesday, March 03, 2004

Changing the Laptop Game

A new generation of laptops is out, with very similar specs. IBM's new T-series, the Sony Vaio Z1 series, and Dell's Inspiron 600m all offer seemingly simple upgrades over previous generations of notebooks:

Sub 5 pounds (4.5 - 5.0 across the three)
Fast processors (1.5-2Ghz Centrinos)
Integrated wireless
Fantastic battery life (~ 4 hours)
Integrated CD-RW/DVD drives
Superb screens -- all three offer 14.1 inch SXGA+ 1400x1050 screens. Just to be clear -- this is almost twice as many pixels (1,470,000) as a high-res 1024x768 screen (786,000)

So what, you say? That's what I thought. I've been using my trusty March 2000-era Vaio F-480 for four years now, and I loved it. It was only its increasing mechanical creakiness from about 400,000 miles of use that finally made me buy a new laptop -- and I had no real expectations of a life-changing experience, because I'm not an early adopter tech-head.

WoW!! Was I surprised?!!

My new machine -- a Vaio Z1VAP -- is phenomenal. The keyboard is better than any I've used, even better than the wonderful one on the F-480. My one (small) beef is that the absence of separate home, end, pgup, and pgdn keys makes fast editing difficult -- these are critical keys for sentence, line, and paragraph selection, and pressing ctrl+fn+down arrow is a hack. But this gripe is utterly washed away by the entirely new usage patterns I'm seeing after just two days. The battery actually, uh, works. I can go places with no adaptor -- the thing is light enough to carry and forget about. The screen is incredible. Not only is the resolution surreally sharp, it's brighter and whiter than any display I've ever seen. I assume that the other two machines are using similar quality LCDs, and the difference from 1024x768 is incredible. You truly can no longer see individual pixels in most situations.

This thing is going to change my working life. I spend hours a day in PowerPoint, and I can now see so much more in nearly every view. I spend hours day in email, and I have screen real estate like I never imagined. I am now actually considering buying a Tablet PC, because for the first time in years (since, oh, 1995, when I got my first laptop) I have actually had my user experience significantly changed by new hardware, and I am now open to the idea that additional value might actually exist in a different and innovative hardware configuration.

Go buy one. You can get "low end" versions of all these machines for about $1300, and anyone who's doing anything remotely productive on their PC will get real benefits right away.



Tuesday, February 17, 2004

Reed Hundt on Big Broadband

Reed Hundt, formerly head of the FCC, wrote a recent thing piece on the "inevitability of big broadband". As you might imagine, his speech lays out that it's far from inevitable -- unless the government does something Real Soon Now.

His argument seems to be for the creation of a single, highly regulated, universal high-speed fiber network across the USA, which would be able to pay for itself by charging regulated fees for voice ($40) internet ($25) etc. He views the parallel creation of multiple competing 'mini-broadband' networks by telcos and cable as a bad thing, and the attempt to support the old copper phone system and old universal access scheme as a bad thing.

Smells like Ma Broadband to me.

The piece is surprisingly short on "why to do it" beyond vague hand-waving about how the network will realize untold benefits. Has anyone actually figured out what to do with broadband networks?

One interesting tidbit is Hundt's summary list of all the old "vs." arguments of the past decade:

Since the beginning of convergence, dated from about 1992 (plus or minus a year), the battle to be the primary medium of at least the next decade - the one we are in now - has raged among various antipodal rivals: content vs. conduit, local vs. long distance, wireless vs. wire, data vs. voice also sort of known as packet vs. circuit, communications vs. computing, network vs. edge, and copper vs. HFC (also known as telco vs. cable). Other, possibly lesser dialectics include satellite vs. terrestrial and broadcast vs. cable. Convergence describes then a clash of networks, businesses, and even cultures.

Monday, February 16, 2004

Microsoft -- winner of the coming convergence collision and the next media empire?

William Safire, whose conservatism is on best behavior when he's focusing on the important issue of media consolidation, has just predicted that Microsoft will become one of the three biggest media companies in the world, acquiring ABC/Disney, NBC, or possibly even both.

Remember, Microsoft is sitting on a $49 billion cash reserve the last time that anyone checked.

Bob Cringely, who is entertaining even when he's not prescient, has an entertaining and possibly prescient claim that Microsoft is laser-focused on owning the next generation of converged computing/entertainment/communications through control of underlying protocols such as Windows Media Player 9. What Bob is probably missing here is that Bill Gates is as likely to fight the next war (media) as the last war (standards). Why worry about owning the standards when you can just own the media (or a large chunk of the media) itself?

Over in the communications side of the converging global future, AT&T Wireless is being sold to either Vodafone or Cingular; and win or lose, one of these two will almost certainly merge or acquire T-Mobil in the U.S. within the next 18-24 months. This will shake the U.S. cellular market (long overdue, pundits say) down to perhaps 3 1/2 companies by 2006 -- uber-dogs Verizon/Sprint (on compatible CDMA) and Vodafone/AT&T/T-Mobil or Cingular/AT&T/T-Mobil; and trailing lap-dog Nextel, plus one other component that somehow loses out of the great mating dance.

All of this is about the same thing -- convergence of communications, computing, and media/entertainment. The stakes are simply all the money in the world. Enjoy the show!

Tuesday, February 10, 2004

Barney Pell Responds on the Value of SNS

My unidentified uber-LinkedIn friend in the previous post is Barney Pell. I sent him a link to the piece, and he's responded thoughtfully and at length. I'll attempt to address some of his points in a forthcoming piece, but on others, he's probably got me pinned. Three cheers for network-mediated pruning of ideas!

Barney writes:

Initial vs intended use: it's definitely a good point.

To paraphrase: Initial use is just playing around, people won't pay for
it, and we have no evidence that people will pay for the intended use or
that the intended use will really deliver value for the users at all.

But:

1. There is still value in giving something away for free until the
network effect kicks in and starts driving the intended value. That's
just basic business strategy in networked information goods. Of course
it is risky (napster) but there have been significant successes.

- paypal was free, and paid users to spread it virally, who did so
without much in the way of real transactions for quite a while (I'll bet
the # of paypal users was huge compared to the # of ebay buyers for
quite a while). Then it wound up becoming a money-generating machine.

- classmates.com and the dating sites (match.com) are money machines.
They all started out free until they built their network. It doesn't
take a huge imagination to see that the social networking sites
(friendster, tribe, orkut, etc) can compete in the dating game once they
hit sufficient scale.

- for that matter, napster might well have made money had it not been
for the lawsuits. I disagree that they failed to monetize completely.
It is quite possible that their time ran out before they managed to
monetize, in which case the lawsuits were a material cause of this
failure (not a red herring as you suggest).

2. That these SNS sites are growing so fast leads one to ask whether
there might in fact be true value even in the initial use (besides
entertainment from building a network). One such value is that of
"organizing and tracking" your friends. People change emails frequently
nowadays (esp personal ones, given spam problems and links to ISPs). If
we all sign up on a SNS, then so long as we update our contact info with
the SNS (which is a condition of use), then we'll be able to continue
to find each other independent of these address changes.

Related to this theme, I have found a recurring pattern on linkedin of
people wanting connections to people they already know but have lost
touch with. And I have been reunited with acquaintances. There really
is power in network connections. It's just like when you meet an old
friend and ask what happened to so and so, and hope that one of you is
still in touch or otherwise tracking the lost acquaintance.

3. If it isn't for the initial value (that of network building and
tracking), we need to ask what these users are expecting to get
eventually. The best time to build a network is when you don't need it.
You only change jobs or need to raise money once in a while, but when
that happens, you really need your network already built. So these
efforts can be viewed as a form of network insurance for future use.

The insurance analogy is actually quite interesting. The ratio of
insurance claimers to insurance payers is probably less than 1%, which
you found so damning with respect to linkedin... Yet people go right on
paying, and it certainly isn't done for mere entertainment value.
Initial costs with the expectation of ultimate benefits are actually
quite standard in traditional markets as well as online services...

And all that was generated at 2AM after a long day's work at NASA. Barney, you really have to start blogging!

Thursday, February 05, 2004

Does Clay Shirky's Howard Dean piece fundamentally question the value of social software?

My previous post discusses Clay Shirky's fantastic analysis of Howard Dean's Internet presence. I've been thinking about the larger implications of this analysis, and I'm beginning to wonder if Clay has actually suggested that many of today's "hot" SNS services are a tempest in a teapot, and that we're wildly over-reacting to them in just the way we did to Howard Dean's seemingly impressive early momentum.

The neatest trick of Clay's analysis is to flip our perceptions of several key Internet-related Dean events on their head:

Shirky quote: "We were right to be excited about this MeetUp, but wrong about the reason, because MeetUp was founded to lower the coordination costs of real world gatherings... Prior to MeetUp, getting 300 people to turn out would have meant a huge and latent population of Dean supporters, but because MeetUp makes it easier to gather the faithful, it confused us into thinking that we were seeing an increase in Dean support, rather than a decrease in the hassle of organizing groups."

Ethan's derived lesson: Old-world metrics (300 people at a political gathering) are not applicable to new-world-driven phenomena.

Shirky analysis: "Dean's campaign was never actually successful. It did many of the things successful campaigns do, of course -- got press and raised money and excited people and even got potential voters to aver to campaign workers and pollsters that they would vote for him when the time came. When the time came, however, they didn't. The campaign never succeeded at making Howard Dean the first choice of any group of voters he faced..."

Ethan's derived lesson: It is clear that key metrics (in this case, votes) are crucial in the old world or the new; but old relationships between different metrics (in this case, people at a political gathering --> translates to votes at the poll in some predictable ratio) are not applicable to new-world driven phenomena.

Now how does this apply to social software?

Well, I can't count the number of times I've heard that XYZ product is "the most successful product launch in history." This is always followed by a comparison of number of units of XYZ sold in some unit of time (months, years) as compared to past wild successess like the television, the telephone, the fax machine, the VCR, etc. For instance, here is a Y2K press release saying that the DVD is the most successful CE launch ever. A bit less impressively, RCA claimed similar historic status for the VideoDisc in 1981.

This old-world metric -- units sold in a given timespan -- is reasonably applicable when we consider one electronic device vs. another. While the number of underlying represented changes is large -- better retail distribution, better media publicity for a new product, more central role of media and electronics in consumer life -- we're still in apples:apples territory when we compare the DVD player, VCR, TV, and radio.

Not so SNS. This is for three distinct reasons:

1) Any free product cannot be compared to a product which is purchased.
2) Any ephemeral product (meaning, can be accessed through other than physical means) cannot be compared to a physical product.
3) Any product in which the initial use is different from the intended use cannot be compared to a single-value-proposition product.

Let's break these down.

First point: No consumer product organization in the world -- Proctor and Gamble, for instance -- would dream of conflating demand for a free, bundled, or drastically-price-reduced offering with the eventual market demand for that same product. And yet, the hype and excitement of new-to-the-world products or services, and investments on the back of those products or services, is often based on exactly such incomparable situations. Napster started out free, and stayed free -- and Napster and its investors (notably Hummer Winblad) never saw a dime for the costs they bore. The lawsuits were a sideshow -- Napster's failure to monetize was fatal.

Second point: We all talk about 'stickiness' and 'switching costs' but we forget that it's as easy to become unstuck as to stick, and as easy to switch away as to switch in. Ephemeral one-click-to-join products may attract many, many members or users quite quickly -- Friendster, LinkedIn, etc. being great examples -- but with joining costs near zero, can we really compare this to the consumer awareness and commitment indicated by, for instance, going out and purchasing a DVD player? Hardly. We in the pro-blogging, pro-SNS, pro-software-network-collective world are fond of arguing that there is inherent value in this easy-to-start-up nature -- which is a true statement. But we can't measure new-world services, with new-world joining costs, in old-world ways. This is the Dean mistake which Clay has so rightly pointed out. Clay's piece cautions us that while we can celebrate the ease of one-click joining and the reduction in transactional costs it represents, we must at the same time discount the perceived value of every one-click-casually-joining member. If we don't, we are guilty of having our conceptual cake and eating its hype, too -- and that is the road to red-faced failure in the unforgiving marketplace.

Third point, and to my mind this is the most worrisome: The initial use of SNS is nothing like the intended or expected eventual use, and the claimed value proposition. Every argument I have seen for SNS, and every funded business model with which I am familiar, suggests that people will use the improved informational flow of SNS to do things, and the desire of people to continue to use SNS to do things will be monetizable by the SNS owner. Contrast this with how we use SNS today? We 'use' (if the word 'use' is appropriate, which is highly debatable) SNS to play connect-the-dots with our friends and acquaintances. If you doubt me, empirically measure the amount of time you spend in connect-the-dots vs. the amount of time you spend searching, leveraging, and deriving value from your dot-connected network. As an associated analysis, think about how easily these services enable 'connect the dots' (which is in fact their barrier to entry, rather than their use case) and how easily they enable use of one's existing network. In nearly every case, I would argue that while it is fairly easy to join and connect-the-dots, few if any of the major services have figured out a compelling user experience for actually deriving value from their service. Initial joining-related activities are easy, and in fact a fun diversion -- which, while worth celebrating (turning one's barrier to entry into an engaging game is a Good Thing) strongly suggests that measuring the number of members in a SNS is completely meaningless; a far more 'vote' like metric is the number of people using the network for actual business of some sort.

These numbers are incredibly, embarassingly, surprisingly small.

I recently got a self-puffing email from LinkedIn. One of my very few contacts on LinkedIn happens to be one of the most-linked people on the entire service, so I have a disproportionately large 'network' for my level of participation. That email said, in part:

"Have you noticed how your network has grown? Since you last logged in [note that this was about one week prior to my receipt of this email], 33,300+ professionals have joined your network, and you now have access to 62,800+ contacts who are linked to you through your 7 existing connections... And every day, over 300 LinkedIn professionals request a referral from someone they know..."

That's a fascinating ratio. 33,000 new people in my network (which represents a large percentage of the LinkedIn network, due to my uber-connected acquaintance) as compared to 300 who daily request referrals. That 300, of course, are unlikely to be 100% successful, which means that actual connections made through those referrals will be an even smaller number.

That's an actual LinkedIn use rate of just 1%, and a presumed success rate of even less than 1% -- which is hardly a testament to the power and value of SNS. Most important is the demonstrable fact that on any given day, 99.5% or more of LinkedIn members are not using the service for its intended purpose, even if they are active as beavers building their social networks. I would bet you any amount of money you care to name that on any given day, and despite our constant carping at their limitations, 90+% of the users of Microsoft Office are using those applications for their core intended purpose, and deriving real value from that use.

This core use vs. initial use distinction also applies to other web-based services, and once again SNS comes out looking the poorer. Ephemeral access or not, Amazon.com could tell when it had a customer. That customer a) gave Amazon money on a given purchase in preference to all other retailers that the customer could have used (see first point); and b) in that initial use, used Amazon for its ultimate and intended purpose -- e-commerce -- in completing that transaction. From that single transaction, usage, and addition of a customer, Amazon knew that one person had understood and benefitted from its core value proposition. In the SNS world, Tribe, Spoke, LinkedIn, and their brethren have no similar assurance.

Clay’s outstanding analysis of Dean and the Internet points us toward this final, most worrisome point. Dean’s Internet supporters did a superb, net-enabled job of finding each other, communicating with each other, creating buzz, and even raising money — but they failed at achieving their core goal and metric, which was to get votes for Dean. If other social Internet phenomenona such as SNS are similarly inept at their intended purpose — be it finding dates, selling stuff, or whatever — they will similarly implode, because all the metrics we’re using to measure their success, such as new users and rate of uptake, are as irrelevant as were the Deaniacs and their ultimately ephemeral movement.
Clay Shirky on Howard Dean and the Internet

I am a regular and avid reader of Corante's Many 2 Many group blog; some of the best in the business, including Clay Shirky, Ross Mayfield, and Dave Weinberger are regular contributors.

Despite the very high level of ongoing content creation, Clay's recent piece on Howard Dean and the Internet stands head and shoulders above the rest. It is an absolutely brilliant piece, and contains key insight after key insight. Below are a few teaser excerpts; rest assured that despite the pithy bromidic nature of my quotations, Clay supports each of these in full and fascinating fashion:

"The press has a way of running fast epidemics, where an idea virus runs its course quickly, leaving everyone inoculated in its wake."

"The first time Dean appeared on our radar was when 300 people showed up for a Howard Dean MeetUp in New York City in early 2003...We were right to be excited about this MeetUp, but wrong about the reason, because MeetUp was founded to lower the coordination costs of real world gatherings... The size of the MeetUp in NYC was as much a testament to MeetUp as to Dean... it created a false sense of broad enthusiasm. Prior to MeetUp, getting 300 people to turn out would have meant a huge and latent population of Dean supporters, but because MeetUp makes it easier to gather the faithful, it confused us into thinking that we were seeing an increase in Dean support, rather than a decrease in the hassle of organizing groups."

"Margaret Mead once said “Never doubt that a small group of thoughtful, committed people can change the world. Indeed, it is the only thing that ever has.” Generations of zealots have tacked these words up on various walls, never noticing that the two systems that run the modern world – markets and democracies — are working right precisely when they defeat these attempted hijackings by small groups."

"Money does not in fact buy votes, as candidates like Michael Huffington and Ron Lauder have shown — you can be very rich and still very lose. In Dean’s case, though, the effect was compounded by two other effects from above. By moving campaign donations online, they made it much easier to donate, so much easier in fact that raising millions from individuals was never the sign of strength we thought it was. (Support isn’t votes.) Like MeetUp, a lot of what the campaign achieved was by lowering the threshold to contributing, which helped create a false sense of strength."

"Since the 1970’s, anyone who has looked at the cultural effects of the internet has picked the same key element: the victory of affinity over geography. The like-minded can now gather from all corners, and bask in the warmth of knowing you are not alone... Voting, though, is the victory of geography over affinity. Deaniacs in NYC could donate money and time, blogging like mad or tramping through the cold to talk to a handful of potential voters, but they couldn’t actually vote anywhere but NYC. Iowa was left up to the Iowans."


Absolutely brilliant stuff. There are more real insights in this single piece than a week of the New York Times' editorial page. I cannot recommend it highly enough.
Spybot Search & Destroy: Remove your spyware

I think that one of the most important uses of blogs is social filtering of information. So when I find a product that I really, really, like, I am going to write about it and link to it, so that you can find it here; and I can participate in the collective web of recommendation and trust that allows us to have some confidence in unknown-to-us vendors and products in the great mishmash that is the Internet today. With some reports of spyware posing as spyware removal tools, this sort of real-person filtering for what works is crucial.

Spybot Search & Destroy is the best spyware removal tool. It's fast, efficient, is updated regularly, and simply works. It combines an easy interface and complete user control, which is harder than you might imagine. Best of all, it's free. Get Spybot Search & Destroy here.

If you don't trust me (heh, heh) PC World says Spybot S&D is great, and PC Magazine says Spybot S&D is wonderful.

This message brought to you by me.

Monday, January 26, 2004

Visualizing Social Networks -- PieSpy

This is a newly released tool for visualizing social networks over IRC chat. What I find most useful about this site is the detailed discussion of both SNS representation and visualization algorithms, at a high enough level to be comprehensible to the non-programmer. If you want to know how the many visual tools such as Touchgraph work, check this out.

Thursday, January 22, 2004

Touchgraph and Amazon

Alex Shapiro over at Touchgraph (proof that Open Source can be truly innovative!) has added an Amazon browser to the way-cool Google browser. For all you teenage girls out there, there's a LiveJournal browser as well. The Amazon implementation is based on 'customers who bought also bought' and the discussion includes a link to a good IEEE article on recommendation algorithms, including collaborative filtering, cluster algorithms, and search-based methods. The article serves as a handy collaborative filtering primer.

Wednesday, January 21, 2004

Social Networks, Information Economics, and Search

SNS has become a big enough field with sufficient internal dialogue that it's easy to forget that the promise of the whole shebang can be reduced to a single sentence:

Reducing informational exchange costs, particularly transactional search costs, through leveraging existing information exchange pathways such as personal relationships.

That's it.

We're reminded of that singular purpose by a new SNS-driven search engine, Eurekaster. Danny Sullivan over at Search Engine Watch has a good article on this new twist on a major field.

And yes, Eurekaster, you should offer your service as a layer built on top of the Google API. That API is a fantastic tool for developers and experimenters; the fact that it acts as an innovation magnet for Google, by which the company can observe --> copy or observe --> acquire promising new technologies is just the price of the piper.

Monday, January 19, 2004

Interest cost and real value

While I am talking about property, I'll throw out another thought. It's conventional wisdom that home-buyers are most sensitive to their monthly payment, and declines in interest rates drive up selling prices in a near-perfect inverse relationship, as lower interest cost is swapped for higher sticker price, leaving the buyer with the same total cost. Call me contrarian, but I think that buying a house at times of historically low interest rates is madness. Why? Because you're fully leveraged. Let's assume that there is a nominal value to a house, and an actual value. The nominal value is the sticker price -- say, $400K for that two-bedroom apartment in Berkeley. Since the vast majority of buyers are on 15- or 30-year mortgages or their functional equivalent, a significant interest cost drives the actual value of the place -- the contracted financial commitment which the buyer undertakes in order to acquire title -- up to something like $700K. In times of low interest rates, this $700K (which is based, conventional wisdom tells us, on the buyer's ability to make 30 years x 12 months x a given monthly payment) might be made up of a nominal/sticker cost of $550K, and an implied interest cost of only $150K. In times of *high* interest rates, that $700K number might be made up of a nominal cost of only $320K, and the same overall total due to a vastly increased interest cost of $380K.

If I buy at a time of high interest rates and low sticker costs, it seems to me that I have nothing but capital gains upside from interest rates, which historically have varied cyclically (and which are due for a massive rise, if our trade deficits and the plummeting value of the dollar are any indication -- but that's another story as well). In addition to my capital gains upside (the nominal--> real value of my place going from $320K to $550K) I have the potential to refinance, getting a low low interest rate on a new mortgage to go with my low nominal cost, giving me a genuine total-cost bargain.

But what if I bought in times of low rates? The nominal value of my property can only fall on a cyclical basis, even if overall rising prices mask this somewhat; and if I've taken a variable-rate mortgage, my interest cost can only go up. In the best case, if I am forced to sell my house for any reason during a high-interest period, I'll lose out on its declined nominal value and avoid (through a fixed rate mortgage) any increase in interest costs. In the worst case, I am faced with a rising monthly payment on an asset whose value is declining precipitously, in an economy which is tanking because interest rates kill business investment. Ouch.

Have I mentioned that I'm feeling very secure being in cash?
Why are Bay Area housing prices still stable or rising?

The New York Times has an article today, the damn-with-faint-praise title of which is Job Losses Slow in Silicon Valley. Seems that we only lost 5% of local jobs in the past year, which is relative cause for celebration after the bloodbath of 2000 - 2003. What is staggering to me is the raw numbers. Peak employment in 2Q 2001 was 1.38 million; now it's 1.18 million, after a loss of 202,000 jobs. Average pay (meaning, I presume, wages; not income, which would have been further inflated by capital gains) was $81.7K in Y2K; it's now $62.4K. A little algegra shows that there was about $112 billion of annual salary in SV in 2000-2001; now there's about $74 billion.

For those scoring at home, that's a decline of more than a third -- 34% -- in local salary from 2000 -- 2003.

Since we are talking wage-earners here, it's not the $2 million options pads in Portola Valley that should be affected, but rather the $500K shoeboxes in Palo Alto and Mountain View. Instead, fueled by historically low interest rates, prices for homes have remained surprisingly robust during the past three years of nuclear winter -- an employment freeze that, if fear-mongering about job outsourcing overseas is true, may never actually end.

Six or nine months ago, you still regularly saw news articles about the 'housing bubble' in America, which laid out a case that people had pulled their money from the stock market and put it into real estate, taking advantage of low rates and upset at the declines in their portfolios. You don't see those articles any more, and yet prices continue to be surprisingly robust. What's going on here?

And yes, I am a frustrated wannabe-homeowner. But boy, does being in cash feel nice! Don't even get me started on the dollar and the 'which currency to be liquid in' question...

Sunday, January 18, 2004

Parties, Food, and Anti-Food

I just threw a dinner party last weekend, and in the post-mortem with my girlfriend, happened upon an interesting concept. The 'dinner party' as formally constructed probably originated around 1900; when the bourgeousie acquired dining tables and formal dining rooms and all the tat necessary to impress the neighbors. The key point of differentiation here is that a 'dinner party' is not for a special occasion like a wedding, funeral, holiday, etc; includes more than just family; and is thrown by people not of the nobility or aristocracy.

Food was then expensive, and manners were formal, so throwing a mini-feast for no particular reason was an impressive thing to do.

Today, dinner parties are all-but-dead; the closest analogue being a BBQ in the back yard during summer. I would suppose that this has a lot to do with cooking skills being on the decline, but I think that in fact the key difference is that food is now cheap, so having a party centered around a meal (when half the people there are likely on a diet) makes very little sense. In fact, most parties I attend these days center around activities -- volleyball in the park and a burrito afterwards; a hike or run in the hills and burgers afterwards; etc. It's as if anti-food has replaced food as the centerpiece of our socializing.

This needs a bit more thought, but it's a dichotomy worth recording in its present inchoate form.

Thursday, January 15, 2004

Visual Representation of Social Networks
Great overview of the visual representation of social networks over the past 30 years. Note to self: Read this in detail!

Friday, December 19, 2003

Hyperdictionary - designed to become a web service?

In my opinion, Hyperdictionary is the best dictionary on the web. Fast, clean, easy to use. And it's dead-easy to build the service into other sites and scripts, since each definition page is simply the word appended to the URL http://www.hyperdictionary.com/search.aspx?define= like this: http://www.hyperdictionary.com/search.aspx?define=petard

Very cool. This seems to be run by a group of Canadians known as Webnox - I don't know much about them, but they've got a great service.
ENUM and SIP

I've decided that I need to understand this in detail, very soon. I don't yet get all the implications, especially in combination with IPv6. ENUM is more or less a registry translation service done by an interesting Quango (Handy UK term - what's that mean?) called Neustar. Here is their whitepaper on ENUM -- I need to find similar discussions of SIP, process, and blog it here. To be continued.
John Battelle's Searchblog

This is an excellent site. On a quick read, I ended up at a really thoughtful discussion of Google. I highly recommend this in conjunction with Tim Bray's Search Series.
Searching gets smart filtering
This is damn cool. Google-enabled 'search my trusted universe' tool. When you start cascading it, you will *really* have something: "Search my trusted universe, and the trusted parties of my trusted universe, unto three degrees of separation. Go forth, my spider."

How to add it to your site:
http://www.alpern.org/weblog/php/blogsearch/mysite.php

How it works:
http://www.alpern.org/weblog/php/blogsearch/writeup.html

I found it on Kevin Werbach's site (he's implemented it) and followed it back to the source. Excellent.

Tuesday, December 02, 2003

Commodity hardware + open source = asterisk PBX

This in and of itself is not a game-changer. This is the equivalent of the "garage hosting" model of ISPs that was prevalent through the early 90s, with Mr. Modem Bank selling to local customers. However, it presages what larger players can do with the same technological building blocks -- after all, those little ISPs led to AOL and Earthlink.

"Asterisk is a complete PBX in software. It runs on Linux and provides all of the features you would expect from a PBX and more. Asterisk does voice over IP in three protocols, and can interoperate with almost all standards-based telephony equipment using relatively inexpensive hardware.

Asterisk provides Voicemail services with Directory, Call Conferencing, Interactive Voice Response, Call Queuing. It has support for three-way calling, caller ID services, ADSI, SIP and H.323 (as both client and gateway). Check the Features section for a more complete list.

Asterisk needs no additional hardware for Voice over IP. For interconnection with digital and analog telephony equipment, Asterisk supports a number of hardware devices, most notably all of the hardware manufactured by Asterisk's sponsors, Digium. "

Software: http://www.asterisk.org/

Hardware: http://www.digium.com/

Tuesday, November 04, 2003

Ray Ozzie and Power to the Edge

Ray says:

Power to the Edge: A new book by Dave Alberts and Richard Hayes - open sourced in its entirety by CCRP.

This book is truly a must-read for anyone interested in decentralization and the social and organizational relevance of shifting power to the edge, whether in a commercial or a defense context. As you read about the technology enablers of the edge, it'll become clear why products such as Groove - as COTS enablers of the fully-networked collaborative environment - have such immediate relevance to the defense community.


Note to self, and world: Read this thing. Yes, all 303 pages. It looks really interesting, and the DoD is funding most of the fundamental research in America right now.

Tuesday, October 28, 2003

Sony: Dot-Bomb Stock?

As Silicon Valley creakily starts to turn around and begin to innovate again, I have been required to fend off the doubts and dismissals of non-local naysayers who came late to the last tech party, and cannot believe that real value is created through the venture process. One of the favorite gambits is to point out the utter collapse of equity valuations even among the "real" companies in the tech sector -- those who had products and revenues, like Cisco or Sun or Inktomi.

This quote in today's NYT article on the Sony layoffs was thus a revelation: "Before today's announcement, Sony's stock closed at 3,960 yen, about 89 percent off its lifetime high of 33,900 yen in March 2000." Hmmm. 89 percent for a device company? 89 percent for a 50-year-old industrial bellwether? 89 percent for a company based 5000 miles from San Jose? Maybe, just maybe, the fault of the bubble was something called the stock market, not someplace called Silicon Valley.

Innovation is happening. These are good times for those who create.

Friday, October 03, 2003

English Attitudes
The English apply the same philosopy to their shoes as they do to their public schools. After seven years of confinement and torment, you end up with something really worthwhile.
The Metaphor and Reality of Silos

Every time someone starts talking about legacy enterprise applications, the term 'silos' jumps into the conversation. Silos are bad. We need to break away from all these old vertical silos to something horizontal, and dynamic, and cool. It's worth remembering a few things about silos:
1) They are designed to hold grain in and keep the rain outside
2) They also keep rats away from the grain
I believe that a new generation of dynamic software will improve the way enterprises operate immeasurably, within the next decade. And I think that whatever solution emerges will make considerable use of silos.

Wednesday, October 01, 2003

Bon Mot
"If you're paying rent in three places and still sleeping on the floor, that's bad management"
-- Prabal Dutta

Thursday, August 07, 2003

Last single person, turn off the Match.com?

It has just dawned on me that if forums like Match really are fundamentally more efficient at putting couples together than were traditional dating methods, we should see a measurable decline in the number of single people in America over the next 10 years as a percentage of the population, and possibly an absolute decline. Aside from some hard core of want-to-be-alone people, most single people are "waiting for the right one" and if you can find the right one faster, their numbers will go down. While there will always be churn and turnover, this should be a measureable artifact. It's kind of like the way that just-in-time inventory systems reduced the amount of stuff sitting in warehouses...

...associated wonder -- did all the Internet job sites reduce the number of unemployed, there during the bubble, through a similar effect? There aren't necessarily enough jobs for people (or people for jobs, for that matter) in the way that men and women are both approximately the same in number, and approximately the same in intent. So the two cases are clearly not directly comparable... but may be analagous.