Showing posts with label RWW. Show all posts
Showing posts with label RWW. Show all posts

April 22, 2010

The Buzz Bots Brouhaha and Other Empty Statistics

On Tuesday, ReadWriteWeb and a number of other tech blogs (including Mashable) stirred themselves into a tizzy around a report from social media measurement firm PostRank, who reported that almost 90 percent of content in Google Buzz, the company's nascent social network, came from news feeds and other non-native sources. Their self-congratulatory conclusions, kicking dirt on the service which hasn't hardly had a chance to grasp for air following body blow after body blow after a stumble out of the gate around privacy, came with phrases saying Buzz had "fallen short of capturing the hearts and minds of the social web", and another, "Much of Buzz’s content may be hand-crafted — it just doesn’t originate on Buzz."

Oh dear! What a terrible thing! Prepare the wake! I would be honored to act as a pallbearer, if given the opportunity.

PostRank's data served only one purpose - to count the total number of entries from all individual services being aggregated by the downstream collection service (Buzz). It is no secret, and I don't think anybody has successfully ever argued against it, that every single network that has the potential to import Twitter entries can easily be dominated by updates from that service. It was true with FriendFeed. It is true with Ecademy. It can be true with Facebook, depending on who you follow, and there is also that potential with Buzz. Twitter entries are light, easy, and update often. Even the most avid bloggers post more entries to Twitter than they do on their own site - as they post their own blog entries in addition to other updates, so it is no surprise that many people see their own feeds dominated by their Twitter. Add in Google Reader shares, which are very well tied in with Buzz, and you can see how this so-called "automated" "bot" behavior can start to ratchet up the stat-o-meter.

But the data (and reporting) was an inch deep, and it very poorly associated the world of automated feeds with the world of automated non-human behavior on other social networks which is nefarious. (Twitter bots, for example)

Aggregation services are built to aggregate data from disparate networks. It's practically their core definition. However, many of them also offer the option to create net new content. Buzz lets you make very detailed blog-like posts if you like. It's a great engine for sharing pictures, or YouTube videos, or simply adding a link from a Web page, and adding your observations. These so-called native entries are what amassed the approximately 10% of all updates seen as non-bot behavior by the survey.

As I mentioned back in March, native entries to Buzz get the most engagement. Some people, such as DeWitt Clinton from Google, make fantastic high quality posts on Buzz, and can get dozens to hundreds of comments. So do native entries from highly-visible people like Robert Scoble, and very active supporters of the community, including svartling and Leo Laporte.

I recognize that the vast majority of my own updates on Buzz are feeds. I love feeds. RSS or Atom, whatever you like, I recognize they are great tools. I will not stop sharing my items from Google Reader on Buzz unless the community goes silent, or tells me to stop - but between my Reader shares and my native blog entries, my own 4 native Buzz entries represent a miniscule 2% of my last 185 Buzz updates (according to Buzz-Charts).

This doesn't mean I haven't supported the platform, and it doesn't mean people are not engaging. In fact, according to the Buzz-Charts site, the last 185 posts received 478 total comments, or between 2 and 3 replies for every single entry. Some got none, and others got much more. But every single share was originated by a human, and every single comment was a so-called native entry.

PostRank surveyed Buzz with the subtlety of a political pollster operating under payment by the Committee to Re-Elect the President. (CREEP) While it may be factual that many people imported their Twitter entries and other automated feeds to Buzz, the Buzz community knows to largely ignore those items, or to not follow those people who are using the service in a poor way.

Real Buzz Data That Shows Real Engagement By Source

Buzz-Charts shows a much different story. In their graph of actual replies (comments) on all entries in Buzz, they found that of 24,584 posts that gained comments, 13,520 of them were native Buzz entries, a full 55%, not the measly 10% created. In second place, Google Reader shares showed 3,978 entries with comments, or 16%. Add on Buzz Mobile's 1,572 active threads (6.4%) and those top three conversation starters represent more than three quarters of all active threads, more than 18,000 of the total 24,000.

In contrast, all those automated Twitter entries that PostRank said was 62% of Buzz? They only had 837 active threads, or just over 3% of active threads. That means more than 90% of all Twitter entries that hit Buzz are ignored. Given you can't really have conversations in Twitter in the same way you can Buzz, and the person's entries already seem disjointed, that's probably just fine.

So we now know the community can self-select and self-filter, and that Buzz users comment on native entries and Google Reader shares more than anything. But let's go back to what a "bot" is anyway.

According to Tweetmeme, that ReadWriteWeb story was retweeted 662 times on Twitter. Is that any different than making a share in Google Reader? Even if a human clicked that button, are those not 662 bot entries into Twitter? In addition, the main RWW feed (@rww) is dominated by entries from TwitterFeed. That very post about Google Buzz bots was fed by a bot to their Twitter account. So the underlying hint of "Buzz loses and Twitter wins" is dead on its face, if this is the measure.

Do I want Twitter updates in my Buzz? Not really. I can go to Twitter for that. Also, in March, I talked about the trend toward disaggregation and putting native data in native sites. There is a huge benefit of getting that right. I took Twitter out of my Buzz for a reason. But Reader shares obviously have a good role in the community, and they are part of my curation activity.

Want to get a real good idea of what is happening on Buzz, with real people? Spend some time with a statistical site that really knows Buzz at http://buzz-stats.appspot.com/. Run your own chart to see if you're as feed-crazy as me, or who the top 100 people are on Buzz for getting replies to their content. PostRank didn't do themselves a service with their shallow report, and neither did the guys who regurgitated their data, and then retweeted it, or let their bots do the dirty work for them.

March 29, 2010

Cliqset, Status.net Support Salmon for Comments Migration



As Marshall Kirkpatrick noted in a post on ReadWriteWeb this morning, the Salmon Protocol project we first introduced back in October of 2009 looks like it is progressing beyond the planning stages, as it has been integrated in two small, but influential, social networking sites, Status.net and Cliqset - both of whom are strong open standards advocates laboring away in the shadows of larger communities.

The Salmon Protocol, which aims to define a standard protocol for comments and gestures (such as likes) to swim upstream to the originating blog, hopes to unify conversations across diverse locations. It's planned for adoption by Google Buzz (See: Designing Buzz for a Google-Free World) and solves the problem that first blew up back in 2008 around fractured conversations.

While many people, including myself, have adapted to a new world whereby conversations take place in a wide array of communities, it would be nice for the content originator to have one centralized location to see all downstream activity.

Evan Prodromou of StatusNet mentions integration of Salmon's digital signatures in a blog post from Sunday, where he also notes support for Activity Streams encoded in RSS 2.0 and Atom. The move, paralleled by deployment at Cliqset, should be what I hope is the start of a lot more announcements, from small and big companies alike, to make Salmon a reality. The next big target, which I am biased in favor of, obviously, would be for Disqus to integrate with Salmon and pull comments upstream to this blog.

I managed to find time to talk with Darren Bounds of Cliqset at SXSW earlier this month. In our quick discussion, recorded on CinchCast, you can find his comments on their support of open standards, and that network's direction - which could hint at why they're an early adopter of Salmon.

March 13, 2010

Users vs. Companies: Conflicts over the Real-Time Web?

If 2009 was the year of real-time Web, with practically every major service finding ways to bring content to its users instantly, 2010 is about optimizing the new real-time world, expanding interoperability between sites, finding more ways for users' content to be discovered, and taking the potential of real-time out of the status world and into the real world. Today, at the South by Southwest Interactive event in Austin, Texas, one panel asked if we were making serious progress in this vision, and if companies, feeling increased competitive pressures, are short-changing users in the process.

Marshall Kirkpatrick of ReadWriteWeb, who moderated the panel, featuring representatives from Collecta, Google, Gowalla and Microsoft, said "the real-time Web is a big, complex and multi-headed beast," adding, "almost as many people you talk to on the subject will give a different perspective."

For most, the real-time Web represents reducing latency from the time updates are published and when they are experienced practically to zero. This can be anything from updates from blogs to downstream aggregators and RSS feed readers, status updates from social networks to other points in the ecosystem, or instant alerts from the Web at large that a saved search you requested has found a positive match.

But one of the existing problems with the real-time Web that has occurred is that despite the focus by many services to solve the same problem, many have done so without delivering true data interoperability - and other services are trying to solve for real-time without having full access to users' public data.

"Back in the day, you couldn't send e-mail from AOL to Compuserve, and today, you can't send data from Google Buzz to Facebook," said Brett Slatkin of Google's App Engine team, and co-author of Pubsubhubbub. "Part of what we are trying to work on is breaking down these barriers that connect to different sites. If I am on Buzz and Marshall is on Identica and Jack is on Twitter, we should all be able to communicate."

Standards have evolved in the real-time Web space, from OAuth to PubSububbub, WebFinger and Salmon (as documented here), but that's not to say there aren't still heated debates over these standards, or even which version of standards should be supported. (See this article for a discussion of OAuth 2.0)

"I try to be a practical person, and when I hear about a family of specifications, it sounds like a family of work," said Dare Obasanjo of Microsoft. "There is clearly a place where we have a common pain that we can work on. There is a bunch of shared pain, and the way you have to get real-time service is to work on APIs, and that is a clear starting point for standards. Pubsubhubbub can help solve that problem, but I get concerned when you have to implement certain specs to solve that problem."

"These specifications we agree on should be useful on their own," answered Slatkin. "When you implement a specification like HTML, you are not buying into an ideology."

As the real-time Web's protocols are debated and deployed, so too does the application of these services. Google Buzz and Facebook have received scrutiny for their aggressiveness in converting assumed private data to public, and Netflix recently canceled an algorithm development contest thanks to concerns of assumed privacy violations.

"When talking about privacy, right now, unfortunately, the social networking market is failing, and they have little incentive to encourage user privacy," said Obasanjo. "I am waiting to see when people find what they thought were private updates as part of trending topics on Google and Bing. Users and companies are in conflict."

Obasanjo gave the example of Twitter needing its users to be public in order to drive value into the system. After all, if users were all private, there would be no trending topics, and thus it is Twitter's best interests for updates to be public. "There is a factor that if a user wants to be private, it subtracts value from the system," he said.

Beyond these concerns, known benefits of the real-time Web are scratching the surface of what could be done with more expanded to real-time data from other sources, it was argued. Slatkin forecast a time where you could query supply chains for inventory and purchase locally instead of from Amazon.com, turning economies of scale on their head. Scott Raymond of Gowalla talked about intersecting real-time Web technologies with geodata to show trending locations and the hot parties of the moment, by decaying the relevance of checkins over time. Jack Moffitt, CTO of Collecta, said a development environment for new tools and applications that leveraged zero latency was becoming "very interesting".

"All these guys are working on realizing the potential right now, working on real-time data," Kirkpatrick said. "Brett Slatkin said it was important people focus on the unforseen future that systems we worked on to support undiscovered use cases - things are going to get real crazy real soon."

Web-wide adoption of RSS and Atom standards has eliminated the problem of publishers providing their data, and tools like Pubsubhubbub are working to get data from one site to another faster. "Polling doesn't scale and you need a push notification to deliver it. It's possible we will have multiple winners, and we have to consider privacy considerations that people won't want their data available to everyone," said Moffitt.

The element of real-time is being layered across the Web, and it seems to be happening even if developers aren't completely in agreement over the tools needed to optimize the experience or if the debates on privacy versus public data are solved. And there's a lot of room for real-time to grow outside of the statusphere and to more traditional markets. The question is can developers provide solutions that don't have users running to the FCC?

January 14, 2010

Can We Live In Private and Demand Companies Be Open?

The growth of social services enables us to share an ever-increasing granularity of our lives with complete strangers who opt in to sample our updates. From traditional blogging to Twitter to location sharing services, and now, Blippy (which displays my purchases online and lets me follow others' spending habits), I can live my life open and transparent, letting you know what I like, what I do, where I go, and how I spend my time. In parallel, we are also asking businesses to be more open. Thanks to SEC regulations and Sarbanes-Oxley, as well as best practices, we expect public companies to tell us how much money they have, how much profit they made, where they made that money, how many employees they have, how much they intend to make next quarter, and a blizzard of other things that fall under the guise of information.

Now, we're even asking some companies to open up their kimono and blog themselves, to respond on a personal level with Twitter and other social networks, or open source their code, and let us know who their investors and partners are, lest we find they may have some bias. So yes, it's easy to see that open, open, open, is the word of the day.

This Week's Target: Facebook

Looking at these trends, it seems clear to me that the overriding move in both personal and business is to remove barriers and enable more visibility into activity. Yet, there are still pockets on the Web who see the data they share as needing to be under lock and key. The most recent kerfuffle was flamed by the often independent Marshall Kirkpatrick, who slammed Facebook's CEO Mark Zuckerberg for moving the once-closed social network toward increased openness, and reduction of what he termed "privacy". While others, notably TechCrunch's Mike Arrington, thought that argument to be overblown, Kirkpatrick followed that post with another line of questioning, asking how we would react if, for example, Google revealed our contacts in GMail and our subscriptions in Google Reader.

In the non-connected life, one was always taught that there were a few things one does not ask about in polite company - including politics, one's salary, and sex. Yet, the Web has made finding out most of those things very obvious for public members of society. We know what the vice presidents of many companies make, and we know the dollar amounts on contracts. We can do searches online and discover what political parties people are registered to, and where they put their money in the last elections and fundraising cycles.

Forget Location. Who Have You Slept With?

Although I was only joking, somewhat, I suggested to a friend this week that a new social networking service should be introduced that updates your stream with your sexual partners, and encounters, complete with their partner history and total number of partners. Given some's willingness to post all their data online, and the rising casual nature of some behavior, this isn't so far out of reach to be completely ridiculous. Just think of what message that would send if you were the "mayor" of multiple places on that particular network?

Yet, before that level of debauchery occurs, we still have to recognize that the trend Zuckerberg and his team are recognizing is true. The first rule of sharing things online is to expect that they will be discoverable by search engines and viewable by anyone. Even the closed, secure, password-only sites are usually potential victims for copy/paste, and are always susceptible to screenshots.

From Anonymous to Open. We're Getting There.

I understand the Internet's roots require a level of privacy, and you see many people still clinging to the antiquated hope of anonymity, assuming one can hide behind a user name and avatar to mask their true identity. I understand the concerns some have that the more data which is shared online makes them more vulnerable offline. But I am seeing prominent people talk out of both sides of their mouth when they claim to push companies to get more open and more transparent, while at the same time, clinging to the hope that we can push our content into a safe place on the Web and consider it "private".

During the week's discussions around Facebook and privacy, Marshall had some solid points, in his comments on TechCrunch, saying that people were joining groups and "fan pages" on the social network that had to deal with sensitive personal issues, such as infertility or friends of same-sex couples. His concern was that this data would be surfaced and revealed in a way that was not desired. In a week of drum-banging and emotion, I thought those examples to be the most reasoned. It makes sense, then, that sites like Facebook make the ability to hide certain affiliations from public view if so desired. If I chose to sign up and follow Fox News, and not want my left-leaning friends on Facebook to know, then I should have that option to hide the details from my page. But to fight against the tide of open in our own lives while demanding the tide of open be ever loosened when it comes to the companies we engage with seems bizarrely silly. Let's pick one and stick with it. If you don't trust the companies, then why should we trust you? What are you hiding?

Of note, and I mentioned this in comments to ReadWriteWeb, the notion of Facebook data being private is in itself the exception rather than the rule. You can see my Twitter stream and who I follow or don't, just as you can with FriendFeed and other networks. I have chosen to share my Google Reader feeds with you on Toluu and encourage others to do the same. The world is trending open and transparent, and you can embrace that or you can fight until you are blue in the face. But choose a side.

December 22, 2009

For All the Gloom Around RSS, Readers Continue to Climb

Skimming many of the leading technology outlets, you would think RSS had given up its ghost, making way for new services, like Twitter. Just this week, ReadWriteWeb claimed the RSS reader market was in "disarray" and continued a "decline". This came after a summer in which TechCrunch IT's Steve Gillmor declared RSS dead and suggested that it "rest in peace" and others ditched RSS for microblogging lists. With the world watching Twitter's top names and their Suggested User List-boosted following counts cross well into seven digits, data from FeedBurner and other sources shows RSS counts climbing - in some cases dramatically - for nearly all blogs, and a number of them also sport reader counts in the millions. While the independent market for RSS readers may be in bad shape, having ceded ground to Google Reader, RSS as a utility is actually growing. It's not going down, not by a long shot.

Prior to ReadWriteWeb's alarmist article, back in August, I arrived at similar conclusions, when I said "stand-alone feed readers are collapsing", highlighting the fact that Google Reader, and followers on FriendFeed, dominated my personal statistics.

That Google Reader and iGoogle have reached a dominant position in the market does by no means indicate the technology's death, but instead its maturity - something that, beyond aggregate reader counts, does not take in to account the fact that RSS is the mechanism for getting data between sites practically everywhere, even if it is not called out as such.

In January of 2008, I highlighted the debut of a new service called Rating Burner, which aims to display the most subscribed to blogs that utilize FeedBurner. The site shows not only a leaderboard, but also the change in subscriber data from day to day.

Data from Rating Burner shows green, positive growth across the board for feeds. And while total counts can always be debated as to their accuracy, sites including TechCrunch (with 4 million subscribers), the YouTube blog (with 2.4 million), Smashing Apps (with 1.5 million) and Simply Recipes (1.5 million) are into seven digits. This parallels the top Twitter accounts, who crest in the 4 million range, crowned by screen celebs Ashton Kutcher, Britney Spears and Ellen Degeneres.

Yes, statistics are statistics. The high numbers for these well-known personalities is suspect, thanks to their inclusion on Twitter's Suggested User List, and it is also expected that some top blogs here are bundled and not read every day - adding to their counts. Also true is the fact that not every blog uses FeedBurner, and thus cannot be tracked as well. Many sites point instead to a raw XML file, keeping control. That said, regardless of the data's perfection, the growth in 2009 for RSS subscribers cannot be questioned.

Using BlogPerfume's Feed Analysis tool, I took many top blogs and plugged in their statistics, to see how they grew in 2009. Running their query for the last 12 months provided 11 months worth of data. (Not perfect, but good enough)

For the charts below, I used three data points to show blogs' subscriber trends:
  • January 22, 2009
  • June 22, 2009
  • December 22, 2009
(Yes, the gap between January and June is 5 months, and the gap between June and December is 6, but you get the idea.)

To run the numbers yourself, simply plug in any FeedBurner enabled feed. (Examples: TechCrunch, Mashable, louisgray.com and I Can Has Cheezburger)


ProBlogger Nearly Double Subscribers in 2009 to ~140k




LOLCats Added 100,000 Readers to 200k+




TechCrunch Doubled from 2 million to 4 million




My Own Stats, Aided by FriendFeed, Quintupled




ReadWriteWeb added 40k from June (Jan. data flawed)




Google's Mac Blog Added 30k Subs to Top 70k.




Mashable Added Almost 150k Subs to Near 350k.


Also - keep in mind that my own personal numbers are inflated thanks to FriendFeed, but most non-personal blogs are not.

So what is the point? The point is that while some services (read: Twitter and Facebook) may be getting many people's attention, and while it is also true that Google Reader has the lion's share of the RSS reader market, the current discussion around RSS being less useful, or less important, than in years past, is flawed, period.

Lest this early adopter sound too much like a curmudgeon, just because something is newer does not automatically make it better, nor does it mean that there will be a rapid mass exodus from the previous technology. Just like Twitter can drive good traffic to a Web site, so too can RSS. Just like Twitter can pass along top content, so too can RSS. Just like top followed accounts can get million-plus audiences, so too can RSS. And both are growing in terms of connections, with no reversal in sight. Both are tools to be used well, and both are being used more than ever. The only "disarray" is in the current thinking.

December 19, 2009

Growing Grumblings on Tech News Don't Address Incentives

If you are the subject of the news, people will judge your actions and how you react to being in the spotlight. If you are the distributor of the news, how you message that news, and how accurately you report that news, will also be dissected. On the Web, especially in our sliver of Silicon Valley, where real time is becoming the standard, analysis of said news is itself happening in real time. From many corners, often from the more technically-oriented folks on the Web, I am seeing discussion around the tech news industry's alleged failings, inaccuracies, and usefulness (or lack thereof). While some of the feedback no doubt has merit, it too comes in simplified form, without offering potential solutions, taking into account how the creators and publishers of this tech news blogosphere are incentivized and rewarded.

On Sunday, Mike Arrington of TechCrunch, as he often does, started a discussion around what he termed "fast food content", saying that "hand crafted content is dead", summarizing a piece that lamented sites which steal content without attribution, and more darkly, sites that employ people to rewrite others' content, without adding anything new or doing "real reporting", the kind one learns in journalism class, or is required to do when working for a "dead tree" newspaper or magazine.

Given Mike's focus, running one of the more widely read tech news sites on the Web, his concerns lie around those who borrow much of his and his writers' content and publish it as their own. But I have also given a lot of thought, especially of late, to the vast number of tech news sites and blogs that are out there covering the same stories, and are jostling amongst each other to beat their competition by a few minutes - opting not to win on quality, but instead, on time. In this case, it's often not another tech blog's news that is being borrowed, but official announcements from companies.

One of the easiest things for tech blogs to do is repeat updates from the official blogs of interesting companies, add a few internal links to previous coverage they have done on that topic, add a paragraph or two of analysis, and hit the post button. I've no doubt done it myself over the last few years, even with this self-awareness, but you can see the process unfold practically every day. Watch for phrases like "According to a post on the official Twitter blog..." or "In an update on Google's blog this morning"... as many of the better-known sites all post their own interpretations of the news that came from the top.

This, in my opinion, is the very definition of the "fast food news" Mike is talking about, and time spent both producing it and consuming it could be put to better use - as in these cases, links could serve just as well as full articles.

I am by no means an ombudsman for tech media and the tech news consumer. I am but one person who takes in a lot of content, and produces a little on my own. But I see a few other areas where the tech news engine is falling short for news consumers, news makers and the news authors themselves.

I believe "fast food news" also can refer to the mass hysteria over making sure every site posts the news that a major browser or a major operating system has issued a point release, or when a popular site has an outage, that the incident becomes front page news for every blog. At some point, given the vast multitude of interesting tech stories, individuals and companies out there, one must take a deep breath and realize that being the 10th site to report that Twitter got hacked last night didn't really add a lot of value to readers.

In fact, when Twitter did get hacked Thursday night, Mike (again) had a solid post that added information, and, as he gained more knowledge of the incident, he updated the same post multiple times throughout the night. Because he was the first to the scene, with real data, his post had meat, while many, many others that followed were just echoes of the obvious.

So why is this happening? There are a few reasons:

First, the advertising model that forces many sites to drive page views and social interactions, through Digg, StumbleUpon, and Twitter retweets, is turning many tech news sites into post mills, staffed largely by inexpensive writers and freelancers. Instead of deep analysis posts that require interviews, backgrounds, and research, these sites are instead home to excerpts from YouTube, polls, user surveys, and whatever happens to be trending on Twitter that day. Quality is exchanged for quantity.

Second, many of these sites operate under the guise that they are the only site their readers see. Just because one major tech site covered a story 30 minutes before doesn't mean they should assume their readers already know. That is why if you do subscribe to many technology blogs, as I do, you can expect the vast majority of them to report the same story around the same time - instead of choosing a specific focus that can set them apart from the competition.

Third, thanks to competition and personal interactions, not every site likes the others. Years of infighting and annoyances, thanks to individual posts, personalities, or business priorities means that some sites really dislike each other. They won't link to one another. They will ban the competition from their user conferences, and when they aren't taking potshots, they will act like the other doesn't exist. Thus, if the competition "breaks" a story, the other will post it anyway, or try to find a wrinkle that makes their own version of events "improved" or invalidating the other.

Fourth, the rise of aggregation sites makes piling on to the news something that is rewarded. If all competitive blogs have covered a major story, many others will follow suit, be it to get into "discussion" on Techmeme, to see TrackBacks on the originating posts, or to come up when the popular terms are searched for on Twitter, Google and other engines.

In essence, the incentives, for the most part, do not tilt in favor of writing unique stories or doing the required research necessary to get a full story, to get quotes from a source, or find data points that back up analysis.

That's why you see people like Alex Payne (of Twitter) complain, saying "Rarely does technology journalism produce informed, correct, relevant, and readable content. This is a sorry and damaging state of affairs." in his rant from March (Towards Better Technology Journalism), and why Marco Ament, the lead developer of Tumblr and Instapaper creator, this week, wrote: "Over the last few years, I have unsubscribed from nearly every tech-news feed. I have never regretted the decision afterward, and I haven’t missed anything important. Tech news needs help. Badly. It’s truly terrible."

Keep in mind that it's not unexpected for the more technical among us to dislike the way their works are interpreted. Engineering distrusting marketing is practically a requirement and a religion. But we know they are somewhat right. As much as we can complain about the public relations industry as a whole, many flaks often find that their offers for reporters to speak with the CEO or an official representative of the company go without interest, either due to time issues or a lacking skill set. It's always a lot easier just to ask for the press release ahead of time, and an embargo date.

In an ideal world, those who are acting as our news filters would take the extra time necessary to ferret out news before its time, would ask those making the news the questions they didn't want to answer, would understand competitive landscapes, and wouldn't worry about getting a post up in a few minutes to hit a quantity threshold, without it first passing a quality threshold.

Lest we think Alex and Marco are the lone cries for help, you can see other comments this week from The Angry Drunk, and from Google's DeWitt Clinton, who posted to Twitter, "Don't worry. Save some time. Your story doesn't need a shred of truth to it. It will be retweeted just the same." in response not to a tech blog story, but a mainstream media piece that had missed the mark. (He later, in contrast, praised Marshall Kirkpatrick of ReadWriteWeb for solid reporting)

Content producers need to make choices in terms of what it is they cover, and where their field of expertise lies. If not breaking the news, or having access to the technology elite, there are many other ways to make your voice heard, through analysis and personal use cases, as well as the option to find new stories. Content consumers too have the choice as to where they get their news. I would hope that those people who are being spoon fed repeats of others' original reporting, or are waiting, jaws agape, for recaps of company blog posts, recognize what it is they are really missing.

Given the low cost structure needed to create content, it doesn't look like there is going to be a painful consolidation any time soon. In the meantime, the system is set up to reward those who publish quickly and pile on - for extra effort doesn't bring home the page views. There are going to be pockets of the Web that harbor original ideas, a focus on quality and data, and there are going to be other places where copying, scraping, and shortcuts are going to rule the day. I know what I hope to be. The question is, can we do our part, as publishers and consumers, to somehow reward those that do things right?

January 15, 2009

RSS Overload: Don't Complain, Do Something About It

By Mike Fruchter of MichaelFruchter.com (Twitter/FriendFeed)

There seems to be a trend lately of posts regarding RSS overload. A lot of people are complaining about being overwhelmed with their Google Reader, and some are even advising for you to stop using your RSS reader altogether. I say, hogwash. Do something about it and take back your Google Reader. Now is the time to reclaim it.

Some suggest to use Twitter and FriendFeed as the alternative. If your scope is limited to one or two particular subject matters, this may be fine. You can easily follow the relevant news sources by following them on Twitter and FriendFeed. The imaginary friend feature on FriendFeed was basically intended for this purpose.

The beauty of the imaginary friend feature is that you do not have to follow that person on FriendFeed. Chances are that person might not even be on FriendFeed, instead all you need is the blog's RSS feed and your set. You could follow that particular news maker/blog on Twitter, but you would be sorting through an already noisy feed of updates from the rest of the people you are following. Yes you could always set up a second Twitter account for just that reason, or you could directly go to that person's Twitter feed for the latest updates. That to me seems like too much work though, and is unnecessary.

Google Reader, for me, is the most effective power tool in my social media arsenal. Why? Simply because I don't have to visit hundreds of websites per day to get the information I seek. It's a competitive advantage when it is used right. Less time spent on numerous websites equals higher productivity. It enables me to work smarter not harder. I consume information at an increasingly high rate, maybe higher than some other people. To get the most of your Google Reader, it requires periodic maintenance. Just as your car requires an oil change every 3,000-5,000 miles, Google Reader is no different. That's the discovery aspect of it. Do I need to even go into the distribution aspect of it, sharing? Perhaps that's a topic for another post.

There is no need to feel overwhelmed by the unread count:

This is just an application. Why are we letting it get the best of us? We feel overwhelmed with the amount of bills we need to pay every month, or the amount of emails we may need to reply to in a timely manner. These things are overwhelming at times. An application that was built to discover and distribute information is a blessing, not our enemy. We see the unread count of 1,000+ items, and automatically anxiety kicks in. We feel like it's game over, we lost, and there is no turning back. The feed reader has won. Without going deep into the human psyche, there is a solution. The solution is called "hide unread counts", a feature that was recently integrated into the recent Google Reader overhaul.

Garbage in equals garbage out:

I'm subscribed to about 800 feeds in Google Reader. Without RSS, I would have never known the existence of these sites, or much less have the time to visit these sites on a daily basis. RSS has enabled me to broaden my horizons like no application has ever done before. Knowledge is power, RSS makes me smarter every single day. Do I really need to be subscribed to all of these feeds, of course not. Initially I would subscribe to every blog I visited that gave me some sort of value. I could easily trim my subscriptions down to 200-300 feeds and get the same value out of my Google Reader. A lot of these feeds are content clones, they simply regurgitate the same breaking news as the next site. At most I need a handful of these sites, primarily 2-3 is enough. I don't mind seeing another site's angle on the same story, and often they will contain more info that was missed or left out from the first site which is breaking the news. It's never a bad idea to get different perspectives on a story.

This is why I have begun to start going through my feeds and deleting the ones who are strictly content clones.

I'm an avid reader of both ReadWriteWeb and Mashable, but for the most part they are both content clones. I check RWW first, as it's a higher caliber of quality and writing, and, sure enough, the same regurgitated content appears on Mashable, and 50 other sites. I have since unsubscribed from Mashable and the other 50 content clones. Nothing personal, it just does not give me any value anymore. Remove the clutter from your Google Reader, there is no reason why you should not. I mention it's good to get different perspectives on a news item. It's often the lesser-known blogs who will give this to me, not the 100 pound gorillas who are competing for pageviews just to get a story published every five minutes. I want quality content, not headlines and 200-300 words of text that equates to a press release with some type of spin put on it.

Productive reading means organization:

Google Reader also allows you to set up folders. Take advantage of this. Create folders and set up a tiering system. Dumping all of your feeds into Google Reader without the use of folders, makes it clutter central. Set up folders for must reads, or folders based on topical interest. You could create a folder system for "daily”, “important”, and “other”. Only you know what will work and what will not work for you. This makes consuming RSS a breeze, and probably will give you a better Google Reader experience as well. If you must keep the clutter, put it into a folder, so that it is out of sight until you are ready for it.

Use what the power readers use, keyboard shortcuts:

This feature is a plus for productivity, especially for those of you with larger amounts of feed subscriptions. Save precious time by quickly exploring your reading list without moving your hand back and forth between your keyboard and mouse. The full list of Google Reader keyboard shortcuts is located here.

Keep a backup OPML file:

I use a site called Toluu just for this purpose. Toluu is a powerful feed discovery service, but it's also good tool for storing rss feeds. I keep my must read feeds only stored at Toluu. When I come across a feed that I must subscribe to, I input it into Toluu first, second comes Google Reader.

When all else fails, reclaim your Google Reader and start from scratch.

In order to do this, you need to have an OPML copy of your RSS feeds. If you already have a Toluu account you are ahead of the game. If not, sign up for their service and start inputting your must read feeds only. Remember to leave the garbage out, there is no need to start from scratch with the same garbage that overwhelmed your Google Reader in the first place. When you have your OPML file, head over to Google Reader and delete everything, so that you have a blank slate. Now you can import your OPML file into Google Reader, and presto you have just reclaimed your Google Reader. From this point on make sure you are using folders, tagging when necessary and most importantly cautious about what you add to Google Reader. Ask yourself is this feed really worth subscribing to, if so, add it to Toluu first, then into the appropriate folder in your Google Reader. Keeping a pristine and productive Google Reader is not easy, even a power Google Reader like myself needs to do a complete cleansing from time to time. I get to this point every 5-6 months or so. Since I have been using folders and organizing my Google Reader, I probably wont need to cleanse it as often, once a year should be suffice. It's all relevant to the amount of information you consume and digest. I tend to be on the excessive side.

If anyone would like an invite to try Toluu, please leave a note in the comments along with your email address, either Louis or myself would be glad to send you an invite.

Read more by Mike Fruchter at MichaelFruchter.com.

October 16, 2008

Hey FeedBurner, Wake Up. You And Google Didn't Talk Last Night.

You would think that as FeedBurner has been further incorporated into the Google monolith, recently incorporating with Google's feedproxy, that its service would be finer tuned and could be trusted to sync up with the Web giant's other products, including iGoogle and Google Reader. On most days, they seem to do a fairly good job, getting feeds out to the various RSS readers, and reporting statistics accurately. But today, like many other days before, the two seemed to walk by one another in the hallway and not make eye contact, because we are once again seeing a decimation of feed counts across the blogosphere, chopping away thousands of subscribers from popular blogs, and for the little guys, taking them down to zero.


My subscriber count plummeted by two-thirds (at least for today)



Coalminersgd wonders if all her subscribers went away.

This miss, one in a series of misses over the last few years, also comes at a time when many are openly voicing concern that FeedBurner is asleep at the wheel, having moved its ping server without telling anyone, and adding delays between people add posts to their site and when they actually hit the RSS feed. Techmeme's Gabe Rivera and ReadWriteWeb's Marshall Kirkpatrick have been among their most vocal detractors. Gabe said yesterday that "Feedburner lameness continues", and at the end of last month, Marshall said that FeedBurner May Not Be Hearing Your Pings.


DearRobot is clearly not happy.

In Marshall's story at the end of September, Steve Olechowski of FeedBurner said "we hear all your pings" and that "both ping servers still work", but that hasn't been the experience for everyone. Gabe said "It's inexcusable," adding "At this point, Feedburner is infrastructure" to the Web, something virtually all bloggers, myself included, use to have their content distributed. In fact, Gabe's response to Steve was quite direct, saying, "you guys broke the blogosphere, and your above verbiage reads like a bunch of evasive hooey."


TimBrownson is frustrated to the point of physical violence.

Whatever the problems are at FeedBurner, they aren't seeming to get any better with time, no matter how many times people like Gabe, Marshall and I bring it up. The company's blog hasn't been updated since May 30th, even though they've been called out for being silent before. (See: FeedBurner Quietly Kills All-Time RSS Feed Stats from February). It's alarming for some that a product that has become infrastructure and is expected to have 100% uptime continues to have such gaps and flaws. Losing one's statistics for a day is essentially meaningless, but it really makes you wonder what's going on over there.

See previous coverage of FeedBurner/Google mismatches:

September 09, 2008

Blogs' Never-Ending Battle of Page Views vs. Conversation

In a perfect blogging world, the very best writers with the very best content would get the most visitors, page views and subscribers. Every visitor would leave comments, send the links to friends, click through ads, and engage in thoughtful dialog with the author. And authors would be more than happy to pass along credit to other blogs for finding stories early, link to lesser-known voices, and admit when they got things wrong. But, alas, this theoretical utopia doesn't exist, and as a result, there's always a gap between what authors expect from readers and vice versa. And this gap can at times send even the best among us muttering to ourselves or launching into screeds when wronged.

The truth is, if you ask just about any blogger who has been active for a while, they could tell you some of their best posts withered into the dustbin of history, while a quick post that took no thought grabbed completely unexpected attention.

A couple examples on either side were visible this weekend:

On the up side: Adam Ostrow of Mashable posted to Twitter:
"looks like I posted one of my most successful (in terms of traffic ... thanks digg) posts ever on 2 hrs of sleep from Vegas hotel room."
On the down side: Marshall Kirkpatrick of ReadWriteWeb also posted to Twitter:
"omg pageviews are SO low on both of the posts I've put up today. dreadful. must write a big one next. i try to do 1 fabulous thing each day"
Adam and Marshall are among the most visible authors to post to their very popular blogs. ReadWriteWeb and Mashable are professional blogs with a staff of reporters, that rely on ad revenue to make money - making the battle for page views much more important for them than for those of us who look at blogging as a hobby, or at least, not the prime source of income.

Whether they receive a small handful of visits, or thousands per day, it's a rare blogger who doesn't look at their statistics, or at least at broad trends that tell which posts were the most popular, and whether visits are trending up and down. For the better part of the last year, I even took to posting my statistics at the beginning of each month, only recently having chosen not to as some people misinterpreted my goals as being promotional, as the numbers increased over time.

But statistics aren't why I blog. (See: Why Do I Blog? An Introspective Look and What I Believe: My 10 Web and Blogging Expectations for more about that.) For me, I like engaging in conversations about technology, trends, and business, and providing commentary, while learning from smart folks around the Web. That's why it's less important to me whether comments take place here or on Friendfeed and other aggregation services, and that's why you don't typically see me begging for Digg votes.

In fact, the only time I ever made the Digg front page, back in April 2007, was when I noted that Google's Earth Day logo was an homage to global warming. It was a post that took maybe 15 minutes, and got a lot more attention than I ever had anticipated. Since then, the closest I ever got to the Digg front page was when in July, I announced the introduction of TweetDeck. It actually reached the precarious top position of "Upcoming" before dying on the vine.

Knowing one's statistics and caring about writing articles that find an audience aren't bad things at all. Seeing which articles are most-widely read, and which topics spur engagement are often key ways to let your readers guide what you should be covering. But when page views drive ad dollars, and income, the entire foundation of why people blog changes - as blogging moves away from conversations and more toward revenue creation.

Following Marshall's comments on Friday, there was a short discussion on FriendFeed that covered the push-pull of conversations versus page views. After I asked if it was "really about pageviews or about getting a good story and discussion", Marshall answered, "it is about good stories and discussion generally - but pageviews are also important. I do this for a living..." which had Svetlana Gladkova of Profy hoping for a long thread on "blogging for a living vs. blogging for passion", which she saw as core to the debate. The debate wasn't settled.

If all ads on all blogs disappeared tomorrow, cutting off the revenue air supply to professional bloggers, it would be interesting to see how many of them would keep going in their spare time. How many of them would change what they cover, or change the way they write headlines, or link to other peers, once money was removed from the equation, assuming they kept writing? Tom Foremski of Silicon Valley Watcher, in a Monday article, quoted Gabe Rivera of Techmeme as saying that in today's competitive landscape where page views are king, that sites like "Techcrunch and the others used to link to each other and now they don't--they only link if they have to." Linking is part of the conversation, something we talked about at some length this time last year, when I said Internal Linking On Some Tech Blogs Is Out of Control.

It seems the only way to take page views out of the equation, and reduce the number of Shouts I get from Digg on a daily basis from authors trying to promote their own blogs' articles, would be to find ways to compensate writers that are not linked to advertising. But trends seem to be going in the opposite direction. Gawker Media has famously offered to pay reporters by the page view, a practice that came under fire from many corners of the Web, but continues, even as those who question the landscape are some of its biggest practitioners. In fact, back in 2006, ReadWriteWeb's Richard MacManus, in an article called Page Views 2.0, wrote, "It's funny that this page views model is at its foundation almost identical to the Dot Com days (bubble 1.0). Drive as many users to your site as humanly possible."

We all know how the Dot Com days and bubble 1.0 ended. We've already debated whether ads and blogs are a good mix. But the idea that conversations and commentary can trump the importance of the almighty page view looks to be losing out. It's no wonder that blogs looking to keep their costs low in a time when users are clicking on ads a lot less than they had hoped are often hiring inexperienced, inexpensive, young journalists looking to take a bite out of old media.

I know I couldn't quit my day job and try to make money from blogging, and I wouldn't want to be a slave to the page view. But for those who lay awake at night designing Google AdWords copy and trying to think of the next big headline that will take Reddit, Digg and Yahoo! Buzz by storm, sending a swarm of readers that send page views through the roof, I wonder if they miss the simpler time when they could write more for themselves and engage with their readers to share a story and ideas, before feeling pushed to get their next article out the door in an assembly line of online copy or finding themselves redesigning the site to optimize for page views and increased ad displays. That's worth having a conversation about.
DISCLOSURE: In addition to his work at Mashable, Adam Ostrow is also the CEO of ReadBurner, where I am an advisor, and hold a small equity position.

August 20, 2008

Why the Embargo Process Is Broken and Why We Still Need It

In the world of public relations, press management and blogging, an embargo sets a date and time by which a story can be written. Often, the embargo date and time coincides with a press release from the company, a Web site refresh, or the product's availability. Assuming all goes well, an embargo restricts all outlets from publishing a story until all is ready, and assuming multiple parties have been briefed, you can expect a waterfall of stories and press coverage to flow in a short period of time.


But, as you know, any time humans are involved, things can go awry, especially, as you see often in the blogosphere, you have a large number of media outlets that cover similar spaces, and a scarcity of topics. The resulting clamor to be heard amongst the noise, when so many different people are writing very similar stories, makes for an environment where the slightest bit of mistrust can turn into a simmering feud, or outright frustrating and finger-pointing, be it at a competitive blog, or the people behind the service being launched. Add in to the mix a rising number of inexperienced writers, prone to mistakes, with high levels of visibility, and this can happen with some regularity.

To start, why would a company ask for an embargo?
    1. To be sure a product would not be pre-announced before it was ready.
    2. To prepare and have enough time to brief all interested parties.
    3. To ensure no favoritism was shown to any media outlet.
Why would media/press/bloggers agree to an embargo?
    1. If they wouldn't agree, the company might not give them the story.
    2. Because an embargo often comes with news ahead of time, allowing time for writing.
    3. The service might have given them an interesting non-standard angle.
At an enterprise company, a media and analyst tour typically consists of a series of face to face meetings, plus conference calls, with an agreed upon date for a press release that coincides with the product's launch. Reporters often are looking for customer references and analyst references to validate the company's claims or add a wrinkle to the story.

For more bare-bones operations, including startups focused in the Web space, face to face meetings are less necessary. Sometimes, a series of e-mails, with potential for a phone call, is all that's needed. That's why you, on most blogs, rarely see quotes from a company's executives or customers, even if they had an extensive beta. Most bloggers, even if they have tested a product themselves, are echoing a press release or e-mail introduction from the service's founder. Again, a date is usually referenced in the e-mail to "go live".

Sounds good. Right? So why do these nicely laid plans fall apart?

On the company side:
    1. Sometimes an embargo is for "everybody except one or two publications", who are allowed to break it.
    2. Sometimes the Web site or company blog can go live before the embargo, in effect, scooping themselves.
    3. Sometimes a story isn't all that much of a secret, and things leak to the point there's no reason for an embargo.
On the media/blog side:
    1. Going first is seen as being "special", even if it's a matter of minutes.
    2. Being first can make the originating blog get more attention and linkage, or prominence on sites like Techmeme.
    3. Some blog management systems aren't fool-proof, enabling stories to go "before their time".

Clearly, you have some juxtaposed issues. The company launching an announcement would benefit from being covered by the most publications as possible, seen by the highest number of people. This is augmented by a need to be seen by publications with a high level of prestige. (Think Wall Street Journal, News.com, eWeek, TechCrunch, etc.) But there's something of a magnetic pull on press or blogs to go early, whether that's at midnight on the day of launch, or by posting five minutes before an embargo is lifted, and simply moving the timestamp, as has been known to happen. Blogs and press publications get a lot of visibility through gaining exclusives, and even if the same announcement has been sent to a wide audience, to hit the "post" button a little early, getting the word out first makes you appear more "in the know".

Whether intentional or not, blogs are rewarded for breaking embargoes, even if it hurts the launching service. And there's rarely any level of repercussion, as competing blogs in the know of the embargo are not naming names.

Of late, I've seen a healthy dose of complaining by some bloggers that other blogs have willingly or unwillingly violated an agreed-upon embargo. Yet, interestingly, it's a rare person who will name the offending party, even after their activity has clearly irked them.

See for instance:
    Svetlana Gladkova of Profy:
    "Very-very angry. Is it impossible to run a blog without breaking embargoes these days???"
    08:23 PM August 18, 2008

    Allen Stern of CenterNetworks:

    "wtf is up with the broken embargoes this past week - 3 today, 5 in the last week - im feeling like busting out a video tonight"
    06:32 PM August 18, 2008

    Marshall Kirkpatrick of ReadWriteWeb:
    "PR just called to say that mainstream media guy broke embargo, lol. you can't trust those mainstream media types with embargoes!"
    02:18 PM August 15, 2008
Notice how even though they claim frustration and anger, nobody says who the offending parties are...

Embargoes serve a real purpose for the company making the announcement. They are there to build time to polish the product, to enable true beta testing, to set up press activity with multiple targets, and to get one's message distributed. Embargoes serve a purpose for the blogging community, for those who choose to follow them, to help guide an editorial calendar, or to be sure you're also talking about a story on the day of its debut. And while some people might wish they disappear, it's not going to happen, so long as companies look to synchronize their internal and external activity.

As we see a rise in the total number of bloggers writing on the same topics, the issue of some sites trying to get out a step ahead of others isn't going to go away. Those that play by the rules and follow the agreed-upon embargoes, are on occasion, going to get burned. But what doesn't help the situation is that nobody is making a list and checking it twice. Why complain if nobody is going to name names? If there are one, two, three or ten blogs that regularly break an embargo, and it's clear there is a pattern, it should be visible, and these sites should be avoided by companies like the plague.

I believe in and honor embargoes. I also love exclusives, and think that there is more than one way to launch a product. But this practice is tried and true, so long as we have more transparency. What disincentive is there for bloggers who break embargoes if nobody steps up?

July 31, 2008

Where You Get Your Tech News Shapes Your Tech Views

By Rob Diana of Regular Geek (Twitter/FriendFeed)

FriendFeed seems to be the source of most of my interesting conversations these days. Sometimes the benefit of FriendFeed is not even the conversation itself, but finding a link to a blog post that I normally would not read. This happened this week when Jesse Stay shared a post to a story on newspapergrl.com. I read a lot of what Jesse shares, but this site is one I had never read. I found the post interesting because it was about tech news and how slow things are:
I just got off the phone with my friend Chris and we talked about how we hardly blog anymore. Also about how nothing seems that exciting in tech lately. It's mostly about Google and the iPhone over and over. Are we just cynical or have things quieted down considerably?
I had no idea that this is what people thought. This was not written during the iPhone hype, this was written a few days ago. So, I decided to look and see what news was posted on Thursday, July 31st.

First, let us look at what TechCrunch had to offer.


Click to Enlarge Image

Out of 16 stories in our selection, 4 were tech financial news, 3 streaming video stories and the remainder (9) were about various sites and their features. For a technology news site, that seems very reasonable.

ReadWriteWeb tends to have more opinion and review posts than TechCrunch and their stories reflect that.


Click to Enlarge Image


You can not tell from all of the headlines, but of the 16 posts, 6 were opinions and reviews. 4 of the posts were about video, image or mobile devices. The remainder were about various sites and their features. Again this is a reasonable breadth of information.

The last "heavy" technical news site I want to look at is Mashable. They tend to be not as news-heavy as TechCrunch, and have more of a social application focus. So, what did they post?


Click to Enlarge Image

Out of Mashable's 16 posts, 5 were about video, audio or images and 10 were opinions or reviews of various sites. Lastly, there was 1 self-promotion post. Given the specific content focus, this is also reasonable. So, we have looked at the 3 popular tech sites that many early adopters read. In order to contrast what a mainstream user might read, I took a look at what stories Yahoo Tech News listed for the day.


Click to Enlarge Image

For Yahoo, we again sampled 16 stories. Of these stories, 5 were financially related, 2 were about cell phones, specifically controlling kids use and cancer risks. 3 of the stories were about server products (VMware, Microsoft "Midori", and SharePoint). 3 more stories were about video games, 2 of which were about WordScraper/Scrabulous. The last 3 stories were the Chinese internet censorship, a Blu-ray player for Netflix, and 6 Ways to Save on Groceries. A simple breakdown does not really show the difference, except for the groceries story. The 3 stories on server products were mostly business related. VMWare giving something away, another product trying to replace SharePoint, and what "Midori" could do for Microsoft.

Most of the stories on Yahoo contain little or no technical detail. You do not see anything about social networks or other social applications. There was no announcement for the SocialMedian release or the redesign of Delicious. So, why is this important? It is important because most people are not reading about the same things that an early adopter is reading. Obviously, there will always be some overlap, but the mainstream users care about very different things. Given the various discussions about passionate users, early adopters and mainstream users, maybe we need to take a step back and think about how we bridge that gap. If you do not agree, then find your most non-technical friend and explain why they need to use Twitter and FriendFeed. Do not be surprised if they ask whether they could find more than 6 ways to save on their groceries.