Thursday, September 28, 2006

In the beginning, there was . . XML

2006 was the 15th anniversary of the web . . so it seemed appropriate to look forward to the next big thing on the internet after Extensible Markup Language (XML) . . , and ask Tim Bray, one of the founders of XML

According to Tim's technology standards mantra "simple beats complex" XBRL is still in early evolution stages with mostly "complex beating simple". He, amongst others, have argued for taming the beast to enable broader and deeper adoption in the mass market - such as an analyst tool for retail investors. Done right, the impact of XBRL could dwarf that of RSS and Atom put together. I’m on board, and anyone who believes “truthful business” isn’t necessarily an oxymoron should be too."

It's not quite clear what tools even relatively savvy retail investors will use, let alone pay for - the trick continues to be finding a fee paying model that's scalable. With the SEC funding $500,000 for these analytical tools and stipulating they must be built as open source products, it leaves a question mark about the commercial viability of building these tools. There are others that argue XBRL will be one of many threads in a sea of financial information that needs to be sifted in a manner that investors' tools can rapidly assimilate and act upon.

As the market shows interest in alternative reporting standards such as Ceres, Enhanced Business Reporting, company blogs and RSS, it's clear there's no silver bullet to this dynamically mutating challenge. Now, we're hearing of Web 3.0? basically the Semantic Web with technologies like RDF, Microformats, GRDDL and ContentLabels being just a few of the newer technologies that will form part of the vocabulary which we will all be rattling off in 2007, just like RSS, tagging and UGC were newer terms that entered the mainstream conversation in 2006. Notice how new VCs and their investments are now showing their faces in this new wave of metadata formation, discovery and more. Aggregate Knowledge Attensa and TouchStone are all new startups focused on getting users attention with metadata and reusing it's value either for advertising or discovery of new information.

Tim was invited to speak at the SEC discussion on Oct 3, and made some interesting observations about the potential behind interactive data (increasingly aliased from XBRL).

quote

Back to Interactive Data · Anyhow, here’s the dream: right now, if you know the name of a company, you can be pretty sure that by visiting www.company-name.com you can find the basics: where the offices are, who the CEO and Directors are, and so on.
I imagine a future in which you can go to xbrl.company-name.com and be pretty sure of finding authoritative machine-readable financial data. And in this picture, Metcalfe’s Law applies in more than one way: not only does the value of the financial data increase as a strong function of how many companies are providing it, but the pressure to join in does too, on those companies who aren’t providing it.

XBRL ain’t perfect; they made no particular effort to hit any 80/20 points, so it’s big and sprawling and taxonomist-ridden and it tries to Solve the Whole Problem. In this particular case, I claim that the information is so valuable that it’s worth fighting through all this and finding a way to make it work.

In this vision, it‘s a whole lot harder for a management team gone bad to turn a decent company into a den of thieves.

Done right, the impact of XBRL oops Interactive Data could dwarf that of RSS and Atom put together. I’m on board, and anyone who believes “truthful business” isn’t necessarily an oxymoron should be too.

quote

Tim was the only person on the illustrious XBRL "experts" panel to admit we're at a starting point and we still have a long way to go to make XBRL part of the "financial plumbing" and efforts to sell XBRL into the pain within organizations in terms of legacy system integration (the holy grail or the holy grave) was a whole different ball game.

OK, back to Tim and his belief that we are now playing with a green field moving from HTML to SGML to XML to, more recently, blogging, ATOM, GData (new RSS), EC2 and GRID. Stay tuned to the adoption of these standards.

Other trends: dynamic languages versus scripted languages . . Overall, he appears to put his bet on the emerging acceptance of ATOM as the next big thing -- "ATOM has the potential to have the same impact as XML," Bray

The Innovators Dilemma

The Innovators Dilemma and The Innovators Solution are two great books to read . . As we approach the precipice to adoption of XBRL, it's worth listening to a scholar on the gotchas to innovation and why the road to success is paved with market boobytraps. The podcast is a little dated . . .March 2004, but simply timeless and one I like to listen to once in a while to remind myself that building a new business is more than just good ideas. wwww.innosight.com is a legacy to his work at HBS.

Tuesday, September 26, 2006

Nuts and Bolts of Electronic Trading



Simply a great link to trends in capital markets. .

What's next . . "stay hungry, stay foolish!"

Change may come in the form of a database, applications or integration solutions. Is this the time for the next wave of Enterprise Application Integration solutions aka EAI 2.0? Some people in the business reporting world claim we are ripe for a breakthrough as big as Visicalc. while other people predict a somewhat daunting future depending on your POV, see googlezon

The humble spreadsheet harnessed the power of the microprocessor to millions of PC users. It was and remains the only significant programming tool used by millions of people who know nothing of simple programming such as compiling, scripting, or even simple looping. It provides a simple method of assembling data sources to create a custom "application". The application is really part of a business process, most often a financial process. A "smart spreadsheet" loaded by tagged data for business processes would be a powerful way to unlock collaboration and process knowledge and mitigate the ever growing costs of regulatory reporting and compliance. Sarbox costs -- be gone!

Here's the raw data . . 2006 will see the number of personal blogs exceed 60 million. The number of new blogs created daily will rise to over 100,000 a day or more than one per second. Howver, many analysts are saying that the relevance, average quality and value of each blog will decline pointing to the stat. that over half of all blogs cease to be active within three months of their creation, and only 13 percent of all blogs are updated more than once a week. However, although some say it will be increasingly difficult to find quality blogs, they miss the point. This is the best spot for CEO blogs that I've been able to find.

"...Growth in the numbers of blogs tracked by Technorati continues to grow briskly. While the doubling of the blogosphere has slowed a bit (every 236 days or so), interest in blogging remains considerable. About 55% of all blogs are active, which means that they have been updated at least once in the last 3 months."

"The integration of blogs and traditional media sites on the web continues. Technorati has put together the top 100 sites that make up "The short head" (as opposed to "the long tail"), which is still predominantly made up of traditional media sites, like The New York Times, Yahoo! News, CNN, and MSNBC."

"By the time you reach the top 5000, blogs have essentially taken over, with very few well-funded mainstream media sites listed." For the full monty of graphics and analysis, check out the full report.

Information disssemination is becoming more fluid and a new form of intermediary will likely emerge: the blog (or ideally somethings that includes the larger world of semantic data) aggregator will emerge, most likely funded by advertising, and specilaized in identifying the best quality content . . segway to a coffee meeting I had with the leaders of Monitor110 (I think that's read Monitor One One Zero, but I could be mistaken) and their recent $11 million financing . . where the FT reported on a seemingly innovative search/news aggregation idea aimed at the financial trading community aka ‘hedge funds.’ Basically it is touted as a revolution in information gathering, digging out the nuggets that exist below the radar screen of the conventional or mainstream press.

They do sound like they have some smart people and decent technology so it may be a useful toy - - however, I suspect the really smart money traders who have known how to search blogs and use RSS readers and tagging and social-bookmarking services etc. will be a bit miffed that any old trader will be (in theory) able to find the same gems of information by paying up for Monitor110’s services. The ground Monitor110 is breaking has been tried before by Clearforest and Relegence and other less known startups for some time now and digital generation traders can mash-up their own intelligent news filters either from scratch or using tools like Netvibes. And beware this space was hyped up by Majestic Research who quickly faultered and fell on their sword.

Edward Hadas over at breakingviews.com (another paywall, but really good analysis site founded by Hugo Dixon) compares it to using the ‘wisdom of crowds’ to trade. ‘Wisdom of crowd’ - mining would be things like Marketocracy and SocialPicks.

Monitor110 is all about finding the needle in the haystack; finding the individual voice or nugget that escapes crowd amplification. Finding the kernel before it becomes a snowball. Beware the paradox of diminishing returns, however: the more people find the needle the more difficult it will be to monetize. Or paraphrasing Dash - ‘if everybody is special, it really just means that nobody is…’ I'm bullish on their assumptions and wish the founders well.

Dash

Is it so far fetched to envisage (a future) Google Money and (a future) iPod converging and delivering the killer app, iMoney - - making investing as cool as turning on a music file. Don't rest on your iPod laurels Steve Jobs - we need your brilliance ("stay hungry, stay foolish")! We are indeed in strange times where innovation is being stimulated by government regulators and accountants. Perhaps their time has come - when was the last major shake up in accounting - double-entry bookkeeping? . . a 1,000 years ago.

Regulatory shove - we're doing it, no really!


Mark the date -- Sep 26, 2006 - change to interactive data is now inevitable. The SEC has turned a corner and truly made it clear that talking up XBRL is not enough, it's now time to walk the talk, and they're putting their money where their mouth is . . . this is now a 10-bagger!

COUNT ONE! SEC announced a $54 million investment to update the commission's EDGAR financial statement filing system to XBRL,

COUNT TWO! An announcement of an SEC roundtable on ”interactive data" to be held October 3.

COUNT THREE! The two announcements came on the heels of news that the SEC's small business roundtable slated on September 29, will focus, in part on, interactive data - the new buzz word for XBRL.

This triple whammy from the SEC is a clear signal that companies will HAVE to be XBRL compliant within a year, since Cox said this morning that the SEC's coding, taxonomy, and technology efforts will be done within a year. All current EDGAR filings will be switched over to XBRL inside of a year, and more telling, the SEC website will be peppered with XBRL software tools to help investors and analysts use the data "in interesting ways," says Cox. The chairman even noted this morning that developers should begin to "exploit" XBRL's potential by writing whiz-bang software and tools for companies, investors, and analysts.

Streamline Enterprise Business Reporting with XBRL Jeff Thompson, Institute of Management Accountants

Thursday, September 14, 2006

We're doing it . . SEC

Business Wire hosted a webinar on Sep 13 inviting the SEC's Corey Booth and others to answer questions on the adoption of interactive data. . . sounds like business-speak is finally "in" and geek-speak is fashionably out in the world of XBRL "XBRL is such an ugly word," Booth. He finally gets it.

Booth went on to to say that he estimates there were about 40 companies that were part of the SEC interactive data program and he expected many more to join. While it's uncertain where and when the tipping point will be reached for mass adoption of interactive data, Booth did say that the SEC is looking at their options more carefully for moving beyond adoption. . . and alluded to several RFPs that were underway to solicit guidance from vendors to help the SEC in their transition to use XBRL data. While it became more and more clear why the SEC is moving on this -- driven by Congress and coupled with SOX initiatives demanding that the SEC review a higher number of company filings - in the order of 40% over a three year span, the value prop. for the issuer still appeared somewhat illusory as depicted by the panel. Without citing specific quantitative benefits, issuers were left with an unsettling feeling that this is yet another "digital" wave about to blow their way with far reaching consequences and obvious benefits somewhere in the business reporting supply chain (oops, sorry for using that overused phrase, but apparently all good XBRL citizens, heady with the XBRL coolaid, are meant to refer to this term and associated visual to the point where we're all supposed to nod in unison, and mutter "...hmm, aha, eureka, yes. .").

While Booth speaks well for the regulator(s), he leaves the issuer with a nervous feeling about their benefits and the need for immediate action. While the regulators represent institutions to be respected and followed outside the US, the financial shinanigans in the US markets in 2001 followed by SarBox attempt to reign the markets in, have left US issuers more than just a little reticent about following US regulators -- when there is a choice.

Kudos to Business Wire and Michael Becker for moderating an excellent discussion and extracting some valuable insights from the panel in a masterly conversational form that made it a pleasure to listen (and even podcast!) in . Thank you, Michael for asking the tough questions and for making it sound so easy.

Monday, September 11, 2006

One view of things to come, . . perhaps


As publishing/syndication methods morph it is interesting to look at where this maybe going. The new syndication methods are all about making content simultaneously available for use (and re-use) for a variety of purposes.

Syndication has its roots in the publication industry where it means "to sell (a comic strip or column, for example) through a syndicate for simultaneous publication in newspapers or periodicals." In recent years, the term has applied to web content, making the same content simultaneously available for multiple purposes. The most common uses of syndicated content are:
  • to provide fresh, up-to-date information (i.e., news headlines, stock prices, weather forecasts, etc.) for incorporation into a web site; or
  • to monitor an existing web site for changes.
One can extend the syndication concept to encompass many additional information reuse possibilities, including:
  • transforming channel content into an outline format (OPML),
  • allowing the content to be manipulated in outline processing tools; subscribing to a channel using KlipFolio,
  • a desktop utility that provides real-time notification of changes to a channel; subscribing to a channel as smart tags in Microsoft Office XP, allowing Office application to automatically hyperlink channel item names appearing in Office documents; direct access to all channel content in XML form,
  • subscribing to a channel using all standard RSS variations (e.g., 0.91, 0.92, 1.0. 2.0) as well as variations that include extensions for content security information;
  • and many more.
Now, from this to . . a lively view of things to come in You Won't Recognize the Capital Markets in 2015 by Sean Park of Dresdner Kleinwort Wasserstein -- who draws on a compelling set of fictitious events that portent change in the capital markets . . . and makes the ominous prediction that as technology advances, investors' needs shift and regulations allow for new business models, sell-side firms will compete directly with stock exchanges sometime during the upcoming decade. Quite entertaining!

Other interesting interviews can be heard at podcasts, in particular, the IBM 2015 survey report.

Wednesday, September 06, 2006

HF Regulator adopts Electronic Filing

2006-08-24

E-Reporting Initiative: Minor Changes in Data Collection Will Provide Significant Positive Benefits

GRAND CAYMAN (Thursday, 24 August 2006)

When the Cayman Islands Monetary Authority's (CIMA) electronic reporting initiative comes on stream in early 2007, for CIMA-regulated funds with a December 2006 year-end, fund managers will submit, in a prescribed manner, their annual reporting requirements using a secure, streamlined and paperless system.

CIMA will then have accurate, electronic data, extracted mostly from audited accounts, for use in reporting aggregate information on the fund industry.

The Authority believes the change in how fund information is filed - not the information itself -- will represent a marked improvement to the submission process.

"The system we are seeking to develop will eliminate redundant data requirements, align reporting to make use of more of the data that regulated entities use for their own purposes, and minimise ad hoc requests from us to those entities, thereby making CIMA more effective in its regulatory oversight," said Mr. Gary Linford, Head of Investment and Securities for the Authority.

He added: "CIMA-regulated funds will not need to file 'information on transactions' as recently reported in an online media, nor is CIMA seeking to increase its prudential regulation of hedge funds beyond the existing regulatory framework. The initiative will allow us to significantly improve our compilation of aggregate statistics on the 8,000 funds regulated in our jurisdiction. This will better enable CIMA to meet the needs of our stakeholders for reliable, representative industry statistics, such as size, growth, change, market share, and investments by and in funds."

Mr. Linford stressed that the information submitted has always been and will continue to be managed in an extremely confidential manner. "In no way will fund- or manager-specific information be made available to the public, only aggregate industry statistics."

Demand for such data from industry and others is high, but as the Authority currently does not have the mechanisms in place to collect the relevant data from manual reports, aggregate statistics on Cayman's fund industry are not available. For example, to report total assets under management by CIMA-regulated funds would currently require manually reviewing each set of audited accounts held in paper form - e-filing will change all of that.

"The industry has been supportive of this move," said Mr. Linford, who added that the Authority undertook a consultation exercise with the private sector to seek input. "Many of the investment managers and service providers with whom we have had dialogue are relieved to hear that meaningful statistics on the industry will soon be available from a credible source."

With the growth rate in CIMA-regulated funds, electronic reporting will also enable the Authority to maintain the appropriate supervisory capacity without a proportionate increase in staff.

Additional details on the objectives and features of the e-filing initiative is available from CIMA.

This brief provides a list of the information to be collected, explains why operators of regulated funds will need to submit the information and confirms the operational benefits of the e-reporting initiative.

Tuesday, September 05, 2006

CFO Blog : Tiny XBRL

CFO Blog: Ron's Rant

The SEC announced today (September 5, 2006) that its next Small Business Forum, scheduled for Friday, September 29, will focus on interactive data. "Interactive data" is a code phrase for XBRL, which, in turn, is code for the Internet-language method of tagging financial data. (In other words, I suppose, it's code for code for code.)

Nudging companies to adopt XBRL has become somewhat of a pet project of Chairman Cox.

As CFO.com reported, the SEC has shied away from requiring companies to adopt XBRL, even though Chairman Cox argues publicly that it would make financial statements easier for investors to use and compare. But that doesn't mean that the chairman doesn't do his share of prodding. An XBRL pilot program has attracted 25 big companies to adopt the method, and Cox has asked software vendors to develop a tool for the SEC's EDGAR online filing system, which is where public company financials are filed.

Now it looks like the chairman is aiming his XBRL prod at smaller companies—the same companies that went to the mat with Cox over Sarbox Section 404, calling for a scaled-back version because they said the internal controls regulations were too onerous and costly for small companies to bear.

Wouldn't small companies be likely to feel the same way about implementing XBRL? By some estimates, it's not a particularly expensive proposition. Yet even with incentives from the SEC, most of the 25 companies willing to volunteer for the XBRL pilot program were large and "appear to have some commercial interest in [XBRL's] widespread adoption," writes Alix Stuart in her article XBR-What?.

So what are the odds that the chairman's prodding will convince smaller companies to jump on the XBRL cattle car? Skeptics abound at companies of all sizes. Witness Comcast controller Lawrence Salva, who told the SEC at a roundtable that "The payback is difficult to quantify," or Bill Ferko, CFO of Genlyte Group, who expressed support for the goal of transparency, but noted "XBRL really doesn't do that much for the company, for the registrant."

Let's see what the small-business contingent has to say on September 29.

Saturday, September 02, 2006

ComplianceWeek, Tigger or Eeyore!

A rather poorly researched op-ed on XBRL by the editor of ComplianceWeek, Scott Cohen. But, still, its of note to read the current journalistic view of XBRL as reported to its readers in the investor relations and governance groups inside public companies.

It was somewhat disturbing and sad to read the editors comments on XBRL. Disturbing for two reasons. Firstly, it points to the lack of serious industry journalistic coverage of the business or market drivers for a global accounting standard that has been incubated, piloted, deployed over the past 8 years with obvious implementation successes, and secondly, it seems to suggest a rather “head in the sand” point of view that “we should let the capital markets work their magic to solve this problem.” It is sad to see the editor of a “compliance” magazine stoop to such dramatic headlines (XBRL Hell!) without doing the minimum required research to better serve its readership. “Eye candy?!” And puzzling nonetheless to see why a magazine focused on compliance issues should be so flippant about “letting the markets work their magic.” But then, it has always been far easier to be an Eeyore than a Tigger!

While the Editor focuses on the SEC program as the driving force and chides the regulators efforts to build a critical mass of early adopters, he completely misses the point in terms of why public companies need to, and will adopt a financial reporting standard that safeguards the accuracy and timeliness of public company disclosures. It is a point often missed when a technology is first hitting the market and is receiving understandable and predictable, anemic market uptake. While the majority of reporting and compliance people in any company are hired to perform a role that deals in very critical and visible information sharing in a very methodical and structured manner, and are managed by CFOs, IROs or Compliance Officers who are charged with policing the release of public information, the manner in which company information gets transported and ultimately consumed suffers systemic problems that are outside the control of these diligent and well meaning reporting groups. It would serve the Editor well to survey a handful of public companies and track the root cause of reporting problems, such as a correction in the media, such as improperly interpreted line items from their footnotes or base financial tables, such as the impact of some company restatements that stem from known poor internal information gathering.

So, while the editor should take comfort in knowing that XBRL is complex because information companies disclosures can be very complex, and XBRL is merely a reflection of a reality that isn’t going to change – highly complex, company specific, industry specific, geographical guided, regulator shaped, accounting rule based public disclosures, he should also be cognizant of the fact that solutions have emerged to make the process of adopting XBRL as easy as updating an Excel worksheet. While it may be useful to “talk XBRL” and rebut the technical issues cited in the editorial, it would serve your readers better to know that the benefits of XBRL are far larger, near term, and early adopters are now shouting louder than the noise that has filled the Eeyore’s camp.

So, let’s address the article specifically.

First, the editor claims that US adoption of XBRL filing has been desultory. Well, one could argue that the SEC reeling under the pressures first to enforce SOX by Congress, then to review and moderate SOX by the market, adopted a more pragmatic position with XBRL by launching a voluntary program. The SEC now has some 30 companies creating XBRL files and is hoping to reach the 100+ company mark in the next 6 months. While the numbers may seem small as a percentage of total company’s filing, they have grown by an order of magnitude in 12 months and continue to rise especially with the new messaging to publish accurately and instantly without burdening the reporting groups with any additional work.

While a tipping point hasn’t been reached, changing market behavior with little immediate benefit until infomediaries and analysts institutionalize XBRL usage will continue to be a challenge. However, beyond compliance to the SEC’s directive, public companies are slowly beginning to understand that they have a reporting problem that directly impacts their company. And, while the SEC has its own analysis problem – the ability to process up to 40% of a million filings a year, public companies are suffering a (possible permanent) downturn in analyst coverage, media journalists improperly retyping facts about their company and investors reading information that has been filtered and massaged by junior data entry operators with little or no quality control.

The Editor cites costs and ROI as a barrier. Misperception. It takes a company accountant 1 or 2 hours (yes, hours!) to review an Excel template as they are starting up to file in XBRL for the very first time. Subsequent filings in XBRL are completely transparent and require no additional work (yes, no additional work!). This new publishing platform – coined EarningsDirect via the Intelligent Financial Statement, developed by an innovative teaming effort by Business Wire and CoreFiling, costs a few hundred to a few thousand dollars depending on the complexity of reporting and pays for itself instantly with the first filing alleviating the inherent problems in transporting the same information to the analysts in the traditional manner.

The Editor would better serve its readers by highlighting a common and widespread external reporting problem for all public companies and understanding how this accounting standard can be used to eliminate these problem – and, to highlight that XBRL continues to be debated inside the financial reporting community to address the broader issues of taking financial information and moving it across the many silos of information consumers both inside companies and its external stakeholders.

Time will tell whether interactive data or some other bold initiative by infomediaries will incent the market to adopt electronic tagging for financial disclosure. Certainly, while the process appears easy, describing business in accounting terms including all the nuances of a specific company and making sure it is consumed consistently is unlikely to be an autopilot operation any time soon - as alluded to by Sun CEO J. Schwartz in his recent blog, although there are some interesting textual analysis and statistical mining "smarts" that may alleviate the problem as demand for electronic tagging catches hold - more later.

As a result of misguided information like this one in ComplianceWeek, IROs/CFOs are still holding on to the view "tell me when I have to do it, and I'll do it, . .otherwise (door slam!)."




John Udells interview podcast Aug 2006

XML for business reporting gains momentum

Two years ago I wrote an unflattering report on XBRL (eXtensible Business Reporting Language), an emerging standard that aims to improve the speed, accuracy, and transparency of business and financial reporting. I applauded the goals, as we all should in the wake of Enron and other scandals, but worried about the complexity of the 151-page XBRL specification, its aggressive use of esoteric features of XML, and its reliance on accounting "taxonomies" defined by committees. I've too often seen these kinds of ambitious efforts stumble and give way to simpler approaches. SGML gave way to XML, for example, and while XML itself offers many advanced features, its most successful application -- RSS -- uses none of them. Would XBRL wind up being used mainly by what one wag called a "master race" of consultants and accountants? [Full story at InfoWorld.com]

In last week's podcast, XBRL's inventor, Charlie Hoffman, assured me I'm not the only one to express these concerns. Just this week, for example, when the SEC announced its Request for Proposal for the development of XBRL-based software, Dave Winer echoed them:

Sounds like the SEC is wanting to re-invent RSS?

Although I felt and to some extent still feel that way about XBRL, I have a much more complete understanding of the issues after researching, recording, and editing the podcast. It runs way longer than the others in my series, almost 70 minutes (edited down from 90), but I think the material warrants that lengthy treatment. Charlie Hoffman doesn't want to reinvent RSS, he wants to reinvent accounting, and he speaks as an accountant not an XML geek.

Friday, August 11, 2006

A conversation with Charlie Hoffman and Brian DeLacey about XBRL

Charlie Hoffman, the director of industry solutions for UBmatrix, is acknowledged as "the father of XBRL" -- the eXtensible Business Reporting Language to which I had a bit of an allergic reaction when I first encountered it a couple of years ago. But when Brian DeLacey, a researcher turned XBRL entrepeneur, suggested that I interview Charlie I jumped at the chance. In this week's podcast the three of us discuss the history of XBRL, its relationship to XML, its goals, its successes, and its challenges.

In next week's InfoWorld column I'll write more about what I learned from this long and fascinating conversation. But in a nutshell, though my criticisms of XBRL's complexity were and are valid -- as Charlie Hoffman admits -- the real story is (as always) much more nuanced. The inherent complexity of accounting standards, the competitive forces at work in the realm of global finance, the regulatory pressure being brought to bear -- these and other factors form the context in which the development of XBRL must be understood.

It's worth noting that while XBRL is a complex beast that makes aggressive use of certain advanced features of XML, Charlie Hoffman isn't (or anyway wasn't originally) an XML geek. He's an accountant who, as you'll hear in this interview, is deeply grounded in the practice of his trade. That makes this story an interesting contrast to the development of many of the web services standards I've studied.

Monday, April 03, 2006

Customizable industrial taxonomies

The subtleties of XBRL are lost in discussions where some vendors are swaying the "uneducated" to lean on the SEC to prescribe fixed taxonomies. This is unrealistic and impossible to conceive given the market driven economies embraced globally with very few exceptions (Cuba?, Gulf countries?) and the power of XBRL to process and compare within these natural variances (extension taxonomies) from one company to the next.

However, it is worth referring to the cry for standardization:

In response to the Friday March 31st posting entitled "XBRL Update" on the AAO Weblog (http://www.accountingobserver.com/blog/), Eric Linder, CFA, replied:

--------------------------------------------------------------------------------
Jack,

I was forwarded your blog entry about software to read XBRL files (http://www.accountingobserver.com/blog/) from several people as our company, SavaNet, you may know is the only company providing an XBRL analysis application, called the SavaNet XBRL Reader, which is free and available now from www.savanet.net. It is actually even more than a "Reader" application because it also performs professional-level security analysis on information in XBRL format. Although there is a large library of over 100 available Form 10-Ks in XBRL which can be accessed through the Reader's file manager, it doesn't support the filings made to the SEC under its voluntary reporting program. As a former Wall Street equity analyst, XBRL International member and leading XBRL software developer, I can tell you exactly what is going on here:

The problem here is that the XBRL documents filed with the SEC use company-specific XBRL taxonomies which do not allow for the automated processing and comparable analysis that has been promised to the marketplace by XBRL. In essence, companies are creating their own report form and then filling it out, which, as any financial analyst (but not accountant) will tell you, eliminates the ability to automatically process and compare the information because much of the information could be tagged differently by different companies. Even though companies all start with the same base industrial taxonomies in their filings under the voluntary trial program, they have the unlimited ability to add new items and re-do the calculation relationships of existing items, which essentially changes their definition and makes them unusable for analysis.

I often find myself explaining to the non-analysts involved in the XBRL effort that the moment even one item is added to a base taxonomy statement it invalidates most of the rest of the statement for automatic processing and analysis because, in most cases, without manually reviewing the new item, one cannot automatically tell where this new amount was taken out of the existing taxonomy. Many do not understand this conservation rule of "financial physics": since all financial statements add up to a total, something added in one place is necessarily taken out of the location where it was expected. And since this location is unknown, proper financial analysis requires that you do not trust any of the information in statements which is equal to or below the calculation level of the extension or alteration made by a company while performing automated analysis. (Note: since most compaies reported statements are very "flat" these corporate extensions or alterations are usually done at the highest levels of calculation relationships which thus invalidates most of the statement.) So, with company-specific taxonomies that alter or extend base industrial taxonomies, analysts need to go back to manual processing and the majority of the benefits of XBRL are lost.

So the SavaNet XBRL Reader doesn't support the unrestricted company-specific taxonomies used by companies in the voluntary reporting program because these files cannot and should not be used for automated financial analysis, comparison and valuation purposes. (and if all you want to do is view the statements, you can get them off EDGAR in prettier html format). The REAL risk to the financial community is actually that some users (or companies) that don't understand financial analysis DO attempt to perform financial analysis on these files because such an application could only practicably use the items from a base taxonomy for its ratio and valuation analysis, without taking extensions into account, which will lead to erroneous results in many cases.

Most non-analyst and non-accountants don't understand that even though, for example, the Operating Income and Revenue line item elements may still be reported using a company-created taxonomy, that the analysis ratio of Operating Margin (defined as Operating Income / Revenue) should not be calculated if there have been extensions or calculation alterations. This is because companies may have moved items in the calculation relationship which change the definition of Operating Income element and/or, if the application tries to adjust for this by referring to the specific elements that they believe should go into their definition of operating income instead, they will still likely receive an erroneous result because the company could have added extension elements that are not in this formula. You can immediately see the problem for erroneous analysis by the unaware and the opportunity for "gaming the system" by companies who can "create their own" uncategorized extensions to hold their undesirables - making them somewhat invisible to the automated analysis that will inevitability result from the use of XBRL.

Luckily, there is a XBRL implementation method that gets everyone (investors, analysts, companies and the SEC) what they need-- and that is called "customizable industrial taxonomies". Under a customizable industrial taxonomy XBRL implementation, all companies in an industry use the same hyper-detailed taxonomy (up to 3,000 elements in all statements and notes) without extension BUT companies can completely alter the presentation of these items and over-write labels to exactly re-create the current presentation of their As Reported set of financial statements. (Actually, company extensions ARE allowable IF they fall outside of the taxonomy calculation relationships, such as for company-specific notes.) So, the financial statements that investors and analysts see using a customizable taxonomy solution appear exactly as the company desires, but the underlying elements are structured for accurate, hyper-detailed analysis and comparison.

So, anyone who is reading this, here is what you need to know about XBRL: XBRL has absolutely enormous potential to solve a great deal of the reporting issues that investors and financial analysts face today and really can be a once-in-a-career advancement in analyst tools, BUT only if it properly implemented using taxonomies with restricted extensions, such as in a "customizable industrial taxonomy" solution. If extensions and calculation relationship over-rides are not restricted, XBRL will provide little to no value to investors and analysts. But even worse than no benefit is, if companies are allowed to extend taxonomies without restriction, XBRL could, in some respects, even make matters worse for investors who may BELIEVE they are getting an accuracy and reliability they are not, and then rely on erroneous automated analysis, and/or companies make use of uncategorized (or generally categorized) extensions to even further game the system.

Friday, March 31, 2006

Congress and better financial reporting

Baker Subcommittee to Advocate Transparency in Financial Reporting

The Financial Services Subcommittee on Capital Markets, Insurance and Government Sponsored Enterprises, chaired by Rep. Richard H. Baker (LA), will convene for a hearing entitled Fostering Accuracy and Transparency in Financial Reporting. The hearing will take place on Wednesday, March 29 at 10 a.m. in room 2128 of the Rayburn building.

Members of the Subcommittee are expected to discuss ways to promote more transparent financial reporting, including current initiatives by regulators and industry.

For the capital markets to operate most efficiently, information about public companies must be understandable, accessible, and accurate. Corporate statements are mathematical summaries meant to convey a company's condition. The four basic documents which must be filed with the U.S. Securities and Exchange Commission (SEC) are at the heart of investor disclosure: the income statement, the cash flow statement, the balance sheet, and the statement of changes in equity.

Among the current initiatives to improve the clarity and usefulness of public company information is a trend away from quarterly earnings forecasting, the use of technology to decrease complexity, and a review of the various accounting standards and how they interact.

Subcommittee Chairman Baker said, "If U.S. markets are to remain on top in an increasingly competitive global marketplace, we need to move away from the complex and cumbersome and explore technological and other methods of enhancing the clarity, accuracy, and efficiency of our accounting system. At the same time, we need to look at whether earnings forecasting and the beat-the-street mentality, which appears to have contributed to some of the executive malfeasance of the past several years, truly serves the best interest of investors or the goal of long-term economic growth."

The corporate scandals several years ago revealed weaknesses in the financial reporting system. While many companies were violating financial reporting requirements, regulatory complexity also may have contributed to some lapses in compliance.

Fraud, general manipulation of statements, and regulatory complexity all contribute to a reduction in the usefulness of financial statements and all may obfuscate the picture of companies' financial health. A number of recent studies have argued against the practice of predicting future quarterly earnings, concluding that the drive to "make the numbers"
can lead to poor business decisions and the manipulation of earnings.

Congress, regulators, and the industry subsequently have assessed financial reporting failures and have reacted with efforts aimed at strengthening the system, including many provisions of The Sarbanes-Oxley Act of 2002.

More recent initiatives by regulators to streamline financial reporting standards and accounting include:

* A Financial Accounting Standards Board (FASB) review of complex and
outdated accounting standards;

* The use of principles-based, rather than rules-based, accounting;

* FASB's continued cooperation with the International Accounting Standards Board on the convergence of accounting standards; and

* The use of eXtensible Business Reporting Language, or XBRL, a computer code which tags data in financial statements. The use of XBRL allows investors to quickly download financial data onto spreadsheets for analysis.

Public Companies have been filing financial statements with the SEC since the passage of the Securities Exchange Act of 1934.


Scheduled to testify:

Panel I

Willis Gradison, Acting Chairman, Public Company Accounting Oversight Board

Robert H. Herz, Chairman, Financial Accounting Standards Board

Scott Taub, Acting Chief Accountant, Securities and Exchange Commission


Panel II

David Hirschmann, Senior Vice President, U.S. Chamber of Commerce

Marc E. Lackritz, President, Securities Industry Association

Colleen Cunningham, President, Financial Executives International

Barry Melancon, President, The American Institute of Certified Public Accountants

Rebecca McEnally, Director of Capital Markets Policy, Center for Financial Markets Integrity, CFA Institute

Tuesday, March 28, 2006

Company DashBoard


Spring is here . .and, here I am in (snowy!) Minneapolis to present XBRL to an investor relations audience with Dan Roberts, Chair XBRL US. Dan's day job is Director, Assurance Innovation (hmm oxymoran if ever I saw one!) at Grant Thornton. Dan led me to appreciate that we must strive to create new ways of challenging ourselves and our children and to that point imparted his love and mastery of the unicycle. Dan is a very engaging speaker and always warms the audience with his personal experiences in life. Also joining us in our lively discussion was Garry Lowenthal for Viper Powersports Inc. (Pinksheets: VPWS). Garry is a very affable guy and has over twenty years of senior operations & finance experience, having served as a CEO, COO, and CFO, with a record of facilitating acquisitions, business launches, IPO’s, reorganizations and turnarounds while driving rapid revenue production. BTW, if you ever want to check out of the rat race and join the life of extreme sports -- you must take a peep at Garry's company and his custom bikes.

As I left my hotel for the meeting, I glanced at the FT headline and lobed a copy into my bag as I headed into the Mpls snowstorm. The headline was timely as Pfizer and others were clearly making noises about earnings forecasts pandering to sell side analysts. Opposition grows to earnings forecasts as Pfizer is the latest group to scrap quarterly guidance. -Financial Times March 13, 2006.

The backdrop to this headline is the recent spate of news about the continued convergence of the buy and sell-sides and the potential economic annihillation of sell side business. Some major bulge bracket firms derive as much as 75% of their revenues from principal transactions today. Firms that have traditionally relied on agency transactions as their bread and butter are now also starting to indicate they will start to leverage their balance sheets to act as principal. The unbundling of research and trading in the domestic equities business along with a greater reliance among publicly traded Investment Banking & Brokerage firms to derive earnings growth through principal transactions will likely lead to a further convergence of the buy and sell-sides of the business. Block trading, the specialist system, and the experienced institutional salesman many become a thing of the past as a result. Sell-side research is likely to continue to become more short-term oriented and even rare. More reason for publicly traded companies to take the initiative and market their companies more aggressively and consistently than before -- hence the real time dashboard or some corollary may become more relevant.

Thursday, March 02, 2006

Pressing the right buttons

The Accountant: February 28, 2006

A multitude of software products and technological tools are on
offer for companies and firms to use as they seek to improve the flow of
information when it comes to crunching and analysing the numbers.
Catherine Woods reports on where a rapidly evolving market is heading

There is a sense among regulators and the audit profession that more
needs to be done to simplify financial reporting for the users and
preparers of accounts, and that technology will have a key role to play
in this process.

David Turner, group marketing director for software provider CODA,
says that when it comes to using technology, accountants have always
been reasonable. He adds: "[They] are also... slightly conservative and
you can't blame them for that. Auditors, I think, are probably lagging
further behind."

Phil Donarthy is business development manager within the
accountants' division for software company Sage. He says that he does
not find accountants reluctant to take up new technology, but he finds
that many are not as progressive as they could be. He notes that there
has not been a sudden shift whereby accountants have become more
receptive to new technologies than in the past. "I think we're at the
stage now where across all walks of life, people are more receptive to
technology and also, let's be honest, technology is getting better," he
says.

Slow uptake

A sluggishness to use new systems is illustrated by the results of
research commissioned by Sage. Nearly 500 accountants, 200 business
start-ups and 2,100 established businesses were polled for the Sage
Accountants Business Collaboration (ABC) survey.

Eighty-one percent of accountants who responded felt they could make
better use of technology to increase the effectiveness of services they
offered. Sixty-nine percent believed it could also cut the cost of
providing those services.

Turner says the development of the internet has had one of the
greatest impacts on the use of technology in the financial services
sector. "You've had a whole new generation of web-based reporting tools
that have since come out," he says.

This greater use of technology across the finance spectrum, he adds,
is being driven less by clients and investors and more by the general
push for efficiency. Turner notes: "Number crunchers want to spend less
time creating the numbers and more time analysing them."

However, Donarthy says that there is another side and this is when
accountants want better systems for their clients: "Many accountants are
still working with clients who bring them bags full of receipts, rather
than actually providing them with any form of formatted data so a lot of
accountants are saying: 'If only my clients could be more efficient, I
could be more efficient and provide them with real insight into their
business.'"

Providing better business insight for the clients, notes Donarthy,
is also better for a firm's bottom line as it means accountants can
"concentrate their efforts on the higher value stuff [and] they can bill
for more". The benefits of technology can be in terms of greater
transparency, which suits today's environment in which resources are
constrained and yet the demands from regulators and stakeholders are
high.

Firms are progressive

Helen Nixseaman, partner in risk assurance services at
PricewaterhouseCoopers UK (PwC), and Steve Maslin, head of assurance
services for Grant Thornton UK, stress that the firms are progressive
when it comes to the use of technology.

Maslin says it is something that the accounting network, Grant
Thornton International, recognised in the early 1990s after member firms
predicted growing commercial and regulatory pressure for there to be
greater consistency. That led to the development of a common audit
methodology. He says: "Certainly, over the past few years... that's
given us huge commercial advantages and enabled us to deal as
efficiently as we can with regulatory demands."

In addition to this audit methodology, a lot of work at Grant
Thornton International has gone into developing software around internal
controls. The network now uses a software product which Maslin says
builds up a database of the sorts of internal controls one would expect
in different organisations. He notes: "That means whenever we're doing
an audit throughout the world, we've got a single methodology for going
around and testing the effectiveness of our clients' internal controls
system."

The software is linked to another product which enables Grant
Thornton International clients to capture their internal controls
electronically. Maslin says this is a format which allows a member firm
to carry out an audit procedure on controls without a client also
documenting it and causing unnecessary duplication.

Regulations and standards

Maslin observes that the latter system has been introduced to help
meet the requirements of the US Sarbanes-Oxley Act and the new
International Standards on Auditing which now require auditors to look
at the design effectiveness of a company's internal controls.

Internal control software, according to Turner, is a part of the
business that has grown rapidly in the US and is also picking up in
Europe. "Either because of Sarbanes-Oxley or because people in Europe
are seeing other legislation coming down the line, they're realising
that they're going to have to get their internal controls nailed down,"
he says.

Another way Grant Thornton International ensures that work being
done conforms to international regulations is through the use of
electronic audit files. Maslin says by working electronically "we can be
confident if we're doing multi-national audits that we're doing the work
to a single set of standards, but adding on to it the individual
standards of any one country". Audit files in the UK, US and Canadian
member firms became electronic around 1999.

PwC has also used electronic working papers for the firm's audit
files for a number of years. Nixseaman says: "It just makes sense in
terms of sharing information, particularly with teams spread around
different locations or even different countries."

The majority of this software at PwC and Grant Thornton has been
developed specifically for the firms. Maslin says Grant Thornton
International has employed a software team of 15 specialists in North
America for the last 15 years to develop and maintain its suite of
products.

It is, he adds, better at present to use an in-house system: "We've
found the software we've developed and maintained ourselves is certainly
a lot more robust and efficient than a lot of the software we've bought
from the commercial market."

An area where PwC is looking to further enhance its use of
technology is audit methodology. Nixseaman suspects that the firm will
move to make more use of the data analytic-type tools referred to by
Turner. These tools fall into two main areas: "One is in analysing
clients' data, so re-performing calculations or carrying out our own
analysis to look for trends or exceptions, and then producing some of
our own reports or graphs.

"The second area where we use it is to look at systems such as
financial systems or ERP [enterprise risk management] systems and to
look at how those have been configured and how segregation of duties has
been set up."

Quicker to use

Turner says these tools are being used by more people now that the
technology is quicker to implement and cheaper. He describes these tools
as occupying a third level of reporting analysis. The other forms are,
he says, rudimentary, online browsing-type and query-type reports which
people would perform within their accounting systems, and reporting
tools which entail taking something like Microsoft Excel spreadsheets
and turning them into more sophisticated reporting tools, or a purely
web-based product which allows accountants to quickly assess something
like profit and loss.

The tools on the third level of reporting analysis, Turner claims,
are what the market is going to want more of in the next three to five
years: "We've just gone through a number of years of focus on big ERP
applications. I think there's a reaction against that towards 'light
technology' - solutions you can implement fast, that will help you
automate your business and that are almost disposable so you can bring
them in, use them and then, if necessary, move on to the next
technology."

Maslin believes there is more the firm can do with regards to
internal control. He says the challenge for Grant Thornton will be to
move the current audit approach "from instead of just enabling it to
meet our regulatory and professional needs to working with clients to
make sure they're using the results of that audit work to actually
improve the efficiency and robustness of their own systems".

He identifies a new technology that electronically picks out the
parts of company reports that are most often examined as another tool
the firm is looking to harness. "It takes electronic financial
statements and every time an investor or analyst looks at the company
accounts, it builds up a profile of what sections of the accounts seem
to be of most interest. That is going to help both issuers and the audit
firms to better understand the needs of investors," Maslin notes.

A simple language

Simplifying financial reporting for all users of the information has
been driver behind eXtensible Business Reporting Language (XBRL),
especially in the US. Greater uptake in the UK is another area of
interest to Grant Thornton International. XBRL is an online system
that works by tagging data within financial information. The tags then
enable automated processing of business information by computer
software. XBRL can process data in different languages and accounting
standards.

In the US, one of the champions of XBRL is Securities and Exchange
Commission (SEC) chairman Christopher Cox who, since taking over from
William Donaldson last year, has consistently promoted the use of the
technology. The SEC is currently running an XBRL voluntary filing
programme which offers companies incentives to take part.

The XBRL project in the UK is not as advanced although Philip Allen,
director of XBRL UK - a consortium which advances the use of XBRL in the
UK - says there has been a huge increase in interest. Allen says the
"massive gain" to be made from the online application in the US is
different to the gains to be made in the UK.

"In the US, the issue is you have a huge number of listed companies
and no-one can really compare or analyse their accounts. Being able to
put it all [into] XBRL will allow people to simplify the analysis of
listed company accounts dramatically," says Allen.

In the UK, he suggests that the main benefit will be to improve a
company's access to credit. XBRL could allow banks to better process and
analyse the accounts they receive periodically from businesses to which
they have loaned money.

Allen comments: "If you have a large bank that actually could look
at a million corporate accounts and compare them all properly in real
time, it would revolutionise the way the bank provides credit. I think
that's probably going to have more of an impact on the UK economy than,
let's say, what listed companies do. They all list in the US anyway so
will be more affected by what the SEC is doing."

Companies House, the official government register of UK companies,
and HM Revenue & Customs (HMRC), are spearheading the British
government's involvement with XBRL. Allen says both have technically got
to the point where they have "solved all the problems relating to the
receipt of XBRL". Companies House now has a live service for receipt of
XBRL, trialling the system for companies which are exempt from audit.

Allen says Companies House and HMRC are "taking it very carefully
and very slowly because they don't want to put a foot wrong on this". He
adds that it is widely acknowledged that listed companies in the UK will
become more familiar with the software.

Accounting firms and software companies, claims Allen, are listening
carefully to the UK government before committing a large amount of
capital in this area: "What they understand is that in practical terms,
this is all going to be driven by government saying: 'This is how you do
your filings.' It's not that they're not interested, it's just that they
could lose a lot of money trying to go too fast on this."

As for when the government is likely to insist that companies must
use XBRL when filing accounts, Allen says: "Lord Carter [head of the
review of HMRC online services] has been undertaking a review of
corporate filing processes and it is expected that he will report at
some point on this. That report will drive how HMRC reacts." In the
meantime, XBRL UK is planning a conference in London during May for
accounting practices and software vendors about the Companies House and
HMRC projects.

Benefits

Just as the UK is keeping track of the US when it comes to interest
in internal controls software, the same trend is expected to happen with
XBRL. Allen believes that the Wall Street community is now starting to
understand what the investment analyst can do with XBRL and the same
equation will occur in the City of London soon, although he acknowledges
that "it is fair to say not many people there have quite got that far
yet". The possible credit benefits of the technology, he adds, are
unlikely to be realised until there is a larger take-up of the
production of XBRL accounts.

When it comes to technology which accountants use daily, the US has
less influence in the UK. Donarthy's theory is that simplifying the
technology he presents to accountants works best: "Accountants want to
do a good job. They're not at all interested in the technical details of
the solution. They're interested in how it can help them provide a
better service to their clients or make their practice as efficient as
possible. That's one thing we've learnt - we have to talk about the
benefits rather than getting hung up by how clever we are with our
latest whiz-bang feature."

Wednesday, March 01, 2006

Folksonomies v. taxonomy

folksonomies + controlled vocabularies


Posted by Clay Shirky
Email This Entry

There’s a post by Louis Rosenfeld on the downsides of folksonomies, and speculation about what might happen if they are paired with controlled vocabularies.

…it’s easy to say that the social networkers have figured out what the librarians haven’t: a way to make metadata work in widely distributed and heretofore disconnected content collections.

Easy, but wrong: folksonomies are clearly compelling, supporting a serendipitous form of browsing that can be quite useful. But they don’t support searching and other types of browsing nearly as well as tags from controlled vocabularies applied by professionals. Folksonomies aren’t likely to organically arrive at preferred terms for concepts, or even evolve synonymous clusters. They’re highly unlikely to develop beyond flat lists and accrue the broader and narrower term relationships that we see in thesauri.

I also wonder how well Flickr, del.icio.us, and other folksonomy-dependent sites will scale as content volume gets out of hand.

This is another one of those Wikipedia cases — the only thing Rosenfeld is saying that’s actually wrong is that ‘lack of development’ bit — del.icio.us is less than a year old and spawning novel work like crazy, so predicting that the thing has run out of steam when people are still freaking out about Flickr seems like a fatally premature prediction.

The bigger problem with Rosenfeld’s analysis is its TOTAL LACK OF ECONOMIC SENSE. We need a word for the class of comparisons that assumes that the status quo is cost-free, so that all new work, when it can be shown to have disadvantages to the status quo, is also assumed to be inferior to the status quo.

The advantage of folksonomies isn’t that they’re better than controlled vocabularies, it’s that they’re better than nothing, because controlled vocabularies are not extensible to the majority of cases where tagging is needed. Building, maintaining, and enforcing a controlled vocabulary is, relative to folksonomies, enormously expensive, both in the development time, and in the cost to the user, especailly the amateur user, in using the system.

Furthermore, users pollute controlled vocabularies, either because they misapply the words, or stretch them to uses the designers never imagined, or because the designers say “Oh, let’s throw in an ‘Other’ category, as a fail-safe” which then balloons so far out of control that most of what gets filed gets filed in the junk drawer. Usenet blew up in exactly this fashion, where the 7 top-level controlled categories were extended to include an 8th, the ‘alt.’ hierarchy, which exploded and came to dwarf the entire, sanctioned corpus of groups.

The cost of finding your way through 60K photos tagged ‘summer’, when you can use other latent characteristics like ‘who posted it?’ and ‘when did they post it?’, is nothing compared to the cost of trying to design a controlled vocabulary and then force users to apply it evenly and universally.

This is something the ‘well-designed metadata’ crowd has never understood — just because it’s better to have well-designed metadata along one axis does not mean that it is better along all axes, and the axis of cost, in particular, will trump any other advantage as it grows larger. And the cost of tagging large systems rigorously is crippling, so fantasies of using controlled metadata in environments like Flickr are really fantasies of users suddenly deciding to become disciples of information architecture.

This is exactly, eerily, as stupid as graphic designers thinking in the late 90s that all users would want professional but personalized designs for their websites, a fallacy I was calling “Self-actualization by font.” Then the weblog came along and showed us that most design questions agonized over by the pros are moot for most users.

Any comparison of the advantages of folksonomies vs. other, more rigorous forms of categorization that doesn’t consider the cost to create, maintain, use and enforce the added rigor will miss the actual factors affecting the spread of folksonomies. Where the internet is concerned, betting against ease of use, conceptual simplicity, and maximal user participation, has always been a bad idea.

Comments (12) + TrackBacks (0) | Category: social software


COMMENTS

1. Simon Willison on January 7, 2005 06:26 PM writes...

Further to your points about, I think a key element of folksonomies that is yet to be fully explored is ways of improving their support for "emergent" vocabularies.

Here's an example: I'm posting a picture of a squirrel on flickr; do I tag it with "squirrel" or "squirrels" for best effect? I can find out which term will be most effective by seeing how many pictures are already tagged with those two terms respectively, and going with the most popular.

At the moment that's a slightly tedious manual process, and one that many people are unlikely to bother with - but if the software offered a seamless interface for doing that (a Google Suggest style popup showing how many images are tagged with that tag as you type for example) people would be far more likely to form and follow a consensus.

I'm confident that there are a lot of things that can be done to improve the quality of folksonomy-produced metadata, without increasing the price (and rendering them useless).

Permalink to Comment

2. Lou Rosenfeld on January 7, 2005 07:29 PM writes...

Clay, interesting comments, but you seem to have missed my point. True, I shared my concerns about folksonomies; I expect you'd agree that they're no panacaea. Nothing is. It'd be silly not to be skeptical about them at this early point in their development.

But I'm also quite skeptical about controlled vocabularies. I've probably read all the same studies you have--perhaps more--detailing their high cost. I spent four years in an LIS program and worked in libraries, so I have a little first-hand knowledge. Oddly, people who attend my IA seminar walk away with the sense that I'm against controlled vocabularies. So shoot, Clay, we actually agree on this point.

But how these two forms of metadata might work together is what's really exciting. (And that's why I used the holistic term "Metadata Ecologies" in my posting's title.) They may be quite complementary, which is wonderful, as salvation lies in neither. I hope we might begin brainstorming how they can work together.

We're not even bringing up how the nature of content, users, and context plays out in all this. Folksonomies might work fine for archives of photos. But I'd prefer that my doctor rely on professional indexing to do his research the next time I'm in urgent care with some strange condition. And I'm hopeful that down the road a medical folksonomy might somehow improve on the performance of MESH headings, thereby increasing my chances of survival.

In the meantime, is there anything else you'd like me to convey to the "‘well-designed metadata’ crowd" at our next meeting (every second Tuesday at the south entrance to Dewey's mausoleum; be there or be uncontrolled)?

Permalink to Comment

3. Jay Fienberg on January 7, 2005 08:23 PM writes...

I'm glad you connected the folksonomy issue to the Wikipedia one, because I think they're similar stories in terms of the battles of loose vs controlled ways of doing things, and how folks who like one or the other tend to react to the other's approach.

But, I think this story of the loose vs controlled battles, however one would tag the two sides, is one that folks like Lou don't fit into so neatly, and that you over reacted to his points.

I think the implication is wrong that folks who practice information architecture automatically fall into some kind of controlled vocabulary metadata control freak category who opposed all wiki folksonomy tag flipsters.

Likewise, I think the implication is wrong that all ordinary folk are, by nature, free tag lovers who'd only desire controlled vocabularies if it got them out of a deal with the devil.

As Lou suggests, there is a whole interesting realm of possibilities wherein both of these approaches are combined and/or co-exist. Even Wikipedia has forms of control--loose vs control is co-existing there.

And, Flickr / del.icio.us have controls in terms of how one can change tags, once they are created--which are controlled vocabulary techniques that (maybe) could actually be removed, IMO, were those folks really committed to folksonomies!

Personally, the most interesting thing to me is creating ways to allow the one approach to evolve into the other, and vice versa, as IMHO, the "right" way is one that can evolve either way, dynamically (e.g., things can be under organized or over organized, and good organization is a dynamic balance between the two).

Permalink to Comment

4. Dave Evans on January 7, 2005 09:38 PM writes...

I think some meta-data will be more controlled that others. Business environment stuff, like "bought by", "owned by", "works for", "funded by", which are the types of tags I'm using in my vizualisation system, are pretty easy to standardize. Tagging "squirrel" is probably good enough for most people without having to worry about plural forms, or black or red squirrels. I wonder if there is a way to self-organize tags against the most popular ones that emerge over time? Changing tags in one fell swoop like in Flickr might be a (scary) good thing, like upgrading software for new features.

Permalink to Comment

5. Rick Thomas on January 7, 2005 09:53 PM writes...

This is a microcosm of the process of language formation. For matters of consensual reality language is fairly fixed. When there's something new to talk about language is fluid and then converges as the subject is understood. The resulting language will always vary by community - English vs. Russian, engineers vs. marketers - because they have different experiences. Bridging communities depends on multi-lingual people using clever tools.

This is also why it's easy for a million bloggers to write quick opinions, but relatively harder to synthesize collaborative works - there is an unavoidable cost of semantic reconciliation.

Evolution uses this algorithm to create life. Start with any found stability. Produce diversity. Choose better stability. Create highly conserved systems along the way.

Permalink to Comment

6. Shannon Clark on January 8, 2005 12:16 AM writes...

It seems to me that there is another, very significent and high "cost" to controlled vocabularies - except in a very few cases, users have to learn (and/or navigate/use other tools) the vocabulary to use it, let alone use it effectively.

i.e. take an extreme example of a library shelving system - it is not at all trivial or obvious to most users (let alone professionals) where a given book "should" be shelved and I assume the process of integrating new/emergent categories is a decidedly non-trivial one. A library shelving system also shows one of the major flaws of many formal metadata systems for many users - they assume an either/or system - i.e. a book can only be in one place at a time, so it is either in one category or another, but not both (at least not physically).

Online there are countless cases when a user, very logically, wants something to be multiplely tagged - i.e. it is both a book business and a book on technology, it is a photo myself as well as a photo containing a monkey etc.

It is also useful to keep in mind why, where, how and for whom users apply metadata (in non-formal situations). Most of the time in most systems users apply none or very little metadata. It is only when doing so ads value very directly for the user that users generally speaking take the time to add metadata.

- blog posts might get metadata if someone wants to make it easier for they themselves to find their own posts. And/or if they have enough readers to assist those readers in finding related posts

- photos may get tagged if someone wants to make it easier for their friends to find specific photos, as well at times to open up photos (ala Flickr) to a wider audience, such as other attendees of the same event.

These fairly adhoc, mostly relatively limited in scope uses of metadata differ very widely from the more formal uses imagined by many people - such as the "Semantic Web" crowd etc. In those cases the assumption is that metadata (and extensive formal metadata at that) is to some degree inherently valuable and useful - but also that it will enable a new class of applications and uses.

I would argue that most of the time the cost of doing all of this tagging, especially the cost of learning the system for tagging (which is more than just learning the names of the tags - it is also learning how to pick and choose between tags, how to search for the "right" tag(s) etc) is vastly higher than most people (or their companies that pay for their time if done in a professional environment) are willing to incur.

Potentially some tools can be built to automate the process - to suggest tags, to apply many of them in a mostly painless and automated way - though all such systems have to guard against inaccurcy as well as the "other" category problem Clay highlights.

In short - an important topic for discussion and one where I pretty much agree with Clay.

Shannon

Permalink to Comment

7. Bill Seitz on January 8, 2005 11:06 AM writes...

I wonder whether folksonomies will just turn into free-text search engines? That's the other extreme of the uncontrolled-vocabulary spectrum...

Permalink to Comment

8. Bill Seitz on January 8, 2005 11:12 AM writes...

Specifying which *contexts* are being discussed seems awfully relevant for discussions like this.

The more coherent (non-diverse?) the "user" "community", the more easily a SharedLanguage can emergence and be maintained...

http://webseitz.fluxent.com/wiki/SharedLanguage

9. pb on January 8, 2005 06:53 PM writes...

Check out this outlandishly clueless call for a "well-designed" web:
http://www.opendemocracy.net/debates/article-8-10-2277.jsp

Not only are all of Thompson's complaints completely wrong, they are the key drivers of the web's crazy success!

Permalink to Comment

10. Edward Vielmetti on January 9, 2005 02:27 AM writes...

This discussion reminds me of the James Fallows NY Times piece on knowledge management where he distinguishes between the "big heap of laundry" approach (= folksonomy) and the "neatly folded PJs" (= taxonomy) approach to handling volumes of information.

Given how much attention people pay to presentation when it comes to materials that they expect to have a big impact or a long lifetime, I can only expect that we'll continue to see both systems in place, sometimes in parallel, as long as there are exclusive categories (Michelin 4-star restaurants) where common-folks opinions aren't the point.

12. bborn on January 21, 2005 10:12 AM writes...

What if the descriptive taxonomy (what this thing is) was open-ended (a folksonomy), but the functional taxonomy (what would you do with this thing) was controlled?

So, say I was bookmarking this post: I could tag it with any words I wanted - tech, library, cataloging, and so on. Those words describe what this item is about in ways that are primarily relevant to me. If they also happen to make sense for someone else, fine.

Then I would also have to choose one or more verbs, words that describe what I want to do with this item. Do I want to read it, save it, comment on it, disagree with it, build something with it, etc.

Sunday, February 19, 2006

Software as a Service Model

Experiential model. . . follow the money . . myths . . .1st wave software service ASP move (capital was free, money poured into the initial setup, long time for the recurring revenue), 2nd wave - permission based financing strategy, very little capital goes in initially to get to the first customers because you can launch the service quite quickly compared to the enterprise model. What used to take 18 months to get "Golden master" with the ASP model takes much less time (6-9 months). CSF Unit economics - prove to land customer is less than than that of what they get back from the customer in the life of the customer. Another myth is that it takes longer to get to profitability. Exit - last 2 -3 quarters, highly visible companies are showing much higher P/E ratios such as Salesforce.com. Here to stay over the next decade.

Economic model based around subscriptions based on a smooth revenue streaming. Strong lifetime revenue. Scale benefits as the company grows. Enables you to compete with the big players very quickly. On demand allows you to play at a very different set of rules.

"Democratization of software"

Love-in not lock-in - really allow people to feel that they are not locked in. . . keeps the vendor honest. Only promise what you have.

Sales model -- free trial, as you use the service, you enter the first tier of pricing. Leak in from the bottom up. Expense-able not approvable business model.

Low upfront costs web 1.0.
Web 2.0 value prop. quite different . . The ability to road test the solution first before they actually have to commit - alignment with the vendors - real asset of this model. Actually get access to a community of users input and it shows up as value. Upgrades are seamless. Allows for much more rapid innovation 2 week release cycle. Most SAS companies have labeled their releases as Fall 06 or Winter 06 - speed of innovation follows a release cycle of 3-4 times a year.

"Power in the cloud" - new value by being part of a community. Creating brand new business opportunities that could only be created in the new cloud. Ability to see when the customer is having problems in aggregate and solve it before it has a larger impact. With traditional software they had no idea how they were using it and neither did the vendor - in SAS gives a dashboard to both sides - usage based feedback -- big metric of success.

Focus NOT on the core - sweet areas of a company -- G2000 customer require more integration compared to SMB customer. Start in the mid tier market space. Emergence of open standards integration can be easily done.

SMB market traditionally very fragmented.

White space. . enterprise mash up space - central easy to use resource, new buyers are looking at new ecosystems IBM, ADP.

'democratization of software'

XBRL v other open specs



XBRL is uniquely positioned compared to other open specifications. It is optimized for the exchange of historical, archival business reporting data. It models data that is hierarchically arranged for drill-down and reported along dimensions of time, entiry and scenario (or context). And, it is independent of specifc industry, regulatory regime or level of detail.

Key themes distinguish XBRL from other standards in the financial arena:
  • Reports, as distinct from transactions. A purchase order, or, more precisely, the sending and acceptance of a purchase order, is a transaction, transactions are the purpose of a whole host of standards from IFX, OFX, ebXML, ACORD and others.
  • Performance data, as distinct from market data. market data tends to be ephemeral and real-time, with pricing being always crucial; performance data is archival and records the history of busness operations and their results. XBRL is about performance data.
  • Entities as distinct from investment instruments. Equities are financial instruments whose underlying value is based on public company entities; an entity is the business itself. XBRL represent detail about entities - not only publicly traded companies, but any business or non-profit entity. Equities and other financial instruments are the subject matter of MDDL, FpML and others.
  • Reporting metadata, as distinct from reporting metadata. Metadata - data about data - is, for the most part, data abouta document. XBRL defines how the individual numbers and facts inside the financial statements and similar documents relate to one another.
In general, XBRL business content can be embedded into any other standard that is related to transactions, market data, instruments, or document metadata. For example, if there were a standard ebXML business process for tax reportng, XBRL could be used as part part of the "payload" of the tax return itself, since it is used to present financial statement level data as well as the ledger of business transactions that are classified into different tax treatments.

XBRL and business performance reporting - makes a distinction between financial information reporting of a set of data and supporting text, versus the data in individual financial transactions. XBRL is focused on providing a rich, detailed, comprehensive standard for representing data used in business reporting.