Saturday, April 07, 2007

Semantic Web is closer to be real, isn't it or is it?

A recent post by technical evangelist Robert Scoble brought a small hype of semantic web again. His article was about the new achievement done by Radar Networks, which was founded by Nova Spivack. The following is quoted by Robert's post.

Basically Web pages will no longer be just pages, or posts. They’ll all be split up into little objects, stored in a database (a massive, scalable one at that) and then your words can be displayed in different ways. Imagine a really awesome search engine that could bring back much much more granular stuff than Google can today. Or, heck, imagine you could view my blog by posts with most inbound links.


Robert have expressed his excitement about watching the semantic web demo with Nova. Although I have not gotten the chance to experience it by myself, I can briefly feel what is going on and which technologies Radar Networks is now employing. This expressed scenario is a certain picture that current semantic web technology can support. Here at our research lab at BYU, we cooperate with some DERI researchers together and work on a project about enabling the semantic web to be real. In fact, the paradigm of our plan is close to the scenario that is expressed by Robert. So I feel familiar when I watch Robert's expressions.

The idea of semantic web has been discussed for years. The web is certainly moving towards holding more and more machine-processable semantics. But until now, the research of semantic web is still mostly limited in labs. What are the reasons?

According to our beliefs, the reason is not due to technical difficulties. Technical difficulties are severe problems, but not the deadly ones. The real difficulty is about who are going to take the control of semantic definitions. Are these definitions controlled by few elites or by the public? This is a grand question.

Many people have predicted the rise of "semantic Google." But I would say that there would never be a "semantic Google." Again, my argument is not due to the technical reasons. The problem is about who is going to be the owner of such a "semantic Google." If it is a US company, I must say sorry to it that countries such as China, Russia, or even European Union would ban it and build their own semantic giants because any major country in this world would not endure its "semantics" being controlled in the hand of another country. This type of threats is intolerant to any independent country.

In fact, we may even not need to raise this problem to the level of nations. Even within individual persons, no one would like to be forced agreeing on the semantics defined by another person. This semantic-definition problem is indeed the most crucial problem to the prevalence of the semantic web.

What is the solution? In fact, the success of Web 2.0 has shown a pragmatic resolution. We need to allow the public to define semantics by themselves. This is the only way that can promise the prevalence of semantic web. Any web user can apply his machine agent to understand the web based on his own understanding. This is what a pragmatic semantic web should deliver. On such a pragmatic semantic web, no "semantic Google" can exist due to the massive diversity of human understanding. A collaborative web search model will replace the current centralized search model. In fact, we have invented a new theory of collaborative web search on the semantic web, and hopefully it could be released soon.

In summary, Nova Spivack and his company Radar Networks have done great achievement on realizing the semantic web. We must congratulate them! Great work, Nova! On the other hand, however, unless they can show that their solution has properly solved the semantic definition problem, the age of semantic web is still not there yet.

Wednesday, April 04, 2007

Story of My Internship at DERI Innsbruck Last Summer


Today the CS department Homepage at BYU posted a story of my intership at DERI Innsbruck last summer. It was indeed a very pleasant experience. DERI Innsbruck is currently one of the best Semantic Web research labs in the world. During this period, I had wonderful working experiences with many DERI people, especially Martin Hepp, Ying Ding, Omair Shafiq, and Jan Henke. This article brings me many wonderful memories about DERI and Innsbruck. If some young Semantic Web researchers are looking for a place to do internship right now, DERI Innsbruck is definately a place they need to try.

Wednesday, March 28, 2007

Online Desktop versus Offline Web


Today there is an interesting post on the Read/WriteWeb: Point/Counterpoint: Which is better, an offline Web App or an online Desktop App? In the article, John Milan and Richard MacManus made a debate on which one is better for the future---offline Web applications or online Desktop applications. This is an interesting article, and many of the comments are worthwhile to read too.

I have also made some comments to the authors. In general, when users work offline, they are on their private life. On the contrary, when users work online, they are on their social life. Both sides, however, are fundamental parts of the human life.

An offline Web application is to take a part of social activities to a private space; and an online Desktop application is to take a part of private bahaviors into a social domain. In our real life, we see both of these phenomena regularly. Most of the time, which one is better does not depend on the method itself. In contrast, it is more about WHO makes the decision. Some people like to let everyone know their privacies; they definately like to use online Desktop applications. Some other, however, like to hide their soical activities as many as possible to be private information; and they will prefer to having offline Web applications. So which one is better? It does not depend on either researchers or developers. The answer can be very much different from one person to another. This is why both strategies have their customs.

Moreover, I would like to analyze this debate from the view of web evolution. By this view, the rise of web applications is a phenomenon of web evolution rather than a pure advancement on web technologies. At the beginning, web and desktop were two comparatively independent environment because web was only a network of newborns (virtually, see my article for the explanation). Desktop applications were crucial for users because they could not hand their assignments to "newborns" to do. The starting of the web evolution cycle (typically the emergence of Web 2.0) begins changing this thought. More and more desktop applications can now be done on the web. This trend is going to continue when web grows to be a network of more and more matured virutal people.

Does this trend result in the death of Desktop applications? Absolutely not. No matter how much we would like to be soical, we still need private spaces where live the Desktop applications. We always have something that is so private that we do not want to share. Even if we have discussed before that how deep we would like to clone ourselves, there is always something that we do not want to be cloned (if ever they become cloneable). Moreover, there are also some social activities that are private. We do not want others to know that we have these social activities. All these reasons make us believe that Desktop applications will not die and some web applications need to go offline.

Human society is complicated. Humans are the most complicated creatures of all. If anyone wants to categorize all human behaviors by a simple model, very likely it may not work well.

Sunday, March 25, 2007

The religionary side of World Wide Web

(Revised at Sept. 29, 2007)
World Wide Web is a religion. But please don't be panic if you are either an atheist or a fundamentalist. World Wide Web is not a "standard" religion. It is a religion-like Atheism. World Wide Web directly addresses the fundamental of all religions --- the seeking to eternity.

Every religion is an attempt to answer one essential question: what does eternity mean to us? The difference among various religions including Atheism is basically the different answers to this question. To Christians the answer is Jesus Christ; to Muslims the answer is Allah; to Jewish people the answer is Jehovah in the Jewish Bible; to Buddhists the answer is transmigration of the soul; to atheists the answer is that there is no eternity, one's life terminates immediately after death without any form of successors. To WWW users the implicit answer is that we can keep ourselves mentally eternal on the Web.

Almost everyone, no matter whether they believe in Gods or which God they believe, agrees on the existence of immortality. There is, however, difference between physical immortality, spiritual immortality, and artificial immortality. Many religions believe in spiritual immortality; our soul can last forever after the death of body. Some religions even seek to the chance of physical immortality; let body live forever. In contrast, both atheists and theists accept certain form of artificial immortality, i.e., transferring one's consciousness (except self-awareness) into an alternative media that can be assumed staying forever. The fundamental difference between artificial immortality and spiritual immortality is that the latter one also includes the resurrection of humans' ultimate identity --- ones' self-awareness.

Humans have experienced artificial immortality in many forms. Writing books, composing musics, producing artifacts, building constructions, and worshiping antecessors are several typical forms to achieve artificial immortality. We regard ourselves being artificially immortal when our consciousness is kept in history and remembered by our posterity through our publications and artifacts. Based these inherited work, later generations can restore part of the precious consciousness of their ancestors. By this sense, these ancestors reach certain level of immortality though they cannot reclaim the ownership over these consciousness (i.e. the restore of self-awareness).

World Wide Web is a revolutionary new way to achieve artificial immortality. In history, artificial immortality was a reachable but expensive goal. Only the greatest work of mankind could be kept; let it alone the requirements of surviving from wars and natural disasters. But the invention of WWW gives normal persons a cheap and convenient way to keep their consciousness in history for long time. This is the first time ever normal people have a chance be remembered by not requiring great thoughts. Anyone can materializes their consciousness on web even if the thoughts are silly. This is poor man's revolution.

By this sense of artificial immortality WWW is like a religion. Religions addict people into them by the promise of providing immortal future. In similar, WWW addicts web users into it by the promise of providing artificial immortality on the Web. This satisfaction of intrinsic nature of mankind is an essential reason (or probably the most essential reason) why World Wide Web engages such a great success in history. Faith and intrinsic desire are the ones that can never be evaluated by money and time.

This faith to artificial immortality is the fundamental driving force to web evolution. In one of my previous posts, I presented that web evolution is a history of incrementally cloning the consciousness of individual web users. The goal of web evolution is to improve this cloning procedure to its deeper level. Ultimately, we may be able to mentally resurrect any web user by the preserved consciousness. If this time could ever come, no one will die with respect to the others. Pitifully, however, everyone of us could still not live longer than our one life, with respect to ourselves.

Friday, March 23, 2007

retrievr: an interesting progress on image search

retrievr is a new image search service that let users find flickr images by drawing rough sketches of them. It is not searched by keywords or meanings. In contrast, it is searched by visual effects. For example, when I upload my own photo to the site, it returns me a set of black-and-white pictures that has similar visual effect as my picture.



I am not sure how this technique can be used in real-world cases. But it is fun to play with it. If retrievr creates a Web-2.0 community and allow users voting and sharing their results, it may greatly help them improve their technique and bring up more creative ideas on how to use this cool technique on real world applications.

Thanks for JurijMLotman pointing this site and its mother site SystemOne to me. They are doing really cool stuffs. The web indeed becomes more and more interesting.

Will the Semantic Web fail? Or not?

(updated Dec. 21, 2007)

A recent post by Stephen Downes has led to quite a few discussion on whether or not the Semantic Web will fail. There are a few supporters, but more opponents.

The center of the debate is on Stephen's premise: the Semantic Web will never work because it depends on businesses working together, on them cooperating. Many Semantic Web researchers decline this assertion by saying that the Semantic Web technologies do not rely on agreements between businesses. This counter-statement is, however, true and false. Certainly, we may say that businesses can develop their own ontologies and store their own data in their own RDF files. In short, they do not have obligations to pre-agree anything. But then the problem is: how would we gain by without agreements? Note that automatic ontology mapping is still a problem that is too hard to be practically solved in the foreseeable future.

I am a Semantic Web researcher and I believe in the future of the Semantic Web. But some of the Stephen's viewpoints are indeed good. Semantic Web is not going to be realized inside the ivory tower. The success of the Semantic Web does not depend on how good our ontology reasoning algorithms have been implemented, or how well our ontology languages have been designed. All of these issues are important, but an even more important one is how we can persuade normal web users starting annotating their own data. Annotated data are the real center of the Semantic Web.

Fortunately, Web 2.0 has already led users to the realm of tagged Web content. This is a realm dreamed by Semantic Web researchers who, however, have never succeeded in bringing the regular Web users to this realm. Now a critical challenge is how we may transfer this Web 2.0 impact into the Semantic Web realm. (Or on its reverse, how the Semantic Web research can be merged into this Web 2.0 phenomenon.) If we could succeed on this challenge, the Semantic Web would be the Web 3.0, 4.0, or x.0. Otherwise, the Web 3.0 might be on another route that starts to run away from this Semantic Web community in the ivory tower. Certainly there will be a web engaged with semantics in the future; but whether it is this Semantic Web that suggested by W3C depends on what Semantic Web researchers are doing at present.

Will this Semantic Web research fail? Or not? If the Semantic Web researchers can be humble enough to admit the shortcomings of the Semantic Web proposal and start to learn from the success of Web 2.0, the Semantic Web will succeed. Otherwise, the Semantic Web will fail because nobody can win a Web battle when they are standing on the opposite side to millions of real-world users because it is these normal users who really decide the future of any Web technologies.

Sunday, March 18, 2007

How deep do we want to clone ourselves?

For individual web users, the evolution of World Wide Web is a history of incremental self-cloning (mentally, of course).

The traditional WWW (or we may call it Web 1.0) allows web users to create homepages to describe themselves. These homepages can show who they are and what they care about. Human readers may understand them. But there is no directly mutual communication between readers and homepages. So basically, Web 1.0 allows users to mentally clone themselves as babies.

Web 2.0 allows users to create personal accounts on various web sites. By subscribing to a web site, such as this one "blogger.com," we can not only post information describing ourselves, but also enjoy community-specific services provided by individual community organizers (such as various blogging facilities here on "blogger.com"). Moreover, these Web-2.0 accounts can actively perform various services for their owners by web feeds and widgets. From the individual point of view, it means that these accounts contain deeper implementation of humans' capabilities. This is why we can say that Web 2.0 allows users to mentally clone themselves as pre-school kids, a higher level of self-cloning mentally.

Tracing this orbit, future web evolution is going to more and more deepen this self-cloning of human web users. Users can materialize their mind, capabilities, and interpersonal relationships at deeper and deeper levels on the web. And these materializations can be behaved more and more close to their real-world copies, who are the original web users. So from one side, these online personal homepages or accounts are the virtual children of the web users. On the other side, they are actually are the mental clones of users themselves.

Now the question is: how deep do we want to clone ourselves? Do we really want to fully clone our thoughts? Will this trend lead to thinking machines? Will the nighmare of terminators become true some day? We don't know the answers yet. But the web is moving towards the answers of these questions.