Welcome to IIT

The Insititute of Informatics and Telecommunications (IIT) focuses on research and development in the areas of Telecommunications, Networks, Web Technologies and Intelligent Systems. The main strategic objective of IIT is to excel in research and innovation required for the development of the knowledge society.

Technology of the month

Large Scale Hierarchical Text classification (LSHTC) Pascal Challenge launched!

The LSHTC Challenge is a hierarchical text classification competition using large datasets based on the ODP Web directory data (www.dmoz.org). Hierarchies are becoming ever more popular for the organization of text documents, particularly on the Web. Web directories are an example. Along with their widespread use, comes the need for automated classification of new documents to the categories in the hierarchy. As the size of the hierarchy grows and the number of documents to be classified increases, a number of interesting machine learning problems arise. In particular, it is one of the rare situations where data sparsity remains an issue despite the vastness of available data. The reasons for this are the simultaneous increase in the number of classes and their hierarchical organization. The latter leads to a very high imbalance between the classes at different levels of the hierarchy. Additionally, the statistical dependence of the classes poses challenges and opportunities for the learning methods.

Read more

© 2012 - Institute of Informatics and Telecommunications | National Centre for Scientific Research "Demokritos"