Prof. Petras erforscht die Infrastruktur und Governance von Forschungsdaten sowie deren Erschließung durch digitale Methoden. Ihr aktueller Schwerpunkt liegt auf der Entwicklung von Standards, Klassifikationen und Werkzeugen zur besseren Auffindbarkeit, Qualität und Nachnutzung von Forschungsdaten — etwa durch das Datenkompetenzzentrum QUADRIGA, das Repositorien-Verzeichnis re3data und Projekte zur automatisierten Extraktion und Modellierung von Metadaten aus Kulturerbe-Beständen. Für Unternehmen und öffentliche Institutionen (Bibliotheken, Archive, Museen, Forschungsförderorganisationen) löst sie damit das Problem der fragmentierten, schwer aufffindbaren Datenbestände und schafft Grundlagen für deren systematische Wiederverwendung. Die Methoden sind relevant für alle Sektoren, die mit großen, heterogenen Datenmengen arbeiten — von der Wissenschaftsverwaltung über Kulturinstitutionen bis zu interdisziplinären Forschungsverbünden.
🔒 Das System hat 93 mögliche Industrie-Partner gefunden — Firmen, Scores und Begründungen sind nur für eingeloggte Nutzer:innen sichtbar. Anmelden
Prof. Vivien Petras
HU-FIS-Profil ↗With the growth of digital libraries and digital library federation (as well as partially unstructured collections of documents such as web sites), a large set of vendors is offering engines for retrieving contents and metadata via search requests by the end user (queries). In most cases these queries are just unstructured fragments of text in a specific language. The first service offered by GALATEAS (LangLog) is focussed on getting meaning out of these lists of queries and it is addressed to library/federation/site managers. Contrary to mainstream service in this field, GALATEAS services will not considered standard structured information of web logs (e.g. click rate, visited pages, user s paths inside the document tree) but the information contained in queries from the point of view of language interpretation. By subscribing LangLog federations administrator and managers will be able to answer questions such as: as Which are the topics which are most commonly searched in my collection, according to a certain language? ; how do these topics relate with my catalogue? ; Which named entities (people, places) are more popular among my users? The second problem addressed by GALATEAS is the one of Cross Language Information Retrieval (CLIR) i.e. the capability of typing a query in one specific language and retrieving documents which are available in different languages. The CACAO consortium is already successfully providing services for indexing and searching over digital libraries and metadata repositories. During commercial exploration for marketing CACAO it emerged that certain institutions prefer to keep indexing and searching at their premises (using their own favourite search engine) and would be perfectly satisfied with a service of plain query translation.The second service offered by GALATEAS (QueryTrans) has the ambitious and innovative goal of providing the first web translation service specially tailored on query translation. Languages addressed by both LangLog and QueryLog are: Italian, French, English, German, Dutch, Modern Arabic and Polish.
Noch keine Publikationen aus OpenAlex zugeordnet.
Measuring is a key to scientific progress. This is particularly true for research concerning complex systems, whether naturalor human-built. Multilingual and multimedia information systems are increasingly complex: they need to satisfy diverse user needs and support challenging tasks. Their development calls for proper evaluation methodologies to ensure that they meet the expected user requirements and provide the desired effectiveness. Large-scale worldwide experimental evaluations provide fundamental contributions to the advancement of state-of-the-art techniques through common evaluation procedures, regular and systematic evaluation cycles, comparison and benchmarking of the adopted approaches, and spreading of knowledge. In the process, vast amounts of experimental data are generated that beg for analysis tools to enable interpretation and there by facilitate scientific and technological progress. PROMISE will provide a virtual laboratory for conducting participative research and experimentation to carry out, advance and bring automation into the evaluation and benchmarking of such complex information systems, by facilitating management and offering access, curation, preservation, re-use, analysis, visualization, and mining of the collected experimental data. PROMISE will: foster the adoption of regular experimental evaluation activities; bring automation into the experimental evaluation process; promote collaboration and re-use over the acquired knowledge-base; stimulate knowledge transfer and uptake. Europe is unique: a powerful economic community that politically and culturally strives for equality in its languages and an appreciation of diversity in its citizens. New internet paradigms are continually extending the media and the task where multiple language based interaction must be supported. PROMISE will direct a world-wide research community to track these changes and deliver solutions so that Europe can achieve one of its most cherished goals.