workshop III. @ code hub zadar

Turning scientific journals into semantic knowledge systems using artificial intelligence

Domagoj Ivanuša [MatchMindz, Zagreb, Croatia]

Mario Ivanuša [Cardiologia Croatica ; University of Rijeka, Faculty of Medicine ; Institute for Cardiovascular Prevention and Rehabilitation, Zagreb, Croatia]

The value of a scientific journal lies not only in its individual articles, but in the body of knowledge they collectively create. Although scholarly publishing has become largely digital, its underlying infrastructure remains primarily focused on storing and retrieving individual documents. Most journals provide access to content through PDF files and keyword-based search, while more advanced publishing platforms additionally publish articles in structured XML format to support interoperability and indexing in international bibliographic databases. However, the full potential of this structured content remains largely untapped. Advances in artificial intelligence (AI) now make it possible to leverage the same infrastructure for a new role of journals – not merely as repositories of publications, but as active knowledge systems.

Using the journal Cardiologia Croatica as a case study, this workshop presents the transformation of more than 2,000 peer-reviewed articles published over the past decade into a semantic knowledge system. The textual content of structured XML records was converted into vector embeddings and stored in a vector database, enabling semantic retrieval across the journal’s entire archive based on meaning and conceptual relationships rather than keywords alone. This allows users to identify related publications regardless of differences in terminology, trace the evolution of scientific topics over time, and uncover relationships among authors, concepts, and publications that remain hidden using conventional search methods. The platform combines traditional and AI-enabled approaches to accessing scientific content. In addition to browsing by author, keyword, and publication year, it enables semantic literature search using natural language. There is a conversational AI agent, based on a Retrieval-Augmented Generation (RAG) architecture, which was trained on retrieving relevant sections of the journal’s corpus and exclusively uses journal insights to generate responses. The same infrastructure also supports AI-generated educational flashcards, interactive clinical case presentations, and chronological visualizations of the development of scientific topics and the contributions of individual publications.

Alongside presenting the system architecture and implementation process, the workshop will discuss the applicability of this approach to other scientific journals, institutional repositories, and digital archives. The broader adoption of this model has the potential to transform traditional literature archives from collections of individual publications into dynamic knowledge infrastructures that provide more personalized access to scientific research, education, and scientific knowledge.

Domagoj Ivanuša is an entrepreneur working at the intersection of technology, product strategy, and artificial intelligence with nearly a decade of experience across startups, Amazon’s EU headquarters, and his own venture studio. His career has focused on translating business and knowledge management challenges into scalable digital products – from developing go-to-market strategies, driving product adoption, and supporting business expansion at Amazon, to leading commercial and marketing strategy at a VC-backed technology startup. Alongside his professional career, he organized innovation programs and served on the jury of Luxembourg’s oldest startup competition. Today, he is the CEO and Co-Founder of MatchMindz, where he develops custom software, AI systems, and digital products across e-health, scientific publishing, business automation, sport management, real estate, tourism, and media. His current focus is on designing AI-native knowledge systems that combine semantic retrieval, structured data, and large language models to make scientific and business knowledge more accessible and reliable.

thursday, 1o September, 9:30 – 10:30 

PARALLEL WORKSHOP SESSION

[Workshop I]

  • AI and copyright (Maja Bogataj Jančič)

[Workshop II]

  • Whose journal is it anyway? Workshop on journal ownership (Milica Ševkušić, Sofie Wennström, and Johan Rooryck)

[Workshop III]

  • Turning scientific journals into semantic knowledge systems using artificial intelligence (Domagoj Ivanuša and Mario Ivanuša)
Check out the full » programme «

Domagoj Ivanuša

MatchMindz » Zagreb, Croatia