سخن سردبیر
عنوان مقاله English
نویسنده English
In recent decades, the dominant discourse in national science and technology policy has consistently underscored concepts such as scientific authority, science diplomacy, the ranking and internationalization of universities and journals, and enhancing the visibility of Iran’s research achievements and scholarly publications. Accordingly, upstream policy documents, strategic roadmaps, and academic promotion regulations have persistently steered researchers and academic institutions toward active participation in the international scientific constellation and the dissemination of scholarship with global resonance; after all, elevating the standing and advancing the internationalization of periodicals and universities across all dimensions guarantees invaluable dividends for all stakeholders within the country’s science and research landscape (Noroozi Chakoli, 2021-a, 2021-b, 2022).
Nevertheless, reflecting on the technical and functional attributes of the nation’s scientific web reveals a conspicuous hiatus between the targets envisioned in policy frameworks and the empirical realities at the infrastructure and operational layers—a gap that warrants critical scrutiny and strategic recalibration. Addressing this issue within the domains of scientometrics, webometrics, and alternative metrics (Altmetrics) is of paramount importance because, in contemporary science assessment paradigms, web-based visibility serves as an indispensable prerequisite for achieving scholarly impact and citability (Noroozi Chakoli, 2011). Today, the global standing of researchers, scientific journals, and universities—all of which constitute the cornerstones and emblems of national scientific stature and authority—is governed more than ever by quantitative and qualitative indicators derived from data networks and web-based presence. In global university ranking systems (including the Webometrics Ranking, as well as citation- and reputation-based systems such as Leiden, Times Higher Education, and QS) (Aguillo et al., 2008), as well as in journal impact calculations, a substantial weight of evaluation matrices is allocated to information retrieval capability, linkability, citation yield, and active engagement within the open scientific information ecosystem. When access to dissemination platforms is severed or impaired, the automated web crawlers and indexers of international abstracting systems and academic search engines (such as Google Scholar) are rendered incapable of indexing and crawling these corpora. Consequently, even the most rigorous scholarly output is omitted from scientometric computations—a vicious cycle that directly results in the decline of university rankings, the diminution of journal impact factors, and the erosion of the nation’s key scientific indicators.
Currently, one of the most critical impediments to the global visibility of Iran’s scientific publications lies in the enforcement of geographic access restrictions (geo-blocking) and the international inaccessibility of major national repositories tasked with organizing, managing, and disseminating bibliographic and scientific data. Authoritative platforms—such as the National Library and Archives of the Islamic Republic of Iran (NLAI), the GANJ system of the Iranian Research Institute for Information Science and Technology (IranDoc), the Scientific Information Database (SID), Magiran, Civilica, and the Comprehensive Humanities Portal—have frequently encountered structural disruptions or full overseas access blockades due to technical contingencies, data protection policies, or network infrastructure constraints. Such restrictions preclude international scholars and academic institutions from accessing the vast corpus of scientific knowledge generated and indexed in Iran—encompassing books, dissertations, journal articles, conference proceedings, and funded research projects. As a result, researchers, libraries, publishers, and scholarly repositories worldwide are deprived of access to the master record—namely, the official, authoritative bibliographic entry and documented identity of Iranian scholarly works—preventing them from linking to these sources in formal citations. Beyond dispossessing the global academic community of a rich knowledge repository, this breakdown drastically diminishes the capacity of national publications to influence international scientific workflows and accrue global citations; a circumstance that denies Iranian scholarship the opportunity for global uptake, citation accrual, and active participation in international knowledge generation. Consequently, an international researcher seeking to identify or cite an Iranian work is inadvertently funneled toward commercial platforms, online bookstores, or alternative, often volatile secondary sources instead of connecting directly to an authoritative national repository. Under such conditions, the nexus between the original scholarly work and its official bibliographic record is severed, substituting a commercial broker or vendor in the citation chain in lieu of a legitimate academic repository. This displacement undermines the bibliographic authority and integrity of Iranian scholarship within the international scientific domain, posing formidable barriers to its documentation, retrieval, and citability.
The gravity of this impediment becomes increasingly pronounced when juxtaposed against an analogous scenario involving the Library of Congress or global union catalogs such as WorldCat. Were international public access to the bibliographic records of these institutions compromised, thereby barring scholars from referencing and citing their cataloged entries, these entities would progressively forfeit their functional utility and international authority; for the prestige of any bibliographic repository hinges not solely on data accuracy and exhaustiveness, but equally upon its continuous accessibility, referential integrity, citability, and the linkability of its records. Viewed through this perspective, the inaccessibility of authoritative national infrastructures—such as the National Library and Archives, the GANJ system of IranDoc, SID, Magiran, Civilica, and the Comprehensive Humanities Portal—cannot be dismissed as a mere technical defect or transient networking glitch. Rather, it represents an infrastructural rupture between Iranian intellectual output and the global discovery, documentation, and citation fabric, eroding the country’s authoritative standing within the global bibliographic information system.
The immediate consequence of this impediment—bearing profound scientometric ramifications—is the systemic marginalization of a major proportion of national research output, the exclusion of global scholars from engaging with primary Iranian literature, and the artificial compression of Iran’s authentic contribution to international scientific discourse. In an era dominated by ubiquitous artificial intelligence and large language models (LLMs), this predicament assumes even more complex and critical dimensions. Web scrapers, web crawlers, and autonomous AI agents tasked with indexing, synthesizing, and training upon human knowledge inevitably bypass these scholarly assets when confronted with impenetrable digital perimeters.
Through this operational dynamic, an empirical scientometric phenomenon emerges in the behavioral analysis of AI models that aligns directly with Zipf’s Principle of Least Effort (Zipf, 1949, as cited in Case & Given, 2016). Driven by programmatic constraints, AI architectures default to the most easily accessible, cost-effective, and retrievable data nodes rather than filtering and synthesizing the most authoritative, rigorous, and authentic scholarly sources. Within this framework, indigenous academic disciplines—most notably Iranian and Islamic history, culture, literature, art, and civilization—bear the heaviest burden. This transpires despite the fact that elevating the prestige and international authority of the Persian language in scholarly discourse has consistently occupied the highest echelon of national strategic directives, particularly under Macro-Strategy 9 of the Comprehensive Scientific Map of the Country (Secretariat of the Supreme Council of the Cultural Revolution, 2011). Nonetheless, the infrastructural isolation and invisibility of domestic repositories deprive a vast volume of seminal Persian-language works and genuine intellectual paradigms of the opportunity to participate in international discourse, exert scholarly influence, and garner global citations. Conversely, in this algorithmic landscape, when an Iranologist, Orientalist, or general user queries an AI system regarding historical or civilizational facts concerning Iran, the algorithmic agents—finding the primary national databases and bibliographic repositories inaccessible—inevitably bypass authentic domestic scholarship. Instead, they retrieve secondary, superficial, imprecise, and occasionally biased materials solely because they reside upon the visible, open web. In practice, this outcome not only impedes the realization of Persian as a language of scientific authority, but also breeds the hazard of unintended or systemic algorithmic distortion targeting national history, culture, and intellectual heritage.
Beyond the challenges of repository accessibility, another foundational impediment impairing the web-based visibility and citability of national scholarship stems from technical web engineering, architectural design, and the formatting conventions and persistence of Uniform Resource Locators (URLs). An examination of the digital ecosystem of Iranian academic websites, scholarly periodicals, and commercial distribution platforms (such as Gisoom, Ketabrah, Taaghche, Fidibo, Ketabnak, Ketabkhoon, Book Agency, and several university presses) reveals structural deficits that adversely impact the discovery and citability of scholarly assets:
1. Character Encoding Challenges and Referential Fragility of Links: A primary technical bottleneck in the digital tracking and sharing of Iranian scholarly works lies in the handling of Persian Unicode characters within URL conventions and web protocol standards. Directly embedding raw Persian strings into URI paths without executing standardized optimization mechanisms triggers percent-encoding, yielding excessively convoluted, unreadable, and brittle links (e.g., %D8%B9%D9%84%D9%85 representing the term Elm [Science]). In addition to completely compromising human readability, this phenomenon hinders link-sharing and direct citation to web-hosted academic texts. It routinely compels researchers to rely on third-party link-shortening services to manage unwieldy URLs; an intervention that introduces an unnecessary and unstable dependency layer, effectively undermining persistent trackability, seamless algorithmic processing, and the global visibility of scholarly works.
2. Link Fragility and Link Rot: A formidable barrier to international citability is the absence of robust, standardized persistent identification and tracking frameworks. Within digital scholarly communication, link persistence is imperative; a citation forged by an academic in 2024 must not terminate in an impassable dead end (broken link) by 2030. The failure to adopt persistent identifiers (PIDs)—such as DOIs and Handles—for monographs, articles, and scholarly resources, coupled with the legacy practice of reassigning a single transient URL to multiple heterogeneous works over time, severely compromises the integrity and citability of academic literature. While the international publishing ecosystem relies on persistent identifiers—such as DOI (for serial literature) and Handle or ISBN (for monographs)—many domestic systems continue to rely exclusively on volatile dynamic web addresses. Unlike a conventional URL that functions merely as a transient locator pointing to a physical storage address, Handle-based systems are grounded in the Digital Object Architecture (DOA), resolving directly to the digital entity or content itself. This architectural independence permits infrastructural restructuring and physical file migrations without breaking the established citation chain (Hisseine et al., 2024). Consequently, the Handle System empowers web administrators to modify directory structures or migrate servers without the risk of destroying scholarly citation trails. The absence of these identifiers not only disrupts digital trackability but also directly impairs visibility, thereby confronting the fundamental evaluative mandate of scientometrics with formidable methodological challenges in tracking international citations.
In global practice, leading platforms such as Google Books and Amazon maintain the perpetual integrity of an item’s identifier across arbitrary infrastructural overhauls, as their systems are anchored to immutable, persistent identifiers. In contrast, on numerous domestic academic websites and publishing portals, the URLs of a monograph or journal article remains tied to ephemeral folder hierarchies or internal database keys. As a consequence, even the most modest architectural modification or server migration precipitates acute link rot—a terminal state wherein the URL documented in a bibliography degrades into a dead link that no longer resolves to the source document (Pishchyk, 2026). This fragility does not merely obstruct scientific retrieval and citation workflows; it deeply compromises the referential validity and academic prestige of national scholarship in the eyes of the international research community.
Recommendations
To navigate toward an optimal state and transition Iran’s scientific web into a dynamic, highly linked, and globally authoritative ecosystem, concerted action across two interrelated tiers- macro-strategic/policy and technical/infrastructural- is recommended:
A) At the Marco-Policy and Scientific Information Management Level
1. Revising Geo-Access Management Protocols: Drawing a rigorous architectural demarcation between confidential enterprise data and open bibliographic metadata and scientific literature; and guaranteeing stable international access to primary scientific databases, national bibliographic repositories, and university portals via modern edge security layers, adaptive rate-limiting, and intelligent firewalls rather than blanket geographical blockades.
2. Formulating and Promulgating Scientific Web and Linked Data Standards: Implementing coordinated policy initiatives by governing authorities (such as the Supreme Council of the Cultural Revolution, the Ministry of Science, Research and Technology, and the Ministry of Health and Medical Education) mandating institutional adherence to international webometrics, semantic web, and Linked Data standards (Hall & O’Hara, 2009) across all universities, research institutes, and academic journals.
3. Integrating Webometrics and Visibility Indicators into Academic Evaluation Schemes: Elevating performance metrics governing link persistence, international open access availability, and automated discoverability within national periodic institutional audits, university rankings, and journal accreditation frameworks.
B) At the Technical Engineering and Digital Publishing Infrastructure Level
1. Institutionalizing Persistent Identifiers (PIDs): Expanding the mandatory implementation of persistent identifiers, notably DOIs for all journal articles, conference papers, and institutional research reports, while systematically binding monographs and book chapters to Handle systems and standard ISBN architectures across institutional portals and publishers to eliminate link rot.
2. Optimizing URI and Path Architectures: Re-engineering the linking conventions of open journal systems (OJS) and academic repositories through standard semantic naming conventions, intelligent transliteration/slugs, or search-engine-friendly numeric keys in place of raw Unicode characters, thereby obviating percent-encoding artifacts and preserving hyperlink structural integrity.
3. Ensuring Compatibility with Web Crawlers and AI Indexers: Structuring institutional repositories with standardized academic metadata schemas (such as Dublin Core, Highwire Press, and Schema.org microdata) across all landing pages to facilitate automated harvesting, semantic parsing, and systematic indexing by academic search engines (e.g., Google Scholar) and autonomous AI retrieval agents.
4. Reinforcing the Master Record Authority within National Bibliographic Systems: Establishing standard programmatic cross-linking between distributed book portals and the definitive, canonical bibliographic record at the National Library and Archives of Iran, ensuring that international citations anchor directly to the verified national registry of record.
Conclusion and Concluding Remarks
An examination of the nexus connecting scientometrics, webometrics, and digital infrastructures demonstrates that elevating national scientific stature and projecting scholarly output onto the global stage can be profoundly accelerated by synergizing upstream strategic policies with cutting-edge scientific web standards. In the present epoch, optimizing maximum accessibility and discoverability serves as the cardinal prerequisite for citability and sustainable scientific impact-a paradigm that guarantees the authentic, rigorous representation of academic achievements within transnational bibliographic indices and scientometric analyses.
Upgrading technical infrastructures in alignment with Semantic Web specifications, alongside fostering secure and open access to primary scholarship, will chart novel pathways for representing the nation’s intellectual heritage and the Persian language across emerging technologies and artificial intelligence models. Consequently, elucidating the multidimensional facets of this domain unfolds an invigorating agenda of novel, constructive research inquiries for scholars, theorists, and investigators within scientometrics and webometrics:
- Through which operational and mathematical models can the expansion of global access to national academic repositories optimize the visibility rates, foreign citations, and bibliometric performance indicators of the country’s universities and researchers?
- How can the rigorous implementation of web accessibility and data engineering principles maximize the retrieval efficacy of large language models in accurately synthesizing primary Persian scholarship and solidifying indigenous scientific authority within global platforms?
- What empirical role does the systemic deployment of persistent identifiers (such as Handle and DOI) across national academic ecosystems play in sustaining the citation chain, mitigating link decay, and enriching the national knowledge graph on an international scale?
- To what extent can the integration of semantic metadata standards (e.g., Dublin Core, Highwire Press, and Schema.org) augment the crawlability, indexing fidelity, and discoverability of Persian scientific output across global academic search engines?
- How can scientometric evaluation frameworks be redefined to assimilate webometric dimensions—such as link persistence, algorithmic discoverability, and web accessibility—as value-adding parameters within institutional accreditation, academic promotion systems, and university ranking methodologies?