Skip to content
Review Open access

A Reliability-Aware Retrieval-Augmented Generation Architecture for Open Language Models in Higher Education Decision Support

Aug 2026 · Computers · Vol 15, pp. 564 · 0 citations · 23 references

TL;DR

The contribution is a conceptual yet technically grounded deployment artifact that connects cloud computing, data science, and higher education governance; the architecture has not yet been empirically validated, and a protocol for future institutional pilots is specified.

Abstract

Open language models are increasingly considered for institutional decision-support tasks in higher education, including policy interpretation, academic advising, administrative summarization, and quality-assurance workflows. However, their reliable deployment requires more than model availability: it depends on cloud-native orchestration, retrieval quality, evidence grounding, refusal behavior, monitoring, and governance controls. Following a design-science research approach, this paper presents an architectural artifact for deploying open language models in higher education decision support. The artifact operationalizes institutional reliability as a multidimensional construct composed of contextual accuracy, answer faithfulness, retrieval quality, refusal adequacy, latency compliance, auditability, and human-review compatibility, and aggregates these into an institutional reliability index. It proposes a reliability-aware retrieval-augmented generation pipeline that integrates governed document ingestion, embedding generation, hybrid retrieval, reranking, evidence-aware generation, confidence-based refusal, human review, audit logging, and post-deployment monitoring. To support reproducibility, the paper compares four deployment configurations and provides an illustrative worked example of the reliability index. The contribution is a conceptual yet technically grounded deployment artifact that connects cloud computing, data science, and higher education governance; the architecture has not yet been empirically validated, and a protocol for future institutional pilots is specified.

Read PDF

Similar papers

Open access Aug 2026

Architecting Reliable Knowledge Retrieval Systems Using Large Language Models

A literature-based architectural framework for reliable knowledge retrieval systems that separates external knowledge management from LLM-based reasoning and generation is developed and indicates that reliable LLM deployment should be treated as an end-to-end architectural problem rather than solely a model-performance...

Bharat Kumar Reddy Karumuri · 0 citations
Open access Aug 2026

Beyond RAGAS: A Compliance-Aware Evaluation Framework for Retrieval-Augmented Generation in Regulated Sectors

Retrieval-Augmented Generation has become the dominant architecture for enterprise knowledge systems, with adoption spanning healthcare, education, and financial services. Existing evaluation frameworks — most notably RAGAS — measure faithfulness, answer relevance, context precision, and context recall. These metrics a...

Ashutosh Rana · 0 citations
Open access Aug 2026

An Analytical Study of Behavior-Aware Retrieval-Augmented Generation Frameworks in Enterprise Software Ecosystems for Optimizing User Navigation and Decision Support

It is demonstrated that incorporating real-time user behavioral context is critical for transforming generative AI utilities into proactive workflow accelerators within complex, data-dense corporate environments.

Sri Charan Chowdary Konidina · 0 citations
Open access Sep 2026

Evaluating Retrieval-Augmented Generation for personal collections: architecture, models and criteria

Preliminary results indicate that core categories of document mediation, such as relevance, completeness, citability and transparency, cannot be fully reduced to computational parameters but, instead, require continuous negotiation between automated models and disciplinary expertise.

Angelo La Gorga, Lorenzo Verna · 0 citations
Review Open access Aug 2026

A Role-Specialized Retrieval-Augmented Generation Framework for Automated Compliance Checking in Structural Engineering

Automated code compliance checking in structural engineering remains difficult because practical systems must balance accuracy, maintainability, and deployment cost. Pure prompting with large language models is prone to hallucination and unstable numerical judgment, conventional retrieval-augmented generation may fail...

X.-Y. Lu, J.-W. Zhu · 0 citations
#small language model Open access Aug 2026

Are reasoning paradigms scale-aware? A cross-paradigm verification of prompting, retrieval, and knowledge-graph scaffolding for small language models

A scale-aware comparative study of reasoning enhancement for SLMs across three major families of methods: prompting-based reasoning, retrieval-based augmentation, and knowledge graph guided scaffolding shows that reasoning-enhancement strategies are not universally transferable across model scales under the evaluated s...

Zhen-Zhen Gu, Jie Liu, Xian Liu · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.