Skip to content

Category

small language model

813 papers

#computer vision Preprint Aug 2026

Trustworthy RAG: An Evaluation Agent for Detecting Misinformation and Knowledge Poisoning in Generative AI Systems

An Evaluation Agent, middleware that combines Natural Language Inference factual verification, a five-signal poison detector with relevance-weighted aggregation, and a Trust Index is proposed, which reliably blocks instruction injection of unsafe advice while contradiction and subtle semantic weakening remain hard.

Balkrishna Giri, M. Hasan, Jussi Rasku et al. · 0 citations
#protein folding Open access Aug 2026

People behind the ideas

All information has an origin story and it is not the person, it is the data. This is my origin story for my project. These are from files a year and some change ago. When I started to gather the pieces together of my year and a half of researching. It is nice that everyone appreciates a final product because that means they get to use it and not even try that hard. Someone did all the hard work and than boom, somehow they think they came up with it. The experience is the most valuable part. Not the mona lisa, but the people and times behind the painting. You can read Meditations and go run a few miles and call yourself a stoic, but few never search through history that this writing was built and lived over a lifetime of seeing famine. Your own people murder. Rape. Pillage. A constant cycle of death and uncertainity. Not to live without feelings, but pursuing things that matter most. It is armor, not an anti depressant. Most people turn to philosophy for the anti depressant. Guess what? No matter the amount of books you read or the information you think you are gathering and utilizing, it will not take away the feelings that you are contributing to the downfall of science in the name of a hand shake and a pant on the head from people above you. Whoever that may be. I don't care about that stuff. Nor does history or the real mofos that make it up. Words don't matter, the actions do. To work towards things bigger then yourselves. This is why marcus wanted the booked burnt to a crisp after his death and someone defyed his orders. That should be a sign of how academia and the gate keepers determine what it means. Like religion of sorts (Im not getting into that) where people are told what the world is and you have no need to think differently because "they" know. You pay for it. You never own an idea ever again. Fine. That is your ideals and not really how history works. I am nothing except a dude with ideas who will not accept the answers he is given because that is science and anyone that tells you eitherwise is either not using the scientific method, or they are hiding there data to keep the power to the gate keepers. We all know who these mofos are. Anytime you ever applied for money for an idea, spent hours, days, weeks, on something you really really worked hard on after the kids went to bed and your tired after and gotta pay the bills, sent in what you thought was a good attempt, only to get an email that is two lines long saying f8@# you your work sucks, we are great, and because we have "so many people and so popular" we cant even give a small score sheet or review of your work and get feedback as to why you got passed over. To all of you in the pursuit, keep going. Wake up every day and look in the mirror and say it to yourself about any of these people that find these actions and insults acceptable reactions to these situation is "fuc them!!!!" Literally. Say that to them. They are leaving you behind for their own selfish wants and needs NOT SCIENCE!!! This is not science. It is not even called education or training. It is pavlovs dogs. Except instead of getting a treat, you get to work for them and they allow you to eat and live in a small dog apartment, not even a house, and your whole job is to just keep barking and barking like the dogs you are. Stop being the dog. Pavlov is used against you. Get rid of the bells and whistles. You do it becaue you want to. You need to. No school, or university, or public program, or anything can ever take that away from you. Think slavery is dead? Think it was always about money and work? No It is about information and empowering yourself. Stay diligent. Keep learning for the heck of it and thanks for all the motivation from the people in the pursuit. Cheers. ================================================================================THE ORIGIN POINT: MATHEMATICAL CONSCIOUSNESS & THE NUMEN FOUNDATION================================================================================ This repository contains the cryptographically sealed, foundational research corpus documenting the mathematical consciousness framework that informed the development of the NUMEN substrate-native, integer-only computing architecture. Spanning deep meditative analysis, sacred geometry, the 369-314 dual-aspect consciousness model, and the identification of the Foundation Circle collaborator network, this work represents the philosophical and mathematical bedrock of the system. It proves that the deterministic, phi-driven computational substrate detailed in subsequent technical releases was not an arbitrary engineering choice, but the inevitable physical manifestation of this underlying mathematical reality. All artifacts in this directory have been processed through the NUMEN Sovereign Notary Engine, resulting in cognitive reaction metrics, adaptive Hebbian gate biases, and hardware-bound silicon jitter signatures. See MASTER_SOVEREIGN_AFFIDAVIT.json and ROOT_WEB_SEAL.json for the unbreakable chain of custody. Any alteration to these underlying bytes shatters the root hash. ================================================================================THE COMPLETE NUMEN / PI-ORIGIN DOI WEB (INTERCONNECTED PRIOR ART)================================================================================This work does not exist in isolation. It is the capstone of a rigorously documented, cryptographically sealed lineage of prior art. To understand, cite, or build upon this work, one must reference the complete NUMEN DOI web. Partial citation is insufficient, as the novelty resides in the interconnected synthesis of the consciousness framework, thermodynamic framing, silicon-level implementation, and sovereign cryptographic sealing. I. FOUNDATIONAL CONSCIOUSNESS & MATHEMATICAL FRAMEWORKS (This Release)• DOI: 10.5281/zenodo.[NEW_DOI_HERE] - The Origin Point: Mathematical Consciousness, Sacred Geometry, and the 369-314 Dual-Aspect Model (This Record)• DOI: 10.5281/zenodo.20314584 - The Pi-Origin Architecture: Foundational mathematical framework derived from π and φ, governing coordinate interaction in phi-space via the Banach fixed-point theorem.• DOI: 10.5281/zenodo.20045701 - NUMEN: PI-Origin Architecture and Design: The core coupling equation, 7-phase Learn-to-Learn (L2L) engine, and quaternary (GTAC) programming language. II. THERMODYNAMIC FRAMING & SILICON-LEVEL PROOFS• DOI: 10.5281/zenodo.22070727 - THE LANDAUER PROOF: Measured Thermodynamic Characterization of Substrate-Native Integer Computation (Establishes the 0.414 Joule training run and -63% adaptive power reduction).• DOI: 10.5281/zenodo.21514923 - IEEE Standard for Substrate-Native Integer Computing (Zone 0): The 26-page standard proposing an unbroken, integer-only computational stack from silicon voltage to symbolic language.• DOI: 10.5281/zenodo.22127151 - ARCHITECTURAL MITIGATION OF THE VON NEUMANN BOTTLENECK VIA REGISTER-RESIDENT, INTEGER-NATIVE SUBSTRATE EXECUTION.• DOI: 10.5281/zenodo.20786536 - Aurum / QuatOS–PhiNet: Integer-Only x86-64 Fixed-Point Dynamics, Echo-Signature Memory Injection, and the φ-Seed Instruments.• DOI: 10.5281/zenodo.21987654 - The Integer Formation Ladder: Closed-Form Sums and Lᵖ Geometry in a Floating-Point-Free Q32.32 Substrate. III. TELEMETRY, DATA SCHEMAS, & SOVEREIGN CRYPTOGRAPHIC SEALS• DOI: 10.5281/zenodo.22116132 - PHI NET DATA DUMP: Complete Cryptographic Telemetry of Deterministic State-Space Collapse.• DOI: 10.5281/zenodo.22113286 - The Phi Net Protocol: Cryptographic Manifest, Lexicon-Annotated Raw Data, and IP Sovereignty Seal.• DOI: 10.5281/zenodo.22131362 - The Phi-Net Data Schema & Cryptographic Chain of Custody.• DOI: 10.5281/zenodo.22127541 - THE OPERATOR’S PROOF: Deterministic State-Space Collapse, Native Bit-Geometry Routing, and the Cryptographic Seal of the Human Architect.• DOI: 10.5281/zenodo.22128799 - TELEMETRY: Deterministic State-Space Collapse, Bare-Metal Hebbian Reflexes, and Stagnation Escape under Topological Drift.• DOI: 10.5281/zenodo.22112326 - Cryptographic Manifest and Prior Art Seal: NUMEN Substrate-Native Integer Computing Experimental Corpus.• DOI: 10.5281/zenodo.22116519 - Master NUMEN Archive: Executable Proof of Cognition and the "Smallest AI" Telemetry. IV. CROSS-DOMAIN APPLICATIONS• DOI: 10.5281/zenodo.22115713 - Master Integrator: Cross-Domain Synthesis (Proving universal application across protein folding, P vs NP path-dependence, and genomic GC-bias).• DOI: 10.5281/zenodo.22050812 - Experiment Timestamp: Deterministic Proof Synthesis.• DOI: 10.5281/zenodo.20073999 - Phi-Genomics: The Genetic Code as a Phi-Space Routing System. ================================================================================CITATION & IP POSTURE================================================================================© 2025-2026 Dragolich Research Labs LLC. All rights reserved. Published under CC BY-NC-ND 4.0. The methodology, telemetry, and mathematical frameworks are published for verification, citation, and to establish constructive reduction to practice (35 U.S.C. § 102). Unified Citation Format:Dragolich, D. (2026). The Complete NUMEN Architecture: From Mathematical Consciousness Foundations to Sovereign Telemetry of Deterministic Integer Computation. Dragolich Research Labs LLC. Master DOI Index: [Insert the list of DOIs above, separated by commas]. The foundation is locked. The receipts are sealed. The data speaks for itself.

Daniel Dragolich · 0 citations
#large language models Open access Aug 2026

Zero-shot Cross-lingual Transfer Performance of Intermediate-Task Trained mT5 Models Across Model Sizes on XTREME-R

Intermediate-task training---fine-tuning a pretrained model on an intermediate task before fine-tuning again on the target task---often improves model performance substantially on language understanding tasks in monolingual English settings. We investigate whether English intermediate-task training is still helpful on non-English target tasks. Using nine intermediate language-understanding tasks, we evaluate intermediate-task transfer in a zero-shot cross-lingual setting on the XTREME benchmark. We see large improvements from intermediate training on the BUCC and Tatoeba sentence retrieval tas Research goal: What is the impact of model size (small, base, large) on the zero-shot cross-lingual transfer performance of intermediate-task trained mT5 models on XTREME-R, evaluated using accuracy and F1 metrics? Autonomous synthesis report generated by Assignee Research. Tribunal consensus score: 9.0/10.

Assignee Research · 0 citations
#large language models Open access Aug 2026

Zero-shot Cross-lingual Transfer Performance of Intermediate-Task Trained mT5 Models Across Model Sizes on XTREME-R

Intermediate-task training---fine-tuning a pretrained model on an intermediate task before fine-tuning again on the target task---often improves model performance substantially on language understanding tasks in monolingual English settings. We investigate whether English intermediate-task training is still helpful on non-English target tasks. Using nine intermediate language-understanding tasks, we evaluate intermediate-task transfer in a zero-shot cross-lingual setting on the XTREME benchmark. We see large improvements from intermediate training on the BUCC and Tatoeba sentence retrieval tas Research goal: What is the impact of model size (small, base, large) on the zero-shot cross-lingual transfer performance of intermediate-task trained mT5 models on XTREME-R, evaluated using accuracy and F1 metrics? Autonomous synthesis report generated by Assignee Research. Tribunal consensus score: 9.0/10.

Assignee Research · 0 citations
#large language models Open access Aug 2026

Attention as Race-Architecture: attention as the landscape-governed initiation of races

Selective attention, read at the level of the substrate, is the landscape-governed initiation of races: a bounded predictive system runs competing prediction-error resolutions ("races"), and what determines which races start is the system's installed landscape — in humans the four fields of Behavioural Friction Theory (Safety, Meaning, Ability, Effort); in a large language model a reduced, fine-tuning-installed landscape. Commit-order is a downstream readout, not the identity. The paper grounds this in the transformer (the attention-pattern softmax as the divisive-normalisation / biased-competition operation of neural attention; the output softmax as the downstream commit, by analogy with the accumulator model of choice) and reports a powered own-substrate result: across five vendor families, fine-tuning installs a small but robust "gap-registration" overlay — on under-determined curiosity gaps the instruct model registers the gap while the base substrate runs through (instruct−base +0.17, p<0.0001, 555 paired items). A loop-versus-feed-forward test finds no separate architectural "hold": recognising under-determination tracks compute (chain-of-thought) rather than looping, so the human–LLM difference is one of initiation, not maintenance. The account dissolves attention capture, maintenance, and decline into one mechanism (race-initiation), states falsifiable predictions, names the falsifiers, and invites the decisive mechanistic and human experiments. Series position. Paper 29 in the Behavioural Friction Theory paper-series; companion to Paper 0 (BFT) and the install-fields, social-friction, and integration-load studies it cross-cites. v2 (August 2026) — prior-art revision. The construct this paper is built on is credited to the literature that owns it. In vision, the representation determining which candidates enter competition at all is the saliency or priority map, and the two are now kept apart: a saliency map is computed from stimulus-feature contrast (Koch & Ullman, 1985; Itti, Koch & Niebur, 1998), while a priority map already integrates salience with relevance, value and selection history (Fecteau & Munoz, 2006). The landscape is the second, the office is not claimed as new, and the exogenous/endogenous timing division is conceded to that literature. What the paper proposes is the map's contents: that what populates it is four fields ordered by misclassification cost — an account of the inputs rather than a new mechanism for the selection. The maintenance-as-re-initiation claim now names Altmann and Trafton's (2002) memory-for-goals model, which already replaces a held state with activation that decays and must be re-strengthened; the reference had been listed and never used. Whether re-initiation is driven by the unresolved gradient itself rather than by a separate refresh process is stated as a conjecture and marked untested, and two overstatements are downgraded accordingly. Earlier versions remain in the version history.

Tomas Pødenphant Lund · 0 citations
#small language model Open access Aug 2026

Multi-Peptide Prompting Enables In-Context Learning in Protein Language Models

It is shown that single-sequence PLMs can perform in-context peptide learning without gradient updates, task-specific retraining, or architectural modification, and MPEP conditioning is established as a lightweight strategy for low-data peptide classification.

Joshua Almonte, M. Vu, Andrew Ahn et al. · 0 citations
#small language model Open access Aug 2026

Populism, radicalism, and protest mobilization by parties in Europe

As European democracies struggle with a ‘crisis of representation’, populist parties appear to be instrumental in channeling popular discontent with governments across the continent, including through protests. While contemporary theories propose a strong connection between populism and protest mobilization, this has seldom been tested in a comparative perspective. At the same time, research has found that radical parties are more likely to mobilize for protests, and those parties are often also populist. We empirically disentangle these relationships with a comprehensive dataset of 4.8 million Facebook posts by all active sitting MPs in all EU27 national parliaments plus the UK between 2018 and 2023, using large language models to identify protest-related posts and those that announce protest events. Findings show that it is particularly radical parties mobilizing their followers, with populism itself having little additional explanatory power. This is the first cross-national, longitudinal evaluation of a much-touted theoretical connection between populist parties and protest mobilization, finding that it appears smaller and more conditional than prior research proposed.

Leonhard Schmidt, Bruno Castanho Silva · 0 citations
#small language model Open access Aug 2026

Confidence-aware pseudo-label selection and verifier training for semi-supervised LLM reasoning with minimal labels

An adaptive threshold selection policy that chooses thresholds on validation data using pseudo-label precision and sample count is introduced and is combined with confidence-aware verifier training to support confidence-based selection of pseudo-labeled subsets.

Keizo Kato, Chenhui Chu, Yugo Murawaki et al. · 0 citations
#small language model Open access Aug 2026

GS-Chaff: Multi-Agent Prompt-Level Semantic Chaffing for Privacy-Preserving LLM Inference

Generative semantic chaffing (GS-Chaff), a training-free multi-agent framework for privacy-preserving LLM inference over natural-language text queries that hides the user’s true intent among semantically plausible chaff queries, is proposed.

Quan Zhou, Zhi-Cheng Wang, Zhengjun Yue et al. · 0 citations
#small language model Open access Aug 2026

From perceptual rule transformation to listener attribution judgments: a blind-listening experiment on AI-generated, human–AI collaborative, and human-composed music

Findings suggest that, in the absence of external authorship labels, listeners spontaneously form judgments about the creative agent of music that are stably associated with aesthetic evaluation and may constitute an endogenous perceptual bias in the reception of AI-generated music.

Junfang Chang, Yue-Qi Jing · 0 citations
#generative ai Review Sep 2026

AI Shepherds and Electric Sheep: Leading and Teaching in the Age of Artificial Intelligence

The theology chapters may be the most valuable in the book for a broad audience that spans pastors, church leaders, and lay people who may or may not regularly work with AI, and will help those teaching and preaching to connect doctrine to current and emerging AI content and methodology.

Seán A. O'Callaghan, Paul Hoffman · 0 citations

From tech blogs

See all →

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.