Skip to content

Chipmink: E ! icient Delta Identification for Massive Object Graphs

· 1 citation · 95 references

TL;DR

A graph-based object store, named Chipmink, that acts like the centralized manager of DBMSs, dynamically induced by partitioning objects into appropriate subgroups (called pods), minimizing expected persistence costs based on object sizes and reference structure is proposed.

View source

Similar papers

#artificial intelligence Preprint Sep 2026

Git4Data: Database-Native Version Control for AI Agents

Git4Data is presented, a database-native version-control layer for agentic workflows that sheds light on how relational databases can better support AI agents through efficient versioning.

Hongshen Gou, Zu-Yu Zhang, Yu-Ze Sun et al. · 0 citations
Jul 2026

The Data World is Not Flat: Efficient Factorized Execution for Relational Systems

A novel code-generating engine with factorization that enables intra-query-parallelized query execution on factorized representations and generates code to overcome their CPU-unfriendly layout, offering a unified and scalable solution for modern workloads.

Stefan Lehner, Thomas Neumann · 0 citations
Preprint Aug 2026

MEMONDEMAND: A Memory Management System for Large-Scale Enterprise Data

On EnterpriseRAG-Bench, MEMONDEMAND outperforms the strongest published LB#1 result at every evaluated scale from 10M tokens through the complete 618M- token collection, and results on FinanceBench, HotpotQA, and FRAMES further show strong performance across financial, multi-hop, and fact-retrieval settings.

Xin-Yuan Song, Bo-Wen Zhu, H. Haque et al. · 0 citations
Book Open access Sep 2026

Anchor: Mitigating GPU Shallow Disruptions with Decoupled Memory

Job failures are frequent in large-scale GPU clusters for LLM workloads, leading to significant resource wastage. The vast majority of these are shallow disruptions (e.g., software errors or updates), where only the worker process crashes while the underlying GPU and OS kernel remain intact. Existing recovery systems,...

Hao-Yi Ma, Shi-Wei Gao, You-Min Chen et al. · 0 citations
#machine learning Preprint Sep 2026

RAILS: Retrieval-Augmented Incremental LLM Clustering at Scale

Using a Large Language Model (LLM) as the clusterer at production scale is hard: prompts cannot hold the entire label space, and per-document serial processing does not deliver the throughput real workloads require. We present RAILS, a retrieval-augmented incremental LLM clusterer that turns clustering into a simple lo...

Armin Oliya, Aleksandra Sawczuk, Radosław Białobrzeski · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.