VLDB Journal

Papers
(The median citation count of VLDB Journal is 3. The table below lists those papers that are above that threshold based on CrossRef citation counts [max. 250 papers]. The publications cover those that have been published in the past four years, i.e., from 2022-08-01 to 2026-08-01.)
ArticleCitations
Threshold queries in theory and in the wild487
An efficient and scalable graph database with built-in temporal support148
Beyond influence: voting theory for opinion maximization140
Generating highly customizable python code for data processing with large language models103
Efficiently Counting Four-Node Motifs in Large-Scale Temporal Graphs39
Optimizing navigational graph queries37
Hypergraph decomposition with intersection bounds34
Transactional panorama: a conceptual framework for user perception in analytical visual interfaces (extended version)31
Missing Value Imputation in Tabular Data Lakes Unleashed: A Hybrid Approach28
Efficient and robust active learning methods for interactive database exploration24
Third and Boyce–Codd normal form for property graphs22
FOSS: A learned doctor for query optimization21
Efficient discovery of arbitrary cycles in large-scale networks20
On Efficient Top-k Empirical Variance Computation: A Once-For-All Progressive Sampling Approach20
BioGITOM: Matching Biomedical Ontologies with Graph Isomorphism Transformer19
Hu-Fu: efficient and secure spatial queries over data federation19
Model reusability in Reinforcement Learning18
On efficient 3D object retrieval17
PTSSP: privacy-preserving top-k spatial keyword similarity query with priority matching17
In-database query optimization on SQL with ML predicates16
Efficient graph embedding at scale: optimizing CPU-GPU-SSD integration16
Hyper-distance oracles in hypergraphs16
A new window Clause for SQL++14
Can large language models be a cardinality estimator? An empirical study13
Learned sketch for subgraph counting: a holistic approach13
Discovering critical vertices for reinforcement of large-scale bipartite networks12
An update-intensive LSM-based R-tree index12
LEON+: towards robust ML-aided query optimization12
Efficient top-k spatial-range-constrained approximate nearest neighbor search on geo-tagged high-dimensional vectors12
GPU-based butterfly counting12
Special issue on the best papers of DaMoN 202012
SQUID: subtrajectory query in trillion-scale GPS database11
DBSP: automatic incremental view maintenance for rich query languages10
P$$^2$$CG: a privacy preserving collaborative graph neural network training framework10
On Querying Historical Connectivity in Large-scale Temporal Graphs10
ByShard: sharding in a Byzantine environment10
The Status-Quo in nested data processing for high-energy physics10
DB-BERT: making database tuning tools “read” the manual10
Multi-constraint shortest path using forest hop labeling10
LIST: learning to index spatio-textual data for embedding based spatial keyword queries9
State Migration in Styx: Towards Serverless Transactional Functions9
DIST: Efficient k-Clique Listing via Induced Subgraph Trie9
Efficient and scalable huge embedding model training via distributed cache management9
Efficient detection of multivariate correlations with different correlation measures8
Efficient and effective algorithms for densest subgraph discovery and maintenance8
A graph pattern mining framework for large graphs on GPU8
Generating adversarial SQL queries for evaluating cardinality estimators8
DumpyOS: A data-adaptive multi-ary index for scalable data series similarity search8
Incremental discovery of denial constraints8
Anytime bottom-up rule learning for large-scale knowledge graph completion7
Tiered-Indexing: Optimizing Access Methods for Skew7
Privacy-Utility Balanced Cooperative Online Matching in Spatial Crowdsourcing7
Efficient discovery of co-movement patterns from video data7
Towards flexibility and robustness of LSM trees7
Eris: efficiently measuring discord in multidimensional sources7
Assisted design of data science pipelines6
AutoML in heavily constrained applications6
A survey on deep learning approaches for text-to-SQL6
Accelerating directed densest subgraph queries with software and hardware approaches6
Survey of window types for aggregation in stream processing systems5
Special issue: modern hardware5
Scalable decoupling graph neural network with feature-oriented optimization5
Morphtree: a polymorphic main-memory learned index for dynamic workloads5
BatchHL$$^{+}$$: batch dynamic labelling for distance queries on large-scale networks5
Join optimization revisited: a novel DP algorithm for join&sort order selection5
A survey on the evolution of stream processing systems5
Editorial for Special Issue: VLDB 20225
A multi-facet analysis of BERT-based entity matching models5
HINT: a hierarchical interval index for Allen relationships5
HPCache: memory-efficient OLAP through proportional caching revisited5
xDBTagger: explainable natural language interface to databases using keyword mappings and schema graph5
Performant almost-latch-free data structures using epoch protection in more depth5
Efficient Algorithms for Uncertain Restricted Skyline Query Processing5
Lamba: A pretrained model for latency prediction over distributed databases5
How good are machine learning clouds? Benchmarking two snapshots over 5 years5
Tee-based key-value stores: a survey5
Efficient Task Assignment for Multi-Workerset Crowdsourcing with Time and Expense Considerations5
FlexpushdownDB: rethinking computation pushdown for cloud OLAP DBMSs4
Identifying similar-bicliques in bipartite graphs4
PINE: Extracting Correlated Token Pairs for Explainable Entity Matching4
Time-topology analysis on temporal graphs4
Data collection and quality challenges in deep learning: a data-centric AI perspective4
An Evaluation of B-tree Compression Techniques4
Scalable lighting-fast temporal indexing4
BQSched$$^{+}$$: A generalizable RL-based scheduler for varying batch concurrent queries4
MSAD: A deep dive into model selection for time series anomaly detection4
Netherite: efficient execution of serverless workflows4
VUS: effective and efficient accuracy measures for time-series anomaly detection4
Editorial: Special Issue for Selected Papers of VLDB 20214
A near-optimal approach to edge connectivity-based hierarchical graph decomposition4
A generic framework for efficient computation of top-k diverse results4
$$\textsf{IACS}^{+}$$: Inductive Attributed Community Search via Learning across Graphs4
Cardinality estimation using normalizing flow3
Ingress: an automated incremental graph processing system3
Density decomposition on large static and dynamic graphs: algorithms and applications3
On topology and time: efficient evaluation for temporal-clique subgraph queries3
Accelerating maximum biplex search over large bipartite graphs3
Temporal graph patterns by timed automata3
SWOOP: top-k similarity joins over set streams3
Similarity-driven and task-driven models for diversity of opinion in crowdsourcing markets3
Towards GPU memory-aware efficient contrastive shapelet learning for unsupervised representation learning in multivariate time series3
AutoCTS++: zero-shot joint neural architecture and hyperparameter search for correlated time series forecasting3
Finding Locally Densest Subgraphs: Convex Programming with Edge and Triangle Density3
MinJoin++: a fast algorithm for string similarity joins under edit distance3
A powerful reducing framework for accelerating set intersections over graphs3
Correction: Verifiable Authenticated Data Structure (V-ADS) for Analytic Queries3
Reliability evaluation of individual predictions: a data-centric approach3
Measuring approximate functional dependencies: a comparative study3
A systematic evaluation of machine learning on serverless infrastructure3
Leveraging user itinerary to improve personalized deep matching at Fliggy3
Static and streaming algorithms for random sampling over spatial range joins3
Table integration in data lakes unleashed: pairwise integrability judgment, integrable set discovery, and multi-tuple conflict resolution3
Flexible grouping of linear segments for highly accurate lossy compression of time series data3
Efficient indexing and searching of constrained core in hypergraphs3
HMI: hierarchical knowledge management for efficient multi-tenant inference in pretrained language models3
Butterfly counting and bitruss decomposition on uncertain bipartite graphs3
Efficient algorithms for reachability and path queries on temporal bipartite graphs3
C5: cloned concurrency control that always keeps up3
Hypergraph motifs and their extensions beyond binary3
0.078035831451416