←

Daily arXiv Papers

Graph Learning · LLM × Graph · Multi-Agent · Science

Showing 6 papers for 2026-09-22

★ Must read

full paper read, not just the abstract

Nothing cleared the bar today. Papers read in full today: 2.

🤗 Hugging Face daily top 5

most upvoted on 2026-09-21

EvoOntology proposes a self-evolving ontology layer to bridge the agent-data gap in data agents that execute natural-language instructions over heterogeneous data sources such as tables, files, and databases. By providing a scalable, adaptive semantic layer that operates beyond static prompts, EvoOntology aims to scale to large data sources and adapt to different agent behaviors.

RUC-DataLab· Hugging Face ·GitHub ★278

CodeMidas presents a pipeline that turns implemented functionality in existing codebases into executable reinforcement learning environments, using source code as the sole task-specific input. This approach aims to scale agentic coding RL by avoiding reliance on development artifacts and by allocating compute to every stage of environment construction.

Xiaomi MiMo· Hugging Face

Code2Skill introduces a fully automated pipeline that transforms selected code units into implementation-anchored, reusable skills for agentic intelligence. By leveraging executable code as evidence, it grounds abstractions and overcomes limitations of trajectory-based synthesis and document-derived skills.

ant-international· Hugging Face ·GitHub ★31

RecreationWorld introduces a five-platform framework for hybrid computer-use agents that autonomously decide when to explore interfaces, implement software, and run and visually verify artifacts. Given a running reference, the agent must discover the behavior and build a faithful implementation with no prescribed workflow, enabling scalable and verifiable hybrid environments.

Qwen· Hugging Face ·GitHub ★40

IntBMoE proposes Integrating Block-Level Conditioning into Expert Composition for Full-Participation Mixture-of-Experts, enabling independent control of participation, execution, and materialization costs. By introducing block-level conditioning, it aims to achieve full token participation without prohibitive compute or memory overhead, addressing the trade-offs of sparse routing and dense output mixing.

Hugging Face

All papers

6
Temporal Generalization and Explanation Stability of Control Flow Graph Neural Networks for Malware Detection
GNN Graph × Science

This paper studies the temporal generalization and explanation stability of Graph Neural Networks over Control Flow Graphs for malware detection. It uses strict temporal splits (train on one period, test on a later one) and analyzes both detection performance and the stability of explanations across two CFG corpora, where each node carries 37 features.

Locally Fair PageRank: Mean-Field Approximation and One-Step Refinement
Graph Theory

We introduce a scalable analytical framework to approximate Neighborhood Locally Fair PageRank and Uniform Locally Fair PageRank, addressing the scalability bottlenecks of exact convergence. The approach uses a group-aware heterogeneous mean-field representation to enable fast, accurate locally fair rankings on large graphs.

Exploiting Residual Reachability for Cross-Model Migration of Graph-Based Indexes in Approximate Nearest Neighbor Search
Graph Learning

This paper studies cross-model migration of graph-based indexes for approximate nearest neighbor search by exploiting residual reachability. It investigates reusing parts of an existing graph index when the embedding model changes to avoid full rebuilds.

Consistent Relexicalization of Clinical Documents using Graph-Based Approach
Graph Learning

This work addresses Consistent Relexicalization of Clinical Documents using a Graph-Based Approach. It aims to mask sensitive information while preserving longitudinal structure, relational coherence, and temporal consistency across records; existing entity-by-entity replacements often introduce cross-time inconsistencies.

Efficient Dense Vector Search within Knowledge Graph Content Embeddings
Knowledge Graph

We propose Efficient Dense Vector Search within Knowledge Graph Content Embeddings, enabling SPARQL engines to natively operate on dense embeddings integrated with KG content for neurosymbolic reasoning. This enables tensor operations in SPARQL and tighter coupling of language models with knowledge graphs.

Graph Memory for LLM Agents: At What Cost? A Comparative Evaluation of Query, Ingest, and Update Performance Across Graph Database Engines
Graph Learning

The study provides a comparative evaluation of graph databases using a large synthetic biomedical property graph (1.02 million nodes, 5.34 million total rows) and a twenty-query workload. It analyzes query, ingestion, and update performance across engines, highlighting differences in query planning, indexing, and data-readiness costs.