Daily Digest — Oct 5
Enrich-on-Graph: Query-Graph Alignment for Complex Reasoning with LLM Enriching
• Enrich-on-Graph (EoG) framework achieves state-of-the-art performance on knowledge graph question answering benchmarks by leveraging large language models to enrich knowledge graphs and bridge the semantic gap with unstructured queries.
arXiv NLP · Knowledge Graphs
Continual Graph Memory for Mathematical Research Agents
• Ansatz successfully closes all ten tasks in the First Proof Second Batch benchmark by utilizing a Continual Graph Memory system to organize long-horizon mathematical proof searches. • The system employs a unified graph architecture alongside an evidence-sensitive curator to manage intermediate facts, plans, and counterexamples across parallel agent explorations.
arXiv AI · Knowledge Graphs
• The Hop-Decayed Influence attack achieves an 88 to 94 percent success rate across HotpotQA and 2WikiMultiHopQA benchmarks while corrupting only 0.016 percent of auxiliary structures in Microsoft GraphRAG and HippoRAG2 architectures.
arXiv AI · Knowledge Graphs
• Asterism extracts observations from hundreds of papers as concept-relation triples and unifies those concepts within a hierarchical ontology. • A field deployment with ten researchers demonstrates that users curate an evidence graph and aggregate observations at various levels of granularity to form theories aligned with their preferences.
arXiv NLP · Knowledge Graphs
Improving Scientific Document Retrieval with Academic Concept Index
• An academic concept index, organized by taxonomy, is introduced to address limitations in general-domain retriever adaptation for scientific documents. • The academic concept index enhances query generation via CCQGen for broader concept coverage and context augmentation with CCExpand for concept-focused snippets.
arXiv AI · Knowledge Graphs
ComplianceNLP: Knowledge-Graph-Augmented RAG for Multi-Framework Regulatory Gap Detection
• ComplianceNLP, a new system integrating a knowledge-graph-augmented RAG pipeline and multi-task obligation extraction, achieves 87.7 F1 on regulatory gap detection, outperforming GPT-4o+RAG by 3.5 F1. • The system processes 9,847 regulatory updates over four months, demonstrating 96.0% estimated recall and 90.7% precision while increasing analyst efficiency by 3.1x.
arXiv NLP · Knowledge Graphs
DeFAb: A Verifiable Benchmark for Defeasible Abduction in Foundation Models
• The DeFAb benchmark, converting knowledge bases into formally grounded instances for defeasible abduction, generates over 372,648 instances from 18 sources, with rule-based solvers achieving 100% accuracy while frontier language models reach a maximum of 65%.
arXiv AI · Knowledge Graphs
• Relational semantics emerge in autoregressive LLMs with sufficient logic-bearing supervision, even in shallow 2-3 layer models, as demonstrated by a controlled Knowledge Graph-based synthetic framework. • Successful generalization to unseen entities in relational tasks aligns with stable intermediate-layer signals within the LLMs.
arXiv NLP · Knowledge Graphs
8 stories