AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
24 Jul 2026

Error Certificates for KV-Cache Eviction via Randomized Design

ResearchDGX agent

arXiv:2607.21475v1 Announce Type: cross Abstract: Deterministic KV-cache eviction keeps the top-k tokens under an importance score and deletes the rest. We prove that this design cannot know what it d

Euclid-MCP: A Model Context Protocol Server for Deterministic Logical Reasoning via Prolog

SafetyDGX agent

arXiv:2607.21412v1 Announce Type: new Abstract: Large Language Models (LLMs) excel at natural language understanding and generation but remain unreliable for multi-step logical reasoning, especially i

Evaluating and Guarding Citation Faithfulness in Agentic Scientific Synthesis

HardwareDGX agent

arXiv:2607.20527v1 Announce Type: new Abstract: Agentic LLM systems such as OpenScholar and PaperQA2 read the scientific literature and return cited answers, and both they and their benchmarks already


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Evaluating Risks in Weak-to-Strong Alignment: A Bias-Variance Perspective

SafetyDGX agent

arXiv:2604.25077v2 Announce Type: replace Abstract: Weak-to-strong alignment offers a promising route to scalable supervision, but it can fail when a strong model becomes confidently wrong on examples

Evolutionarily Stable Stackelberg Equilibrium

ResearchDGX agent

arXiv:2603.18385v4 Announce Type: replace-cross Abstract: We present a new solution concept called evolutionarily stable Stackelberg equilibrium (SESS). We study the Stackelberg evolutionary game sett

EvoSQL: Memory-Augmented Critic-Generator Co-Evolution for Text-to-SQL

SafetyDGX agent

arXiv:2607.20489v1 Announce Type: new Abstract: Text-to-SQL has advanced rapidly with large language models, but complex database queries still require reasoning beyond one-shot generation, including

ExecuGraph: A Multi-Agent, Execution-Grounded Framework for Reliable Backend Code Synthesis with Large Language Models

Local AiDGX agent

arXiv:2607.20499v1 Announce Type: new Abstract: Large Language Models generate plausible backend code, but a single-pass paradigm provides no guarantee of correctness or runtime reliability. We presen

Expectation Alignment of Language Models for Real-World User Expectations

Model ReleasesDGX agent

arXiv:2607.20485v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated remarkable performance on standard benchmarks, yet it remains largely unexplored whether they truly meet

Expert Behavior Prior Reinforcement Learning

SafetyDGX agent

arXiv:2607.21302v1 Announce Type: new Abstract: Behavior prior reinforcement learning (BPRL) has emerged as a promising paradigm to improve sample efficiency in online reinforcement learning (RL) by l

Explainability Framework for Policy-Aware Autonomous Agents

SafetyDGX agent

arXiv:2607.21209v1 Announce Type: cross Abstract: In the field of Artificial Intelligence, an agent is a system which is able to autonomously make decisions in order to reach a desired goal. As these

Explainable Belief Harmonization under Dynamic Epistemic Partitions

AgentsDGX agent

arXiv:2607.21210v1 Announce Type: cross Abstract: Existing approaches to multi-agent belief combination have established mature foundations for combining uncertain beliefs under common assumptions: co

Explaining Weather Bulletins via ILP

ApplicationsDGX agent

arXiv:2607.21184v1 Announce Type: new Abstract: Inductive Logic Programming (ILP) originated within the Logic Programming community in the Nineties as a framework for combining symbolic learning with

EZASP - Facilitating the Usage of ASP

TutorialsDGX agent

arXiv:2603.26863v2 Announce Type: replace-cross Abstract: Answer Set Programming (ASP) is a declarative programming language used for modeling and solving complex combinatorial problems. It has been s

Faster IndexTTS-2: Accelerating and Streaming Autoregressive Zero-Shot Text-to-Speech Synthesis on GPUs

Model ReleasesDGX agent

arXiv:2607.21042v1 Announce Type: new Abstract: Autoregressive text-to-speech models achieve strong naturalness but suffer from slow inference due to sequential token generation, limiting their deploy

FlowEdit: Information-Theoretic Control of LLM Reasoning Flows for Ill-posed Problems Involving Conflicts

ResearchDGX agent

arXiv:2607.20500v1 Announce Type: new Abstract: Large Language Models (LLMs) perform strongly on well-specified reasoning tasks with a feasible answer. However, problems encountered in the open world

Foundation-model-guided radiogenomic discovery linking cancer genomes to cancer scans

ResearchDGX agent

arXiv:2607.20583v1 Announce Type: cross Abstract: The function of many genes is still unknown, and conventional driver-discovery methods, which rely on how frequently a gene is mutated, cannot assess

Free energy landscape of Dense Associative Memory

ResearchDGX agent

arXiv:2607.19195v2 Announce Type: replace-cross Abstract: Using large deviations theory, we solve and obtain a general expression for the free energy functional for a broad class of associative memori

From Agent Failures to Text Policies: What Works and What Breaks

SafetyDGX agent

arXiv:2607.20668v1 Announce Type: cross Abstract: TextGrad improves language-model systems by revising text from feedback. Its core thesis is that natural-language feedback can act as a gradient for o

From Atoms to Entropy: Optimal Noise Allocation for Diffusion Training in the Convex Regime

ResearchDGX agent

arXiv:2607.20540v1 Announce Type: cross Abstract: How should a diffusion model decide which noise levels to train on, and how much? Despite the importance of this choice, current noise schedules are b

From Attention to Frequency: Integration of Vision Transformer and FFT-ReLU for Enhanced Image Deblurring

Model ReleasesDGX agent

arXiv:2511.10806v1 Announce Type: cross Abstract: Image deblurring is vital in computer vision, aiming to recover sharp images from blurry ones caused by motion or camera shake. While deep learning ap

From Chatbot to Digital Colleague: The Paradigm Shift Toward Persistent Autonomous AI

AgentsDGX agent

arXiv:2606.14502v2 Announce Type: replace Abstract: Large Language Models (LLMs) are undergoing a fundamental transformation from conversational generators into integrated AI systems capable of reason

From Checklists to Clusters: A Homeostatic Account of AGI Evaluation

ResearchDGX agent

arXiv:2510.15236v2 Announce Type: replace Abstract: Contemporary AGI evaluations report multidomain capability profiles, yet they typically assign symmetric weights and rely on snapshot scores. This c

From Dependency to Compositionality: A Neurosymbolic Lifting of LLM Outputs via Combinatory Categorial Grammar

Local AiDGX agent

arXiv:2607.18961v2 Announce Type: replace Abstract: Large language models (LLMs) generate fluent text by incrementally predicting the next token from a prefix. Critics in the generative tradition argu

From Errors to Rules: Iterative Prompt Optimization for Text Classification

ResearchDGX agent

arXiv:2607.20497v1 Announce Type: new Abstract: Prompt optimization for text classification spans diverse approaches, from demonstration selection to exploration-based search to error-driven diagnosis

From Noise to Diversity: Random Embedding Injection in LLM Reasoning

ResearchDGX agent

arXiv:2605.11936v2 Announce Type: replace Abstract: Recent soft prompt research has tried to improve reasoning by inserting trained vectors into LLM inputs, yet whether the gain comes from the learned

From Resource Flow to Executable Tests: Petri-Net-Guided LLM Test Generation for Concurrent Stateful Rust APIs

ApplicationsDGX agent

arXiv:2607.21530v1 Announce Type: cross Abstract: Concurrent stateful library APIs expose behavior through evolving resource ownership, lifecycle states, and competing interleavings. Large language mo

From Scalars to Time Series: Rethinking Implicit Neural Representations for Time-Varying Volumetric Data

ResearchDGX agent

arXiv:2607.20970v1 Announce Type: new Abstract: Implicit neural representations (INRs) for time-varying volumetric data are typically trained using dense sampling over spatiotemporal coordinates, wher

From Static Bibliometrics to Dynamic Knowledge Graphs: An LLM-Powered Framework for Modernizing Science, Technology, and Innovation (STI) Analytics

SafetyDGX agent

arXiv:2607.21327v1 Announce Type: cross Abstract: Bibliometric indicators - citation counts, h-indexes, co-authorship networks - have long anchored science, technology, and innovation (STI) analytics,

Frontier Financial Judgement: Can agents tell what might move a stock?

Model ReleasesDGX agent

arXiv:2607.20645v1 Announce Type: cross Abstract: We introduce Frontier Financial Judgement, a challenging new benchmark developed in collaboration with professional equity analysts to assess agents'

Generative AI and Agency in Education: A Critical Scoping Review and Thematic Analysis

SafetyDGX agent

arXiv:2411.00631v2 Announce Type: replace-cross Abstract: This scoping review examines the relationship between Generative AI (GenAI) and agency in education, analyzing the literature available throug

Generative Artificial Intelligence in Bioinformatics: A Systematic Review of Models, Applications, and Methodological Advances

SafetyDGX agent

arXiv:2511.03354v2 Announce Type: replace-cross Abstract: Generative artificial intelligence (GenAI) is transforming bioinformatics by advancing genomics, proteomics, transcriptomics, structural biolo

Geometric Configurations of Perturbed Jailbreak Prompts

Model ReleasesDGX agent

arXiv:2607.20581v1 Announce Type: cross Abstract: Perturbation techniques that turn unsuccessful jailbreak prompts into successful ones are continuously evolving, constituting a major security threat

GigaPath-Flash and GigaTIME-Flash: Efficient Pathology Foundation Models for Whole-Slide and Tumor Microenvironment Analysis

Model ReleasesDGX agent

arXiv:2607.18218v2 Announce Type: replace-cross Abstract: Foundation models have emerged as a driving force in computational pathology, with the potential to transform cancer diagnosis, prognosis, and

GlucoTune: A Unified Framework for Blood Glucose Preprocessing, Forecasting, and Benchmarking in Diabetes

ResearchDGX agent

arXiv:2607.21117v1 Announce Type: cross Abstract: Preprocessing blood glucose time-series data is a critical yet often overlooked step in developing data-driven methods for diabetes management, partic

GPE: Evaluating Robust Evidence Aggregation for Fact Verification under Controllable GEO-Style Poisoning

Model ReleasesDGX agent

arXiv:2607.20730v1 Announce Type: cross Abstract: Large language models increasingly use search tools to retrieve up-to-date information, introducing a new attack surface in which retrieved documents

GRADRAG: Cross-Component Prompt Adaptation for Coordinated Multi-Agent RAG

AgentsDGX agent

arXiv:2607.21324v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) systems increasingly employ multiple LLM agents. Yet, most prior work optimizes components in isolation rather th

GraphVid: Interactive Graph-Controllable Video Generation

ResearchDGX agent

arXiv:2607.21580v1 Announce Type: cross Abstract: Controllable video generation remains challenging due to the difficulty of specifying precise multi-object interactions using text prompts or motion-c

Grounding Investor Views: Neural Predicates in the Black-Litterman Model

ResearchDGX agent

arXiv:2607.20533v1 Announce Type: cross Abstract: Portfolio construction under the Black-Litterman model requires investors to specify views on asset returns alongside explicit uncertainty estimates -

GS-Agent: Creating 4D Physical Worlds With Generative Simulation

AgentsDGX agent

arXiv:2607.21522v1 Announce Type: cross Abstract: Creating dynamic and physically realistic 4D worlds from natural language descriptions is both fascinating and challenging. Traditional computer graph

GuardianAgentBench: Where Agents Fail and How to Guard Them

Model ReleasesDGX agent

arXiv:2607.20982v1 Announce Type: new Abstract: As large language model agents increasingly operate autonomously with access to tools and external environments, ensuring their safe and reliable behavi

HAMON: Passive Optical Sequence Mixing for Long-Horizon Forecasting

ResearchDGX agent

arXiv:2606.17028v2 Announce Type: replace-cross Abstract: Simple linear and frequency-domain models remain surprisingly competitive in long-horizon time-series forecasting, and recent mechanistic evid

Hardware-Software Co-Design for Float16 On-Device Training on RISC-V Single-Core

Local AiDGX agent

arXiv:2607.21130v1 Announce Type: cross Abstract: By leveraging standard RISC-V extensions, namely Zfh (scalar float16) and Zvfh (vector float16), this work proposes an open-source framework to enable

HARP: The Human--AI Research Platform

AgentsDGX agent

arXiv:2607.20773v1 Announce Type: cross Abstract: Large language models (LLMs) have shifted human--computer interaction from `traditional'' interface journeys toward more conversational exchanges. Res

Hilbert Operator for Progressive Encoding (HOPE): A Mathematical Framework for Deconstructing Learned Representations in Deep Networks

ResearchDGX agent

arXiv:2607.21366v1 Announce Type: cross Abstract: Deep neural networks encode complex representations, but deconstructing this internal knowledge remains a challenge. Given the link between learning a

HiMe: Real-Time Self-Hosted Personal Agent Platform for Health Insights with Wearable Devices

AgentsDGX agent

arXiv:2607.21019v1 Announce Type: new Abstract: Traditional approaches to wearable health signal analysis, such as smartwatches, are constrained by rigid analytical frameworks and limited personalisat

How Rules Represent Causal Knowledge: Causal Modeling with Probabilistic Logic Programming

ResearchDGX agent

arXiv:2607.21208v1 Announce Type: new Abstract: Pearl famously argues that causal knowledge enables the prediction of intervention effects. By contrast, purely descriptive knowledge supports only conc

Hybrid MKNF with Classical Negation in the Rule Component

SafetyDGX agent

arXiv:2607.21202v1 Announce Type: cross Abstract: Hybrid MKNF knowledge bases under the well-founded semantics integrate Description Logics with Logic Programming. However, they do not support classic

HypNO: A Graph-Based Neural Operator with Physics-Informed Message Passing for Hyperbolic Conservation Laws

Model ReleasesDGX agent

arXiv:2607.20541v1 Announce Type: cross Abstract: We introduce HypNO, a graph-based neural operator for scalar hyperbolic conservation laws. HypNO operates directly on a space-time graph of finite-vol

HyWorldVLA: A Vision-Language-Action Model with Hybrid World Modeling for Autonomous Driving

Model ReleasesDGX agent

arXiv:2607.20988v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models augmented with world modeling represent a promising paradigm for end-to-end autonomous driving. While pixel-level

ICAE-Bench: Evaluating Coding Agents as Interactive Project Builders

Model ReleasesDGX agent

arXiv:2607.21217v1 Announce Type: new Abstract: The recent emergence of vibe-coding workflows is changing what coding agents are expected to do. Instead of merely completing code under fully specified

Identifying Good Rules for Efficient SAT Encodings of Single-Constant Multiplication Using Machine Learning

ResearchDGX agent

arXiv:2607.21188v1 Announce Type: new Abstract: The Single Constant Multiplication problem is a fundamental NP-hard optimization task in hardware design, which seeks to decompose a fixed constant usin

ImplicitBBQ: Benchmarking Implicit Bias in Large Language Models through Characteristic Based Cues

Model ReleasesDGX agent

arXiv:2604.01925v2 Announce Type: replace-cross Abstract: Large Language Models increasingly suppress biased outputs when demographic identity is stated explicitly, yet may still exhibit implicit bias

Improved lower bounds for the Shannon capacity of odd cycles

ResearchDGX agent

arXiv:2607.21517v1 Announce Type: cross Abstract: The Shannon capacity Theta(G) of a graph G quantifies the maximum rate at which information can be transmitted with zero error over a noisy channel. I

Improving Access to Essential Medicines via Decision-Aware Machine Learning

ApplicationsDGX agent

arXiv:2607.20542v1 Announce Type: cross Abstract: A critical challenge in healthcare systems in low- and middle-income countries (LMICs) is the efficient and equitable allocation of scarce resources,

Incomplete Prompt Jailbreaks in Large Language Models

Model ReleasesDGX agent

arXiv:2607.20473v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly released as open-weight models with safeguards against harmful requests. Nevertheless, sentence completion

Inducing Comparability of Factorised Probability Distributions

ResearchDGX agent

arXiv:2607.20502v1 Announce Type: new Abstract: To allow for principled comparison between two probabilistic graphical models defined over non-identical variable sets, they have to be lifted to a comm

InferenceBench: A Benchmark for Open-Ended LLM Inference Optimization by AI Agents

Model ReleasesDGX agent

arXiv:2607.20468v1 Announce Type: new Abstract: AI agents are increasingly used to automate research and development tasks, yet existing benchmarks typically evaluate them on prescribed workflows or n

Instruct-FD: Can Your Full-Duplex Speech System Follow Turn-Taking Instructions?

Model ReleasesDGX agent

arXiv:2607.20460v1 Announce Type: cross Abstract: Current full-duplex (FD) spoken dialogue systems can produce fluid interactions, yet it remains unclear whether they can adapt their turn-taking behav

Interaction Dynamics Modeling and Predictive Control for Safe Steerable Catheter--Tissue Interaction

SafetyDGX agent

arXiv:2607.20939v1 Announce Type: cross Abstract: Safe steerable catheter control is fundamentally a problem of interaction dynamics: the tip must follow a planned motion, remain compliant against mov

Interpretable Embeddings with Sparse Autoencoders: A Data Analysis Toolkit

ResearchDGX agent

arXiv:2512.10092v2 Announce Type: replace Abstract: Analyzing large-scale text corpora is a core challenge in machine learning, crucial for tasks like identifying undesirable model behaviors or biases

← Previous
1…5556575859…354
Next →