AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
Human
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
24 Jul 2026

Agent-Guided Relational Concept Discovery: Toward Interpretable Surgical Margin Assessment

AgentsDGX agent

arXiv:2607.21437v1 Announce Type: new Abstract: Deep learning models can effectively use Rapid Evaporative Ionization Mass Spectrometry (REIMS) data for surgical margin assessment. However, their clin

Agentic coding without the cloud: evaluating open-weight large language models on longitudinal data preparation tasks

Model ReleasesDGX agent

arXiv:2607.21482v1 Announce Type: new Abstract: Large language models (LLMs) and agents are now widely used tools in code development, with data typically sent to third-party cloud-based models. Their

Agentic Context Management: Solving Agent Memory and Cost by Treating Them as Lifecycle and Architecture Problems

AgentsDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.21503v1 Announce Type: new Abstract: Production AI agents' failures are less often due to an inability to reason well and more often because they cannot manage what is in their reasoning co

Agentic Designer: Progressive Multi-Agent Collaboration for Structure-Aware Interior Layout Generation

Model ReleasesDGX agent

arXiv:2607.20866v1 Announce Type: new Abstract: Generating realistic interior furniture layouts that strictly adhere to architectural constraints (e.g., walls, doors, and windows) remains a fundamenta

Agree on the Model, Verify the Inference: GKR Protocols for HND-Based Transformer Inference

ResearchDGX agent

arXiv:2607.21162v1 Announce Type: new Abstract: Outsourced Transformer inference exposes clients to model substitution and incomplete execution, while direct replay removes the computational benefit o

AI Assistants Overassist

Model ReleasesDGX agent

arXiv:2607.21306v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as tutors and thought partners, helping users reason through problems. While guidance from AI assis

AI-Driven Multi-Hop Relay Selection for Smart Urban NR-V2X Networks via Learning-to-Optimize Graph Neural Networks

ResearchDGX agent

arXiv:2607.20554v1 Announce Type: new Abstract: Reliable and low-latency NR-V2X communications are essential for smart mobility in dense urban environments. However, limited Road-Side Unit (RSU) densi

AI-Driven Surrogate Models for Predicting Electrode-Scale Discharge Behavior in Lithium-Ion Batteries

ResearchDGX agent

arXiv:2607.20577v1 Announce Type: new Abstract: Physics-based simulations are essential for understanding the electrode-scale discharge behavior of lithium-ion batteries (LIBs) but suffer from prohibi

AINTMA: Agentic AI Architecture for Autonomous Test Management with Generative Intelligence, Secure Cloud Communication and Adaptive Quality Analytics

AgentsDGX agent

arXiv:2607.20452v1 Announce Type: new Abstract: Modern software quality assurance demands intelligent, autonomous systems capable of adaptive decision-making across distributed cloud environments. Thi

AISE-Bench: A Full-Cycle Curated Benchmark for Information Seeking on Academic Knowledge Graphs

Model ReleasesDGX agent

arXiv:2607.20498v1 Announce Type: new Abstract: Large language models (LLMs) augmented with tools are emerging as autonomous agents capable of using Web engine, APIs, and code to solve complex, long-h

Algorithmic Approaches to Sequential Decision-Making and Social Epistemology

ApplicationsDGX agent

arXiv:2607.20636v1 Announce Type: cross Abstract: As humans, we face many decisions that require us to choose between sticking to something and giving up. This thesis uses algorithmic tools to derive

An Analytically Trained Variational Surrogate for Quantum Phase Estimation on NISQ Hardware

Model ReleasesDGX agent

arXiv:2607.20943v1 Announce Type: cross Abstract: Quantum Phase Estimation (QPE) is a foundational algorithm for molecular ground-state energy estimation, but its deep circuit requirements make direct

An Evaluation Framework for Structured Audio Captions Validated by Controlled Perturbations

ResearchDGX agent

arXiv:2607.21424v1 Announce Type: new Abstract: Recent advancements in automated audio captioning (AAC) have shifted from monolithic sentence generation toward structured formats that explicitly disen

An LLM-Driven Workflow for Automated Process Control Strategy Generation and Tuning from Dynamic Process Models

Model ReleasesDGX agent

arXiv:2607.21292v1 Announce Type: new Abstract: We present a structured large-language-model-driven workflow for automated multi-variable control design from dynamic process models. The workflow decom

Animation, Verification and Visualisation of Prolog Transition Systems with ProB

ResearchDGX agent

arXiv:2607.21192v1 Announce Type: cross Abstract: ProB is a Prolog-based model checker, animator and constraint solver for high-level formal specifications. One can also use ProB to animate transition

Answer-then-Edit: Reasoning Skeleton Editing for Anti-Distillation with Preserved Utility

ResearchDGX agent

arXiv:2607.20440v1 Announce Type: cross Abstract: Proprietary large language models (LLMs) entail substantial intellectual and financial investment, making them valuable intellectual property (IP). Ho

Anti-Goal Reasoning: Rethinking the Theory of Goal Reasoning in Non-Axiomatic Logic

ResearchDGX agent

arXiv:2607.20902v1 Announce Type: cross Abstract: Goal reasoning in Non-Axiomatic Logic (NAL) explains how an adaptive system derives means for realizing desired events under insufficient knowledge an

Anti-Periodic Positional Encoding: Mobius Boundary Conditions Make In-Context Retrieval Reliable

ResearchDGX agent

arXiv:2607.21405v1 Announce Type: new Abstract: Mobius RoPE is a rotary positional encoding built on the anti-periodic frequency ladder heta_i=pi(2i+1)/N: every rotation plane advances by an odd multi

Anticipate Before Acting: Future-State-Conditioned Vision-Language Navigation

SafetyDGX agent

arXiv:2607.18042v2 Announce Type: replace-cross Abstract: End-to-end vision-language navigation (VLN) with causal vision-language models maps instructions and egocentric observations directly to actio

Approximate Quantum State Preparation Through Proximal Policy Optimization

SafetyDGX agent

arXiv:2607.21121v1 Announce Type: cross Abstract: In this work, a quantum architecture search framework for approximate quantum state preparation (QSP) is proposed. QSP is a challenging task, since th

AppWorld-UL: Benchmarking Diverse Agent-User Interactions for Tool-Use

Model ReleasesDGX agent

arXiv:2607.20536v1 Announce Type: new Abstract: Tool-use agents that address day-to-day digital tasks such as ordering groceries must not only operate applications, but also interact with the user, e.

ArbiGraph: Arbitrarily Scalable Verifiable Task Graphs for Evaluating Context Management

Model ReleasesDGX agent

arXiv:2607.20764v1 Announce Type: new Abstract: We introduce ARBIGRAPH, a benchmark generator for evaluating whether tool-assisted language agents can retain, update, compose, and discard task-relevan

ARCO: Adaptive Rubrics with Co-Evolution for Multi-Step LLM-Based Agents

SafetyDGX agent

arXiv:2606.21262v2 Announce Type: replace Abstract: Reinforcement learning for multi-step LLM agents often relies on scalar rewards that indicate success but cannot explain why a trajectory is good or

Are Diversity Metrics Measuring Diversity? A Capability-Controlled Audit of Majority-Vote Gain in LLM Ensembles

ResearchDGX agent

arXiv:2607.20768v1 Announce Type: cross Abstract: Majority voting over LLMs is widely assumed to benefit from diversity, and diversity measures are used to choose which models to combine. We ask wheth

Are Single-Token Sparse Autoencoder Features Causally Necessary? Layer-Depth and SAE-Family Effects

Model ReleasesDGX agent

arXiv:2607.20596v1 Announce Type: cross Abstract: Sparse autoencoder (SAE) features are used to interpret and steer large language models, yet whether a feature's causal role is stable across SAE fami

AREX: Towards a Recursively Self-Improving Agent for Deep Research

Model ReleasesDGX agent

arXiv:2607.21461v1 Announce Type: new Abstract: Deep research requires agents to find answers that jointly satisfy multiple constraints. Discovering such answers is costly, whereas verifying a candida

Artificial Epanorthosis: Why large language models overuse a classical rhetorical figure, and how to mitigate it

TutorialsDGX agent

arXiv:2607.21498v1 Announce Type: cross Abstract: A rhetorical figure that Cicero and Quintilian catalogued two thousand years ago reappears, systematically, in the text of large language models: epan

ASTRA-Net: Anatomy-Specific Transfer and Representation Alignment for Drug-Induced Sleep Endoscopy Segmentation

SafetyDGX agent

arXiv:2607.21370v1 Announce Type: new Abstract: Quantitative drug-induced sleep endoscopy (DISE) requires reliable airway boundaries at specific anatomical levels. Pixel-level DISE annotations are sca

AsymVerify at SemEval-2026 Task 6: Asymmetric Confidence-Gated Verification for Political Evasion Detection

ResearchDGX agent

arXiv:2607.20439v1 Announce Type: new Abstract: Political evasion is difficult to detect because evasive answers often appear cooperative while avoiding concrete commitment. We present AsymVerify, a c

Attention-based Experience Replay Framework for Continual Learning of Agnostic Time Series Forecasting Models

ApplicationsDGX agent

arXiv:2607.20493v1 Announce Type: new Abstract: Deep learning has led to remarkable progress in artificial intelligence, particularly in robotics, imaging and sound processing. However, a major limita

Attention Degradation, Function Token Anchoring, and the Limits of Attention-Based Intervention in Large Language Models

Model ReleasesDGX agent

arXiv:2607.20524v1 Announce Type: new Abstract: Mean cross-positional attention degradation is widely reported in transformer interpretability, yet whether it causally limits contextual retrieval rema

Attribution Markets: A Fisher-Market Formulation for Fractional Credit Assignment Between Planned Tasks and Performed Actions

Model ReleasesDGX agent

arXiv:2607.20694v1 Announce Type: new Abstract: Personal and organizational planning systems maintain two records that drift apart: what was planned (a task's effort budget) and what was done (a logge

AttriMem: Attribution-Guided Process Feedback for Agent Memory Learning

SafetyDGX agent

arXiv:2607.21106v1 Announce Type: new Abstract: Effective memory is crucial for LLM agents, yet constructing it effectively remains challenging. A memory-construction policy decides what information t

AUCH-Net: Action Unit-Based Consistency-Aware Hypergraph Network for Cross-Domain Few-Shot Facial Expression Recognition

ResearchDGX agent

arXiv:2607.21004v1 Announce Type: new Abstract: Recently, cross-domain few-shot facial expression recognition (CF-FER) has received considerable attention. However, the performance of existing CF-FER

Auditing Evidence Use in Medical LLM Diagnosis

Local AiDGX agent

arXiv:2607.20848v1 Announce Type: new Abstract: Medical LLMs are often evaluated by whether they select the correct diagnosis, but diagnostic accuracy alone does not show whether the model used the ca

Auditing Provenance Sensitivity in LLM Agent Action Selection

Local AiDGX agent

arXiv:2607.20827v1 Announce Type: new Abstract: LLM agents choose tools and arguments from context that mixes user requests, tool outputs, retrieved records, memory, and untrusted text. Evidence can b

Automated Synthesis and Adversarial Validation of Executable Causal Research Pipelines

Model ReleasesDGX agent

arXiv:2607.21173v1 Announce Type: new Abstract: While automated research systems promise to accelerate empirical analysis, they are prone to silent failures: instances in which analysis code executes

Automatic knot selection in smooth additive models

ResearchDGX agent

arXiv:2607.21083v1 Announce Type: cross Abstract: B-spline regression constitutes a widely used framework for nonparametric modeling. The performance of this methodology depends on specifying the numb

Autonomous disproofs of the sum-product conjecture over mathbb R with GPT-5.5 Pro

Model ReleasesDGX agent

arXiv:2607.20525v1 Announce Type: new Abstract: OpenAI's recent disproof of the Erdos unit distance conjecture marked a milestone for AI in mathematics. It also inspired another breakthrough: a human

Autonomous Topology Mutation: Safe Runtime Restructuring for Multi-Agent LLM Systems with Capability, State, and Shadow Invariants

Model ReleasesDGX agent

arXiv:2607.20488v1 Announce Type: new Abstract: Multi-agent LLM frameworks typically fix their team topology at boot time. When an individual agent becomes overloaded at runtime, for example by mixing

AXIS: A Growable Community-Driven Data Engine for Scalable Robot Manipulation

Model ReleasesDGX agent

arXiv:2607.21588v1 Announce Type: new Abstract: Learning effective robot manipulation policies requires diverse, high-quality demonstrations, yet existing data pipelines are often difficult to scale b

Axolotl3D: a Unified Framework for Faithful 3D Shape Completion

SafetyDGX agent

arXiv:2607.20660v1 Announce Type: new Abstract: Recent 3D generative models produce high-quality geometry from a single image using large-scale priors and diffusion architectures. However, they assume

Backpropagation-Free Test-Time Adaptation for Lightweight EEG-Based Brain-Computer Interfaces

ApplicationsDGX agent

arXiv:2601.07556v2 Announce Type: replace-cross Abstract: Electroencephalogram (EEG)-based brain-computer interfaces (BCIs) face significant deployment challenges due to inter-subject variability, sig

Barzilai-Borwein Fails Superlinear Convergence on an Open Set of Quadratics for Every Dimension ngeq 4

ResearchDGX agent

arXiv:2607.21579v1 Announce Type: cross Abstract: Barzilai--Borwein (BB) method has shown strong practical performance in continuous optimization, yet its convergence dynamics remains poorly understoo

BasketEvent: Understanding Who Did What and When in Basketball Videos

ResearchDGX agent

arXiv:2607.21267v1 Announce Type: new Abstract: Comprehensive basketball video understanding requires resolving not only what event occurs, but also who is responsible and when the key evidence appear

Bayesian uncertainty estimation improves clinical decision making in medical AI agents

AgentsDGX agent

arXiv:2607.20582v1 Announce Type: cross Abstract: Machine learning models for medical image analysis typically lack a reliable measure of confidence, limiting their use in ambiguous or atypical cases.

Belief Propagation in LLM World Models: Measuring Strategic Information Bias with Prediction Markets

SafetyDGX agent

arXiv:2607.20441v1 Announce Type: new Abstract: Every information ecosystem produces beliefs that shape strategic decisions. Both human analysts and AI systems inherit the blind spots of their informa

Benchmarking Large Language Models on Multi-Sensor Physical Hazard Assessment

Model ReleasesDGX agent

arXiv:2607.20476v1 Announce Type: new Abstract: We present an empirical benchmark evaluating how five large language models assess multisensor physical hazard data. Testing 60 scenarios across three c

Benchmarking the Personalization Capabilities of Large Language Models

Model ReleasesDGX agent

arXiv:2607.20471v1 Announce Type: new Abstract: Personalization, the act of varying a message to induce action from a specific receiver while keeping sender, channel, and time fixed, has a long tradit

Benchmarking Unlearning for Vision Transformers

Model ReleasesDGX agent

arXiv:2602.20114v2 Announce Type: replace-cross Abstract: Machine unlearning (MU) refers to the post-training capability to remove (the influence of) training examples that are incorrect, biased, or l

Best-of-Evidence: Best-of-N Selection under Partial Verification

ResearchDGX agent

arXiv:2607.20950v1 Announce Type: new Abstract: BoN improves model outputs by sampling several candidates and selecting one with a proxy score, but it assumes that complete candidates can be evaluated

Beyond Episodic Evaluation: Memory Architectural Bottlenecks in Sequential Embodied Question Answering

ApplicationsDGX agent

arXiv:2607.21571v1 Announce Type: new Abstract: Embodied question answering (EQA) is traditionally evaluated under an episodic formulation, where agents solve each task independently and reset interna

Beyond Heavy Log Curation: Perplexity-Based APT Detection via Unsupervised, Context-Augmented Language Models

ResearchDGX agent

arXiv:2607.20832v1 Announce Type: cross Abstract: Advanced Persistent Threats (APTs) remain difficult to detect because only a small fraction of events in large-scale logs are attack-related, and inve

Beyond Independent Optimization: Compression, MoE Routing, and Quantization Interactions in Multimodal Edge Intelligence

ResearchDGX agent

arXiv:2607.20981v1 Announce Type: new Abstract: Efficient multimodal inference is increasingly constrained not only by model quality or FLOP count, but also by the cost of preserving, moving, routing,

Beyond Liars' Bench: The Impact of Lie Typology, Depth, and Sparsity on Deception Detection in LLMs

Model ReleasesDGX agent

arXiv:2607.20479v1 Announce Type: new Abstract: Training probes to detect deceptive outputs from large language models is still an open problem. Recent work has demonstrated that detection probes fail

Beyond SBDD: Geometric Deep Learning in Polypharmacology and Multi-target Drug Design

TutorialsDGX agent

arXiv:2607.20550v1 Announce Type: cross Abstract: The traditional 'one drug, one target' paradigm of structure-based drug design (SBDD) frequently proves inadequate for treating multifactorial disease

Beyond Sufficiency: Time Series Explanation with Counterfactual Necessity

ApplicationsDGX agent

arXiv:2607.21573v1 Announce Type: cross Abstract: Faithful explanations of time-series classifiers should identify subsequences that are not only sufficient to preserve a black-box model's prediction,

Beyond Sycophancy: Structured Resistance and Compliance in LLM Moral Reasoning

SafetyDGX agent

arXiv:2607.21558v1 Announce Type: new Abstract: Building socially calibrated large language models, which can learn from others without simply yielding to them, requires more than reducing sycophancy

Black box behavioural modelling: Predicting human activity schedules with a deep conditional generative approach

ResearchDGX agent

arXiv:2512.04223v2 Announce Type: replace Abstract: Modelling the complexity and diversity of human activity scheduling behaviour is inherently challenging. We demonstrate ActVAE, a deep conditional-g

Boltzmann generators for amorphous particle systems

ResearchDGX agent

arXiv:2512.16607v2 Announce Type: replace-cross Abstract: Sampling configurations in thermodynamic equilibrium is a long-standing challenge in statistical physics. Boltzmann generators address this pr

← Previous
1…163164165166167…998
Next →