AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
15 Jul 2026

Jetson-PI: Towards Onboard Real-Time Robot Control via Foresight-Aligned Asynchronous Inference

Model ReleasesDGX agent

arXiv:2607.12659v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have achieved impressive performance on diverse embodied tasks. However, deploying VLA models on low-power onboard

Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes

Model ReleasesDGX agent

arXiv:2607.12924v1 Announce Type: new Abstract: In this paper, we study Reinforcement Learning in Parametrized Action Markov Decision Processes (PAMDP), where each decision consists of a symbolic acti

LakeQuest: A Three-Domain Benchmark for Grounded Question Answering across Data Lakes

Model Releases

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2607.12310v1 Announce Type: cross Abstract: While modern question answering (QA) systems excel on clean, schema-aligned corpora, real-world knowledge is rarely so neatly packaged. Answering ques

Learning-based Probabilistic Load Forecasting with Post-hoc and In-model Uncertainty

ResearchDGX agent

arXiv:2607.12730v1 Announce Type: cross Abstract: Smart-building load forecasters are often trained offline on dense, multivariate, high-frequency data, but deployment may provide only hourly, feature

Learning in Curved Weight Space:Exponential-Linear Weight Reparameterization for Improved Optimization

Model ReleasesDGX agent

arXiv:2607.09967v2 Announce Type: replace-cross Abstract: Many neural networks operations have a multiplicative nature rather than additive: halving or doubling a norm are analogous relatively but req

Learning to Discretize: Diffusion-Based Adaptive Mesh with Spectral Guidance

TutorialsDGX agent

arXiv:2607.11974v1 Announce Type: cross Abstract: Most neural partial differential equation (PDE) surrogates learn how fields evolve after a grid has already been chosen. However, before any operator

Learning When to Trust in Contextual Social Bandits

SafetyDGX agent

arXiv:2603.13356v2 Announce Type: replace Abstract: Robust reinforcement learning typically assumes that feedback sources are either globally trustworthy or corrupted within a fixed global budget. We

Less Experts, Faster Decoding: Cost-Aware Speculative Decoding for Mixture-of-Experts

Model ReleasesDGX agent

arXiv:2607.12696v1 Announce Type: cross Abstract: Sparse Mixture-of-Experts (MoE) models have become an important approach for scaling Large Language Models (LLMs), but their inference efficiency depe

Line-Anchored Feedback Cuts Token Costs and Improves Correctness in AI Code Editing

Model ReleasesDGX agent

arXiv:2607.12713v1 Announce Type: cross Abstract: Generated tokens are a direct driver of the cost, latency, and energy of generative AI (GAI) code editing. We show the format of feedback is a lever o

LLMs Can See the Smoke but not the Fire: Evaluating Abductive Reasoning with Elenchos

ResearchDGX agent

arXiv:2607.12733v1 Announce Type: new Abstract: Large language models (LLMs) excel at pattern recognition and text generation, but their capacity for abductive inference - inferring latent hypotheses

Look, Focus, Act: Efficient and Robust Robot Learning via Human Gaze and Foveated Vision Transformers

Model ReleasesDGX agent

arXiv:2507.15833v3 Announce Type: replace-cross Abstract: Human vision is a highly active process driven by gaze, which directs attention to task-relevant regions through foveation, dramatically reduc

Lost in Visual Translation: A VLM-Assisted Perceptual-Semantic Coherence Framework for EEG-to-Image Reconstruction

SafetyDGX agent

arXiv:2607.12364v1 Announce Type: cross Abstract: EEG-to-image evaluation should distinguish visual fidelity from recoverable meaning. Yet EEG-derived reconstructions are blurry, distorted, and low-de

LP Mining with LP2Graph: A Use Case for Railway Rescheduling

ResearchDGX agent

arXiv:2607.11980v1 Announce Type: new Abstract: Like many optimization-driven domains, railway rescheduling relies on Mixed-Integer Linear Programming (MILP), yet the field's modeling knowledge is sca

MAG: A Web-Agent Benchmark and Harness for Multimodal Action and Guide Generation

Model ReleasesDGX agent

arXiv:2607.10079v2 Announce Type: replace Abstract: Digital Adoption Platforms (DAPs) are embedded overlays widely used on web systems to guide users through operations inside a page, helping them get

Mathematics of Data Science

ResearchDGX agent

arXiv:2607.11938v1 Announce Type: cross Abstract: This book is about the mathematical foundations of data science. 1. Introduction 2. Curses, Blessings, and Surprises in High Dimensions 3. Singular Va

MaxSAT-Based Feedback for Guiding Vision-Language Models in Sudoku

TutorialsDGX agent

arXiv:2607.12711v1 Announce Type: new Abstract: Vision--Language Models (VLMs) have recently demonstrated promising performance on structured visual reasoning tasks, including grid-based puzzles. Howe

MemOps: Benchmarking Lifecycle Memory Operations in Long-Horizon Conversations

Model ReleasesDGX agent

arXiv:2607.12893v1 Announce Type: new Abstract: Long-term memory has become a foundational capability for LLM-based agents that accompany users across extended, multi-session interactions. Existing be

Mind the Gap: Promises and Pitfalls of Hierarchical Planning in LeWorldModel

ResearchDGX agent

arXiv:2607.12547v1 Announce Type: cross Abstract: We investigate whether temporal hierarchy can improve LeWorldModel on long-horizon goal-conditioned control. We introduce Hi-LeWM, an extension that f

Mistake gating leads to energy and memory efficient continual learning

SafetyDGX agent

arXiv:2604.14336v2 Announce Type: replace Abstract: Synaptic plasticity is metabolically expensive, yet animals continuously update their internal models without exhausting energy reserves. However, w

Mitigating The Effect of Class Imbalance in Data with Hierarchical and Dependable Structure

ResearchDGX agent

arXiv:2607.11994v1 Announce Type: cross Abstract: Classifying cybersecurity vulnerabilities using the Common Weakness Enumeration (CWE) taxonomy is challenging due to extreme class imbalance and stron

Mobility-Aware Cache Framework for Scalable LLM-Based Human Mobility Simulation

ApplicationsDGX agent

arXiv:2602.16727v2 Announce Type: replace Abstract: Simulating large-scale human mobility is fundamental to understanding population movement patterns and supporting real-world geospatial applications

Modeling Story Expectations: A Generative Framework using LLMs

ResearchDGX agent

arXiv:2412.15239v4 Announce Type: replace-cross Abstract: Consumers' engagement with stories is shaped by their expectations about what will happen next, yet modeling these forward-looking beliefs ove

Multi-Perspective Agentic Program Repair via Code Property Graphs and Temporal Execution Graphs

Model ReleasesDGX agent

arXiv:2607.12605v1 Announce Type: cross Abstract: Large language models (LLMs) have improved automated program repair (APR), but two limitations remain. First, raw execution traces are often too large

Navigating the Mirage: A Dual-Path Agentic Framework for Robust Misleading Chart Question Answering

SafetyDGX agent

arXiv:2603.28583v2 Announce Type: replace-cross Abstract: Despite the success of Vision-Language Models (VLMs), misleading charts remain a significant challenge due to their deceptive visual structure

NeSy-Route: A Neuro-Symbolic Benchmark for Constrained Route Planning in Remote Sensing

Model ReleasesDGX agent

arXiv:2603.16307v2 Announce Type: replace Abstract: Remote sensing underpins crucial applications such as disaster relief and ecological field surveys, where systems must understand complex scenes and

OmniPMNet: Bridging discrete and gridded PM10 forecasts via omni-query neural processes

Local AiDGX agent

arXiv:2607.11896v1 Announce Type: cross Abstract: Forecasting particulate matter (PM10) requires both station-scale accuracy and continuous spatial fields, especially during severe dust storms. Chemic

On-Device Deep Research at 4B: Exposure Bounds Faithfulness, Retrieval Bounds Coverage

Local AiDGX agent

arXiv:2607.12257v1 Announce Type: new Abstract: On-device research agents search a corpus, read sources, and write a cited brief on a personal laptop. Whether their citations are faithful, and at what

Ontology-Amplified Distillation and Contextuality Auditing for Sovereign Enterprise Language Models: A Combined Proof-of-Mechanism and Negative-Results Method Study

Model ReleasesDGX agent

arXiv:2607.11948v1 Announce Type: new Abstract: Regulated financial institutions operating under data-residency rules need tenant-owned language models that can run inside the institution's perimeter.

OOD-RL-Bench: A Benchmark Framework for Out-of-Distribution Detection in Reinforcement Learning

Model ReleasesDGX agent

arXiv:2607.12523v1 Announce Type: cross Abstract: Reliable reinforcement learning (RL) agents must maintain operational integrity amidst sensor malfunctions, dynamic disturbances, and slow environment

Operationalising Multi-Dimensional Evaluation for Conversational Agents: A Scalable, Governed Pipeline with Selective Re-evaluation and Model Benchmarking

SafetyDGX agent

arXiv:2607.12085v1 Announce Type: new Abstract: Evaluating retail conversational agents requires methods beyond lexical-overlap metrics to assess intent alignment, factuality, helpfulness, clarity, to

Optimal Adaptive Market Making: A Theoretical Framework for High-Yield Liquidity Provision in Perpetual Futures Markets

Model ReleasesDGX agent

arXiv:2607.11888v1 Announce Type: new Abstract: We develop a rigorous theoretical framework for optimal market making in perpetual futures markets with zero maker fees. We model the market maker's pro

Optimization Is Not All You Need

Model ReleasesDGX agent

arXiv:2607.11977v1 Announce Type: new Abstract: In 2019, OpenAI released two million GPT-2 outputs-ungrammatical, half broken-to aid the detection of machine-generated text. The alignment that produce

PalmClaw: A Native On-Device Agent Framework for Mobile Phones

Local AiDGX agent

arXiv:2607.13027v1 Announce Type: cross Abstract: Large Language Model (LLM) agents have moved beyond generating responses to executing multi-step tasks by calling tools, observing the results, and it

Partial Identification with Multiple Nonlinear Measurements of a Latent Regressor

ResearchDGX agent

arXiv:2607.12219v1 Announce Type: cross Abstract: We study linear regression when the regressor is latent and observed only through multiple noisy measurements, each a smooth but possibly nonlinear fu

PFAdapter: Hierarchical LoRA Decomposition for Personalized Federated MLLMs

Model ReleasesDGX agent

arXiv:2607.12111v1 Announce Type: cross Abstract: Agentic AI systems are reshaping communications and networking by deploying autonomous intelligent agents capable of collaborative learning while main

Physics-Informed Structure Anchoring With Capture-Aware Prototype Calibration for Cross-Environment RF Fingerprinting

Model ReleasesDGX agent

arXiv:2607.09760v2 Announce Type: replace-cross Abstract: Radio frequency fingerprint identification (RFFI) exploits transmitter-specific hardware imperfections as physicallayer identity cues for Inte

PixelLoop: Shortcut Topological Navigation with Pixel-Level Loops

ApplicationsDGX agent

arXiv:2607.12811v1 Announce Type: cross Abstract: Although topological mapping and navigation have been studied extensively, the specific role and downstream effect of loop closures in purely topologi

PluRel: Synthetic Data unlocks Scaling Laws for Relational Foundation Models

TutorialsDGX agent

arXiv:2602.04029v2 Announce Type: replace-cross Abstract: Relational Foundation Models (RFMs) facilitate data-driven decision-making by learning from complex multi-table databases. However, the divers

PM-Bench: Evaluating Prospective Memory in LLM Agents

Model ReleasesDGX agent

arXiv:2607.12385v1 Announce Type: new Abstract: A significant challenge in agentic AI is prospective memory: the ability to execute an intention at a specific future cue or state while other activitie

Practical Judgment, Virtue, and Intuition in the Use of Opaque AI-Enabled Systems

AgentsDGX agent

arXiv:2607.12755v1 Announce Type: cross Abstract: AI-enabled systems are seeing increasing deployment across numerous domains, with many being 'black boxes' with respect to core functions and capabili

PRISM Edit: One Vector for All Temporal Answers

Model ReleasesDGX agent

arXiv:2607.11327v2 Announce Type: replace-cross Abstract: Model editing keeps large language models (LLMs) up to date without retraining, but temporal facts expose a limitation of the prevailing locat

Propheticus: Machine Learning Framework for the Development of Predictive Models for Reliable and Secure Software

ResearchDGX agent

arXiv:1809.01898v2 Announce Type: replace-cross Abstract: The growing complexity of software calls for innovative solutions that support the deployment of reliable and secure software. Machine Learnin

QDEvo: A Multi-Objective Quality-Diversity Framework for Automated Heuristic Design

TutorialsDGX agent

arXiv:2607.11916v1 Announce Type: cross Abstract: The integration of Large Language Models (LLMs) with evolutionary computation has emerged as a powerful paradigm for automated heuristic design in com

Quantification of Credal Uncertainty: A Distance-Based Approach

TutorialsDGX agent

arXiv:2603.27270v2 Announce Type: replace Abstract: Credal sets, i.e., closed convex sets of probability measures, provide a natural framework to represent aleatoric and epistemic uncertainty in machi

QwenPaw-Data: Bridging Facts, Methodology, and Execution for Autonomous Enterprise Data Analytics

AgentsDGX agent

arXiv:2607.11019v2 Announce Type: replace Abstract: Enterprise data analysis is emerging as a distinct frontier for autonomous agents. Compared with general-purpose interaction and software engineerin

RCWT: Measuring Task-Budget Displacement from Coordination Content in LLM Calls

Model ReleasesDGX agent

arXiv:2607.12216v1 Announce Type: cross Abstract: Multi-agent and memory-augmented LLM systems often place coordination content, shared state, prior discussion, tool outputs, summaries, and role instr

Real-time fall detection based on vision for low-power edge platforms

Model ReleasesDGX agent

arXiv:2607.12909v1 Announce Type: cross Abstract: Falling detection is vital for elderly care and intelligent surveillance; however, prevailing vision-based approaches predominantly frame it as static

Real-Time Model Checking for Closed-Loop Robot Reactive Planning

Local AiDGX agent

arXiv:2508.19186v2 Announce Type: replace-cross Abstract: Reactive obstacle avoidance methods often cause agents to become trapped in local minima, because they can often only reason one step ahead (i

ReflectWorld-MM: An Entity-Oriented Multimodal Memory System for Open-Ended Video Streams

ApplicationsDGX agent

arXiv:2607.09759v2 Announce Type: replace-cross Abstract: Building assistants that can continually watch the world, remember what they see, and reason over their accumulated experience is a long-stand

ReLope: KL-Regularized LoRA Probes for Multimodal LLM Routing

TutorialsDGX agent

arXiv:2603.24787v2 Announce Type: replace Abstract: Routing has emerged as a promising strategy for balancing performance and cost in large language model (LLM) systems that combine lightweight models

Removable Defects: The Economics and Limits of Deliberate Deficiency

SafetyDGX agent

arXiv:2607.11983v1 Announce Type: cross Abstract: A specialist tolerates blind spots that a generalist does not. Usually this is treated as a cost to be minimized. We treat it as a design variable: a

Representation and Reference Selection in Training-Free Synthetic Image Attribution

ResearchDGX agent

arXiv:2607.12052v1 Announce Type: cross Abstract: Synthetic image attribution aims at identifying the generator responsible for a given AI-generated image. Training-free reference-based attribution me

Representing and Generating Levels Over Time through Playtrace Reconstructive Partitioning

ResearchDGX agent

arXiv:2607.12097v1 Announce Type: new Abstract: Video games are a dynamic medium experienced over time. While there are many Procedural Content Generation (PCG) approaches for generating video game le

Reproducible Reservoir Computing with Thermally Driven Superparamagnets: Controlling Temperature Sensitivity

Model ReleasesDGX agent

arXiv:2607.12840v1 Announce Type: cross Abstract: Unconventional computing systems must demonstrate robust performance under real-world environmental conditions to enable practical deployments. We hav

RepTran: Search-Based Repair of Transformer Models

ResearchDGX agent

arXiv:2607.11193v2 Announce Type: replace-cross Abstract: To ensure the overall quality of AI-enabled software, not only traditional software components but also AI components need to be tested and re

Research Novelty in Information Systems Journals After ChatGPT: Differences Across Institutional Language Contexts

ResearchDGX agent

arXiv:2603.22510v2 Announce Type: replace-cross Abstract: Large language models are increasingly used in scholarly work, yet it remains unclear whether their productivity gains are accompanied by chan

Resist and Update: Counterfactual Report Coordinates for Incentive-Compatible LLMs

Model ReleasesDGX agent

arXiv:2607.12985v1 Announce Type: new Abstract: Aligned language models routinely misreport under non-evidential incentive pressure: they agree with a confident user or overstate certainty even when t

Rethinking Reward Models for Multi-Domain Test-Time Scaling

ResearchDGX agent

arXiv:2510.00492v3 Announce Type: replace Abstract: The reliability of large language models (LLMs) during test-time scaling is often assessed with external verifiers or reward models that distinguish

Rethinking the Evaluation of Harness Evolution for Agents

Model ReleasesDGX agent

arXiv:2607.12227v1 Announce Type: new Abstract: We revisit the evaluation of automatic harness evolution for LLM agents. Existing harness evolution methods use unit test cases to search for harness co

RippleBench: Capturing Ripple Effects Using Existing Knowledge Repositories

Model ReleasesDGX agent

arXiv:2512.04144v3 Announce Type: replace Abstract: Targeted interventions on language models, such as unlearning or model editing, aim to modify specific information, but their effects often propagat

← Previous
1…7172737475…358
Next →