AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
7 Aug 2026

MicroEvo: Knowledge-Guided LLM Sampling for Efficient Microarchitecture Design Space Exploration

SafetyDGX agent

arXiv:2608.06183v1 Announce Type: new Abstract: Microarchitecture design space exploration suffers from expansive search spaces and expensive PPA evaluation, leaving only a small simulation budget for

Mind the Gaps: Mixture-of-Minds for Human Simulation

Model ReleasesDGX agent

arXiv:2608.06115v1 Announce Type: new Abstract: Predicting how a population will answer a new question is a long-standing goal. Statistical methods succeed at the level of the mass but falter at the l

MM-ISTS: Cooperating Irregularly Sampled Time Series Forecasting with Multimodal Vision-Text LLMs

SafetyDGX agent

arXiv:2603.05997v2 Announce Type: replace-cross Abstract: Irregularly sampled time series (ISTS) are widespread in real-world scenarios, exhibiting asynchronous observations on uneven time intervals a


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Multi-Agent Reinforcement Learning for Online Traffic Scheduling in Time-Sensitive Application

SafetyDGX agent

arXiv:2608.05346v1 Announce Type: cross Abstract: Time-sensitive networking (TSN) is increasingly integrated into mobile edge computing (MEC) to support applications with stringent latency requirement

Multi-Agent Transformer for Queue-Level XR Traffic Scheduling in TSN Networks

AgentsDGX agent

arXiv:2608.05340v1 Announce Type: cross Abstract: Time-Sensitive Networking (TSN) and Mobile Edge Computing (MEC) hold strong potential for enabling ultra-reliable low-latency communication for time-s

Multivariate Time Series Forecasting needs Cross Variable Loss

ResearchDGX agent

arXiv:2608.05742v1 Announce Type: cross Abstract: Multivariate time series forecasting presents unique challenges because future variables often co-evolve under shared system dynamics. While existing

NavTrust: Benchmarking Trustworthiness for Embodied Navigation

Model ReleasesDGX agent

arXiv:2603.19229v2 Announce Type: replace-cross Abstract: There are two major categories of embodied navigation: Vision-Language Navigation (VLN), where agents navigate by following natural language i

Negotiating Risk Boundaries in AI for Policing Through Mixed-Stakeholder Deliberation

SafetyDGX agent

arXiv:2608.05418v1 Announce Type: new Abstract: AI tools are being increasingly adopted in policing in the UK and worldwide. Racial bias is a known and well-documented risk, yet representatives of aff

Nonvisual Classification of Ground-Condition by Artificial Proprioception in an Amoeba-Inspired Autonomous Walking Robot

AgentsDGX agent

arXiv:2608.05684v1 Announce Type: cross Abstract: Nonvisual classification of ground condition based on a multimodal sensing approach was investigated for an amoeba-inspired autonomous walking robot.

Once a Response, Always a Response: Detecting LLM-generated Text via Latent Prompt Restoration

ResearchDGX agent

arXiv:2608.05741v1 Announce Type: cross Abstract: Large language models (LLMs) can generate fluent and convincing text at scale, creating growing risks for misinformation dissemination, educational mi

One Leak Away: How Pretrained Model Exposure Amplifies Jailbreak Risks in Finetuned LLMs

Model ReleasesDGX agent

arXiv:2512.14751v3 Announce Type: replace-cross Abstract: Finetuning pretrained large language models (LLMs) has become the standard paradigm for developing downstream applications. However, its secur

One Qubit Can Beat One Bit: Quantum Advantage for Post-Training Quantization

ResearchDGX agent

arXiv:2608.05240v1 Announce Type: cross Abstract: One-bit post-training quantization represents each weight using only its sign, requiring all deployment contexts to share the same binary weight matri

Online Reasoning Calibration: Test-Time Training Enables Generalizable Conformal LLM Reasoning

ResearchDGX agent

arXiv:2604.01170v2 Announce Type: replace-cross Abstract: While test-time scaling has enabled large language models to solve highly difficult tasks, state-of-the-art results come at exorbitant compute

OPERA: Operator-residual feedback for reliable autonomous optical experiments with language-model agents

AgentsDGX agent

arXiv:2608.05990v1 Announce Type: new Abstract: Autonomous agents choose actions using scores that may not reflect experimental success. We developed OPERA, an operator-residual framework for optical

OrchestraBench: Evaluating Multi-Agent Orchestration Failure Modes, Recovery, and Decomposition Quality

Model ReleasesDGX agent

arXiv:2608.05263v1 Announce Type: new Abstract: Multi-agent orchestration frameworks are moving from demos to production, yet benchmarks typically report task accuracy without diagnosing why a pipelin

Otter: A Time-Aware, History-Conditioned Human Chess AI

Model ReleasesDGX agent

arXiv:2608.05206v1 Announce Type: new Abstract: Otter is a 15.3M-parameter human chess AI that predicts human move selection by modeling play as a time-aware, sequential process rather than treating e

PaDoc: Layout-Grounded Parallel Decoding for Document Parsing

HardwareDGX agent

arXiv:2608.06146v1 Announce Type: new Abstract: End-to-end document parsers provide a unified interface, but serialize page layouts and regional contents into one autoregressive sequence. This formula

Path Planning of Cleaning Robot with Reinforcement Learning

SafetyDGX agent

arXiv:2208.08211v2 Announce Type: replace-cross Abstract: Recently, as the demand for cleaning robots has steadily increased, therefore household electricity consumption is also increasing. To solve t

PD-GS: Phoneme-Driven 3DGS for Audio-Driven Talking Heads

SafetyDGX agent

arXiv:2608.05218v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) enables fast, photorealistic talking-head rendering, yet accurate lip articulation remains elusive: mouth motion is often o

Personalized Deep Research Query Refinement with Graph-Scaffolded Evidence Grounding

SafetyDGX agent

arXiv:2608.05876v1 Announce Type: new Abstract: User requests serve as research specifications for deep research agents, shaping what evidence to seek and how to synthesize it. In personalized deep re

Perturbation Sensitivity at Convergence: A Simple Signal for Identifying Spuriously Correlated Samples

ResearchDGX agent

arXiv:2608.05419v1 Announce Type: cross Abstract: Models trained by empirical risk minimization on data containing spurious correlations achieve high average accuracy while failing on subpopulations w

PhaseCoder: Microphone Geometry-Agnostic Spatial Audio Understanding for Multimodal LLMs

Model ReleasesDGX agent

arXiv:2601.21124v2 Announce Type: replace-cross Abstract: Current multimodal LLMs process audio as a mono stream, ignoring the rich spatial information essential for embodied AI. Existing spatial audi

Pixel-TTS: Image based Text Rendering for Robust Text-to-Speech

ResearchDGX agent

arXiv:2606.14750v2 Announce Type: replace-cross Abstract: Recent advances in pixel-based text modeling show that representing text as images enables models to exploit visual cues for language understa

Plausible Patients, Impossible Populations: Auditing Epidemiological Fidelity in Large Language Model Mental Health Simulations

Model ReleasesDGX agent

arXiv:2604.17359v2 Announce Type: replace-cross Abstract: Language models asked to simulate psychiatric patients produce cases that survive inspection one at a time and populations that match no real

Poli-Bias: Understanding and Measuring Large Language Model Biases in International Political Conflicts

SafetyDGX agent

arXiv:2608.06123v1 Announce Type: new Abstract: Measuring political bias in large language models (LLMs) remains challenging as it can manifest through subtle differences in framing, argumentation, an

Position: It's Time to Optimize LLMs for Self-Consistency

Model ReleasesDGX agent

arXiv:2608.05188v1 Announce Type: cross Abstract: Despite ever-increasing sophistication in language model (LM) pre- and post-training pipelines, many important failures persist: models overcondition

Post-Hoc Trajectory-Risk Certification for Modular LLM-Based Security Agents

AgentsDGX agent

arXiv:2608.05199v1 Announce Type: cross Abstract: Autonomous security agents operate as staged pipelines, such as classifying network traffic and then attributing attacks to a specific technique. Spli

Posture and Sustainment Optimization Under Adversarial Uncertainty

ResearchDGX agent

arXiv:2608.05256v1 Announce Type: new Abstract: Pre-commitment posture, the assignment of military assets to theater locations before conflict scenarios resolve, is a critical and formally unsolved pr

PRISM: Distribution-Gated Flow Matching for Controllable Unpaired Image Translation

Local AiDGX agent

arXiv:2608.06240v1 Announce Type: cross Abstract: Unpaired image-to-image translation must decide, per image, what to change and what to preserve without paired supervision. Many diffusion-based unpai

PRISM: Priority-aware Rubric Internalization via Structured Multimodal Data Synthesis

ApplicationsDGX agent

arXiv:2608.05249v1 Announce Type: cross Abstract: Real-world multimodal instructions often bundle multiple requirements with unequal importance, yet most multimodal training data still reduce instruct

ProDVI: Programmatic Dynamics Priors for Value Network Initialization

ResearchDGX agent

arXiv:2608.06015v1 Announce Type: cross Abstract: Deep Reinforcement Learning (RL) is notoriously sample inefficient. One contributing factor is that RL agents are typically initialized from scratch,

Project2Task: Graph-Guided Project-Level Planning for Autonomous Research

Model ReleasesDGX agent

arXiv:2608.05225v1 Announce Type: new Abstract: Research agents can increasingly search literature, propose hypotheses, generate code, run experiments, and draft manuscripts from a single topic. Howev

Quality Diversity for Reliable Data Driven Time-Use Optimization

ResearchDGX agent

arXiv:2608.05230v1 Announce Type: cross Abstract: The daily allocation of the finite 24-hour time budget is strongly associated with physical, mental, and cognitive health. While predictive models can

QuanTiMedAI: Quantum-Enhanced Time-Series Model guided by Agentic AI for Cardiac Arrest Mortality Prediction

AgentsDGX agent

arXiv:2608.06294v1 Announce Type: new Abstract: Cardiac arrest remains one of the most lethal conditions encountered in intensive care units. Despite the growing availability of electronic health reco

RA-CAD: Learning Post-Execution Critique for State-Aware Text-to-CAD Generation

SafetyDGX agent

arXiv:2608.05714v1 Announce Type: new Abstract: Text-to-CAD generation translates natural-language design intent into editable and executable parametric computer-aided design (CAD) codes, reducing the

RealityBridge: Bridging Editable 3D Gaussian Splatting Driving Simulations and Real-World Videos

Local AiDGX agent

arXiv:2606.16278v3 Announce Type: replace-cross Abstract: Long-tail hazardous scenarios are essential for safety-oriented autonomous driving, yet they are difficult to collect at scale. Editable 3D Ga

Recursive Synthesis for Long-Horizon Terminal Tasks

Model ReleasesDGX agent

arXiv:2608.05466v1 Announce Type: new Abstract: High-quality long-horizon training data for terminal agents is expensive to produce, often costing hundreds to thousands of dollars per task, because ea

Reducing belief in conspiracy theories as they unfold using large language models

ResearchDGX agent

arXiv:2608.06151v1 Announce Type: cross Abstract: The emergence of conspiracy theories in the wake of major events is a significant societal challenge. Here we test whether conversational dialogues wi

Refining Over Resampling: Test-Time Self-Correction for LLM Reasoning

ResearchDGX agent

arXiv:2608.05643v1 Announce Type: new Abstract: Test-time scaling improves LLM reasoning by using additional inference compute, but wider sampling alone can suffer from diminishing returns: new rollou

Relay, Don't Route: Adaptive Population Handoff for Cost-Efficient LLM-Driven Evolution

ResearchDGX agent

arXiv:2608.05651v1 Announce Type: cross Abstract: Large language model (LLM)-driven evolution has shown promise for program search and algorithm discovery, but relying on strong models throughout long

Resourced Authority A Mechanism-Design Model for Participatory Governance of Deployed AI Agents

SafetyDGX agent

arXiv:2608.06353v1 Announce Type: cross Abstract: We give a formal mechanism design model for the continuous participatory governance of a deployed AI agent. The mechanism is built on the principle th

Revisiting Black-Box Model Ownership Verification through Information Theory

ApplicationsDGX agent

arXiv:2409.06130v2 Announce Type: replace-cross Abstract: Modern machine learning models require substantial computational resources and data to train, making them valuable intellectual property. Mode

Runtime Observability for Heterogeneous Attention Memory

Model ReleasesDGX agent

arXiv:2608.05863v1 Announce Type: new Abstract: Modern models no longer keep a plain KV cache: latent caches, learned sparse selectors and recurrent states each carry the model's memory in a different

SafeDivertor: Faithful Divertor Heat Flux Reconstruction from Macroscopic Plasma State Signals via Time-Frequency Prior Exploitation

Model ReleasesDGX agent

arXiv:2608.05669v1 Announce Type: cross Abstract: Divertor heat-flux analysis is essential for understanding plasma-wall interactions and protecting plasma-facing components in magnetic-confinement fu

Schema-Guided Hierarchical Information Extraction and Semantic Evaluation Using Generative AI

Model ReleasesDGX agent

arXiv:2608.06167v1 Announce Type: new Abstract: We present a schema-based framework for extracting complex, structured information from unstructured text documents using generative AI, followed by aut

SCP-NL2TL: Selective Conformal Prediction with Semantic Verification for Natural Language to Temporal Logic Specifications

SafetyDGX agent

arXiv:2608.05439v1 Announce Type: new Abstract: Translating natural language instructions into machine-interpretable formal specifications enables robots and autonomous systems to plan, reason, and fo

Search-Aided Joint Agent-Environment Reinforcement Learning for Robust Lifelong Multi-Agent Path Finding with Rotations

SafetyDGX agent

arXiv:2608.05588v1 Announce Type: cross Abstract: Lifelong Multi-Agent Path Finding (LMAPF) requires repeatedly planning collision-free paths for agents that continuously receive new goals upon reachi

Search2Skill: Skill Distillation Beyond Knowledge Boundaries Via Rubric-Based Reinforcement Learning

AgentsDGX agent

arXiv:2608.05245v1 Announce Type: new Abstract: Reusable skills, which encapsulate the procedural knowledge required to solve real-world professional tasks, offer LLM-based agents a path toward self-e

SearchAuditor: Auditing and Attributing Failures in Long-Horizon Search Agents

Model ReleasesDGX agent

arXiv:2608.05212v1 Announce Type: new Abstract: Deep search agents tackle challenging questions through long-horizon web interactions, a process that is both complex and fragile: small reasoning error

Seeing Is Not Deciding: Can Multimodal LLMs Act as Effective CEOs?

Model ReleasesDGX agent

arXiv:2608.05864v1 Announce Type: new Abstract: Large language models are increasingly applied as autonomous decision-making agents. However, in executive business decisions, existing benchmarks are l

Shapes from Examples: Foundations of Shape Learning in Recursive SHACL

ResearchDGX agent

arXiv:2607.27934v2 Announce Type: replace Abstract: SHACL shapes enable data graph validation, making automatic shape learning essential for knowledge graph applications. We investigate the well-known

Shaping Human-AI Interactions to Provide Improvement Pathways and Balance Competing Objectives

TutorialsDGX agent

arXiv:2608.05710v1 Announce Type: new Abstract: When an AI system is deployed, the individuals who use and or are evaluated by it form beliefs about how the system operates and use those beliefs to st

Signal or Spurious Cue? A Randomized Audit of Survey-Country Metadata in LLM Social Inference

ResearchDGX agent

arXiv:2608.06085v1 Announce Type: new Abstract: Survey-country metadata can improve an LLM's forecast of an individual response when informative, yet the same cue may redirect the forecast when assign

Simulator-Grounded Large Language Models for Industrial Causal Reasoning: Tool-Use, Structured Injection, and Plant-Portable Retrieval for Wastewater Treatment Decision Support

Model ReleasesDGX agent

arXiv:2608.05151v1 Announce Type: cross Abstract: Wastewater operators need answers grounded in how their plant's variables interact and how fast effects propagate, not in generic pretraining text, wh

SkillHEX: Improving Agent Skills via Hypothesis-Driven Autonomous Exploration and Exploitation

Model ReleasesDGX agent

arXiv:2608.05628v1 Announce Type: new Abstract: Although agent skills equip LLMs with reusable procedural knowledge, manual maintenance suffers from high costs, unscalability, and misalignment. Real-w

SkillMemo: Expert-guided Skill Memory Framework for Compositional Embodied Manipulation

Model ReleasesDGX agent

arXiv:2608.05970v1 Announce Type: cross Abstract: Embodied visuomotor models, including Diffusion Policy (DP) and Vision-Language-Action (VLA) models, have demonstrated promising performance on roboti

SkillTrace: Multi-Trace Provenance Auditing for LLM-Agent Skill Reuse

AgentsDGX agent

arXiv:2608.05204v1 Announce Type: new Abstract: LLM-agent ecosystems are rapidly growing around reusable skills: mixed-modality packages of metadata, natural-language instructions, code, tools, refere

SkillTV-Bench: Benchmarking How Well Judges Perform on Skill-Augmented Agentic Execution

Model ReleasesDGX agent

arXiv:2608.05573v1 Announce Type: new Abstract: LLM agents increasingly execute long-horizon tasks through tool use and environment interaction, shifting evaluation from final-response scoring to veri

SkillZip: Contract-Preserving Graph Compression for Scalable Agent Skill Libraries

AgentsDGX agent

arXiv:2608.05604v1 Announce Type: cross Abstract: Large Language Models (LLMs) increasingly act as agents whose procedural knowledge is stored in reusable skill packages and loaded at inference time.

Small Foundation Models of Human Cognition and Behaviour

ResearchDGX agent

arXiv:2608.05224v1 Announce Type: new Abstract: Large language models fine-tuned on human behavioural data have emerged as general-purpose cognitive proxies, but the scale this requires, and whether t

← Previous
1…2425262728…354
Next →