AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
5 Aug 2026

Optimising for Flourishing: Flourishing Metrics and Return on Flourishing as Success Criteria for Artificial Intelligence and Post-AGI Economic Systems

SafetyDGX agent

arXiv:2608.00151v2 Announce Type: replace-cross Abstract: Current evaluation frameworks for artificial intelligence focus mainly on capability, safety, and proxies such as adoption, engagement, effici

OR-Agent: Bridging Evolutionary Search and Structured Research for Automated Algorithm Discovery

AgentsDGX agent

arXiv:2602.13769v3 Announce Type: replace Abstract: Automating heuristic design in complex, experiment-driven domains requires more than iterative mutation of solution algorithms. Current LLM-based ev

Output-Aware Rotation for INT2 KV-Cache Quantization

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2608.02691v1 Announce Type: cross Abstract: The key-value (KV) cache has become a major memory and bandwidth bottleneck in long-context large language model inference, making ultra-low-bit quant

PACE: Adaptive Budget Allocation for Time-Efficient Embodied Planning

Model ReleasesDGX agent

arXiv:2608.03034v1 Announce Type: cross Abstract: Reasoning-enhanced large language models have achieved remarkable improvements in planning tasks, yet their deployment in embodied systems remains imp

PAIChecker: Uncovering and Checking PR-Issue Misalignment in SWE-Bench-Like Benchmarks

AgentsDGX agent

arXiv:2607.28587v2 Announce Type: replace-cross Abstract: SWE-bench-like benchmarks are widely used for evaluating LLM's issue resolution capability. They typically follow a common construction pipeli

PASE: Leveraging the Phonological Prior of WavLM for Low-Hallucination Generative Speech Enhancement

ResearchDGX agent

arXiv:2511.13300v1 Announce Type: cross Abstract: Generative models have shown remarkable performance in speech enhancement (SE), achieving superior perceptual quality over traditional discriminative

Patient-centered data science: an integrative framework for evaluating and predicting clinical outcomes in the digital health era

AgentsDGX agent

arXiv:2408.02677v2 Announce Type: replace-cross Abstract: This study proposes a novel, integrative framework for patient-centered data science in the digital health era. We developed a multidimensiona

Pattern over Pixels: Measuring Pattern Completion Bias in Multimodal Code Generation

Model ReleasesDGX agent

arXiv:2608.03691v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) are increasingly used to translate webpage screenshots into front-end code, but repeated UI patterns may sway

Permission Denied: Policy-Graded Evaluation of Coding Agents in Hardened Environments

SafetyDGX agent

arXiv:2608.02670v1 Announce Type: cross Abstract: Coding agents increasingly run inside organizations whose security controls (scoped credentials, restricted egress, read-only filesystems, non-root ex

PhyAI: Real-Time Physical AI at the Edge, Scalable Rollouts in the Cloud

SafetyDGX agent

arXiv:2608.03682v1 Announce Type: new Abstract: Physical AI policies require inference throughout their lifecycle, including model evaluation, cloud reinforcement learning rollout, edge GPU serving, a

pi-Attention: Online Efficient Sparse Transformers for Long-Context Modeling

ResearchDGX agent

arXiv:2511.10696v3 Announce Type: replace-cross Abstract: Sparse attention is crucial in long-context Transformers, which restricts each token to a limited neighborhood and thereby reduces the quadrat

PI-Mem: Pushing Long-Context Reasoning to 3.6M Tokens with Parallel-Iterative Memory

Model ReleasesDGX agent

arXiv:2608.03048v1 Announce Type: cross Abstract: Long-context reasoning remains a critical bottleneck for large language models, as recent recurrent-memory approaches face two inherent challenges: se

Pin Once, Swap Light: Subspace-Aligned Centroid-Residual Training for Efficient Ultra-LoRA Serving

Model ReleasesDGX agent

arXiv:2608.03579v1 Announce Type: cross Abstract: Modern multi-tenant Low-Rank Adapters (LoRAs) serving systems concurrently host tens to hundreds of LoRA adapters. Though powerful, this introduces a

Pivot-Centric Trajectory Prediction: Bridging Long Horizons via Dynamical Guidance

Local AiDGX agent

arXiv:2608.03521v1 Announce Type: cross Abstract: Forecasting precise future motion of surrounding agents is essential for reliable autonomous vehicles. However, as the demand for longer prediction ho

PLAN: Parallel Liquid-Inspired Approximation Network for Efficient Representation Learning in Flexible Job Shop Scheduling

Model ReleasesDGX agent

arXiv:2608.03041v1 Announce Type: cross Abstract: Deep reinforcement learning (DRL) approaches for flexible job shop scheduling (FJSP) heavily rely on attention-centric architectures to achieve state-

Policy Fragmentation or Institutional Alignment? Institutional Governance of AI in Universities and Business Schools

SafetyDGX agent

arXiv:2608.03584v1 Announce Type: new Abstract: Artificial intelligence (AI) is rapidly transforming high-skilled domains, requiring higher education institutions (HEI) to balance the teaching of foun

Predictive Set Theory: A Generative Framework for Cognitive Architecture with Operationalized Core Mechanisms

ResearchDGX agent

arXiv:2608.02704v1 Announce Type: new Abstract: Predictive processing theories portray the brain as a hierarchical prediction engine that minimizes prediction error, yet they lack operational definiti

Preferred, Not Safer: Pairwise Preference Is a Poor Proxy for Clinical Safety

SafetyDGX agent

arXiv:2608.02617v1 Announce Type: cross Abstract: We evaluate whether clinician pairwise preferences provide a reliable signal of clinical safety in large language model (LLM) evaluation using expert

Principles of Robot Autonomy

AgentsDGX agent

arXiv:2608.03496v1 Announce Type: cross Abstract: Autonomous robots are moving rapidly from research labs into everyday life - on roads, in the air, in warehouses, and in space. Robot autonomy is no l

PRISM: Powerful Time Series to Image (TS2I) Representations for Multivariate Anomaly Detection

ApplicationsDGX agent

arXiv:2608.03926v1 Announce Type: cross Abstract: Time series anomaly detection (TSAD) underpins applications in predictive maintenance, finance, and cloud computing, however performance remains sensi

PRISMA: Improving the Accuracy-Latency Frontier of Diffusion-based PDE Solvers Using Physics-Informed Spectral Attention

Model ReleasesDGX agent

arXiv:2512.01370v2 Announce Type: replace-cross Abstract: Diffusion-based solvers for partial differential equations (PDEs) are often bottle-necked by slow gradient-based test-time optimization routin

Privacy-Preserving AI Verification via Minimal Information Disclosure

SafetyDGX agent

arXiv:2608.02774v1 Announce Type: cross Abstract: AI verification crosses a trust boundary: a verifier must learn enough to establish an authorized claim, yet the same evidence can reveal sensitive de

PRIVEE: Privacy-Preserving Vertical Federated Learning Against Feature Inference Attacks

ResearchDGX agent

arXiv:2512.12840v2 Announce Type: replace-cross Abstract: Vertical Federated Learning (VFL) enables collaborative model training across organizations that share common user samples but hold disjoint f

ProPRL: Property-Aware Prerequisite Relation Learning in Educational Knowledge Graphs

ApplicationsDGX agent

arXiv:2608.03006v1 Announce Type: new Abstract: Prerequisite relation learning is central to adaptive instruction, yet existing methods often formulate it as conventional link prediction, limiting the

PULSE: An Executable Contract Language for Spatiotemporal Knowledge Graph Engineering

SafetyDGX agent

arXiv:2608.02630v1 Announce Type: new Abstract: Knowledge graph engineering often distributes accepted state, observations, constraints, processes, and hypothetical scenarios across artifacts whose co

Quantifying Hallucinations in Language Language Models on Medical Textbooks

Model ReleasesDGX agent

arXiv:2603.09986v3 Announce Type: replace-cross Abstract: Hallucinations, the tendency for large language models to provide responses with factually incorrect and unsupported claims, is a serious prob

Quo Vadis, World Modeling?

SafetyDGX agent

arXiv:2608.02713v1 Announce Type: cross Abstract: Continually improving agents require dynamic interaction feedback beyond static supervision, yet direct real-environment interaction is costly, slow,

Reachability Is Not Realization: Tracing the Sources of LLM Benchmark Gains

Model ReleasesDGX agent

arXiv:2608.03219v1 Announce Type: new Abstract: Benchmark gains are often treated as evidence of greater LLM capability. Yet the same gain can reflect different changes in model behavior. A model may

Rectify Then Diffuse: Disentangling Concepts Before Denoising Trajectory Unfolds

Model ReleasesDGX agent

arXiv:2608.03135v1 Announce Type: cross Abstract: Text-to-image diffusion models can generate individual concepts well, but they often omit or merge concepts incorrectly with multiple concepts. We tra

ReflectRL: Learning from Golden Negative Trajectories via Reflective-to-Direct Reasoning

SafetyDGX agent

arXiv:2608.03972v1 Announce Type: new Abstract: On-policy training has emerged as a powerful post-training paradigm for improving the reasoning capabilities of large language models, and is often enha

Rethinking Modality Reliability in Multimodal Sentiment Analysis with Incomplete Observations

SafetyDGX agent

arXiv:2608.03611v1 Announce Type: new Abstract: Multimodal Sentiment Analysis (MSA) integrates text, audio, and vision to infer human affect, yet real-world multimodal observations are often incomplet

Rethinking Self-Evolving Agent Skills: Feedback Dynamics over Multiple Rounds

Model ReleasesDGX agent

arXiv:2608.02636v1 Announce Type: cross Abstract: Self-evolving skill systems promise to improve agents by turning execution feedback into persistent skill updates without changing the underlying mode

Reversing Arrows in Large Language Models

Model ReleasesDGX agent

arXiv:2608.03512v1 Announce Type: new Abstract: Large language models (LLMs) have achieved strong performance on text-to-knowledge graph generation and related tasks. Nevertheless, it is still unclear

Risky Business: Measuring The Faithfulness-Safety Tension

Model ReleasesDGX agent

arXiv:2608.03745v1 Announce Type: new Abstract: Chain-of-Thought (CoT) reasoning offers a promising window into model monitoring. However, monitoring relies on faithfulness, i.e., the model output str

Robust Counterfactual Policy Optimisation via Nondeterministic Causal Models

SafetyDGX agent

arXiv:2608.02893v1 Announce Type: cross Abstract: Counterfactual inference approaches for sequential decision-making typically assume deterministic causal models, where all randomness stems from laten

Route-Align-Verify for Functional Correctness in Code Generation

Model ReleasesDGX agent

arXiv:2608.03341v1 Announce Type: cross Abstract: Large language models (LLMs) have substantially improved code generation, yet achieving strong functional correctness remains difficult, especially fo

Rubrics as Privileged Information for Open-Ended Generation

Model ReleasesDGX agent

arXiv:2608.02948v1 Announce Type: cross Abstract: On-policy self-distillation (OPSD), where a single model acts as both student and teacher with different contexts, has shown promise in verifiable dom

S^3: Improving Agent Safety through Multi-Stage Defense

Model ReleasesDGX agent

arXiv:2608.02683v1 Announce Type: cross Abstract: Large Language Model (LLM) agents rely on multi-stage agentic workflows, with stages such as memory, planning, and tool execution, to accomplish compl

SAGE: Semantic Explainability of Attention-Based Survival Models in Computational Pathology

Local AiDGX agent

arXiv:2608.02803v1 Announce Type: cross Abstract: Attention-based multiple instance learning (ABMIL) is the predominant approach for slide-level prediction in computational pathology, yet its attentio

SAT-Edge-Agent: Hardware-in-the-Loop Edge-Agent Orchestration for Onboard Satellite Intelligence

Local AiDGX agent

arXiv:2608.03728v1 Announce Type: new Abstract: Onboard satellite intelligence requires a task layer that translates mission intent into local tool calls, exposes execution state, and returns machine-

Scaling an Autoregressive Transformer for Single-Cell Generation

ResearchDGX agent

arXiv:2608.02961v1 Announce Type: cross Abstract: We study a self-supervised generation task for single-cell gene expression vectors: given a set of vectors from a cell type, we aim to generate additi

SciRet: A Compute-Aware Empirical Study of Retrieval and Reranking for Scientific RAG

Model ReleasesDGX agent

arXiv:2608.03860v1 Announce Type: cross Abstract: We introduce SciRet, a compute-aware empirical study of retrieval-augmented generation for scientific question answering over CORD-19. Rather than pro

Screenshots or Tools? Eliciting Tool Use and Managing Multimodal Context in Hybrid GUI-MCP Computer-Use Agents

Model ReleasesDGX agent

arXiv:2608.03327v1 Announce Type: new Abstract: Hybrid computer-use agents can act through screenshots or call text tools. We find that having a tool available does not settle which way the effect goe

Search, Inspect, Fetch: Exploiting Boolean Retrieval for Deep-Research Agents

AgentsDGX agent

arXiv:2608.02751v1 Announce Type: cross Abstract: Existing deep-research agents use a search-visit workflow that retrieves and reads whole pages, without considering the addressable structure that web

SeaSlides: Semantic Abstraction Layer for Agentic Slide Generation

Model ReleasesDGX agent

arXiv:2608.03298v1 Announce Type: new Abstract: Agentic presentation generation must preserve source content, maintain coherent visual design, render specialized objects, and produce usable artifacts.

Secure AI Watermarking Framework for IP Protection in Multi-Tenant Cloud Platforms

ResearchDGX agent

arXiv:2608.02656v1 Announce Type: cross Abstract: The Secured data safe guard transaction with multi-tenant environments run on private-protected authenticate platforms runs by secured handed environm

Security-First Evaluation of Text-to-Terraform: Benchmarking LLMs and SLMs for Secure IaC Generation

Model ReleasesDGX agent

arXiv:2608.02672v1 Announce Type: cross Abstract: Cloud misconfiguration remains a leading cause of security incidents, yet whether LLMs and SLMs can generate security-compliant Infrastructure-as-Code

Self-Guided Adaptive Safety Alignment: Synthesizing and Internalizing Guidelines in Reasoning Models

SafetyDGX agent

arXiv:2511.21214v4 Announce Type: replace-cross Abstract: Explicit safety policies can improve reasoning-model safety, but their effective coverage may lag behind evolving jailbreak strategies. We stu

Self-Organising Digital Circuits

SafetyDGX agent

arXiv:2608.02606v1 Announce Type: new Abstract: Fault tolerance in classical computing has traditionally relied on static strategies like hardware redundancy and error-correcting codes. Biological sys

Self-Supervised Representation-Guided Generative Dataset Distillation

SafetyDGX agent

arXiv:2608.03218v1 Announce Type: cross Abstract: Dataset distillation compresses a large training set into a compact synthetic set while retaining its downstream utility. Most existing methods target

Separating quantum circuits from classical LLMs

ResearchDGX agent

arXiv:2608.03962v1 Announce Type: cross Abstract: Modern large language models - transformers and diffusion language models - are built around two canonical algorithmic tasks: prediction and generatio

Shaping Wind-Tunnel Airflow for Unmanned Aerial Vehicles using Online Learning

ResearchDGX agent

arXiv:2608.03378v1 Announce Type: cross Abstract: The development and testing of advanced aerial robots require experiments in controlled environments with tailored airflow profiles. This paper presen

Shielding for Higher-Order Safety

SafetyDGX agent

arXiv:2608.03662v1 Announce Type: new Abstract: Safety shields are runtime enforcement mechanisms that restrict the actions of a controller to guarantee safety. Classical shields are usually synthesis

Shorter Reasoning, Earlier Answers? An Evaluation of Reasoning Interfaces

Model ReleasesDGX agent

arXiv:2608.03401v1 Announce Type: cross Abstract: Large language models often reason at length before answering, increasing cost and latency. Prompts and trained settings can shorten this reasoning, b

Should We Type or Talk to LLM Agents? A Comprehensive Study of Voice and Keyboard Input Perturbations

ResearchDGX agent

arXiv:2608.03970v1 Announce Type: new Abstract: Human input reaches language models by typing or speaking, and each channel leaves a distinct signature: orthographic noise for keyboards; for voice, di

Single Canonical Prompts Underestimate LLM Safety's Surface-Form Sensitivity

Model ReleasesDGX agent

arXiv:2608.02665v1 Announce Type: cross Abstract: A benchmark score is a measurement instrument, yet most benchmarks read each item at a single canonical surface form. We ask whether that reading is f

SKILL-KD: Contrastive Skill Distillation for LLM Agents

Local AiDGX agent

arXiv:2607.28048v2 Announce Type: replace Abstract: Skill-based prompting has become a practical mechanism for improving large language model (LLM) agents, yet existing skill acquisition methods often

SkillTrace: Traversing a Query-Skill Graph for Composable LLM Agents

ResearchDGX agent

arXiv:2608.02356v2 Announce Type: replace Abstract: Large language model agents increasingly solve complex tasks by composing reusable skills from a library. To address this, the key challenge is not

SMOPD: Multi-Reward Reinforcement Learning via Specialize-and-Merge Online Policy Distillation

SafetyDGX agent

arXiv:2608.03092v1 Announce Type: cross Abstract: We aim to improve model performance in multi-reward reinforcement learning training process. Existing Group reward-Decoupled Normalization Policy Opti

Socially Grounded Agentic AI: Coordinating Plural Perspectives through Social Theory

SafetyDGX agent

arXiv:2608.03910v1 Announce Type: new Abstract: As AI systems are deployed across increasingly diverse social contexts, alignment can no longer be framed as the optimization of a single, unified set o

← Previous
1…3334353637…354
Next →